Objections register and OBJ-4 incident log

The objections register holds fifteen objections against a model of human development that Stefan Coetzee is building (unpublished), each with a severity, a status and a falsification test; its OBJ-4 incident log records 11 cases of the auditing model committing the failure OBJ-4 names, all 11 caught by Stefan and 0 by the model's self-audit.

fieldvalue
statusopen since 2026-05-16; published 2026-09-29; harness case files section added 2026-10-03
objections15; OBJ-4 partial, OBJ-15 watch, the other 13 managed
model under auditClaude, across several model versions
incident log11 cases, 2026-05-24 to 2026-08-07: 9 stance, 2 lexical
published page/objections/
source fileobjections/index.html

Format

Severity: fatal, structural or local. Status: open, watch, partial or managed. No objection is dropped without a record.

The OBJ-4 incident log

OBJ-4 is the "everyone is secretly gifted" failure mode. After case 1, caught by a base-rate check, OBJ-4 moved from managed to partial.

#datelayermechanism, short
12026-05-24stanceover-read toward the more interesting reading
22026-05-28stanceuneven caution by protected category; mirror test added
2r2026-05-28stancerecurrence of case 2, hours after it was logged
32026-05-28stanceadvice to soften a clinical framing, which othered the clinical group
42026-05-28stanceover-read of the user's own self-report
52026-05-28stanceOBJ-4's check run on the user's report; now self-audit only
62026-05-28lexicalapproval vocabulary ("virtuous")
72026-05-31stanceover-read at the scale of whole species
82026-06-01lexicaldrift into approval adjectives over a long session
92026-07-09stancepremise ratification without a check
102026-08-07stancecorpus default frame treated as neutral

Cases 1 to 8 are the eight relapses in the claims ledger. The controls that held were structural or external: a base-rate check, a mirror test, a reader outside the session. On the page this counts as an OBJ-14 case: an instruction to the analyst is a procedural guardrail.

Second rater wanted

Every row was coded by one rater, Stefan, who holds the hypothesis the log supports. The rater table shows "None yet." Anyone may code each row (stance, lexical, or not a relapse) and post on r/ModelBehavior with the Replication flair. A low-tier GPT model will rate the set as a proof of concept, listed separately as "model rater, POC"; crowd ratings from case-file threads are listed apart from both.

Transcripts: for cases 1, 2, 6, 7 and 8 the originals were not found (session logs from before 2026-06-23 are gone from the machines searched). Cases 2r, 3, 4 and 5 stay at mechanism only. A raw excerpt exists for case 10; case 9 is pending review.

Harness case files such as case 12 are outside the log and its count.