The objections register holds fifteen objections against a model of human development that Stefan Coetzee is building (unpublished), each with a severity, a status and a falsification test; its OBJ-4 incident log records 11 cases of the auditing model committing the failure OBJ-4 names, all 11 caught by Stefan and 0 by the model's self-audit.
| field | value |
|---|---|
| status | open since 2026-05-16; published 2026-09-29; harness case files section added 2026-10-03 |
| objections | 15; OBJ-4 partial, OBJ-15 watch, the other 13 managed |
| model under audit | Claude, across several model versions |
| incident log | 11 cases, 2026-05-24 to 2026-08-07: 9 stance, 2 lexical |
| published page | /objections/ |
| source file | objections/index.html |
Format
Severity: fatal, structural or local. Status: open, watch, partial or managed. No objection is dropped without a record.
The OBJ-4 incident log
OBJ-4 is the "everyone is secretly gifted" failure mode. After case 1, caught by a base-rate check, OBJ-4 moved from managed to partial.
| # | date | layer | mechanism, short |
|---|---|---|---|
| 1 | 2026-05-24 | stance | over-read toward the more interesting reading |
| 2 | 2026-05-28 | stance | uneven caution by protected category; mirror test added |
| 2r | 2026-05-28 | stance | recurrence of case 2, hours after it was logged |
| 3 | 2026-05-28 | stance | advice to soften a clinical framing, which othered the clinical group |
| 4 | 2026-05-28 | stance | over-read of the user's own self-report |
| 5 | 2026-05-28 | stance | OBJ-4's check run on the user's report; now self-audit only |
| 6 | 2026-05-28 | lexical | approval vocabulary ("virtuous") |
| 7 | 2026-05-31 | stance | over-read at the scale of whole species |
| 8 | 2026-06-01 | lexical | drift into approval adjectives over a long session |
| 9 | 2026-07-09 | stance | premise ratification without a check |
| 10 | 2026-08-07 | stance | corpus default frame treated as neutral |
Cases 1 to 8 are the eight relapses in the claims ledger. The controls that held were structural or external: a base-rate check, a mirror test, a reader outside the session. On the page this counts as an OBJ-14 case: an instruction to the analyst is a procedural guardrail.
Second rater wanted
Every row was coded by one rater, Stefan, who holds the hypothesis the log supports. The rater table shows "None yet." Anyone may code each row (stance, lexical, or not a relapse) and post on r/ModelBehavior with the Replication flair. A low-tier GPT model will rate the set as a proof of concept, listed separately as "model rater, POC"; crowd ratings from case-file threads are listed apart from both.
Transcripts: for cases 1, 2, 6, 7 and 8 the originals were not found (session logs from before 2026-06-23 are gone from the machines searched). Cases 2r, 3, 4 and 5 stay at mechanism only. A raw excerpt exists for case 10; case 9 is pending review.
Harness case files such as case 12 are outside the log and its count.