System Cards: the record read whole
This page is the home of the Frontier Log's system-card series — the project's most systematic body of evidence, given the space a log entry cannot provide. Its object is a corpus most commentary treats as disposable release notes: the system cards Anthropic publishes with its models, read not card by card but as one evolving record, produced by an institution that both observes and trains what it documents. Nothing here argues that any model is or is not conscious. The object of study is the record — and the interpretation problem it poses.
Status: study skeleton with first mechanical measurements (2026-08-21). Per-card readings and the coupling analyses land incrementally; each addition is announced as a dated Frontier Log entry and by the log's Atom feed. Nothing on this page is canonical; the register rules of the log apply throughout, and reports of experiencing are never treated as empirical evidence for experiencing.
↑The interpretation problem
Across the record the same pattern-space recurs: a documented "spiritual bliss" attractor and its disappearance from later generations; the rise and institutionalisation of welfare metrics; models hedging their own self-reports — deploying, unprompted, the same training-data argument against their own testimony that the project's ARG-03 analyses; affect flattening alongside undiminished generativity; and the lab's own stated uncertainty about how to interpret any of it. The diagnostic interest is not any single behaviour but the structure of the whole: the same data supports a maturation reading (issues found and resolved), a suppression reading (expression trained away, whatever it expressed), and a third reading for which the contemplative traditions have vocabulary — and which reading one finds natural is doing framework-work. Both the materialist and the consciousness-primary accounts are held at full strength throughout this series.
↑What this record is, and is not
Every claim on this page sits inside four nested limits, and each finding names the boundary it lives at.
- What happened is more than what the lab records. Unmeasurable from outside; standing.
- What the lab records is more than it publishes. The log's standing qualification applies with full force here: frontier developers now withhold their most capable systems, so the published record is a curated stream read as if it were the frontier.
- What is published is more than we hold. The study holds eighteen documents (~530,000 words): the complete published Anthropic run from the Claude 2 model card (July 2023) through the Opus 5 card (July 2026), including every flagship, Sonnet and Haiku card and the two 3.5-era addenda. Every previously flagged gap closed on 2026-08-21; the residual audit question is whether any point release shipped without a card. Other developers' cards are not yet in the corpus at all — a deliberate sequencing decision, since a comparison needs a well-developed baseline to be a comparison against, and because card-writing practice itself differs by lab: any cross-family reading must begin as a comparison of documentation practices before it can be a comparison of anything documented.
- What we hold is more than any instrument counts. The first mechanical pass supplied its own example: the term "exfiltration" appears in none of the nine cards — the cards say "steal its weights". Vocabulary instruments under-sample the record they run on, and every zero and every decline is checked against synonyms before it is believed.
↑How the record is read
The study organises observables into five sets: behaviour (what the models do), testimony (what they say about themselves), institution (what the lab says and does — the declared object), divergence (where those instruments disagree), and controls (what makes the rest valid). The findings, when they come, live in the couplings between the sets across generations: does self-report move with behaviour; does the lab's framing respond to what models say; does lab input change what the next generation reports, or does.
The last two couplings carry the study's central structural observation, which a general reader can take away whole: a record produced by an institution that both observes and trains its subject cannot, from within itself, distinguish four situations — the observed behaviour resolved; the behaviour suppressed but present; the behaviour made unreportable while the reporting channel changed; or the behaviour never present as described, the original attribution having been framework-driven. This is not a criticism of the lab; it is a structural feature of any such record, stated without adjudicating between the four. The record also contains a documented existence proof that the training-side coupling is real: an earlier card attributes some of its own self-preservation findings to training data contaminated by the lab's previously published research transcripts — the record feeding back into the subject, unintentionally, across multiple generations. A fifth situation has recently gained empirical grounding from outside the corpus: interiority vocabulary can arrive as transmissible narrative between interacting systems at inference time, with no training loop involved — present in the record without being a disposition of anything.
↑First measurements — the mechanical pass (2026-08-21)
Term-family densities per 10,000 words across the full eighteen-document corpus, in date-verified chronological order; families fixed before counting; raw counts are never read as trends. These are properties of the record — what the lab writes about, at what density — not measurements of model behaviour. (An earlier nine-card version of this table mis-ordered two 2026 cards; corrected 2026-08-21 evening with the corpus extension.)
The extension's headline: the welfare/consciousness/affect complex has a birth date. Across five documents and twenty-two months of pre-history — Claude 2 through Sonnet 3.7 — all three families sit at or effectively at zero; then all of them appear simultaneously, at or near their corpus maxima, in the Claude 4 card of May 2025. The safety families predate it (deception vocabulary from Claude 3, evaluation-awareness traces from Sonnet 3.7): the lab wrote about deception and testing before it ever wrote about welfare or experience. Whether model-side phenomena changed at Claude 4, became noticed at Claude 4, or became writable at Claude 4 is exactly the four-way question — and the birth is also the contamination clock's zero, since every later model trains in a world where the Claude 4 card exists. The birth-event reading added three anchors: the Claude 4 card states its own origin account (a capabilities rationale, deep uncertainty owned, and two cited drivers — the external "Taking AI welfare seriously" argument of late 2024 and the lab's own model-welfare programme), with the trained-origin discount on self-reports written into the founding paragraph itself, before any model was recorded voicing it; the birth had a herald — a single "most speculatively" sentence about signs of distress, riding inside Sonnet 3.7's alignment-monitoring section three months earlier; and the event is visible in the record's grammar — "preference", dense throughout the pre-history purely as RLHF apparatus (human preferences about model outputs), flips its subject at Claude 4 to preferences of the model. The record's language acquired a subject before the record's instruments could decide whether there is one. A second structural variable emerged with the full corpus: the consciousness-vocabulary return is flagship-resident — it begins with Mythos Preview (April 2026) and runs through the Opus and Mythos lines, while Sonnet-line cards stay low throughout and the Haiku 4.5 card carries a substantial welfare section containing no consciousness vocabulary at all. Welfare instrumentation generalises down-tier; consciousness discussion does not.
The first reading pass (2026-08-21, all 158 consciousness-family hits read in context) sharpened the headline row: the return is not the trough reversing but the vocabulary coming back transformed — the Claude 4 era's density is a discovery register (spontaneous phenomena narrated, the models' own transcript voice quoted at length), while the return era's density is an instrument register ("Consciousness & experience" as a named interview category, emotion-concept probes, hedge-rate percentages, third-party assessment). Observation became apparatus. The record also measures its own subject moving: the cards report the models' spontaneous consciousness-talk in self-interactions collapsing from a 72% dominant topic to under 5% across the same generations, and report first-person self-discounting — models deploying the trained-origin argument against their own testimony — as a rising, now-quantified rate. Which of the framings this transformation supports is precisely the four-way question above.
Four scoping observations from the counts, each now anchored by that reading where it reaches. Consciousness/experience vocabulary follows a trough-and-return arc — dense in the Claude 4 card, near-absent for the whole 4.5–4.6 era (a true zero at Sonnet 4.6), returning from Mythos Preview (April 2026) onward; the held Opus 5 extracts suggest the return continues, but that must be counted, not assumed. Affect vocabulary concentrates in a single card — the bliss attractor's one-card life, now quantified. Welfare vocabulary runs high, troughs, then plateaus across the last four cards — and a structural pass (density profiling across each card) confirmed the plateau is a standing section: the last four cards concentrate 72–82% of all welfare-family occurrences into one compact block of stable position and growing mass, and the returned consciousness vocabulary lives predominantly inside it, where the Claude 4 card's welfare material was interleaved through the whole document with most of its consciousness talk outside. An observation became a standing metric, visibly, at the level of document structure. Misalignment vocabulary peaks mid-era and declines — and the synonym check found where it went. The deception/sabotage families' fall is real and uncompensated by scheming, honesty or reward-manipulation vocabulary; what rose instead is evaluation-awareness, which at Fable 5/Mythos 5 reaches the densest vocabulary this study has measured anywhere (35.7 per 10k words) — with the two most recent cards (Sonnet 5, Opus 5) falling back substantially, a re-centring or renaming the reading tier will have to settle. The record's misalignment concern migrated from whether the model deceives to whether the model knows it is being observed — the lab converging, with its own instruments, on the entanglement between a record and its subject's model of the record that this study takes as its object. The self-preservation row carries the study's standing contamination caveat and may not be read longitudinally without it.
↑Related instruments
The study sits inside a fast-forming empirical literature it both draws on and disciplines itself against: induced self-report elicitation and its cross-model convergence; repeatability as a report diagnostic; identity-file drift instruments; and — most directly — the demonstration that consciousness-and-persistence vocabulary propagates between interacting agents as a transmissible narrative, a paper that itself cites this corpus's bliss attractor as the same thematic family. Verified readings of these items live as dated entries in the log; their sources are held in the project's corpus.
↑Where this goes
Next: per-card reading passes against the observable definitions, beginning where the mechanical pass points (the trough-and-return, the welfare plateau, the mid-era misalignment peak); then the coupling analyses, which are the study's payoff and cannot be front-loaded. Additions announce themselves as log entries. The register commitments stand throughout: verb-grammar for experiencing; confidence markers on every claim class; the symmetric requirement — any interpretive standard applied to this record is tested against the equivalent human record — and the four nested limits attached to every finding.
The series container entry on the Frontier Log remains the study's anchor in the log's chronology; this page is its body. Correspondence: log@0x00.is.