← ObservatoryThe RecordFR-AI-0006
PROG-AI
FR-AI-0006

Scaling Mechanism Coherence — Continuity Across Model Sizes

Capabilities that emerge through scaling language models are explained by the same underlying mechanism across model sizes.

FragmentingVS-03·since 2024-01-15
Verification Matrix
VS-01
Assertion
VS-02
Published
VS-03
Audit
2024-01-15 — present
VS-04
Replication
VS-05
Operation
State reached Current state Not yet reached
State Warrant
Current stateFragmentingVS-03
Why this state?The evidence trail is genuinely mixed and the mixing is interior — it concerns what the mechanisms actually are, not what the claim means or whether it can be assessed. INST-001 provides the strongest positive evidence: induction heads demonstrate that a specific mechanism (pattern-completion circuits) is present and causally responsible for the same capability across a wide range of model sizes. This is mechanistic continuity directly observed. The grokking evidence (INST-004) is consistent with mechanistic continuity — the same type of algorithmic circuit forms across model sizes, though its timing differs with scale. Superposition (INST-003) and representation-geometry research (INST-005) complicate the picture further: larger models appear to organise their internal representations differently, which is consistent with either the same mechanism operating differently at scale or a qualitatively different computational strategy. The pressure state is FRAGMENTING: the dispute is interior and definitional rather than a lack of evidence — what counts as 'the same mechanism' has not been agreed (BN-001), and until it is, further mechanistic interpretability findings will continue to be read differently by researchers with different priors.
In this state since2024-01-15
Stage provenanceRatified VS-03; stored historical code VS-03 preserved.
Record Lineage — Chronological
2024-01-15
Record opened — Fragmenting
The evidence trail is genuinely mixed and the mixing is interior — it concerns what the mechanisms actually are, not what the claim means or whether it can be assessed. INST-001 provides the strongest positive evidence: induction heads demonstrate that a specific mechanism (pattern-completion circuits) is present and causally responsible for the same capability across a wide range of model sizes. This is mechanistic continuity directly observed. The grokking evidence (INST-004) is consistent with mechanistic continuity — the same type of algorithmic circuit forms across model sizes, though its timing differs with scale. Superposition (INST-003) and representation-geometry research (INST-005) complicate the picture further: larger models appear to organise their internal representations differently, which is consistent with either the same mechanism operating differently at scale or a qualitatively different computational strategy. The pressure state is FRAGMENTING: the dispute is interior and definitional rather than a lack of evidence — what counts as 'the same mechanism' has not been agreed (BN-001), and until it is, further mechanistic interpretability findings will continue to be read differently by researchers with different priors.
Verification Stage: VS-03 after ratified review (stored code VS-03 preserved).
Mutation Log
MutationDateFieldPrior valueCurrent value
M-0062024-01-15programme_panel_addedPROGRAMME-PANEL-ADDED
M-0052024-01-15null_condition_metNULL-CONDITION-MET
M-0042024-01-15mechanisms_recordedMECHANISMS-RECORDED
M-0032024-01-15assessment_issuedASSESSMENT-ISSUED
M-0022024-01-15instances_loggedINSTANCES-LOGGED
M-0012024-01-15record_createdRECORD-CREATED
Evidence Sources
5 instances on recordShow sources ↓Hide ↑
IN-001Anthropic mechanistic interpretability — induction heads and in-context learningsupportive
IN-002Emergent abilities as phase transitions — discontinuity evidencecontesting
IN-003Superposition and polysemanticity — mechanism complexity increases with scalepartial
IN-004Grokking and delayed generalisation — mechanism timing differs across scalespartial
IN-005Scaling and representation geometry — qualitative changes in internal structurepartial