AI systems can autonomously conduct scientific research that produces novel, correct discoveries.
"Autonomously conduct" lacks an agreed boundary. The claim requires autonomous research conduct, but the boundary between autonomous AI research and AI-assisted human research is contested. All current leading examples involve human-framed problems solved autonomously. Whether the claim requires only autonomous problem-solving (satisfied by GNoME and FunSearch) or also autonomous problem-identification (not yet demonstrated) is the critical definitional gap. This is the fifth lexical bottleneck in the corpus — and notably the second in PROG-AI within two records, following the same pattern identified at FR-AI-0006.
Novelty assessment is itself a research task. The claim requires that discoveries be novel, but establishing novelty requires surveying the accessible scientific literature — which is itself an incomplete and poorly indexed object. For fast-moving fields, a result that appears novel may have been anticipated in preprints, conference talks, or unpublished work. For large, old literatures, a result that appears novel may rediscover forgotten work. Novelty is not directly measurable from the discovery alone; it requires a comparison to the state of knowledge, which is itself uncertain. This is a measurement validity bottleneck of the same type as FR-BT-0002 BN-001: the measurement tool (literature survey) may not reliably track the thing it purports to measure (genuine novelty).
Autonomous problem identification with verified novel correct results. The resolution path is a demonstration where an AI system identifies a previously unrecognised scientific problem, generates hypotheses about it, designs or conducts experiments, and produces results that are independently verified as correct and novel — without a human specifying the problem space. FunSearch and GNoME satisfy part of this; the problem-identification component is the remaining gap. Several AI research systems in development are explicitly targeting this boundary. The attractor is clearly defined and closer than analogous attractors in other records — the current evidence is within one component of satisfaction.
Does "autonomously conduct scientific research" require autonomous problem identification, or is autonomous problem-solving within human-framed domains sufficient? BN-001 cannot close until this is resolved. The claim's satisfaction hangs on this distinction.
Raised 2024-01-15INST-005 is the sixth occurrence of anticipatory institutional evidence and the first within PROG-AI. Does it fit the existing taxonomy of act types (commercial commitment, regulatory preparation, community standards tightening), or does institutional reorganisation constitute a fourth act type? The Broad Institute and EMBL restructuring is neither a commercial contract nor a regulatory act — it is a scientific workflow redesign. This may be relevant to a fourth act-type option within that developing taxonomy.
Raised 2024-01-15BN-002 (novelty assessment as a measurement validity bottleneck) is structurally similar to FR-BT-0002 BN-001 (biological age measurement validity). Both are cases where the measurement tool may not reliably track the thing it purports to measure. Two occurrences of this specific bottleneck structure across two programmes. Has measurement validity as a distinct resistance/bottleneck type now reached watchlist elevation?
Raised 2024-01-15| Mutation | Date | Field | Prior value | Current value |
|---|---|---|---|---|
| M-013 | 2026-08-01 | assessment_issued | AS-001 | AS-002 |
| M-012 | 2026-08-01 | instance_appended | IN-007 | IN-008 |
| M-011 | 2026-07-17 | instance_appended | — | IN-007 |
| M-010 | 2026-07-14 | vector_corrected | neutral--constrained-autonomy-boundary-untouched | NEUTRAL |
| M-009 | 2026-07-14 | instance_appended | — | IN-006 |
| M-008 | 2026-07-09 | reference_corrected | — | REFERENCE-CORRECTED |
| M-007 | 2026-07-09 | description_restored | — | DESCRIPTION-RESTORED |
| M-006 | 2026-07-09 | description_reordered | — | DESCRIPTION-REORDERED |
| M-005 | 2024-01-15 | programme_panel_added | — | PROGRAMME-PANEL-ADDED |
| M-004 | 2024-01-15 | mechanisms_recorded | — | MECHANISMS-RECORDED |
| M-003 | 2024-01-15 | assessment_issued | — | ASSESSMENT-ISSUED |
| M-002 | 2024-01-15 | instances_logged | — | INSTANCES-LOGGED |
| M-001 | 2024-01-15 | record_created | — | RECORD-CREATED |