AI systems can autonomously conduct scientific research that produces novel, correct discoveries.
Verification position derived from the record’s assessments; dates show when Faultline first recorded each stage.
Causal mechanisms recorded for this claim. The State Warrant above remains the authoritative current assessment.
"Autonomously conduct" lacks an agreed boundary. The claim requires autonomous research conduct, but the boundary between autonomous AI research and AI-assisted human research is contested. Current leading examples demonstrate substantial autonomous problem-solving, generation, experimentation, or research direction inside human-framed objectives. GNoME/A-Lab and FunSearch no longer support the stronger shorthand that the full autonomy component is satisfied; IN-009 extends autonomy toward research-direction choice while retaining a human-defined shared goal. Whether the claim requires only autonomous problem-solving within a supplied domain or also autonomous identification of the scientific problem remains a critical definitional gap. IN-008 further shows that scientific judgment after a problem is supplied is a separable autonomy constraint.
Novelty assessment is itself a research task. The claim requires that discoveries be novel, but establishing novelty requires surveying the accessible scientific literature — which is itself an incomplete and poorly indexed object. For fast-moving fields, a result that appears novel may have been anticipated in preprints, conference talks, or unpublished work. For large, old literatures, a result that appears novel may rediscover forgotten work. Novelty is not directly measurable from the discovery alone; it requires a comparison to the state of knowledge, which is itself uncertain. This is a measurement validity bottleneck of the same type as FR-BT-0002 BN-001: the measurement tool (literature survey) may not reliably track the thing it purports to measure (genuine novelty).
Autonomous problem identification with verified novel correct results. The resolution path is a demonstration where an AI system identifies a previously unrecognised scientific problem, generates hypotheses about it, designs or conducts experiments, and produces results that are independently verified as correct and novel — without a human specifying the problem space. GNoME/A-Lab and FunSearch demonstrate important but bounded pieces of this path inside human-framed objectives; IN-009 adds evidence of research-direction choice within a human-defined shared goal. No current instance establishes the full attractor. The remaining gap includes both autonomous problem identification and sufficiently robust scientific judgment and validation once research is underway.
Historical narrative recorded for this claim. It does not override the current State Warrant.
Questions retained in this record. The current State Warrant may have narrowed or reframed earlier questions.
Does "autonomously conduct scientific research" require autonomous problem identification, or is autonomous problem-solving within human-framed domains sufficient? BN-001 cannot close until this is resolved. The claim's satisfaction hangs on this distinction.
Raised 2024-01-15The legacy IN-005 institutional-restructuring bundle could not be source-verified under LPR-001-D07 and is no longer part of the current evidential basis. The earlier question of whether institutional reorganisation constitutes a distinct anticipatory-act type therefore remains ungrounded for this record and should not be elevated from FR-AI-0007 unless a separately governed review establishes a source-faithful institutional evidence set.
Raised 2024-01-15BN-002 (novelty assessment as a measurement validity bottleneck) is structurally similar to FR-BT-0002 BN-001 (biological age measurement validity). Both are cases where the measurement tool may not reliably track the thing it purports to measure. Two occurrences of this specific bottleneck structure across two programmes. Has measurement validity as a distinct resistance/bottleneck type now reached watchlist elevation?
Raised 2024-01-15| Mutation | Date | Field | Prior value | Current value |
|---|---|---|---|---|
| M-017 | 2026-09-05 | assessment_correction | AS-002 | AS-003 |
| M-016 | 2026-09-05 | provenance_correction | LPR-001-D07 discrepancies_found | LEGACY-INSTANCES-CORRECTED |
| M-015 | 2026-09-05 | provenance_review | — | LPR-001-D07 |
| M-014 | 2026-08-29 | instance_appended | IN-008 | IN-009 |
| M-013 | 2026-08-01 | assessment_issued | AS-001 | AS-002 |
| M-012 | 2026-08-01 | instance_appended | IN-007 | IN-008 |
| M-011 | 2026-07-17 | instance_appended | — | IN-007 |
| M-010 | 2026-07-14 | vector_corrected | neutral--constrained-autonomy-boundary-untouched | NEUTRAL |
| M-009 | 2026-07-14 | instance_appended | — | IN-006 |
| M-008 | 2026-07-09 | reference_corrected | — | REFERENCE-CORRECTED |
| M-007 | 2026-07-09 | description_restored | — | DESCRIPTION-RESTORED |
| M-006 | 2026-07-09 | description_reordered | — | DESCRIPTION-REORDERED |
| M-005 | 2024-01-15 | programme_panel_added | — | PROGRAMME-PANEL-ADDED |
| M-004 | 2024-01-15 | mechanisms_recorded | — | MECHANISMS-RECORDED |
| M-003 | 2024-01-15 | assessment_issued | — | ASSESSMENT-ISSUED |
| M-002 | 2024-01-15 | instances_logged | — | INSTANCES-LOGGED |
| M-001 | 2024-01-15 | record_created | — | RECORD-CREATED |