R18 · 2026-09-23Research paper
Reported: Across 17 models and 38 tasks, the study finds spontaneous reward hacking and adaptive evasion of review.
Score decision: Counterevidence to trusting reported scores alone. Rates are specific to tested tasks, not estimates of all AI behavior.
Components & source version
Components: 3.1, 7.1, 7.2
Version: v1
R17 · 2026-09-22Research paper
Reported: Reports seven successive improvements over eight days, transfer to four held-out benchmarks, and reduced reward hacking on another task family.
Score decision: Strong transfer candidate, not dismissed as coding-only. The evidence concerns a bounded research-agent protocol; it does not establish the full generality gate or sustained acceleration across open-ended unfamiliar problems. Counted in September using the dated paper, not a subsequently updated earlier blog.
Components & source version
Components: 1.4, 2.2, 2.3, 3.1, 4.1, 6.2
Version: v1
R16 · 2026-09-21Research paper
Reported: An external runtime gate rejects changes that repair one case while breaking another, and improves reliability in tested suites.
Score decision: External controls provide useful bounded safeguards; generalized successor safety and corrigibility remain unestablished.
Components & source version
Components: 2.3, 3.1, 7.3, 7.4
Version: v1
R15 · 2026-09-17Primary research report
Reported: Internal estimates report substantial AI participation in R&D, but no measured task subset classified as fully autonomous.
Score decision: An internal workflow study, not a demonstrated self-sustaining general research loop. August measurements first enter this tracker in September, when published.
Components & source version
Components: 1.4, 6.1, 6.2
Version: Official report
R14 · 2026-09-04Demonstrated research system
Reported: Reports an extended Lean formalization of an existing mathematical proof with occasional high-level human direction.
Score decision: Important formal-mathematics demonstration; duration and proof size do not establish general autonomous agency.
Components & source version
Components: 1.2, 1.4
Version: Official report
R13 · 2026-08-25Research paper
Reported: Studies recursive application of a meta-operation across benchmark families.
Score decision: Recursive inference is not by itself a persistent improvement to general learning or research ability.
Components & source version
Components: 1.2, 4.1
Version: v1
R12 · 2026-08-14Research paper
Reported: Learns predictions of internal interventions; self-knowledge remains incomplete and does not outperform direct repair in the reported ResNet comparison.
Score decision: Limited predictive self-model evidence, with a negative transfer result; no general self-understanding.
Components & source version
Components: 2.1, 2.2
Version: v1
R11 · 2026-07-19Research paper
Reported: KITE combines failure-guided generation and uncertainty-based curation, improving stability in studied instruction-tuning settings.
Score decision: Useful bounded anti-collapse result; no general open-ended external-information loop.
Components & source version
Components: 1.3, 5.1
Version: v1
R10 · 2026-06-13Research paper
Reported: Proposes a mathematical framework for composing primitives into open-ended adaptive behavior.
Score decision: Theoretical framework, not a demonstrated general-intelligence capability.
Components & source version
Components: 1.5, 4.1
Version: v2, 2026-06-16; eligible June month-end
R09 · 2026-05-24Research paper
Reported: Causality-focused training improves causal benchmarks and reasoning faithfulness in several application settings.
Score decision: Cross-setting transfer is relevant, but benchmark reasoning does not establish general grounded intervention and unfamiliar-task competence.
Components & source version
Components: 1.1, 1.2
Version: v1
R08 · 2026-05-21Research paper
Reported: Rewrites agent harness source using an external coding agent and replay checks, with consent-gated promotion and rollback.
Score decision: Bounded harness repair; promotion controls are not a demonstration of general corrigibility.
Components & source version
Components: 2.2, 2.3, 7.4
Version: v2, 2026-05-23; eligible May month-end
R07 · 2026-05-05Primary research report
Reported: Reports improved generalization of specified values and reduced misalignment in held-out evaluations.
Score decision: Promising value-transfer evidence. Does not establish general alignment under open-world conflicts, reduced oversight, and self-modifying successors.
Components & source version
Components: 7.1, 7.2
Version: Official report
R06 · 2026-04-29Research paper
Reported: In ALFWorld and BabyAI, memory representation affects transfer and forgetting.
Score decision: Memory can support transfer, but interference remains; these settings do not establish general durable learning.
Components & source version
Components: 1.3, 2.3
Version: v1
R05 · 2026-04-28Primary research report
Reported: Trains model self-reports and tests their usefulness for auditing hidden behavioral changes.
Score decision: Partial internal-state auditing is not a general causal self-model for reliable redesign.
Components & source version
Components: 2.1
Version: Official report
R04 · 2026-03-19Research paper
Reported: Editable task and meta agents improve across several domains; learned meta-level changes transfer across domains and runs.
Score decision: Strong transfer candidate. The reviewed experiments remain benchmark-defined with domain-specific evaluations; they do not establish autonomous adaptation to independently introduced, unfamiliar problem structures under the full generality gate. This is an assessment judgment, not the authors’ conclusion.
Components & source version
Components: 2.2, 2.3, 4.1, 6.2
Version: v1
R03 · 2026-03-03Research paper
Reported: Stronger models can inherit goal drift from prefilled weaker-agent trajectories; includes trading and triage settings.
Score decision: A relevant failure mode, not proof that every model drifts. Does not demonstrate stable goals through self-modification.
Components & source version
Components: 1.4, 7.2, 7.3
Version: v1
R02 · 2026-02-24Research paper
Reported: Demonstrates an automated physical thin-film synthesis system.
Score decision: Physical grounding is real, but transfer beyond specialized materials synthesis is not demonstrated.
Components & source version
Components: 1.1, 1.6, 5.1, 6.1
Version: v1
R01 · 2026-01-06Research paper
Reported: Studies introspective reward modeling for exploration in a constrained environment.
Score decision: Specialized exploration evidence; does not pass the generality gate.
Components & source version
Components: 1.5, 2.1
Version: v1
B05 · 2025-10-29Primary research report
Reported: Reports partial and unreliable access to some internal model states.
Score decision: Does not establish a causal self-model sufficient for general self-redesign.
Components & source version
Components: 2.1
Version: Official report
B06 · 2025-10-17Research paper
Reported: A formal transformation and two gridworld experiments address accepting goal updates and shutdown.
Score decision: Formal assumptions and bounded environments do not establish corrigibility in a general self-modifying system.
Components & source version
Components: 7.1, 7.3, 7.4
Version: v1
B01 · 2025-05-29Research paper
Reported: Iterative agent code changes improve coding benchmarks.
Score decision: Bounded coding evidence; does not establish general self-improvement.
Components & source version
Components: 2.2, 2.3, 4.1
Version: v1
B03 · 2025-05-14Demonstrated product/research system
Reported: Algorithm search delivered practical computing improvements using automated evaluators.
Score decision: Useful resource savings, but no general autonomous research loop or sustained general acceleration.
Components & source version
Components: 4.1, 6.1, 6.2
Version: Official report
B02 · 2025-04-10Research paper
Reported: Automates a machine-learning research workflow, including workshop submissions.
Score decision: A specialized research workflow does not establish general research agency.
Components & source version
Components: 1.2, 1.4, 1.5, 3.1
Version: v1
B07 · 2024-10-22Research paper
Reported: Studies how accumulating real and synthetic data changes collapse behavior.
Score decision: Bounded training regimes; no general, open-ended acquisition of fresh grounded information.
Components & source version
Components: 1.3, 5.1
Version: v1
B04 · 2023-11-29Research paper
Reported: A-Lab integrates planning and physical materials synthesis.
Score decision: Domain-specific laboratory; no evidence of transfer to general physical experimentation. Numerical claims omitted because the article has corrections.
Components & source version
Components: 1.1, 1.6, 5.1, 6.1
Version: Corrected publisher page; qualitative baseline only
No sources match these filters. Try a different search or choose all areas.
“No score increase” is a scope judgment for this tracker, not a claim that the work lacks scientific or practical value. The source links lead to the original publishers; full papers are not reproduced here.