All topicsSTATE OF AI REPORT.

Is AI already helping researchers build better models?

AI is already helping build better AI. According to Anthropic’s internal index, Claude led 26% of measured model R&D work in August, up from under 1% in February. Researchers set the tasks and supervise execution. Sustained, fully autonomous recursive improvement remains to be demonstrated.

Source: Anthropic: measuring the pace of AI development.

Evidence you can use

Claude’s share of measured model R&D

February-August 2026

Claude’s share of measured model R&D
PeriodShare led by Claude
February 2026SourceUnder 1%
August 2026Source26%

Anthropic’s internal index measures work inside one lab, under human supervision. It does not measure the share of all AI research automated or establish that AI can run a research agenda on its own.

Sources: Anthropic: measuring the pace of AI development.

Autonomous research and scientific judgment

For many, Karpathy's autoresearch makes part of that loop tangible: an agent edits training code, runs five-minute experiments, and keeps improvements overnight on one GPU. This automates a useful part of research, although sustained, fully autonomous recursive improvement remains to be demonstrated.

What I want to see next is agents developing scientific taste: choosing experiments, recognizing promising directions, and knowing when to abandon a familiar approach. As I argued in Can AI learn scientific taste?, learning that judgment may require the alternatives, failures, and decisions that papers leave out. The ambition is an AlphaGo-like shift in scientific strategy.

Frequently asked questions

Answers drawn from the report and the sources below.

Who leads the AI frontier in 2026?

The report describes a three-lab race between Anthropic, OpenAI, and Google. Its snapshot places Anthropic first on Artificial Analysis’s Intelligence Index and Google first on Arena’s ranking of answers people prefer. The leader depends on the measure, and these rankings change.

Source: State of AI Report 2026.

Sources and dates

2026 report snapshot. Preview revised 2026-10-07. Individual data periods and source checks are listed below. This is not a claim that every source was updated on that date.

  1. Anthropic: measuring the pace of AI developmentFebruary-August 2026. Retained from the launch essay.
  2. Anthropic: recursive self-improvement2026. Retained from the launch essay.
  3. Karpathy: autoresearchRepository, changing over time. Retained from the launch essay.
  4. Nathan Benaich: Can AI learn scientific taste?2026. Author analysis.
  5. State of AI Report 2026, slide 6: Chinese open-weight models overtook American ones in AI research papers in 20262026 report snapshot. Read against the report PDF on 2026-10-07. Study-specific limits retained.
  6. State of AI Report 2026, slide 17: Can agents produce a top-tier research paper? No, but they can do its engineering.2026 report snapshot. Read against the report PDF on 2026-10-07. Study-specific limits retained.
  7. State of AI Report 2026, slide 23: As task benchmarks saturate, RSI evidence is moving inside the labs2026 report snapshot. Read against the report PDF on 2026-10-07. Study-specific limits retained.
  8. State of AI Report 2026, slide 36: But who benchmarks the benchmarks?2026 report snapshot. Read against the report PDF on 2026-10-07. Study-specific limits retained.
  9. State of AI Report 2026, slide 41: Long-horizon coding rankings change with the task and the evaluation budget2026 report snapshot. Read against the report PDF on 2026-10-07. Study-specific limits retained.
  10. State of AI Report 2026, slide 42: High scores can hide unfinished scientific analyses and desk work2026 report snapshot. Read against the report PDF on 2026-10-07. Study-specific limits retained.
  11. State of AI Report 2026, slide 43: METR needs harder tasks to reliably measure the strongest models2026 report snapshot. Read against the report PDF on 2026-10-07. Study-specific limits retained.
  12. State of AI Report 2026, slide 48: Generative video goes real time and lets a streamer steer it!2026 report snapshot. Read against the report PDF on 2026-10-07. Study-specific limits retained.
  13. State of AI Report 2026, slide 105: AI in education: the best tutor is not a helpful assistant2026 report snapshot. Read against the report PDF on 2026-10-07. Study-specific limits retained.
  14. State of AI Report 20262026 report snapshot. Report PDF. See individual slide references for the expanded analysis.

Cite this page

Benaich, Nathan. “Is AI already helping researchers build better models?.” State of AI Report 2026. Published 2026-10-08.