All topicsSTATE OF AI REPORT.

Reasoning becomes the center of AI progress

In 2025, progress increasingly came from what a model did before giving its answer: taking more steps, checking its work, and trying alternatives. OpenAI retained a narrow frontier lead, but DeepSeek, Qwen, and Kimi made open-weight models credible alternatives on reasoning and coding.

Questions in this section

Who led the AI frontier in the 2025 report?What changed with reasoning models?Did Qwen replace Llama everywhere?Were reasoning models reliable on unfamiliar tasks?

Thinking becomes another way to scale

OpenAI’s o1 helped establish inference-time computation as a way to improve difficult answers. Instead of relying only on a larger training run, a system could spend more computation on a particular problem. Reinforcement learning with verifiable rewards then gave models feedback from answers that could be checked automatically, such as a mathematical result or a passing program test.

Thinking becomes another way to scale - 2025 report, slide 30
Thinking becomes another way to scale. 2025 report, slide 30 (PDF page 31)

The open-weight ecosystem changes hands

The report found Qwen accounting for more than 40% of new monthly model derivatives on Hugging Face, while Llama’s share had fallen to about 15%. That measures what developers were building on, rather than downloads, revenue, or all AI usage. Together with stronger DeepSeek and Kimi models, it showed China becoming a major source of capable, adaptable models.

The open-weight ecosystem changes hands - 2025 report, slide 45
The open-weight ecosystem changes hands. 2025 report, slide 45 (PDF page 46)

Better answers still need better evaluations

Olympiad-level mathematics and advances in formal theorem proving showed substantial capability gains. But the report also documented failures from distracting facts and small changes to problem wording. Studies disagreed on whether reinforcement learning created new reasoning abilities or made existing successful paths easier to sample. A strong score on familiar tasks did not settle that question.

Better answers still need better evaluations - 2025 report, slide 33
Better answers still need better evaluations. 2025 report, slide 33 (PDF page 34)

Evidence you can use

Open-weight model derivatives on Hugging Face

2025 report snapshot; new monthly model derivatives

Open-weight model derivatives on Hugging Face
MeasureReported valueDefinition and source
QwenMore than 40%Share of new monthly derivatives2025 report, slide 45 (PDF page 46)
LlamaAbout 15%Share at the report snapshot2025 report, slide 45 (PDF page 46)
Llama, late 2024About 50%Earlier comparison in the report2025 report, slide 45 (PDF page 46)

These shares refer to new monthly model derivatives on Hugging Face. They do not measure all model deployments, revenue, downloads, or the share of AI research. Approximate values are preserved as reported.

Frequently asked questions

Sources and dates

Historical snapshot published October 9, 2025. This web edition was prepared on October 10, 2026 from the online deck and original launch posts. Findings and forecasts retain their original time frame.

  1. 2025 report, slide 12 (PDF page 13)Original 2025 report. Slide numbers printed in the deck are one lower than PDF page numbers because the cover is unnumbered.
  2. 2025 report, slide 18 (PDF page 19)Original 2025 report. Slide numbers printed in the deck are one lower than PDF page numbers because the cover is unnumbered.
  3. 2025 report, slide 21 (PDF page 22)Original 2025 report. Slide numbers printed in the deck are one lower than PDF page numbers because the cover is unnumbered.
  4. 2025 report, slide 22 (PDF page 23)Original 2025 report. Slide numbers printed in the deck are one lower than PDF page numbers because the cover is unnumbered.
  5. 2025 report, slide 30 (PDF page 31)Original 2025 report. Slide numbers printed in the deck are one lower than PDF page numbers because the cover is unnumbered.
  6. 2025 report, slide 32 (PDF page 33)Original 2025 report. Slide numbers printed in the deck are one lower than PDF page numbers because the cover is unnumbered.
  7. 2025 report, slide 33 (PDF page 34)Original 2025 report. Slide numbers printed in the deck are one lower than PDF page numbers because the cover is unnumbered.
  8. 2025 report, slide 45 (PDF page 46)Original 2025 report. Slide numbers printed in the deck are one lower than PDF page numbers because the cover is unnumbered.
  9. State of AI Report 2025: online slides
  10. Nathan Benaich: The State of AI Report 2025Air Street Press, October 9, 2025.
  11. Welcome to State of AI Report 2025Original website launch post, October 9, 2025.

Cite this page

Benaich, Nathan. “Reasoning becomes the center of AI progress.” State of AI Report 2025. Historical report snapshot; web edition prepared 2026-10-10.