AI safety in the 2023 report
Evidence you can use
AI safety in the 2023 report
Historical snapshot: October 2023. Dates and populations are specified per row.
| Measure | Reported value | Definition and source |
|---|---|---|
| UK safety institution | Frontier AI Taskforce | Institution discussed in the October 2023 report, before the later AI Safety Institute.2023 report, slide 143 (PDF page 143) |
| US security initiative | NSA AI Security Centre | Initiative announced in September 2023, as described by the report.2023 report, slide 143 (PDF page 143) |
| RLHF limitations | Oversight, reward mismatch, generalization | Categories of fundamental problems summarized by the report; not a quantitative risk score.2023 report, slide 149 (PDF page 149) |
The report combines institutional developments with technical studies. Their presence is not evidence of a shared safety threshold. Attack results depend on the selected models, prompts, and evaluation conditions.
Sources and dates
Historical snapshot published October 12, 2023. This web edition was prepared on October 10, 2026 from the online deck and original launch posts. Findings and forecasts retain their original time frame.
- 2023 report, slide 143 (PDF page 143)Original 2023 report. Printed slide numbers match PDF page numbers in this edition.
- 2023 report, slide 148 (PDF page 148)Original 2023 report. Printed slide numbers match PDF page numbers in this edition.
- 2023 report, slide 149 (PDF page 149)Original 2023 report. Printed slide numbers match PDF page numbers in this edition.
- State of AI Report 2023: online slides
- Nathan Benaich: The State of AI Report 2023Air Street Press, October 12, 2023.
- Welcome to State of AI Report 2023Original website launch post, October 12, 2023.
Cite this page
Benaich, Nathan. “AI safety in the 2023 report.” State of AI Report 2023. Historical report snapshot; web edition prepared 2026-10-10.