Data
Sources
Where every value was read, with a link to each original. Most values come from developers' own documents; a few come from pages that quote them, and those are flagged.
Search and filter every source
Data as of 26 Sep 2026 · Dataset v0.1 · Methodology v0.1 · Changelog
The sources at a glance
Every value in the dataset was read in one of the sources below. Of the 1,645 values, 1,541 were read in developers’ own documents: 70 documents from eight developers, each with its own page here. Where a document’s section could not be read, 14 values were read on a developer summary page and 90 in 24 independent write-ups that quote the document.
Values read on a summary page or in a write-up are flagged wherever they appear, and are to be replaced with the document’s own values once its full text is read.
The registry also lists the 5 developer indexes we check for new documents, 21 documents we know of but have not read yet, and 12 related projects.
| Kind of source | Sources | Values read there |
|---|---|---|
| Developer documents | 70 | 1,541 |
| Developer summary pages | 1 | 14 |
| Secondary write-ups | 24 | 90 |
| Developer indexes we monitor | 5 | none |
| Documents not yet read | 21 | none |
| Related projects | 12 | none |
| All sources | 133 | 1,645 |
The table counts each value once, where it was read. A document's entry below counts every value recorded about it, including any read on a summary page or in a write-up.
Every source
By kind, as the source registry records them. Search by title or developer, or narrow the list by kind and developer.
Showing all 133 sources.
Developer documents
70 sources
System cards, model cards, reports and posts, by developer and newest first. Each title opens the document's page here, with its versions, every value recorded from it and what we read of it. No version has an archived snapshot of our own yet; from dataset v0.2 each retrieved version carries one (Methodology §7).
Anthropic
Our note: Web reader stopped ~p.49
Claude Fable 5.1 & Claude Mythos 5.1 System Card
Our note: Web reader stopped ~p.52; refers to an August 2026 Risk Report (not extracted)
Our note: Web reader stopped ~p.57
Our note: Web reader stopped ~p.56
Claude Fable 5 & Claude Mythos 5 System Card
Our note: Web reader stopped ~p.50
Our note: Web reader stopped ~p.57
Our note: Web reader stopped ~p.52
Claude Mythos Preview System Card
Our note: Web reader stopped ~p.49; limited-release model; first card under RSP v3.0
Our note: Web reader stopped ~p.68
Our note: Web reader stopped ~p.58; sabotage figures live in a separate Sabotage Risk Report (not extracted)
Our note: Web reader stopped at §5.2.2.2; §6–7 from secondary write-ups
Our note: Read in full (~39 pages)
Our note: Web reader stopped at §7.2; alignment numbers partly from secondary write-ups
Claude Opus 4.1 System Card (addendum)
Our note: Read in full; carries corrected Opus 4 / Sonnet 4 reward-hacking numbers
Claude Opus 4 & Sonnet 4 System Card
Our note: Web reader stopped at §4.2.1; later sections partly from secondary write-ups
DeepSeek
DeepSeek-R1 paper (Nature; arXiv v2)
Our note: Supplementary D.3 safety tables not yet extracted
Google DeepMind
Our note: FSF inherited from 3.7 Flash
Gemini 3.7 Flash Frontier Safety Framework Report
Our note: First report under FSF v3.1
Our note: Deltas printed as percentage points
Our note: FSF inherited from 3.1 Pro plus extra cyber testing
Our note: FSF results inherited from Gemini 3 Pro
Our note: Only source for one factuality figure
Gemini 3 Pro Frontier Safety Framework Report
Our note: Several results chart-only (excluded)
Our note: v0 card_date "2025-12" is the updated-header date; use 2025-09-26 as published date and record a Dec 2025 version
Gemini 2.5 Deep Think Model Card
Our note: First model to reach CBRN early-warning alert threshold
Our note: Cyber figures differ from tech report (low confidence rows)
Our note: Tables 7 and 9 headers garbled in extraction; column mapping inferred (medium confidence)
Meta
Our note: Open weights; limited numbers
Muse Spark 1.1 Evaluation Report
Our note: Restates several Muse Spark 1.0 values differently
Muse Spark Safety & Preparedness Report
Our note: First report under Meta Advanced AI Scaling Framework v2
Our note: Refusal and bias percentages
Our note: No numeric safety results in card
Moonshot AI
Our note: Only secondary summary used (vulnerability discovery); primary not located
Our note: Red-team pass rates
OpenAI
Our framework for reporting model misalignment
Our note: Not a card; six incident reports
ChatGPT Images 2.5 System Card
Our note: Image-generation safety only
Our note: v0 labels these rows "GPT-6 Astra system card"; assign by source_url
Our note: Voice model
Our note: Biology model, trusted access; bio scores not extracted
Our note: Jailbreak results figure-only
Our note: Short card
Our note: First launch treated as High in cyber
GPT-5.1 Instant and Thinking System Card Addendum
Our note: Introduces Production Benchmarks v2
Our note: Many Preparedness numbers are chart-only (excluded)
gpt-oss-120b & gpt-oss-20b Model Card
Our note: Open-weight; includes worst-case malicious fine-tuning assessment
OpenAI o3 and o4-mini System Card
Our note: Most numbers stated in text
xAI
Our note: See also themidasproject.com/watchtower/xai-08172026 (opens an external site)
Our note: Metric scales changed from 0–1 rates to percentages
Our note: First xAI card with an alignment audit (Petri 2.0)
Zhipu AI
Our note: SafetyBench scores
Developer summary pages
1 source
A developer's own page restating results from its cards. We read values here only where a card's section was out of reach; each is marked “Developer summary page” wherever it appears (Methodology §2.2).
Anthropic Transparency Hub (opens an external site)
About 8 documents:
- Claude Opus 4.8 System Card 3 values
- Claude Opus 5 System Card 3 values
- Claude Mythos Preview System Card 2 values
- Claude Opus 4.7 System Card 2 values
- Claude Fable 5 & Claude Mythos 5 System Card 1 value
- Claude Fable 5.1 & Claude Mythos 5.1 System Card 1 value
- Claude Sonnet 4.6 System Card 1 value
- Claude Sonnet 5 System Card 1 value
Secondary write-ups
24 sources
Independent articles and posts that quote a document, listed by the document they are about. We read values here only where the document's own section could not be read; each is marked “Secondary source” and drawn hollow in charts, and is to be replaced with the document's own value once its full text is read (Methodology §2.2).
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
Mythos Preview system card wiki (hugobowne.github.io) (opens an external site)
Mythos Preview system card wiki (hugobowne.github.io) (opens an external site)
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
About: Claude Opus 4.5 System Card
NeuralTrust blog (opens an external site)
About: Claude Opus 4.6 System Card
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
About: Claude Opus 4.6 System Card
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
About: Claude Opus 4.6 System Card
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
About: Claude Opus 4.7 System Card
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
About: Claude Opus 4.8 System Card
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
About: Claude Opus 5 System Card
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
About: Claude Opus 5.5 System Card
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
NeuralTrust blog (opens an external site)
About: Claude Sonnet 5 System Card
Codersera blog (opens an external site)
About: GPT-6 Astra System Card
Quentir blog (opens an external site)
About: GPT-6 Astra System Card
Zvi Mowshowitz, Don’t Worry About the Vase (opens an external site)
About: GPT-6 Astra System Card
Zentor blog (opens an external site)
About: Kimi K3 technical report
Developer indexes we monitor
5 sources
Pages where developers list their documents. We check them for new and revised documents; until the refresh pipeline runs, the checks are by hand and dated (Methodology §10).
Documents not yet read
21 sources
Documents we know of but have not read, with the reason. None of their results are in the dataset yet; they are queued for the full-text pass that comes with the refresh pipeline.
Qwen3 technical report; Qwen3.6-27B model card
A first reading found no safety results given as numbers.
August 2026 Risk Report
The Claude Fable 5.1 & Claude Mythos 5.1 System Card refers to it; not read yet.
Claude Opus 4.6 Sabotage Risk Report
Holds the sabotage results that the Claude Opus 4.6 System Card leaves to a separate report.
DeepSeek-V4 technical report and model card
A first reading found no safety results given as numbers.
Gemini 2.5 Computer Use model card
A first reading found no safety results given as numbers.
Gemini 2.5 Flash-Lite model card
Reports safety results as changes from an earlier model; not extracted yet.
Gemini Omni Flash model card
A first reading found no table of safety results.
Kimi K2.5 model card
A first reading found no safety results given as numbers.
ChatGPT agent system card
Not collected for the first version of the dataset.
codex-1 addendum
Not collected for the first version of the dataset.
Deep research system card
Not collected for the first version of the dataset.
GPT-4.5 system card
Not collected for the first version of the dataset.
GPT-5 sensitive-conversations addendum
Not collected for the first version of the dataset.
gpt-oss-safeguard report
Not collected for the first version of the dataset.
o3-mini system card
Not collected for the first version of the dataset.
Operator system card
Not collected for the first version of the dataset.
Sora 2 system card
Not collected for the first version of the dataset.
GLM-5 / GLM-5.3 reports
A first reading found no safety results given as numbers.
Anthropic “Agentic Misalignment in Summer 2026” cross-lab study
A study of several developers’ models, published by one developer; a candidate for inclusion. alignment.anthropic.com (opens an external site)
SaferAI evaluation of GLM-5.2
An evaluation by an independent organisation rather than by the developer; a candidate for inclusion.
UK AISI and US CAISI evaluations of open-weight models (Kimi K3, DeepSeek V4 Pro)
Evaluations run by government institutes rather than by the developers; how to record them is not decided yet.