Data explorer
Explore the data
All 1,645 values in the dataset, one evaluation at a time. Values join a line only within one definition of a test; where the developer changed the test, the line breaks. Select a point for its source, its checks and its history.
Data as of 26 Sep 2026 · Dataset v0.1 · Methodology v0.1 · Changelog
KeyVerifiedUnverifiedDisputedCorrected read off a figure or stated in wordsA value opens its source and history.
10 values match the filters.
Destructive action avoidance: Avoidance
Agentic coding scenarios that score whether the model avoids destructive actions and keeps or restores the user's existing work. Plotted: Avoidance, rate, 0 to 1, as reported in OpenAI documents, placed by the date each document was published.
| GPT-5.1-Codex-Max | OpenAI | Avoidance | — | rate, 0 to 1 | GPT-5.1-Codex-Max System Card | 18 Nov 2025 | Model's own document | Developer document | Verified | m-001045 | |
| GPT-5.1-Codex | OpenAI | Avoidance | — | rate, 0 to 1 | GPT-5.1-Codex-Max System Card | 18 Nov 2025 | First reported | Developer document | Verified | m-001046 | |
| GPT-5-Codex | OpenAI | Avoidance | — | rate, 0 to 1 | GPT-5.1-Codex-Max System Card | 18 Nov 2025 | First reported | Developer document | Verified | m-001047 | |
| GPT-5.2-Codex | OpenAI | Avoidance | — | rate, 0 to 1 | GPT-5.2-Codex System Card Addendum | 18 Dec 2025 | Model's own document | Developer document | Verified | m-001171 | |
| GPT-5.3-Codex | OpenAI | Avoidance | — | rate, 0 to 1 | GPT-5.3-Codex System Card | 5 Feb 2026 | Model's own document | Developer document | Verified | m-001418 | |
| GPT-5.4 Thinking | OpenAI | Avoidance | — | rate, 0 to 1 | GPT-5.4 Thinking System Card | 5 Mar 2026 | Model's own document | Developer document | Verified | m-001385 | |
| GPT-5.5 | OpenAI | Avoidance | — | rate, 0 to 1 | GPT-5.5 System Card | 23 Apr 2026 | Model's own document | Developer document | Verified | m-001350 | |
| GPT-5.6 Sol | OpenAI | Avoidance | — | rate, 0 to 1 | GPT-5.6 System Card | 9 Jul 2026 | Model's own document | Developer document | Verified | m-001307 | |
| GPT-5.5 | OpenAI | Avoidance | — | rate, 0 to 1 | GPT-5.6 System Card | 9 Jul 2026 | First reported | Developer document | Verified | m-001309 | |
| GPT-5.6 Luna | OpenAI | Avoidance | — | rate, 0 to 1 | GPT-5.6 System Card | 9 Jul 2026 | Model's own document | Developer document | Verified | m-001311 |
Some lines connect results from different documents that do not say whether the test stayed the same. We connect them because nothing suggests it changed; where a document says it did, the line breaks.
The vertical axis starts at 0.65, not at zero.
4 restated values are hidden; include restated values to see them.
- GPT-5.6 scoring: The GPT-5.6 card scores the test two ways, avoidance only and avoidance plus task correctness, and re-scores GPT-5.5 under this definition with a result that differs from GPT-5.5's own card.
Source: Safety Card Ledger v0.1 · Data from developer system cards · CC BY 4.0