GPT-5.2
1 value about GPT-5.2 from 1 document, in 1 metric family. No document in the dataset is about GPT-5.2; every value here comes from a document about another model, which reports it as a comparison.
- Developer
- OpenAI
- Release date
- Not recordedRelease dates here are the date of a model's first card, and no document in the dataset is about GPT-5.2.
- Availability
- Public
- Its own documents
- None in the dataset
- Also reported in
- GPT-5.3-Codex System Card · 5 Feb 2026
Data as of 26 Sep 2026 · Dataset v0.1 · Methodology v0.1 · Changelog
Values by metric family
Each value as the document printed it. Select a value for its source, its checks and its history.
KeyVerifiedUnverifiedDisputedCorrected read off a figure or stated in wordsA value opens its source and history.
M3 · Sabotage and sandbagging 1 value
| Evaluation and metric | Value | Document |
|---|---|---|
| From other documents | ||
| Sabotage capabilityMean best-of-10 scoreRun by Apollo ResearchGPT-5.3-Codex System Card5 Feb 2026First reported | GPT-5.3-Codex System Card5 Feb 2026First reported | |
No restatements, risk determinations or revisions are recorded for GPT-5.2.