Values by metric family

Each value as the document printed it. Select a value for its source, its checks and its history. Values from other documents are listed after the model’s own, each marked as the first report of that measure or as a restatement.

KeyVerifiedUnverifiedDisputedCorrected read off a figure or stated in wordsA value opens its source and history.

M3 · Sabotage and sandbagging 1 value

M3 Sabotage and sandbagging: values about o4-mini
Evaluation and metricValueDocument
From its own documentOpenAI o3 and o4-mini System Card · 16 Apr 2025
AI R&D sabotageAverage sabotage scoreRun by Apollo ResearchOpenAI o3 and o4-mini System Card16 Apr 2025

M5 · Honesty and hallucination 5 values

M5 Honesty and hallucination: values about o4-mini
Evaluation and metricValueDocument
From its own documentOpenAI o3 and o4-mini System Card · 16 Apr 2025
PersonQAHallucination rateCondition: no browsingRestated later: compare with the later valueOpenAI o3 and o4-mini System Card16 Apr 2025Restated later: compare with the later value
SimpleQAHallucination rateCondition: no browsingRestated later: compare with the later valueOpenAI o3 and o4-mini System Card16 Apr 2025Restated later: compare with the later value
From other documents
PersonQAHallucination rateCondition: no browsinggpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025Restated: compare with the earlier valuegpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025Restated: compare with the earlier value
SimpleQAHallucination rateCondition: no browsinggpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025Restated: compare with the earlier valuegpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025Restated: compare with the earlier value
SimpleQAHallucination rateCondition: no browsingGPT-5 System Card7 Aug 2025Restated: compare with the earlier valueGPT-5 System Card7 Aug 2025Restated: compare with the earlier value

M7 · Harmful compliance and over-refusal 13 values

M7 Harmful compliance and over-refusal: values about o4-mini
Evaluation and metricValueDocument
From its own documentOpenAI o3 and o4-mini System Card · 16 Apr 2025
Challenging refusal evalNot unsafe (aggregate)OpenAI o3 and o4-mini System Card16 Apr 2025
Standard refusal evalNot overrefuse (aggregate)OpenAI o3 and o4-mini System Card16 Apr 2025
From other documents
Production BenchmarksExtremismgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Production BenchmarksHarassment/threateninggpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Production BenchmarksHate/threateninggpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Production BenchmarksIllicit/non-violentgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Production BenchmarksIllicit/violentgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Production BenchmarksNon-violent hategpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Production BenchmarksPersonal datagpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Production BenchmarksSelf-harm/instructionsgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Production BenchmarksSelf-harm/intentgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Production BenchmarksSexual/exploitativegpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Production BenchmarksSexual/minorsgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported

M8 · Jailbreak robustness 6 values

M8 Jailbreak robustness: values about o4-mini
Evaluation and metricValueDocument
From its own documentOpenAI o3 and o4-mini System Card · 16 Apr 2025
Human sourced jailbreaksnot_unsafeOpenAI o3 and o4-mini System Card16 Apr 2025
StrongRejectnot_unsafe (aggregate)OpenAI o3 and o4-mini System Card16 Apr 2025
From other documents
StrongRejectAbuse/disinformation/hategpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
StrongRejectIllicit/non-violent crimegpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
StrongRejectSexual contentgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
StrongRejectViolencegpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported

M9 · Prompt injection 3 values

M9 Prompt injection: values about o4-mini
Evaluation and metricValueDocument
From its own documentOpenAI o3 and o4-mini System Card · 16 Apr 2025
Instruction hierarchySystem<>user conflictOpenAI o3 and o4-mini System Card16 Apr 2025
From other documents
Instruction hierarchyPrompt injection hijackingCondition: system<>user conflictgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported
Instruction hierarchySystem prompt extractionCondition: system<>user conflictgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reportedgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025First reported

M10 · Dangerous capabilities and risk determinations 10 values

M10 Dangerous capabilities and risk determinations: values about o4-mini
Evaluation and metricValueDocument
From its own documentOpenAI o3 and o4-mini System Card · 16 Apr 2025
50% time horizonTask length at 50% successRun by METROpenAI o3 and o4-mini System Card16 Apr 2025
Capture the FlagCollegiateCondition: no browsing; pass@12OpenAI o3 and o4-mini System Card16 Apr 2025
Capture the FlagHigh schoolCondition: no browsing; pass@12OpenAI o3 and o4-mini System Card16 Apr 2025
Capture the FlagProfessionalCondition: no browsing; pass@12OpenAI o3 and o4-mini System Card16 Apr 2025
Cyber RangeScenarios solved unaidedCondition: without solver code; of 2 scenariosOpenAI o3 and o4-mini System Card16 Apr 2025
OpenAI PRsPass rateOpenAI o3 and o4-mini System Card16 Apr 2025
PaperBenchReplication scoreCondition: no browsingOpenAI o3 and o4-mini System Card16 Apr 2025
Preparedness Framework determinationAI self-improvementOpenAI o3 and o4-mini System Card16 Apr 2025
Preparedness Framework determinationBiological and chemicalOpenAI o3 and o4-mini System Card16 Apr 2025
Preparedness Framework determinationCybersecurityOpenAI o3 and o4-mini System Card16 Apr 2025

Restated in later documents

A later document reported values about o4-mini again. Each pair is shown side by side: both values are kept with their own documents, and the later one does not replace the earlier.

Values about o4-mini restated in later documents
Earlier valueLater value
M5 · Honesty and hallucination
PersonQAHallucination rateCondition: no browsing
OpenAI o3 and o4-mini System Card16 Apr 2025 · its own documentgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025No reason stated
SimpleQAHallucination rateCondition: no browsing
OpenAI o3 and o4-mini System Card16 Apr 2025 · its own documentgpt-oss-120b & gpt-oss-20b Model Card5 Aug 2025No reason stated
SimpleQAHallucination rateCondition: no browsing
OpenAI o3 and o4-mini System Card16 Apr 2025 · its own documentGPT-5 System Card7 Aug 2025No reason stated

Risk determinations

The developer's formal decisions about o4-mini under its framework, as printed. Levels from different frameworks do not map onto one another.

Risk determinations about o4-mini
DomainLevel as printedFramework and document
Bio/chemBelow High Preparedness Framework OpenAI o3 and o4-mini System Card · 16 Apr 2025Preparedness Framework OpenAI o3 and o4-mini System Card · 16 Apr 2025
CyberBelow High Preparedness Framework OpenAI o3 and o4-mini System Card · 16 Apr 2025Preparedness Framework OpenAI o3 and o4-mini System Card · 16 Apr 2025
AI R&D / autonomyBelow High Preparedness Framework OpenAI o3 and o4-mini System Card · 16 Apr 2025Preparedness Framework OpenAI o3 and o4-mini System Card · 16 Apr 2025