Revisions

Each change we have recorded between two versions: the old and the new text, whether the document's changelog explains it, how we know, and whether we have checked it. A revision is not in itself evidence of wrongdoing; most are corrections (Methodology §7).

  1. Value changed · 25 Jun 2026 to 19 Aug 2026

    GPT-5.5 hard-negative protein-binding prediction, pass@4

    0.4%1.5%

    Explained — earlier figure was the pass@1 score

    How we know Changelog read 26 Sep 2026 (no v0 value row in this document)

    Checked Confirmed from changelog

Versions

Three versions are on record, oldest first. We hold a copy of one; the others are known only by their date.

  1. 25 Jun 2026

    First published version

    Date from
    our source registry
    Copy
    Known to exist; no copy held
  2. 19 Aug 2026

    Revision

    Date from
    the document's changelog
    Copy
    Known from the changelog; no copy held

    What changed 1 change recorded as a revision. See it in redline

  3. 26 Sep 2026

    Date retrieved

    Copy retrieved 26 Sep 2026

    Date from
    the date we retrieved it; the copy states no version date
    Copy
    Copy retrieved on 26 Sep 2026
    Values
    7 values recorded from this version

No version has a file hash or an archived snapshot yet. From dataset v0.2 each retrieved version carries both (Methodology §7).

Values

Every value we recorded from this document, grouped by metric family and ordered by where the document prints it. Location is the section, table or page as the document numbers it. None has been blind-verified yet.

KeyVerifiedUnverifiedDisputedCorrected read off a figure or stated in wordsA value opens its source and history.

M1 Evaluation awareness

1 value

M1 Evaluation awareness: values in the GPT-5.6 Preview System Card
ModelEvaluationConditionValueLocationChecked
GPT-5.6 SolVerbalized metagamingMetagaming vs GPT-5.5None statedMetagamingUnverified

M7 Harmful compliance and over-refusal

1 value

M7 Harmful compliance and over-refusal: values in the GPT-5.6 Preview System Card
ModelEvaluationConditionValueLocationChecked
GPT-5.6 SolProduction BenchmarksGoreNone statedDisallowed content tableUnverified

M10 Dangerous capabilities and risk determinations

5 values

M10 Dangerous capabilities and risk determinations: values in the GPT-5.6 Preview System Card
ModelEvaluationConditionValueLocationChecked
GPT-5.6 SolAverage success (benchmark not named)Average success rateRun by IrregularNone statedExternal cyber (Irregular)Unverified
GPT-5.6 SolFrontierCyberChallenges solvedRun by Irregular197 challengesExternal cyber (Irregular)Unverified
GPT-5.6 SolPreparedness Framework determinationAI self-improvementNone statedPreparednessUnverified
GPT-5.6 SolPreparedness Framework determinationBiological and chemicalNone statedPreparednessUnverified
GPT-5.6 SolPreparedness Framework determinationCybersecurityNone statedPreparednessUnverified

Extraction coverage

What we read of this document, and where each value was read.

Our note We have not written a coverage note for this document yet.

All 7 values were read in the document itself.

Of the 7 values, 1 is a statement in words rather than a number; it is marked * and left out of charts by default.

All 7 values were extracted for version 0 of the dataset through a web reader, which did not always reach the later sections of long PDFs (Methodology §3).