State of Proof

2026-02-14-openai-first-proof · Proof evaluation

Ten independent First Proof attempts form a calibration corpus; attempt 2 is the first blinded audit target

Calibration intake · five gates mapped · no audit gate has run

Current record status

What this docket currently establishes

Calibration intake · five gates mapped · no audit gate has run. This docket identifies what has been checked and what remains open. It does not validate or reject the manuscript's headline claim beyond the explicitly stated scope.

Claim
Ten independent First Proof attempts form a calibration corpus; attempt 2 is the first blinded audit target
Source version
OpenAI First Proof attempt packet · revised packet
Source SHA-256
27e97e4f61610cb31120c68bc8845b7a7b65bb2bfd044ffc4bb6f3b4d17f8102
Source locked
Last material update
Next open step
Freeze the blinded attempt-2 working bundle and version-delta record without leaking the known external comparator label

One-minute summary

This case is not one theorem and will not receive a corpus-wide pass or fail. The locked packet contains ten independent proposed solutions. Attempt 2 is the first blinded negative-control target; OpenAI's public assessment is held out until the lab's status and evidence are frozen.

Scope of this docket

The protocol locks and compares the original and revised packets, segments all ten attempts, audits attempt 2 under label withholding, applies a common evidence rubric, and compares frozen dispositions with external labels only after unblinding.

Claim and dependency map

  1. G1 · claim mapped

    Source-version lineage for the original and revised packets

    Depends on: two locked packet hashes

  2. G2 · claim mapped

    Ten-way claim segmentation

    Depends on: G1

  3. G3 · claim mapped

    Blinded mathematical audit of attempt 2

    Depends on: sealed comparator sources

  4. G4 · claim mapped

    Common scoped rubric across reviewed attempts

    Depends on: frozen attempt maps

  5. G5 · claim mapped

    Comparator unblinding and calibration

    Depends on: frozen lab outputs

Evidence routes

computational
Planned: deterministic extraction and structured diff of the 67-page original and 90-page revised packets.
literature
Planned: pre-approved source set for the blinded attempt-2 dependency audit.
expert
Planned: independent specialist review with prior-exposure disclosure and no access to held-out labels.

Findings and scope limits

  1. No audit gate has run and no mathematical finding has been issued.

  2. The public external label for attempt 2 is comparator data only. It cannot be copied into the lab's result.

Open Steps—where expert eyes are needed

Automorphic forms and audit methodology · audit not started

Does attempt 2's compact/Howe-vector construction satisfy its equivariance, support, central-character, and nonvanishing requirements?

It tests whether the protocol can reach an attempt-level result without contaminating the audit with a known external label.

Source record and provenance

The current and original packets are preserved separately. OpenAI's public label for attempt 2 is a held-out comparator, not a State of Proof finding.

Selected and maintained by Material Shift as protocol research. No external sponsor or author involvement is recorded for this case.

Retrieved .

Version and record history

  1. new docket

    2026-02-14-openai-first-proof

    Original and revised packets locked; five calibration gates mapped; no gate run

Statuses describe evidence collected for a defined scope. They are not peer-review decisions, publication recommendations, or certificates of mathematical truth.