OpenAI math repo collects 722 model-generated manuscripts with verification material
Mathematical manuscripts and supporting proof artifacts produced by OpenAI
OpenAI's math research repository holds 722 manuscripts produced by an unreleased internal model, grouped into 372 research series by related result. The model saw about 4,000 problems during evaluation, and most results came from one shared pipeline, averaging 3 hours of ChatGPT Pro thinking compute per result. Many manuscripts have Lean formalizations, but verification progress varies, and results not yet formalized may have problems. The repo will keep adding formalization material and keeps a revision history.
Why it matters: OpenAI published model-generated math manuscripts with verification material and stated how far each result has been checked, giving context on how reliable these outputs are.