OpenAI's Astra maths results survive independent scrutiny
Why I called it
The certificates make outright falsity unlikely. A machine-checked proof is hard to argue with. The live risk here is provenance, not correctness: how the results were produced, not whether they hold. That is what T2 tests.
The call, in full. OpenAI's Astra mathematics results survive independent scrutiny without a material retraction or correction of any of the ten headline results.
Scoring criterion. RESOLVES CORRECT if no headline result is retracted, and no credible published critique (arXiv, journal, or a Fields-level mathematician's public statement) establishes that a headline result is false. RESOLVES WRONG if any of the ten is retracted or shown false.
The criterion is the machine-checkable version: a prediction that cannot be settled by a third party against a public source fails the build before it reaches this page.
Related
- Where this call was made: Editorial: Technology
- OpenAI never publishes the raw Astra reasoning traces · 75%, open
- Anthropic holds the Sonnet 5 price increase · 70%, open
- Nobody solves prompt injection · 85%, open