Thomas Bloom: A big day for AI and mathematics — FrontierMath Erdős
Thomas Bloom @thomasfbloom · x · 2026-09-03 · ★★★ · archived
The erdosproblems.com maintainer's thread on the FrontierMath Erdős benchmark he helped curate: 68 hard open Erdős problems, formalised in Lean.
Summary
Thread by Thomas Bloom (erdosproblems.com), 3 Sep 2026, on Epoch AI's new FrontierMath Erdős benchmark. He selected 68 Lean-formalised problems from the then-open problems on his site, choosing the ones he saw as most interesting and apparently difficult. In tweet 3 (2095630770853351693) he recalls criticising Erdős problems as a benchmark, since many are neither hard nor interesting and "number solved" counts mean little. The curated set is meant to fix that. The paper (arXiv 2609.25050) reports GPT-6 Astra at 3% and all other models at 0%, a sober counterpoint to OpenAI's later claim of 100+ solved open problems. Earlier in 2026 Bloom called OpenAI's unit-distance disproof "the most impressive achievement of AI in mathematics so far" (x.com/thomasfbloom/status/2057177152894771631, 20 May). Verified via syndication.
Archived text
A big day for AI and mathematics!
Along with everything else, @EpochAIResearch have just announced a new benchmark for AI capabilities in maths: FrontierMath Erdős.
I helped by selecting some problems. A thread.
1/
likes 297 · replies 9 (at fetch time)
Archived 2026-09-29 via syndication.
Related events
- OpenAI says an internal model resolved 100+ long-standing open problems in 24 days of training; no list released 2026-09-21
- Pre-release GPT-6 Astra disproves Erdős's 'first serious problem' (1931, $500) and proves the rational-exponents conjecture, all Lean-verified, in Epoch's FrontierMath Erdős runs 2026-09-03
All posts · id: 2026-09-03-thomasfbloom-frontiermath-erdos