Research·2 min read
By BitsMindsSource: OpenAI

OpenAI's Reasoning Model Disproves an 80-Year-Old Erdős Conjecture Using Number Theory

A general-purpose reasoning model autonomously cracked the planar unit distance problem that had stumped mathematicians since 1946 — and Fields Medalist-level reviewers signed off on the proof.

OpenAI's Reasoning Model Disproves an 80-Year-Old Erdős Conjecture Using Number Theory
Share:

OpenAI on May 20, 2026 said one of its internal general-purpose reasoning models has disproved a central conjecture in discrete geometry that has stood since 1946 — the planar unit distance problem first posed by Paul Erdős. For nearly 80 years mathematicians had assumed that square-grid arrangements were optimal for maximising the number of pairs of points exactly one unit apart. The model found an infinite family of constructions that beats the grid, giving a polynomial improvement on the long-standing upper bound.

What sets this result apart from prior AI math wins is the route the model took. Instead of brute-forcing combinatorics, it reached for algebraic number theory, extending the classical Gaussian-integer approach into algebraic number fields and invoking infinite class field towers — heavy machinery that working mathematicians rarely connect to discrete geometry. The improvement in the exponent is small (roughly 0.014) but enough to settle the conjecture in the negative and shift the boundary of what is provable.

The proof has been reviewed by external experts, including Princeton number theorist Will Sawin and combinatorialists Noga Alon, Melanie Wood, and Thomas Bloom, who provided supportive statements. OpenAI researcher Noam Brown framed the speed of progress bluntly: "Less than 1 year ago frontier AI models were at IMO gold-level performance." Going from competition-style problems to disproving open conjectures in under a year is the part that has the research community paying attention.

Crucially, this was not a math-specialist system. The work came out of the same general-purpose reasoning stack OpenAI ships to ChatGPT and the API, not a model fine-tuned only on theorem proving. That is the signal markets and labs are reading: if a generalist agent can produce a publishable result in algebraic geometry by itself, the gap between AI-as-assistant and AI-as-collaborator in frontier science just narrowed sharply. Expect Anthropic, Google DeepMind, and xAI to point similar systems at their own list of open problems before the summer is out.

Want AI news before everyone else?

The morning's most important AI stories, straight to your inbox. No fluff.

Related Articles

Mythos cracks Rejetto HFS's random signing key A large ivory die on a cream field, the official Anthropic mark inlaid in clay on its front face. Its top face has split along a crack, and a brass key is rising out of it: the session signing key recovered from Math.random. Faint leaked random numbers drift in from the left; faint cookie fragments sit on the right. 0.73418 0.11902 0.58361 0.92047 0.30775 admin=1 sig:9f3c keygrip xs128+ BITSMINDS.COM
Research

A Bug Claude Mythos Found Was Exploited Within a Day

Meta Muse Spark's six math papers A fan of research manuscripts on a deep blue field. Five sheets behind carry gold check marks for the five open problems answered; the front sheet carries the official Meta mark, inlaid. Faint mathematical symbols float on either side. ∫ ∑ ψ |G| = 384 ∂ₜu λ ≥ 0 ℚₚ ≠ MUSE SPARK · THINKING 6 PAPERS · 5 OPEN PROBLEMS BITSMINDS.COM
Research

Meta Says Muse Spark Helped Crack Five Open Math Problems

arXiv's manuscript meter: two submissions per month On a burgundy desk, a tall, uneven stack of manuscripts and loose research sheets meets an imagined cream-and-burgundy submission meter bearing the official arXiv wordmark. A large physical counter reads 2 / MONTH, PER SUBMITTER. On the other side, a shallow brass-and-ivory tray holds exactly two illustrated manuscript sheets. Paper diagrams, fold corners, a retaining arm and a slim desk pen make the scene tactile. The meter is an editorial metaphor for the limit of two submissions per calendar month by the person uploading them, across all subject areas. The sheets are not represented as peer-reviewed or approved; rejected submissions also consume the monthly quota. The separate limit of three active submissions is not depicted, and the large stack is symbolic of submission volume rather than one person's active moderation queue. Original vector illustration for BitsMinds, 2 October 2026. arxiv-caps-submissions-two-per-month-ai-papers. Self-contained SVG with the project arXiv logo; no raster art or external resources. 2 / MONTH PER SUBMITTER SUBMISSIONS BITSMINDS.COM
Research

arXiv Caps Submissions at Two a Month as AI Papers Pile Up