TrueSeeker AI · Verified claim report Case 115d3f6f40 · 2026-08-18

§ Claim under review · Capability

"OpenAI revealed it had produced solutions to 10 long-standing mathematics problems, some of which had confounded academics for decades, unsettling and dividing mathematicians about AI's impact on their field."

Circulating claim, as submitted.

Verdict

Mostly accurate

Confidence

Medium
§

Summary

This is mostly accurate. OpenAI did publish an announcement on August 1, 2026 describing ten new results on problems in mathematics and theoretical computer science that had been open for a long time, and it released a long manuscript plus machine-checkable Lean proof files on GitHub that anyone can run. At least one of the problems had stood since 1999 and another bound had not improved since 1978, so the phrase about decades holds for part of the set. Two things the short summary leaves out matter. First, the work came from an internal model that has not been released, so no outside researcher can test it. Second, named mathematicians publicly accused OpenAI of failing to credit recent published work that key steps rely on, and OpenAI quietly softened its original claim that these problems had seen no progress for at least a decade. The results have not been peer reviewed, and whether each formal proof matches the problem as mathematicians understand it is still being checked. The description of mathematicians as unsettled and divided is well supported by the public record.

§

The readings

key figures from the evidence
27 years

years soficity question remained open before disproof claim

$2,000 USD

claimed token cost, but covers only successful runs not all attempts

§

Why this verdict

The primary artifact exists and was retrieved: OpenAI's 1 August 2026 post and the openai/ten-proofs Lean repository both confirm ten results on long-open problems, and at least two of them concern questions untouched since 1999 and 1978, so "confounded academics for decades" holds for part of the set. The claim's operative proposition is that OpenAI revealed such results and that mathematicians were unsettled and divided, and every cited source supports that, including the critics, who dispute the framing while accepting that a real result was produced. I considered "Source exists but framing is misleading," because the sentence carries OpenAI's own duration language without noting that specialists contested it and OpenAI revised it, but the claim hedges with "some of which" and correctly attributes the assertion to OpenAI, so the omission simplifies without reversing meaning. I rejected "Accurate" for that same omission plus the mathematics versus theoretical computer science compression, and I rejected "Credibly reported but unconfirmed" because this is a named public announcement with published artifacts, not anonymous sourcing. Confidence is Medium rather than High because the producing model is an unreleased internal version that no external party can test, peer review is pending, and the novelty of at least two results is under active named dispute. As of 2026-08-18.
§

Evidence

OpenAI published a post on 1 August 2026 titled "Ten advances in mathematics and theoretical computer science," stating that an internal, unreleased version of a model family it calls Astra produced new results on ten problems. The listed results include a disproof of Connes's rigidity conjecture, new arithmetic circuit and formula lower bounds for computing the permanent including a bound of order n⁴/log n, an exponential parallel repetition theorem for general two-player quantum games, and polynomial-factor hardness of approximation for the closest vector problem , plus Ehrhart's volume conjecture and multicolor Ramsey numbers . The accompanying repository states that it contains Lean 4 formalizations of the results, including improved asymptotic upper bounds on sphere-packing density reaching the Cohn-Elkies threshold, stronger upper bounds for binary and spherical codes, and a construction of a non-sofic group . Press accounts report a 249-page manuscript, reasoning walkthroughs, Apache 2.0 licensing, and a claimed token cost of roughly $2,000 at OpenAI's own API rates.

At least one of the ten concerns a question open for decades: Mikhail Gromov introduced the concept of soficity in 1999, and for 27 years no mathematician proved or disproved whether every countable group must be sofic . Press reporting also describes the first improvement to the general upper bound on high-dimensional sphere-packing density since 1978 .

The community reaction is documented and divided. An independent group theorist published a same-week paper stating that the key technical ingredient crucially builds on work of Kun and Kun-Thom, contrary to the claim in the OpenAI announcement of 1 August that there had been no progress on the main result for at least a decade . He separately wrote that the proof crucially relies on work of Kun and Kun-Thom for the technical Proposition 2.3, applied to a construction involving elementary matrices over the binary Leavitt algebra and Thompson's group V . Another group theorist wrote that the non-sofic existence result answers a genuinely impressive question, but the "no progress for at least a decade" framing is dishonest because considerable theory had recently been building up . Scientific American reported that two of the most prominent results incorporate preexisting ideas from recent literature without proper citation, contradicting OpenAI's initial press release language, and that OpenAI has since updated the language , with a named mathematician saying OpenAI is running roughshod over the work of others who came before, and arguing his own research was effectively plagiarized . Institutional friction predates the announcement: in June the International Mathematical Union endorsed the Leiden Declaration, which warns that AI companies are using published research without consent, bypassing peer review, and threatening the integrity of proof and attribution .

§

Findings

✓ What's accurate 5

  • OpenAI did publicly announce, on 1 August 2026, results on ten long-open problems, and the primary post and Lean repository exist and are public.
  • The count of ten is correct, and the results are listed individually with named conjectures.
  • "Some of which had confounded academics for decades" is supported for at least part of the set. The soficity question dates to Gromov in 1999, and the general sphere-packing exponent had not been improved since 1978.
  • The claim correctly attributes the assertion to OpenAI ("OpenAI revealed it had produced"), rather than asserting independently verified fact.
  • The characterization of mathematicians as unsettled and divided is well supported by the public record, including named misconduct allegations, an independent rebuttal preprint, an essay on arXiv calling for organized resistance, and the IMU's prior endorsement of the Leiden Declaration.

≈ What's misleading 4

  • Omitted qualifier: the claim says "10 long-standing mathematics problems." OpenAI's own title is "Ten advances in mathematics and theoretical computer science," and several results sit in theoretical computer science and quantum complexity rather than mathematics proper. It also says "advances" where the claim says "solutions"; some results are improved bounds rather than closed questions. This is a compression rather than a reversal, but it slightly overstates both scope and finality.
  • Omitted qualifier: the sentence carries no indication that OpenAI's "no progress for at least a decade" framing was contested by named specialists and was subsequently revised by OpenAI, nor that at least two flagship results are alleged to rest on uncited recent work. A reader takes "confounded academics for decades" as uncontested, when the duration framing is precisely what mathematicians disputed.
  • Demo to product conflation: the sentence attributes the work to "OpenAI" with no mention that the producer is an internal, unreleased model that no external party can access or test. Nothing in the sentence signals that the capability cannot currently be exercised or checked by anyone outside the company.
  • Marketing as evidence: the underlying facts, apart from the Lean files, come from the vendor's own announcement about its own unreleased model. The claim reproduces that framing. The attribution to OpenAI partly mitigates this, which is why it is a framing gap and not an accuracy failure.

? What's uncertain 5

  • Whether all ten formal Lean statements faithfully encode the open problems as the community understands them. That judgment requires specialists and was still in progress as of 2026-08-18.
  • Whether the results survive peer review. None had been through a refereed venue at the time of writing.
  • How much of each result is genuinely new versus assembled from recent human work. This is the live dispute, and it is not resolved.
  • The precise division of labor between the model and OpenAI staff in producing the manuscripts and formalizations.
  • I did not retrieve The Verge article directly. The claim text and caption were verified against aggregator reproductions of its opening paragraphs, which match the claim wording exactly, but I could not read the full article or confirm how it qualifies the announcement later in the piece.
Distortion flags exaggeration omitted qualifier demo to product conflation marketing as evidence unreleased as released
§

Sources

10 of 10 linked to records
[1]

OpenAI, "Ten advances in mathematics and theoretical computer science," 1 August 2026

primary vendor announcement of record
https://openai.com/index/ten-advances-in-mathematics/ ↗
[2]

GitHub repository openai/ten-proofs, Lean 4 certificates accompanying the ten results

primary vendor artifact
https://github.com/openai/ten-proofs ↗
[3]

F. Fournier-Facio, "A torsion-free non-sofic group," arXiv:2608.02025, 4 August 2026

primary preprint by an independent named group theorist responding directly to the OpenAI result
https://arxiv.org/html/2608.02025v1 ↗
[4]

Fournier-Facio research page, note published 1 August 2026 on the non-sofic construction

primary named academic, self-published
https://www.fffmaths.com/research ↗
[5]

Scientific American, "OpenAI's latest math breakthroughs commit research misconduct, experts say"

secondary named-outlet science journalism
https://www.scientificamerican.com/article/openais-latest-math-breakthroughs-commit-research-misconduct-experts-say/ ↗
[6]

The Next Web, "OpenAI says its next model, Astra, has solved ten open problems in mathematics"

secondary named-outlet tech journalism
https://thenextweb.com/news/openai-astra-model-ten-math-proofs-non-sofic-groups ↗
[7]

SiliconANGLE, 2 August 2026 report including the International Mathematical Union's June 2026 endorsement of the Leiden Declaration

secondary named-outlet tech journalism
https://siliconangle.com/2026/08/02/openais-astra-solves-10-long-open-math-problems-publishes-proofs/ ↗
[8]

R. Appenzeller, Mathstodon post on the non-sofic result and OpenAI's framing

secondary named academic commentary
https://mathstodon.xyz/@ra/117026682431560858 ↗
[9]

"The crisis of AI-generated mathematics," arXiv:2608.02859

primary unrefereed working literature
https://arxiv.org/html/2608.02859v1 ↗
[10]

The Verge article underlying the post, "The AI takeover of mathematics has begun" by Robert Hart

secondary named-outlet journalism, text seen only via aggregator reproductions of the opening paragraphs
https://holo.fyi/threads/the-ai-takeover-of-mathematics-has-begun.117395/ ↗
How links are chosen. A source is linked only when the address comes from the investigation's own retrieval or from a registry lookup (PubMed, Crossref) that matches the citation's title and year. Author lists shown as registry-verified come from the registry record, not from the report text. Citations that cannot be matched are labeled, never guessed.
This is one case on the record See the full case, browse the archive, and search every checked claim on TrueSeeker AI Open on ai.trueseeker.com →