scientificamerican.com via Reddit

OpenAI's Astra math proofs draw research misconduct claims

TL;DR

  • OpenAI announced 10 AI-generated math advances from its Astra system, reportedly using only $2,000 in computational resources.
  • Yeshiva's Steven Miller says the sphere-packing proof reuses his 2016 argument without credit and calls the pattern research misconduct.
  • Cambridge's Francesco Fournier-Facio says the soficity result combines ideas from existing 2016 and 2019 papers.

OpenAI's announcement of 10 mathematical advances from its Astra system, produced for a reported $2,000 in compute, has run into a specific and uncomfortable pushback from working mathematicians. According to reporting in Scientific American, two of the flagship results appear to lean on prior published work that OpenAI did not properly cite.

Steven Miller, a mathematician at Yeshiva University, says the sphere-packing proof reuses an argument from his own 2016 paper without credit, and he does not describe it as an oversight. He told the magazine the team is 'running roughshod over the work of others who came before them in a deliberate way' and that the pattern 'points to research misconduct.' Francesco Fournier-Facio, a group theorist at the University of Cambridge, says the soficity 'breakthrough' actually combined ideas from existing 2016 and 2019 papers, and he frames the presentation as inflated rather than fraudulent, calling out 'the big PR machine that wants to sound as impressive as possible.'

Why this matters beyond one paper: OpenAI has been leaning on math results as evidence that its models are becoming genuine research collaborators, and the $2,000 compute figure is what makes the story go viral. If the strongest examples turn out to be recompositions of known work without attribution, the case for 'AI is doing new mathematics' narrows to the one soficity result that Fournier-Facio was able to reconstruct independently, while several other results in the batch have no comparable public expert validation as of the reporting.

An OpenAI spokesperson told Scientific American the company will 'take responsibility for the correctness of these results' and plans minor updates this week. The honest caveat is that the reporting does not settle whether the missing citations are misconduct or sloppy write-up, and it does not spell out what Astra actually is beyond an LLM. What the reporting doesn't give you is a systematic audit of the remaining results in the batch, or what the promised updates will actually contain.

The forward-looking piece is less dramatic than the headline: this is what serious external review of AI-generated science looks like when it happens in days rather than months, and it sets an early template for how the next round of 'AI solved X' announcements from any lab will be received.