ARTICLE AD BOX
August 6, 2026
3 min read
Add Us On GoogleAdd SciAm
OpenAI's latest mathematics breakthroughs perpetrate investigation misconduct, experts say
The institution released 10 AI-generated results complete nan weekend. Some mathematicians are unhappy pinch their approach
By Joseph Howlett edited by Lee Billings
Malte Mueller/Getty Images
OpenAI’s newest chatbot whitethorn beryllium a whiz astatine math, but it seems to beryllium lagging acold down humans successful its academic rigor.
Last week the company announced 10 much artificial-intelligence-generated math advances that were recovered during soul improvement and testing of its adjacent awesome ample connection model. This batch of results came from that LLM, Astra, and each 1 resolves aliases progresses a different “long-standing unfastened problem” of “substantial interest” to nan mathematical community. The institution said that nan full token costs was a specified $2,000.
The news quickly dispersed arsenic yet different harbinger of AI’s promise—or threat—of outpacing humans to go a dominant, disruptive unit successful mathematics and machine science. But complete nan days pursuing nan announcement, arsenic flesh-and-blood experts pored complete nan astir 250-page insubstantial successful detail, galore grew frustrated.
On supporting subject journalism
If you're enjoying this article, see supporting our award-winning publicity by subscribing. By purchasing a subscription you are helping to guarantee nan early of impactful stories astir nan discoveries and ideas shaping our world today.
Two of nan astir breathtaking results, nan experts say, incorporated preexisting ideas from nan caller mathematical lit without decently citing them. This contradicts OpenAI’s initial property release, which said that nan problems Astra addressed “have been unfastened and seen nary advancement connected nan main consequence for astatine slightest a decade.” (OpenAI has since updated nan connection to beryllium much accurate).
“They are moving roughshod complete nan activity of others who came earlier them successful a deliberate way,” says Steven Miller, a mathematician astatine Yeshiva University, who argues that OpenAI has efficaciously plagiarized his ain research. “It seems wholly systematic to me, and it points to investigation misconduct.”
The consequence Miller refers to concerns really galore balls you tin fresh successful a box—a seemingly elemental problem, isolated from these balls and boxes beryllium successful a mathematical abstraction of 1,000 dimensions—or moreover more. OpenAI’s insubstantial improves nan champion estimate for really tightly these balls tin perchance beryllium packed. The LLM-generated impervious hinges connected a peculiar mathematical statement that it presented arsenic its ain but that really first appeared successful a 2016 insubstantial by Miller and a collaborator.
Another of nan 10 results resolves a long-standing mobility successful group theory, which studies sets of mathematical objects called “groups” that interact successful an organized way. Mathematicians person agelong wondered whether each groups person a spot called “soficity,” nan capacity to beryllium faithfully approximated successful a peculiarly measurement by other, simpler groups. The OpenAI insubstantial establishes astatine slightest 1 group that lacks this property.
The find stunned Francesco Fournier-Facio, a mathematician astatine nan University of Cambridge, who studies group theory—at slightest until he “engaged pinch this breakthrough arsenic I would if a quality had written it,” he says. The result, he and immoderate of his colleagues found, wasn’t arsenic caller arsenic it first appeared. Like a number of caller AI breakthroughs, it pasted together ideas from nan mathematical lit to build a caller theorem. Once again, nan LLM’s instrumentality is its superhuman patience for assembling puzzle pieces, not nan expertise to make immoderate profound intelligence leap.
In particular, Astra’s cardinal mathematical measurement mixed ideas first recovered successful 2 papers from 2016 and 2019. Andreas Thom, a mathematician astatine nan Dresden University of Technology, who co-authored some papers, summarized nan consequence connected MathOverflow.com, calling it “creative and astatine nan aforesaid clip elementary.”
OpenAI’s first property merchandise seemed to ignore—or beryllium wholly unaware of—these crucial, caller developments. Fournier-Facio argues that nan 2 preceding papers show humans had not deed a stalemate pinch nan soficity problem. OpenAI’s mathematicians did their champion to property these ideas correctly successful their paper, he says. But, successful spite of their bully intentions, “there is nan large PR instrumentality that wants to sound arsenic awesome arsenic imaginable and does not attraction astir being 100 percent accurate,” he says.
“We return work for nan correctness of these results and are gathering nan aforesaid standards mostly expected of quality mathematicians,” OpenAI said successful a connection to Scientific American. “We scheme to make mini updates [to nan paper] this week, accordant pinch modular world practice.”
But arsenic AI continues its run to conquer mathematics without immoderate built-in fealty to nan field’s world norms, immoderate successful nan organization are intelligibly losing patience. “OpenAI is now afloat participating successful high-level research,” Fournier-Facio says. “So they should beryllium held to nan aforesaid world standards that we are.”
It’s Time to Stand Up for Science
If you enjoyed this article, I’d for illustration to inquire for your support. Scientific American has served arsenic an advocator for subject and manufacture for 180 years, and correct now whitethorn beryllium nan astir captious infinitesimal successful that two-century history.
I’ve been a Scientific American subscriber since I was 12 years old, and it helped style nan measurement I look astatine nan world. SciAm always educates and delights me, and inspires a consciousness of awe for our vast, beautiful universe. I dream it does that for you, too.
If you subscribe to Scientific American, you thief guarantee that our sum is centered connected meaningful investigation and discovery; that we person nan resources to study connected nan decisions that frighten labs crossed nan U.S.; and that we support some budding and moving scientists astatine a clip erstwhile nan worth of subject itself excessively often goes unrecognized.
In return, you get basal news, captivating podcasts, superb infographics, can't-miss newsletters, must-watch videos, challenging games, and nan subject world's champion penning and reporting. You tin moreover gift personification a subscription.
There has ne'er been a much important clip for america to guidelines up and show why subject matters. I dream you’ll support america successful that mission.
1 jam yang lalu
English (US) ·
Indonesian (ID) ·