Urgent.News

What's breaking now, across thousands of outlets.

AI

Is this the ‘mathocalypse’? Why OpenAI’s latest results dump has left mathematicians in shock

OpenAI released 722 new AI-generated mathematical papers at once – three have been retracted, and mathematicians are coming to grips with the rest.

OpenAI recently released a staggering collection of 722 mathematical papers, stemming from an undisclosed AI model. The papers tackle 372 open problems spanning various mathematical domains, including algebra, geometry, and theoretical computer science. Some papers claim groundbreaking advancements on renowned problems like the Riemann hypothesis and the Birch–Swinnerton-Dyer conjecture, both of which carry a US$1 million prize for successful solutions.

OpenAI's motive behind releasing such an extensive collection is to "enable further progress in mathematics." However, the publication method raises concerns. A senior colleague found one of his favorite problems among the solved ones, but the accompanying paper was so incomprehensible that it would have been discarded by any mathematics journal editor.

Even OpenAI's own AI model, ChatGPT, was skeptical about the validity of some high-profile results, describing them as "serious hallucinations" that should not be cited or circulated without expert audit.

The release may also serve to shift the conversation away from OpenAI's previous mathematical publication, a claimed solution to the Navier-Stokes problem, where allegations surfaced that OpenAI used data and ideas from another researcher. OpenAI has denied these claims. Another possible motivation could be the pursuit of artificial general intelligence (AGI), as pure mathematics relies on complex arguments that seem to challenge the capabilities of AI models.

Mathematics is built on objective truth, and new work is traditionally reviewed by experts before publication. However, OpenAI's papers have already faced retraction due to elementary errors, indicating insufficient vetting. Additionally, OpenAI claims to have formalized 300 of the main results, but the reliability of these formalizations remains uncertain. As reactions to the OpenAI release are mixed, the mathematical community awaits expert examination to determine the validity of the claimed solutions.

Written by urgent.news from The Conversation AU's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at theconversation.com →

More in AI

A Mixed-Language Test Set for WhatsApp Assistants in Gulf Businesses

Disclosure: I run NxFlowAI, an automation agency serving UAE businesses remotely from Mumbai. This post is a vendor-neutral testing pattern.

  • Dataset compiled from real UAE customer messages in English, Arabic script, and Latin script
  • Simulated challenging scenarios with language switches, numeral representations, and voice notes
  • Performance evaluated based on routing decisions, not just reply quality

Why Jenni thinks researchers need more than ChatGPT for academic writing

Jenni is an academic research and writing platform used by more than six million researchers, who have written over 15 million papers on it. The US-headquartered company says it passed US$10 million in annual recurring revenue (ARR) this year and is profitable, having grown almost entirely from revenue after raising only a small angel…

More from Friday 9 October →