722 manuscripts, 372 hard math problems, with each result averaging roughly three hours of ChatGPT Pro usage—that's the scale of what OpenAI has just disclosed. But what's really intriguing is the method: according to what OpenAI told Scientific American, nearly every single paper was produced using one prompt handed to one AI agent.
All of the manuscripts have been uploaded to GitHub, verified by OpenAI's newly formed advisory group, the Advisory Group on Mathematics and Artificial Intelligence (AGMAI). Last month, OpenAI had already signaled its intentions, claiming its model "solved more than 100 long-standing open problems spanning most areas of mathematics." This release essentially lays out the concrete list. The company stated that the results were produced by a ChatGPT pilot model not yet released to the public.
Among the headline results OpenAI claims are a solution to the four-dimensional Kakeya conjecture, improvements to several key computer algorithms, and progress on the Riemann hypothesis—any one of which, if confirmed, would be a major event in the math world. In its announcement, the company wrote that it is starting with a GitHub repository release, complete with mechanisms for paper revisions and citations, "while still seeking other community-hosted options that meet the advisory group's guidelines," and promised that future releases will be more complete in terms of citation standards, mathematical exposition, and presentation.






