Hundreds of results from a single prompt
OpenAI has published hundreds of new results on major math problems. According to the company, most of these results were produced from a single prompt given to a single AI agent.
Why it matters
Mathematics is a common yardstick for measuring how well AI models reason. By releasing the results in bulk, OpenAI opens its work up to outside scrutiny over both the model's limits and where it actually helps.
Details remain thin
The announcement did not say which problems are covered, what the accuracy rate is, or whether the results went through peer review. So it is not yet clear how much of the output counts as new mathematical contribution.



