Looks like 20% of the results are disproofs/counterexamples.
This is itself a disproof of people who said that the recent math breakthroughs are concentrated among counterexamples because models are only good at (brute force) search
A breakdown of OpenAI's released internal-model math results shows about 73 disproofs and counterexamples, roughly 20% of the total. The author argues this counters claims that recent math breakthroughs are concentrated in counterexamples because models are only good at brute-force search.
Looks like 20% of the results are disproofs/counterexamples.
This is itself a disproof of people who said that the recent math breakthroughs are concentrated among counterexamples because models are only good at (brute force) search
We’re releasing a broad range of new mathematical results produced by an internal frontier model. We’ve been consulting with the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study, and we have drawn on their advice and public recommendations to inform how we release these results. https://github.com/openai/mathView quoted post on X
Source: wh · x.comPublished · added here