Epoch AI’s FrontierMath benchmark now includes roughly 50 unsolved research-level mathematics problems in its Open Problems track, and AI systems have managed to crack exactly three of them. That’s a 6% success rate against questions that have stumped professional mathematicians for years.
The benchmark represents one of the most ambitious attempts to measure whether AI can do more than regurgitate textbook solutions and can actually contribute to the frontier of human knowledge in pure mathematics.
What FrontierMath actually tests
The Open Problems component draws from active research areas including combinatorics, number theory, algebraic geometry, and topology. These aren’t exercises with known answers sitting in the back of a textbook. They’re genuine open questions that working mathematicians haven’t been able to resolve.
One of the three problems AI has solved involved a Ramsey-style hypergraph problem, a category of combinatorial mathematics that deals with the conditions under which order must appear within chaos.














