GPT-6 Astra clears the last unsolved FrontierMath Tier 4 problem
GPT-6 Astra has solved the final FrontierMath Tier 4 problem that had never previously been cracked by any AI system, pushing Epoch AI to label the tier "saturated." OpenAI reported a 97.6% score for Astra on Tier 4, short of a perfect run in a single evaluation. But under Epoch AI’s cumulative method, saturation means every problem in the tier has been solved at least once across attempts by different models over time. Astra supplied the missing piece by answering the only remaining unsolved question. The result marks a sharp shift from the benchmark’s early days. When Tier 4 launched on July 11, 2025, the best score on the board was about 5%. FrontierMath itself was created in November 2024 to keep advanced models from quickly exhausting older math benchmarks such as GSM8K and MATH. Epoch AI worked with more than 60 mathematicians, including Fields Medal winners Terence Tao, Timothy Gowers, and Richard Borcherds, to build original problems. After an audit, Epoch released a v2 update in June 2026 that fixed 12 Tier 4 questions and removed 7, leaving 43. Even after those changes, model performance kept climbing: GPT-5.6 Sol reached 83.0%, Claude Fable 5 hit 90.2%, and GPT-6 Astra posted 97.6%. FrontierMath has now shifted toward harder settings, including Open Problems and FrontierMath Erdős, where Astra solved 2 of 68 tasks.








