Brief IA

Claude Fable 5 from Anthropic Outpaces GPT-5.5 on FrontierMath

🤖 Models & LLM·Tom Levy·

Claude Fable 5 from Anthropic Outpaces GPT-5.5 on FrontierMath

Claude Fable 5 from Anthropic Outpaces GPT-5.5 on FrontierMath
Key Takeaways
1Claude Fable 5 from Anthropic outperforms GPT-5.5 by 13 points on difficult FrontierMath problems.
2The model achieves an accuracy of 87% on levels 1 to 3 and 88% on level 4 of FrontierMath.
3In 2026, Opus 4.5 from Anthropic scored less than 10% on level 4, showing a significant improvement.
💡Why it mattersThese advancements strengthen Anthropic's position in the field of mathematical AI, in the face of competition from OpenAI.
Le brief IA que lisent les pros

Le brief IA que les pros lisent chaque soir

Les 7 actus IA du jour, décryptées en 5 min. Gratuit.

Inclus dès l'inscription : notre sélection des meilleurs guides & comparatifs IA.

Choisis ton rythme

Gratuit · Pas de spam · Désabonnement en 1 clic

📄
Full Analysis

The Claude Fable 5 model from Anthropic has recently demonstrated its superiority over OpenAI's GPT-5.5 by achieving remarkable scores on the FrontierMath benchmark. With a lead of 13 points, Fable 5 excels in the most challenging problems of this renowned test.

According to Epoch AI, Claude Fable 5 achieves an accuracy of 87% on levels 1 to 3 of FrontierMath and climbs to 88% on level 4, the most complex. This performance marks a significant advancement for Anthropic, whose previous model, Opus 4.5, had scored less than 10% on the same level at the beginning of 2026.

In comparison, OpenAI's GPT-5.5 model shows about 75% accuracy on level 4, although the GPT-5.6 version is already in development. All these models have been evaluated according to Epoch AI's standard scale, which requires maximum reasoning effort.

FrontierMath is recognized as one of the most demanding benchmarks for mathematical reasoning in artificial intelligences. The progress made by the models from Anthropic and OpenAI is not limited to benchmarks: recently, an OpenAI model solved a long-standing Erdős problem, and Claude Mythos also achieved this feat, illustrating the growing capabilities of these technologies.

Brief IA — L'actualité IA en français

L'essentiel de l'actualité de l'intelligence artificielle, décrypté et expliqué chaque jour.