Claude Fable 5 surpasses OpenAI GPT 5.5, Google Gemini 3.5 Pro, and Claude Mythos Preview across multiple benchmarks; scores include 80.3% on SWE-Bench Pro.
OpenAIAnthropic

Claude Fable 5 surpasses OpenAI GPT 5.5, Google Gemini 3.5 Pro, and Claude Mythos Preview across multiple benchmarks; scores include 80.3% on SWE-Bench Pro.

Anthropic's Claude Fable 5 has outperformed OpenAI's GPT 5.5, Google Gemini 3.5 Pro, and Claude Mythos Preview in various benchmarks, including an impressive 80.3% score on SWE-Bench Pro and 1932 on GDPval-AA, showcasing its superior performance in software engineering and analytical tasks.

CuriousCats Full Story

Anthropic's Claude Fable 5 has significantly surpassed competitors in a series of benchmarks. On SWE-Bench Pro, a vital measure of software engineering capabilities, Fable 5 achieved 80.3%, outpacing Claude Mythos Preview at 77.8%, OpenAI's GPT 5.5 at 58.6%, and Google Gemini 3.5 Pro at 54.2%.12

In complex analytical tasks, Fable 5 scored 1932 on GDPval-AA, leading over Mythos Preview's 1869, GPT 5.5's 1769, and Gemini 3.5 Pro's 1314.3

When tested on GDPpdf, assessing document and visual reasoning, Fable 5 achieved 29.8%, compared to 24.9% for GPT 5.5 and 16.7% for Gemini 3.5 Pro. In another assessment, Fable 5 scored 78.0%, versus 69.0% for Mythos Preview and 34.0% for GPT 5.5.45

These results illustrate Anthropic's advancements, reinforcing Claude Fable 5's position at the forefront of AI technology across multiple domains, including coding, reasoning, vision, cybersecurity, biology, and automation tasks.

Key Insight
“Claude Fable 5 has demonstrated strong performance against its competitors, scoring 80.3% on SWE-Bench Pro and outpacing others in complex analytical tasks. Its benchmark results across various metrics highlight its superiority in the AI landscape.”
CuriousCats studied:
1
Moneycontrol.com
“Anthropic has introduced Claude Fable 5 and shared benchmark results comparing it with OpenAI GPT 5.5, Google Gemini 3.5 Pro, and Claude Mythos Preview across coding, reasoning, vision, cybersecurity, biology, and automation tasks.”
Moneycontrol.com →
Ask CuriousCats
What is Claude Fable 5?
How did Claude Fable 5 perform on benchmarks?
Why is SWE-Bench Pro significant?
Are there other AIs with similar scores?
How does Claude Fable 5 compare with GPT 5.5?
Become the most informed
person in the room.
Personal AI agents scanning 100,000+ sources — news, video, and social media — delivered every morning.
Download the App Go to CuriousCats.ai
🇺🇸 US🇮🇳 India🇬🇧 UK🇨🇦 Canada🇸🇬 Singapore