---
title: "Fable 5 vs GPT 5.5: Anthropic's model dominated every benchmark, then the government pulled it"
url: https://daily.dev/posts/fable-5-vs-gpt-5-5-anthropic-s-model-dominated-every-benchmark-then-the-government-pulled-it-vqslsqsud
source_url: https://thenextweb.com/news/anthropic-fable-5-vs-openai-gpt-5-5-benchmark-comparison
type: article
source: "The Next Web"
published: 2026-06-14T10:34:28.941Z
updated: 2026-06-17T15:50:11.096Z
tags: ["ai", "llm", "openai", "anthropic", "ai-regulation"]
reading_time: 4
upvotes: 34
comments: 14
language: en
---

> ## Documentation Index
> Fetch the complete documentation index at: https://daily.dev/llms.txt
> Use this file to discover all available pages before exploring further.

# Fable 5 vs GPT 5.5: Anthropic's model dominated every benchmark, then the government pulled it

**[The Next Web](https://daily.dev/sources/tnw)** · 4 min read · 34 upvotes · 14 comments

## Summary

Anthropic's Fable 5 model outperformed OpenAI's GPT 5.5 on every major AI benchmark — including a 22-point lead on SWE-Bench Pro (80.3% vs 58.6%) and a 98 Elo point lead in Code Arena — but was pulled offline by a US government export control directive just three days after launch on June 9. The government cited a jailbreak vulnerability, which Anthropic disputes as minor and achievable by GPT 5.5 without bypasses. GPT 5.5 now holds the top spot by default, despite being the inferior model. GPT 5.5 does offer a cost advantage at half the price of Fable 5, and performs comparably on Terminal-Bench 2.0. Fable 5's return depends on ongoing negotiations with the government over the export control classification.

## Full article

daily.dev links to this article rather than hosting it. Read it at the original source: <https://thenextweb.com/news/anthropic-fable-5-vs-openai-gpt-5-5-benchmark-comparison>

## Community discussion

Top comments from developers on daily.dev.

**@jibonkrishnaroy** · 0 upvotes

> Omg,
>
> Thanks for this information

**@petermrozek** · 0 upvotes

> What about the "AI benchmark of shame" - the ARC-AGI-3 benchmark? Last time I've checked the best score (by Opus) was 1.5% (it's not a typo). 😉 I haven't seen Fable in there at all (they probably didn't make it in time).

---

Tags: [#ai](https://daily.dev/tags/ai), [#llm](https://daily.dev/tags/llm), [#openai](https://daily.dev/tags/openai), [#anthropic](https://daily.dev/tags/anthropic), [#ai-regulation](https://daily.dev/tags/ai-regulation)

[View this post on daily.dev](https://daily.dev/posts/fable-5-vs-gpt-5-5-anthropic-s-model-dominated-every-benchmark-then-the-government-pulled-it-vqslsqsud)
