Table of Contents
“We call on the world to pause.” — Anthropic, June 4, 2026
“Fable 5 is now available at $50/million output tokens.” — Anthropic, June 9, 2026
“Fable 5 must be taken offline immediately.” — US Government, June 12, 2026
Three statements. Ten days. Every possible contradiction packed into one story.
TL;DR
Anthropic co-signed a paper calling for a conditional global pause on frontier AI development on June 4. Five days later they launched Fable 5, their most powerful public model. Four days after that, the US government forced it offline citing a known jailbreak and national security concerns. Pause call to government shutdown: 10 days.
What Happened
June 4 — The Pause Paper
When AI Builds Itself, co-authored by Anthropic staff and co-founder Jack Clark, made the case for caution:
- Claude writes 80%+ of commits in Anthropic’s own codebase
- Engineers are shipping 8x more code than their 2021–2025 baseline
- Model task horizon doubles every 4 months (Opus 3: 4-minute tasks → Opus 4.6: 12-hour tasks)
- Jack Clark’s personal estimate: 60% probability of recursive self-improvement by 2028
The paper called for governments and AI labs to coordinate a “conditional global pause”—halt frontier training if capability crosses certain thresholds.
Simultaneously, Anthropic was in active IPO preparation with a $47B annual revenue run rate.
June 9 — Fable 5
Five days after the pause call, Fable 5 launched. It is Anthropic’s strongest public model to date: SWE-bench Verified 88.6%, enterprise and science-optimized, priced at $50/million output tokens. Mythos 5 (restricted access) launched alongside it.
Five days after “we call for a pause.”
June 12 — Shutdown
At 5:21 PM ET on June 12, a US export control directive arrived. Reason: a known jailbreak could bypass Fable 5’s safety controls. Shutdown effective immediately, globally—including foreign nationals at US offices. Prorated refunds issued.
Anthropic’s response: the jailbreak was “relatively simple” and also worked on GPT-5.5; the government action was a “misunderstanding.”
Days since launch: 4.
Why This Matters
Safety messaging and shipping pressure are simultaneously true—and in tension. Anthropic can genuinely believe AI poses existential risk and genuinely need to ship to stay competitive. The paper and the launch don’t have to be cynical moves. But holding both positions at once doesn’t resolve the tension—it just makes it more visible.
Reactive governance can’t keep pace. A shutdown order arriving four days after launch, triggered by a reported jailbreak, is not a monitoring system—it’s an alarm that went off after the door was already open. With task horizons doubling every four months, governance needs architectural rethinking, not faster manual review.
“Conditional pause” needs an implementation spec. Who determines when the threshold is crossed? What’s the actual pause mechanism? Without answers, it’s a moral statement, not a policy. The paper describes what should happen in principle; what’s missing is how it would actually work.
Technical Angle
The specific jailbreak that triggered the government shutdown wasn’t disclosed publicly, but Anthropic’s framing—“relatively simple, also works on GPT-5.5”—reveals something important: the government’s risk threshold is not the same as the industry’s.
Anthropic considered the jailbreak low-severity by comparison. The government considered it high-severity in absolute terms. As models gain more autonomous capability, “what can this jailbreak actually do” gets a more serious answer over time. A jailbreak that enables minor misuse on a chatbot becomes more dangerous on a model that can autonomously run 12-hour agentic tasks.
What to Watch
- Under what conditions does Fable 5 come back online?
- Does GPT-5.5 face equivalent scrutiny for the same jailbreak?
- Does the IPO timeline slip given the regulatory friction?
- Will the pause paper get any concrete implementation from governments or standards bodies?
References
Answers come from this article only. Click any prompt below or open the chat at the bottom right.
🇺🇸 English
Three sentences. Ten days apart. And together they tell one of the strangest stories in AI this year.
June 4th, Anthropic: "We call on the world to pause." June 9th, Anthropic: "Fable 5 is now available, fifty dollars per million output tokens." June 12th, the US government: "Fable 5 must be taken offline immediately." Every possible contradiction, packed into a week and a half.
Let's walk through it, because each beat is wilder than the last.
Start with June 4th — the pause paper. It's called *When AI Builds Itself*, co-authored by Anthropic staff and co-founder Jack Clark, and it's genuinely a caution document. The headline numbers are striking. Inside Anthropic's own codebase, Claude now writes more than eighty percent of the commits. Their engineers are shipping roughly eight times more code than their baseline from the 2021 to 2025 era. And here's the metric that should make you sit up: the length of task a model can handle on its own is doubling roughly every four months. To put that in human terms — Opus 3 could reliably chew through about a four-minute task. Opus 4.6? Twelve-hour tasks. That's the curve. And Jack Clark's personal estimate is a sixty percent chance of recursive self-improvement — AI meaningfully improving AI — by 2028.
So the paper asks governments and labs to coordinate what it calls a "conditional global pause." Basically: if capability crosses certain thresholds, you halt frontier training. Sober, careful, a little alarming.
Now here's the twist sitting right underneath it. At that exact moment, Anthropic was deep in IPO preparation, running at a forty-seven billion dollar annual revenue run rate. Hold that thought.
Because five days later — June 9th — they launched Fable 5. Their most powerful public model to date. On SWE-bench Verified, a real-world coding benchmark, it scores 88.6 percent. It's tuned for enterprise and science work, priced at fifty dollars per million output tokens, and it shipped alongside a restricted-access sibling called Mythos 5. Five days after "we call for a pause," they shipped the most capable thing they'd ever released to the public.
And then June 12th, 5:21 PM Eastern. A US export control directive lands. The reason: a known jailbreak could bypass Fable 5's safety controls, and they're citing national security. Shutdown, effective immediately, worldwide — including for foreign nationals sitting in US offices. Refunds prorated and issued. Four days after launch, the door slams shut.
Anthropic's response? Essentially a shrug. They said the jailbreak was "relatively simple," that it also worked on GPT-5.5, and that the whole government action was a "misunderstanding."
So what do we actually make of this? Let me pull out the threads that matter.
First — and this is the uncomfortable one — the safety messaging and the shipping pressure can both be completely sincere. Anthropic can genuinely believe AI is an existential risk *and* genuinely need to ship to survive as a company. The pause paper and the product launch don't have to be cynical. But believing both at once doesn't dissolve the tension — it just makes it impossible to hide. You published the warning and the product in the same week.
Second — reactive governance simply can't keep up. Think about what that shutdown order actually was: an alarm that went off four days *after* the model was already out in the world, live, being used. That's not a monitoring system. That's someone noticing the door was already open. And when task horizons are doubling every four months, manual review that arrives days late isn't a slow version of the right answer — it's the wrong architecture entirely.
Third — that phrase "conditional pause" needs an actual instruction manual. Who decides the threshold has been crossed? What is the literal mechanism to halt training? Who enforces it across competing labs and competing countries? Without those answers, a conditional pause is a moral statement, not a policy. The paper describes what *should* happen in principle. What's missing is how it would ever work in practice.
Now, the technical angle, because there's something subtle here. The specific jailbreak was never disclosed publicly. But listen to how the two sides framed it. Anthropic: "relatively simple, works on GPT-5.5 too" — meaning, low severity, no big deal. The government: severe enough to pull it globally, immediately. That gap tells you the industry and the government are measuring risk on completely different scales.
And here's why the government might not be wrong. A jailbreak on a chatbot gets you some bad text. Annoying, limited. But the *same* jailbreak on a model that can autonomously run a twelve-hour agentic task — that's a different question entirely. When the model can actually go do things in the world for half a day on its own, "what can this jailbreak accomplish" gets a much scarier answer. The capability curve doesn't just make models more useful. It makes every unpatched hole more dangerous.
So a few things worth watching from here. Under what conditions does Fable 5 come back online — and what changes before it does? Does GPT-5.5 get the same regulatory scrutiny for the same jailbreak, or does enforcement land unevenly? Does the IPO timeline slip now that there's real regulatory friction on the table? And does that pause paper ever get turned into something concrete by a government or a standards body — or does it stay a well-written PDF?
Let me leave you with the three things to actually hold onto.
One: sincerity and self-interest can coexist. Anthropic warning about AI risk while shipping frontier models isn't necessarily hypocrisy — but running both at full speed makes the contradiction visible to everyone.
Two: governance that arrives four days after launch isn't governance, it's an autopsy. With capability doubling every few months, the whole approach needs rethinking, not just faster paperwork.
And three: a "conditional pause" without a spec — without a threshold, a mechanism, and an enforcer — is a sentiment, not a safeguard. The hard part was never agreeing that we might need to stop. The hard part is deciding who gets to pull the brake, and whether anyone would actually listen when they do.
Ten days. Three sentences. And not one of them has been resolved.
🇹🇼 中文
上週,Anthropic 才站上舞台,苦苦懇求全球的 AI 實驗室——大家一起在前沿 AI 開發上,裝一個「協調式的煞車踏板」吧。理由是:他們擔心模型正危險地逼近所謂的「遞迴自我改善」,也就是 AI 開始能自己改自己、然後越改越強。
結果呢?才過了一週,他們就把那顆煞車丟進碎木機,油門一路踩到底——發布了號稱人類史上最強的 AI 模型,Claude Fable 5。
這集我們不聊那些玩笑,只把真正的技術重點拆給你聽:Fable 5 到底是什麼、它和 Mythos 5、還有 Opus 4.8 差在哪、以及,為什麼偏偏挑這個時間點推出。
先講最關鍵、也最容易被行銷話術蓋過去的一個事實:Fable 5 和 Mythos 5,其實是同一個底層模型。
差別只有一個字——「口罩」。
Mythos 5 是那個「mythos 等級」的原始模型,只開放給受控存取。而 Fable 5,用影片裡的講法,是被「安全地做了額葉切除」、給一般使用者用的版本。它外面套了一組 classifier 模型,會盯著你送出的每一則請求。
一旦你的請求踩進這幾個領域——資安、生物、化學,或者模型蒸餾——這則請求就會被當場攔下、核掉,改由 Claude Opus 4.8 來回答。
所以你可以想像那個流程:你丟出一個 query,先進到 classifier 檢查站。是一般任務,就交給 Fable 5 這個底層模型處理;但只要碰到那四個敏感領域,立刻轉手給 Opus 4.8。
換句話說,Fable 5「更弱」的地方,不是它的能力本身,而是它願不願意在敏感領域出力。這也帶出一個很有意思的副作用:像 DeepSeek、Kimi 這類中國模型,短時間內沒辦法靠蒸餾,弄出一個開源的 Fable 級模型——因為你才剛動念想蒸餾,classifier 就先把你擋在門外了。
接下來講價格,這裡藏著很聰明的一手。
Fable 5 跟幾週前才發布的 Opus 4.8 比,最直接的差別就是——貴一倍。Opus 4.8 是每百萬 output token 二十五美元,Fable 5 直接翻到五十美元。
行銷手法也很有戲:如果你「現在」手上有付費的 Claude 方案,你可以一路免費用 Fable 5,用到六月二十二號。過了這天,Fable 5 就從方案裡拿掉,之後你想用,只能按 token 另外計費。用「限時可用」硬生生催出 FOMO,把人推去訂閱。這一步,很準。
那實際評價呢?在軟體工程圈,Fable 5 上線初期的口碑相當正面。影片裡最強的一個背書,來自 Bend 的作者——Bend 是一個給 GPU 用的程式語言。他形容這是自己的「奇異點時刻」:Fable 5 直接把他的程式碼從頭掃過一遍,然後實作出大幅的效能改進。
至於那些「Fable 屌打 GPT-5.5」的 coding benchmark,影片自己都用「trust me, bro」的口吻標註了——意思就是,這些數字先聽聽就好,別當定論。原始素材根本沒給出具體分數,所以我這邊也不幫它編。
好,那最核心的疑問來了——為什麼是現在?
把「呼籲全球暫停」跟「推出史上最強模型」擺在同一週看,那個違和感是很強的。影片給的一個解讀框架是:上市時間點。這一週被形容成「歷史性的一週」,SpaceX 週五要 IPO,而 Anthropic 自己,也正走在公開上市的路上。
於是問題就變成:Fable 5 到底是真正的技術突破,還是 IPO 前,用來把數字衝漂亮的一步棋?影片作者的態度是——保持懷疑,自己實測看看,而不是照單全收。另外補一個題外話:Sam Bankman-Fried 曾經是 Anthropic 的早期投資人,一度持有大約百分之八的股份。
那我們收個尾。
Fable 5 這件事,剝掉行銷包裝之後其實很乾淨,記住三點就好。
第一,它不是一個全新的底層模型,它就是 Mythos 5 加上一層 classifier 口罩。
第二,這裡講的「更安全」,具體意思是:敏感領域的請求會被降級、丟給 Opus 4.8 回答,順便把蒸餾這條路堵死。
第三,定價翻倍加上限時開放,是非常明確的商業與行銷設計,不是巧合。
一邊喊全球踩煞車、一邊推出史上最強模型,這中間的張力是真實存在的——而且它不會因為模型名字叫 Fable、叫「寓言」,暗示自己「不是真的」,就這樣消失。所以最後留給你的問題也很簡單:這是奇異點的前奏,還是又一輪 hype cycle 而已?
Tags
Related Articles
Anthropic's Wildest Ten Days: IPO, Pause Call, Fable 5, Government Shutdown
In ten days, Anthropic filed an IPO, called for a global AI pause, launched Fable 5, and watched it get forced offline by the US government—all contradictions compressed into one week.
J-lens: Anthropic's New Interpretability Tool for Reading Claude's Inner Thoughts via a 'Global Workspace'
Anthropic proposes J-lens, an interpretability tool that captures the 'verbalizable' representations inside a Transformer, and uses it to show that Claude contains a privileged subspace analogous to the neuroscientific 'global workspace' — a small set of vectors that broadcast, drive reasoning, respond to external steering, and even leak signals during deception and evaluation awareness.
Claude Opus 4.8: What "Lying Machine No More" Actually Means
Opus 4.8's headline improvement is a 4x reduction in the probability of letting code flaws pass silently—plus Dynamic Workflows for parallel subagents and Effort Control for cost tuning.