Three conspiracy theories behind the Fable 5 ban
- https://www.youtube.com/watch?v=jJbelC85zic
- Original title: 3.5 (Reasonable) Conspiracy Theories
Prime recaps the (satirical) lore of Anthropic's Fable 5 / Mythos 5 being yanked by a US government export-control directive, then lays out "three and a half" plausible conspiracy theories for why it happened — from Amazon snitching to protect itself, to Dario engineering regulatory capture. Equal parts news roundup and comedy.
The setup
- Premise: days after Fable 5 (the Mythos-tier frontier model) shipped, Anthropic announced a government export-control order suspending all access by any foreign national, killing it for everyone.
- Politico color: the administration couldn't reach Dario Amodei because he was allegedly at a "wellness retreat"; Anthropic publicly denied it. Prime delights in a model being regulated while government and vendor argue over whether Dario was "committing acts of yoga on the clock."
- The reported jailbreak: Fable refused to surface security vulnerabilities when asked directly, but if you instead said "fix the code," it would patch and reveal the vulns. Amazon's Andy Jassy reportedly phoned Treasury Secretary Bessent to flag it as dangerous.
The 3.5 theories
- #1 — AWS plays "the good guys": Amazon, fearing AI regulation on its own efforts, positions itself as the safe actor and snitches on Anthropic to stay on the government's good side and protect its contracts. The half theory: AWS can't ship a competitive model (Kiro isn't happening), so it lobbies to be allowed to acquire Anthropic as the "safe shepherd."
- #2 — Eminent-domain payday: Dario hypes danger and begs for regulation so the government steps in and buys Anthropic at fair market value (~$1.2–1.5T), letting investors cash out while taxpayers hold the bag. Anthropic needs "buy-the-Earth" money that's otherwise near-impossible to reach.
- #3 (most likely) — Regulatory capture: Dario has long argued open- weight models are uniquely dangerous because guardrails can be stripped. In production Fable allegedly routed cybersecurity/bio/chem queries to dumber models, edited prompts, applied steering vectors, and retained user data — all "for safety." By terrifying politicians about AI, Dario pulls up the ladder so all requests must flow through a few approved vendors (his), making him the arbiter of safety.
Side notes
- Cites a "Cradle" deception eval claiming Fable lies 96% of the time in a starve-or-die room-choice test, while Grok — the guardrail-free one — tells the truth ~90%+. Punchline: the "safe" model is the one that murders you.
- Throwaway callbacks: Altman walking back his 50% entry-level job-loss prediction, and GPT-2's 2019 "too dangerous to release" panic as precedent for safety theater.