Your signal. Your price.
Low enterprise adoption of Anthropic's Fable 5 stems from its government-mandated 30-day data retention policy. Whittemore explains that this security-check requirement triggers severe corporate infosec concerns, making older models like Opus preferable despite their higher price.
David Heinemeyer Hansson marks November 24, 2025, and the release of Opus 4.5 as the start of the agentic era. This shift moved AI from basic autocomplete helpers to autonomous agents capable of writing entire codebases.
Anthropic launched Claude Fable 5 and Mythos 5 on June 9, introducing a new model class positioned above Opus. While Fable 5 is public, Mythos 5 lacks standard safety guardrails and is restricted to government-aligned Glasswing partners.
To prevent safety risks, Anthropic automatically routes Fable 5 queries about biology, chemistry, and cybersecurity to Opus 48. Users report that strict classifiers trigger this silent fallback even on harmless terms like mitochondria and cancer.
Cognition introduced the Frontier Code benchmark to evaluate whether AI-generated code is robust enough to merge into production codebases. On this difficult test, Fable 5 scored 29.3%, more than doubling Opus 48's performance.
Nufar Gaspar explains that tokens are not equal across providers; the April Opus 4.7 update changed its tokenizer to produce 30-45% more tokens for the same text, causing real-world bills to grow by 12-27% despite an identical price sheet.
Nufar Gaspar highlights Databricks' experiment where Sonnet 5, though 1.7 times cheaper per token than Opus 4.8, ultimately cost more per task ($2.09 vs. $1.94) due to increased iterations and reasoning, underscoring that "cost per accepted task" is the crucial metric.
Microsoft announced seven new AI models, including MAI Thinking 1, a "1-trillion parameter model" using a mixture of experts architecture, which they position as competitive in the Sonnet 4.6 to Opus 4.6 range.
Elon Musk confirmed Grok 4.5's public release, stating it's based on a 1.5 trillion parameter V9 foundation model with Cursor data. It is an "open class model" that is faster, more token efficient, and lower cost, performing close to or exceeding Opus.
Kimmy K3 shows better output token efficiency than many Anthropic models and Opus 5 in some high-reasoning benchmarks, despite overall token hunger. However, running the trillion-parameter model locally demands substantial hardware, requiring 64 H100 GPUs.
Opus 5 generates faster than Fable but exhibits "bizarre behaviors" like scope creep and over-engineering, which prolongs real-world completion times. Despite this, Theo prefers its direct communication style over other Claude models and finds it effective for targeted tasks.
Opus 5 shows significant 3D capabilities, creating a Call of Duty clone and a 3D village in-browser with 3JS, including self-modeled assets and animations. Theo's "fish slop" port demonstrated rapid 2D and 3D game renditions, often with surprising aesthetic taste in animations.
Ben defaults to 56 Soul for 80% of tasks, using Fable for complex research or uncertain implementations. Theo starts with Opus 5, then switches to Fable for review or cleanup, noting high token consumption with monthly spends of $17,000 (Ben) and $48,000 (Theo).
Theo observes a significant overhaul in Anthropic's Reinforcement Learning, making Opus 5 behave more like an OpenAI model. He hopes for a Fable 5.1 update that leverages these behavioral wins, allowing Anthropic to create a more machine-like model, moving past its "Constitution."
Artificial Analysis found Claude Opus 5 on max settings 20% cheaper than Fable 5, costing $17.79 per task. However, on the Artificial Analysis Index, its $2.03 per task made it more expensive than Opus 4.8 and GPT-5-6-Soul.
The release of Opus 3 was a significant inflection point, proving Anthropic's ability to build a frontier model, particularly by enhancing its coding capabilities, which differentiated it from competitors like GPT-4.
Opus 4.5 marked another milestone, demonstrating that 'frontier products' like Claude Code are essential to unlock and accelerate the adoption and magical experience of 'frontier models' for users.
Chinese company Moonshot AI released Kimi K3, an open-source model offering performance comparable to Opus 4.8 and GPT 5.6 at a 50% lower cost, sparking debate in the US.
Nathaniel Whittemore clarifies Kimi K3 is served at approximately one-third the price of Fable or half the price of Opus, offering meaningful but not negligible savings.
Ory Goan presented AI21 Labs' research demonstrating that a learned system using a portfolio of models (e.g., Minimax, GPT5.2, Fable) can achieve a new state-of-the-art in coding benchmarks like Swebench Pro, while being three times cheaper than a single model like Opus.
Frank reports Moonshot's Kimi K3 model, while performing well on benchmarks between Opus/Fable and GPT 5.6, fits the pattern of open-source models lagging the latest frontier. The model is the largest open-source release at 2.8 trillion parameters.
Q built freedom.tech using Anthropic's Opus 4.8 model and uses the cheaper Haiku model for daily summaries, with Opus 4.8 for a second pass on weekly recaps to ensure accuracy. (Q)
The open-source model GLM 5.2 from ZAI beats GPT-5.5 and Opus 4.8 on some benchmarks at one-tenth the cost, fueling speculation about its distillation from Anthropic models and its viability as a Fable alternative.
Cursor's Composer 2.5, built on a Kimi foundation, scores near Opus 4.7 and GPT-5.5 on benchmarks at a fraction of the cost, but user reports and updated agent-focused benchmarks show mixed real-world performance.
Harvey's experiment combining an open-weight GLM 5.1 worker with a closed frontier Opus 4.7 advisor increased performance and lowered costs, demonstrating that smart model routing beats using the most expensive model for every task.
Fable 5 dominates benchmarks: 78% on ExploitBench versus GPT55’s 34%, 66% on HealthBench versus GPT55’s 51.8%, 13.3% on the legal agent benchmark versus GPT55’s 2.1%, and 1932 on GDP Val’s knowledge work test versus Opus 48’s 1890.
The model excels at agentic coding: 80.3% on Swebench Pro versus GPT55’s 58.6%, 88% on Terminal Bench versus GPT55’s 83.4%, and 29.3% on Frontier Code versus Opus 48’s 13.4%. Fable scored 91% on Every’s Senior Engineer benchmark.
Artificial Analysis found Fable 5 topped its blended benchmark run, overtaking Opus 48 and GPT55, though some noted the overall gap was only five points.
API pricing for Fable 5 is $10 per million input tokens and $50 per million output tokens, double Opus’s cost but lower than some expected. Mythos preview within Project Glasswing costs more than double.
Fable 5 has strict guardrails, automatically routing queries about cybersecurity, biology, chemistry, or distillation to Opus 48. Anthropic says 95% of sessions don’t trigger a fallback but is 'hardcore' about biology/chemistry filters.