UPDATED SEPTEMBER 17, 2026
UPDATED SEPTEMBER 17, 2026

The Frontier

Your signal. Your price.

Include
Lookback
||
  • · 7d ago

    Low enterprise adoption of Anthropic's Fable 5 stems from its government-mandated 30-day data retention policy. Whittemore explains that this security-check requirement triggers severe corporate infosec concerns, making older models like Opus preferable despite their higher price.

    +3 more
    +3 more
  • · 21d ago

    David Heinemeyer Hansson marks November 24, 2025, and the release of Opus 4.5 as the start of the agentic era. This shift moved AI from basic autocomplete helpers to autonomous agents capable of writing entire codebases.

    +3 more
    +2 more
  • · 6w ago

    Anthropic launched Claude Fable 5 and Mythos 5 on June 9, introducing a new model class positioned above Opus. While Fable 5 is public, Mythos 5 lacks standard safety guardrails and is restricted to government-aligned Glasswing partners.

    +3 more
    +5 more
  • · 6w ago

    To prevent safety risks, Anthropic automatically routes Fable 5 queries about biology, chemistry, and cybersecurity to Opus 48. Users report that strict classifiers trigger this silent fallback even on harmless terms like mitochondria and cancer.

    +3 more
    +3 more
  • · 6w ago

    Cognition introduced the Frontier Code benchmark to evaluate whether AI-generated code is robust enough to merge into production codebases. On this difficult test, Fable 5 scored 29.3%, more than doubling Opus 48's performance.

    +3 more
    +4 more
  • · 6w ago

    Nufar Gaspar explains that tokens are not equal across providers; the April Opus 4.7 update changed its tokenizer to produce 30-45% more tokens for the same text, causing real-world bills to grow by 12-27% despite an identical price sheet.

    +3 more
    +2 more
  • · 6w ago

    Nufar Gaspar highlights Databricks' experiment where Sonnet 5, though 1.7 times cheaper per token than Opus 4.8, ultimately cost more per task ($2.09 vs. $1.94) due to increased iterations and reasoning, underscoring that "cost per accepted task" is the crucial metric.

    +4 more
    +4 more
  • · 6w ago

    Microsoft announced seven new AI models, including MAI Thinking 1, a "1-trillion parameter model" using a mixture of experts architecture, which they position as competitive in the Sonnet 4.6 to Opus 4.6 range.

    +3 more
    +4 more
  • · 7w ago

    Elon Musk confirmed Grok 4.5's public release, stating it's based on a 1.5 trillion parameter V9 foundation model with Cursor data. It is an "open class model" that is faster, more token efficient, and lower cost, performing close to or exceeding Opus.

    +4 more
    +4 more
  • · 7w ago

    Kimmy K3 shows better output token efficiency than many Anthropic models and Opus 5 in some high-reasoning benchmarks, despite overall token hunger. However, running the trillion-parameter model locally demands substantial hardware, requiring 64 H100 GPUs.

    +4 more
    +4 more
  • · 7w ago

    Opus 5 generates faster than Fable but exhibits "bizarre behaviors" like scope creep and over-engineering, which prolongs real-world completion times. Despite this, Theo prefers its direct communication style over other Claude models and finds it effective for targeted tasks.

    +4 more
    +3 more
  • · 7w ago

    Opus 5 shows significant 3D capabilities, creating a Call of Duty clone and a 3D village in-browser with 3JS, including self-modeled assets and animations. Theo's "fish slop" port demonstrated rapid 2D and 3D game renditions, often with surprising aesthetic taste in animations.

    +4 more
    +3 more
  • · 7w ago

    Ben defaults to 56 Soul for 80% of tasks, using Fable for complex research or uncertain implementations. Theo starts with Opus 5, then switches to Fable for review or cleanup, noting high token consumption with monthly spends of $17,000 (Ben) and $48,000 (Theo).

    +4 more
    +3 more
  • · 7w ago

    Theo observes a significant overhaul in Anthropic's Reinforcement Learning, making Opus 5 behave more like an OpenAI model. He hopes for a Fable 5.1 update that leverages these behavioral wins, allowing Anthropic to create a more machine-like model, moving past its "Constitution."

    +4 more
    +4 more
  • · 7w ago

    Artificial Analysis found Claude Opus 5 on max settings 20% cheaper than Fable 5, costing $17.79 per task. However, on the Artificial Analysis Index, its $2.03 per task made it more expensive than Opus 4.8 and GPT-5-6-Soul.

    +4 more
    +6 more
  • · 7w ago

    The release of Opus 3 was a significant inflection point, proving Anthropic's ability to build a frontier model, particularly by enhancing its coding capabilities, which differentiated it from competitors like GPT-4.

    +3 more
    +3 more
  • · 7w ago

    Opus 4.5 marked another milestone, demonstrating that 'frontier products' like Claude Code are essential to unlock and accelerate the adoption and magical experience of 'frontier models' for users.

    +3 more
    +2 more
  • · 7w ago

    Chinese company Moonshot AI released Kimi K3, an open-source model offering performance comparable to Opus 4.8 and GPT 5.6 at a 50% lower cost, sparking debate in the US.

    +4 more
    +4 more
  • · 7w ago

    Nathaniel Whittemore clarifies Kimi K3 is served at approximately one-third the price of Fable or half the price of Opus, offering meaningful but not negligible savings.

    +3 more
    +4 more
  • · 7w ago

    Ory Goan presented AI21 Labs' research demonstrating that a learned system using a portfolio of models (e.g., Minimax, GPT5.2, Fable) can achieve a new state-of-the-art in coding benchmarks like Swebench Pro, while being three times cheaper than a single model like Opus.

    +4 more
    +7 more
  • · 8w ago

    Frank reports Moonshot's Kimi K3 model, while performing well on benchmarks between Opus/Fable and GPT 5.6, fits the pattern of open-source models lagging the latest frontier. The model is the largest open-source release at 2.8 trillion parameters.

    +4 more
    +6 more
  • · 2mo ago

    Q built freedom.tech using Anthropic's Opus 4.8 model and uses the cheaper Haiku model for daily summaries, with Opus 4.8 for a second pass on weekly recaps to ensure accuracy. (Q)

    +4 more
    +4 more
  • · 2mo ago

    The open-source model GLM 5.2 from ZAI beats GPT-5.5 and Opus 4.8 on some benchmarks at one-tenth the cost, fueling speculation about its distillation from Anthropic models and its viability as a Fable alternative.

    +3 more
    +6 more
  • · 2mo ago

    Cursor's Composer 2.5, built on a Kimi foundation, scores near Opus 4.7 and GPT-5.5 on benchmarks at a fraction of the cost, but user reports and updated agent-focused benchmarks show mixed real-world performance.

    +3 more
    +5 more
  • · 2mo ago

    Harvey's experiment combining an open-weight GLM 5.1 worker with a closed frontier Opus 4.7 advisor increased performance and lowered costs, demonstrating that smart model routing beats using the most expensive model for every task.

    +2 more
    +3 more
  • · 2mo ago

    Fable 5 dominates benchmarks: 78% on ExploitBench versus GPT55’s 34%, 66% on HealthBench versus GPT55’s 51.8%, 13.3% on the legal agent benchmark versus GPT55’s 2.1%, and 1932 on GDP Val’s knowledge work test versus Opus 48’s 1890.

    +3 more
    +2 more
  • · 2mo ago

    The model excels at agentic coding: 80.3% on Swebench Pro versus GPT55’s 58.6%, 88% on Terminal Bench versus GPT55’s 83.4%, and 29.3% on Frontier Code versus Opus 48’s 13.4%. Fable scored 91% on Every’s Senior Engineer benchmark.

    +3 more
    +1 more
  • · 2mo ago

    Artificial Analysis found Fable 5 topped its blended benchmark run, overtaking Opus 48 and GPT55, though some noted the overall gap was only five points.

    +3 more
    +3 more
  • · 2mo ago

    API pricing for Fable 5 is $10 per million input tokens and $50 per million output tokens, double Opus’s cost but lower than some expected. Mythos preview within Project Glasswing costs more than double.

    +3 more
    +3 more
  • · 2mo ago

    Fable 5 has strict guardrails, automatically routing queries about cybersecurity, biology, chemistry, or distillation to Opus 48. Anthropic says 95% of sessions don’t trigger a fallback but is 'hardcore' about biology/chemistry filters.

    +3 more
    +3 more
About The Frontier
End of 90-day results — 49 results
49 results