changelogSakana Ai

Sakana AI Claims Fugu Ultra v1.1 Router Beats Anthropic's Fable 5 Without Including It in the Pool

TL;DR

Sakana AI has released Fugu Ultra v1.1, an update to its multi-model router, claiming performance gains up to 7.9 points over v1.0 and results that beat Anthropic's Fable 5 despite Fable 5 not being part of the router's model pool. All benchmark figures come from Sakana itself and remain unverified.

2 min read
0

Sakana AI has released Fugu Ultra v1.1, an update to its AI model router that distributes each query across a pool of publicly available top-tier language models. According to Sakana, the new version delivers performance gains of up to 7.9 points over v1.0, with the largest improvements on ProgramBench and TerminalBench 2.1.

The company claims Fugu v1.1 now beats Anthropic's Fable 5 across most benchmarks — a notable claim given that Fable 5 is not included in Fugu's selection pool of underlying models. Sakana has not published the full benchmark tables or independent verification for these results. All performance figures currently come from Sakana's own testing.

What Fugu does

Fugu Ultra works as a router rather than a standalone model: it evaluates each incoming query and dispatches it to whichever model in its pool it judges best suited to handle that specific request, then returns the result. This approach lets Sakana claim state-of-the-art performance by leveraging whichever frontier model currently performs best on a given task category, rather than training a single foundation model from scratch.

Pricing remains unchanged at $5 per million input tokens and $30 per million output tokens. Sakana says it takes roughly two weeks of training and evaluation before any new top-tier model gets added to the router's pool — which explains, in the company's telling, why a model as recent as Fable 5 hasn't been incorporated yet despite the router allegedly outperforming it.

Sakana has published a technical report describing the router's architecture, though the underlying selection and orchestration mechanisms have not been independently audited.

New in this update

Fugu v1.1 adds a Claude Code-compatible endpoint, allowing developers to call the router directly from the terminal alongside existing coding workflows. Fugu has been available since launch on platforms including OpenRouter and Vercel.

The first version of Fugu received a lukewarm reception. Reviewers and users flagged high token consumption, slow response times, and inconsistent output quality — criticisms Sakana has not directly addressed in this update's announcement beyond the claimed benchmark improvements.

Sakana AI continues to exclude the EU and EEA from Fugu's availability, citing GDPR and unspecified "EU-specific regulations."

What this means

Router-based products like Fugu sidestep the cost and complexity of training frontier models by instead orchestrating access to models built by others — but that also means their competitive claims are entirely dependent on which models sit in the pool and how quickly new ones get added. A two-week onboarding lag means Fugu's benchmark comparisons will regularly involve outdated competitor lineups, which raises questions about how meaningful a claim of beating an excluded model actually is.

Until Sakana publishes verifiable benchmark data or third parties replicate the results, the performance claims for v1.1 should be treated as marketing rather than established fact. The persistent complaints about token usage and latency from v1.0 are also unresolved in this release, which matters more for developer adoption than head-to-head benchmark wins against models not actually available for direct comparison.

Related Articles

changelog

Anthropic reverses course, makes Claude Fable 5 permanent on subscription plans

Anthropic announced July 18 that Claude Fable 5 will remain available on subscription plans, reversing its previous decision to make the model API-only. Max and Team Premium subscribers will receive access at 50% of standard limits starting July 20, while Pro and Team Standard users get a one-time $100 credit.

changelog

Anthropic launches rupee pricing for Claude in India at ₹2,000/month, its second-largest market

Anthropic has begun displaying rupee-denominated pricing for Claude subscriptions in India, its second-largest market after the US with 5.8% of global usage. Claude Pro is priced at ₹2,000 ($21) monthly when billed annually, compared to $17 in the US, with Indian prices including local taxes.

changelog

Anthropic extends Claude Fable 5 access through July 19 amid GPT-5.6 Sol competition

Anthropic has extended Claude Fable 5 access on all paid plans through July 19, 2026, marking another extension of the advanced model's availability. The extension comes after OpenAI released GPT-5.6 Sol, which is classified in the same Fable/Mythos model tier.

changelog

US lifts export controls on Claude Fable 5, Anthropic to restore access July 1

Anthropic will restore access to Claude Fable 5 on July 1, 2026, after the US Department of Commerce lifted export controls that forced the company to disable the model on June 12. The controls were imposed after Amazon researchers allegedly demonstrated that specific prompts could elicit information useful for cyberattacks.

Comments

Loading...