AIToday
Large Language ModelsAI Coding AssistantsTHE DECODERPublished: Jul 25, 2026, 01:00 JST

Sakana's Fugu Ultra v1.1 claims to beat Claude's Fable 5 without using it

Sakana's Fugu Ultra v1.1 claims to beat Claude's Fable 5 without using it

3 Key Points

  1. What happened

    Sakana AI released Fugu Ultra v1.1, a router that distributes queries across a pool of top-tier AI models. The company claims performance gains of up to 7.9 points over v1.0, with the largest improvements on ProgramBench and TerminalBench 2.1. Sakana reports that v1.1 outperforms Anthropic's Fable 5 on most benchmarks, despite Fable 5 not being part of the router's selection pool.

  2. Why it matters

    Fugu Ultra is positioned as a way to get superior performance by intelligently routing requests across existing models rather than building a single large model. Pricing remains at $5 per million input tokens and $30 per million output tokens. However, all claims come from Sakana itself, and no independent third-party verification exists yet—a limitation that matters given Fugu's initial version faced criticism for high token usage, slow speed, and poor results.

  3. What to watch

    Sakana says it takes about two weeks of training and evaluation before a new top-tier model is added to the pool. The update adds a Claude Code-compatible endpoint for terminal use. Fugu is available on OpenRouter and Vercel, but Sakana does not serve the EU or EEA, citing GDPR and EU-specific regulations.

Ask the AI about this article →

Summaries like this, in your inbox every morning.

Context & Analysis

Fugu Ultra v1.1 represents Sakana's attempt to position AI routing—intelligent distribution of queries across multiple models—as a competitive alternative to building or fine-tuning a single large model. The claimed 7.9-point performance gain over v1.0, particularly on ProgramBench and TerminalBench 2.1, suggests optimization is occurring; however, the lack of independent verification is a significant caveat. The fact that Sakana claims v1.1 outperforms Fable 5 without including it in the pool could indicate either genuine architectural efficiency or that the benchmarks themselves may not be representative of real-world use.

The first Fugu version encountered skepticism due to high token usage, slow response times, and disappointing results—issues that a 7.9-point gain may address, but only if verified externally. The addition of a Claude Code-compatible endpoint signals Sakana's intent to integrate more deeply into developer workflows. The two-week evaluation window for adding new models to the pool suggests Sakana has a process to keep Fugu current as new models emerge, though this also means the pool is not static and claims about v1.1's performance may shift as models are added or removed.

FAQ
How much does Fugu Ultra v1.1 cost?
Pricing is $5 per million input tokens and $30 per million output tokens.
How long does it take to add a new model to Fugu's pool?
Sakana says it takes about two weeks of training and evaluation before a new top-tier model is added to the pool.
Where is Fugu available?
Fugu is available on platforms like OpenRouter and Vercel. Sakana does not serve the EU or EEA, citing GDPR and EU-specific regulations.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Amodei's 'pacing the frontier' call splits software from chipsYahoo Finance AI · 2h ago
  • NC State's Amanda Cullen questions AI bans as universities splitFortune AI · 2h ago
  • AI agents cut office work: パナソニック コネクト sheds 78.8万時間AINOW · 2h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleWPP's £500M restructure plan faces culture hurdle as workers resist return-to-office