Skip to content
AI Primer
release

Sakana launches Fugu Max to route requests across specialist models

Sakana says Fugu Max dynamically routes requests across open-weight and specialist models at two to six times lower cost than elite models. In the same release, the company says Fugu Ultra v2 led five of eight hard evaluation suites.

3 min read
Sakana launches Fugu Max to route requests across specialist models
Sakana launches Fugu Max to route requests across specialist models

TL;DR

  • Fugu Max and Fugu Ultra v2 divide Sakana Fugu between cost and peak performance, with SakanaAILabs' announcement describing Max as the outward move on the cost-capability frontier and Ultra as the upward one.
  • Fugu Max is priced at $2 per million input tokens and $6 per million output tokens, according to SakanaAILabs' pricing post, which also claims performance within striking distance of elite models at two to six times lower cost.
  • Fugu Ultra v2 reported a 74.3 DeepSWE score and 48.3 on Chartography, versus 27.3 for Opus 5, in SakanaAILabs' benchmark post.
  • Sakana frames the swappable agent pool as protection against vendor lock-in and service cutoffs in hardmaru's launch post.

Sakana's release post says users already on Fugu can select Max or Ultra v2 with a one-line parameter change. The official Max benchmark chart places its claimed wins alongside single models in a similar price range.

One architecture, two operating points

Sakana says Max and Ultra v2 use the same orchestration architecture, optimized for different objectives. Max routes toward the least expensive capable option; Ultra v2 is tuned for complex, multi-step work where raw output quality takes priority.

Fugu Max routing and price

Fugu Max expands the model pool with open-weight and specialist models, including NVIDIA's Nemotron family. Sakana says the router selects the leanest model capable of a task rather than sending every request to one fixed model.

  • Price: $2 per million input tokens and $6 per million output tokens.
  • Relative pricing claim: 40% to 60% lower output pricing than Sonnet 5, GPT 5.6 Terra, and Kimi K3, according to Sakana's release post.
  • Reported wins: best overall score on Terminal Bench 2.1, GPQAD, AA-LCR, GDP.pdf, AutomationBench, and SWEFish. Sakana identifies SWEFish as its own coding-challenge evaluation.
  • Frontier claim: a cost-performance efficient position on seven of ten benchmarks.

Fugu Ultra v2 benchmark claims

Ultra v2 is Sakana's higher-capability setting. The company reports best or joint-best results on five of eight suites, GDP.pdf, Chartography, SWEFish, DeepSWE, and Toolathon, and a top-two placement on seven.

  • Chartography: 48.3, compared with Opus 5 at 27.3 and Fable 5 at 29.5 on Sakana's published Ultra v2 chart.
  • DeepSWE: 74.3. Sakana says it exceeds models costing three to five times more per token.

Training cutoff and pool exclusions

Sakana lists an August 28, 2026 training cutoff for Ultra v2 and says Fable 5, Fable 5.1, and GPT-6-Astra are absent from its agent pool. The release post names those exclusions but does not publish a complete model-and-version inventory for the remaining pool.

Further reading

Discussion across the web

Where this story is being discussed, in original context.

On X· 2 threads
TL;DR1 post
One architecture, two operating points1 post
Share on X