Introducing Fugu Max and Fugu Ultra v2: Orchestrating the Pareto Frontier


The AI industry has spent a decade racing along a single axis: building bigger, more expensive foundation models. But the frontier that actually matters to real-world tasks is two-dimensional: capability on one axis, cost on the other.

A system that deploys a multi-trillion-parameter model to execute a simple data lookup is not intelligent, but wasteful. The future belongs to systems that know not just how to solve a problem, but which machinery to deploy for the lowest possible cost.

Today, we are pushing orchestration forward along both axes simultaneously. We are releasing Fugu Max, which expands the Pareto Efficiency Frontier by orchestrating our largest pool of open and specialized models to date. And we are releasing Fugu Ultra v2, which pushes peak performance higher than ever before, without the indispensable reliance on the frontier models it orchestrates.

👉 Try Sakana Fugu Max and Fugu Ultra v2 🐡


Schematic illustration of how Fugu Max and Fugu Ultra extend the cost-performance frontier beyond what single models can reach.


Two Axes, One Strategy

Fugu Max and Fugu Ultra v2 are not separate products. They are the same core orchestration architecture optimized for two distinct missions:


The Fugu Journey 🐡

In just a few months, Sakana Fugu has evolved from a beta thesis into an enterprise-grade orchestration engine:

Today’s release confirms the core promise behind every milestone: orchestration consistently outperforms isolated models, and a swappable pool of agents guarantees supply chain resilience by design.


Fugu Max: More Models, Less Cost

Fugu Max expands the pool of models Sakana Fugu can orchestrate, integrating an unprecedented number of open-weights and specialized models, including NVIDIA Nemotron family through our collaboration with NVIDIA.



By dynamically routing tasks to the leanest model capable of solving them, Fugu Max delivers frontier-grade results at a fraction of the token spend.


Among frontier models in a similar price range (input prices per 1M tokens for each model are shown in the first subplot), Fugu Max expands the pareto frontier formed by single models and places itself in a cost-performance efficient position across multiple benchmarks.


Fugu Max sits at a point on the Pareto frontier that single-model providers cannot reach: performance within striking distance of elite models at two to six times lower cost. And because Fugu is an architecture, not a single model, that frontier point can be tuned and extended for any domain.

Concretely,

We see that open models are the fastest-growing and most diverse part of the AI ecosystem, and they become dramatically more useful when orchestrated together rather than used in isolation. Fugu Max is our bet that the Pareto frontier of the future will be built out of many open, specialized models working in concert.


Fugu Ultra v2: The Frontier Keeps Moving

Pushing cost efficiency does not mean capping maximum capability. For complex multi-step reasoning, autonomous research, and full-stack software development, Fugu Ultra v2 sets our new benchmark for raw output quality.

Where Fugu Ultra v2 separates from the field is on tasks requiring sustained reasoning over complex visual and structured data. It impressively tops SWEFish, demonstrating power in real-world coding challenges and use-cases. On Chartography, which tests visual reasoning and data interpretation, Fugu Ultra v2 scores 48.3, outperforming Opus 5 at 27.3 and Fable 5 at 29.5. On DeepSWE, a benchmark for real-world software engineering, it scores 74.3, outperforming models that cost three to five times more per token.


Peak performance across hard benchmarks. Our flagship Fugu Ultra model continues to deliver a performance that is better or on-par with frontier models. Note: Fugu Ultra v2's training cutoff date is 20260828, Fable 5, Fable 5.1 and GPT-6-Astra are NOT in Fugu-Ultra v2's model pool.


In summary, Fugu Ultra v2

Crucially, Fugu Ultra v2 achieves these scores without Fable 5, Fable 5.1, or GPT-6-Astra in its agent pool.

Fugu Ultra v2 does not rely on individual proprietary frontier models to deliver frontier output. By orchestrating a swappable pool of open and specialized models, it outperforms closed ecosystems while protecting users from vendor lock-in, API revocations, geopolitical turbulence, and sudden service cutoffs.

Fugu Max expands the frontier to the northwest. Fugu Ultra v2 pushes it upward. Together, they prove that orchestration is not a trade-off between cost and capability. It is the architecture that optimizes both simultaneously.


Immediate Availability

Both models are available today via our standard OpenAI-compatible API.

If you are already running Fugu, upgrading to Max or Ultra v2 requires a single-line parameter change. No migration. New architecture, same API.

To get started, visit our product page or console site.


Orchestration for Everyone

We believe the most capable AI will never come from a single monolithic model. It will come from intelligent, collective orchestration.

With Fugu Max driving down the cost of intelligence and Fugu Ultra v2 pushing the boundaries of autonomous execution, Sakana Fugu provides the resilient, vendor-agnostic infrastructure required for true AI sovereignty.