Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vercel AI Gateway logo

Model details

Fugu Ultra v2

Fugu Ultra v2 sits in Sakana's fugu family as a performance-oriented tier built for hard, high-stakes questions. According to Sakana's positioning on the listing page, the v2 release coordinates a deeper pool of expert agents than its siblings, with the explicit goal of squeezing more answer quality out of difficult reasoning tasks in exchange for a higher compute and cost footprint. That framing makes it less of a general-purpose workhorse and more of a specialist that a team would route to when an ordinary model is likely to fall short.

In practice, the model is aimed at workflows where reasoning depth, tool use, structured output, and visual input all matter on the same request, and Sakana markets it as ready for long-context, multi-step problem solving with implicit caching to help tame the cost of repeated prompts. The trade-off is real: Fugu Ultra v2 is positioned as the priciest option in the family, so it is best suited to selective use on critical prompts rather than as a default backend. Teams evaluating it should weigh its expert-agent coordination and reasoning focus against the steeper cost relative to lighter fugu variants.

Vercel AI Gatewaysakana/fugu-ultra-v2fugu

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
sakana/fugu-ultra-v2
Release date
Sep 10, 2026
Last updated
Sep 10, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$5.00
Output token cost
$30.00

Limits

Output tokens
1,000,000 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare Fugu Ultra v2 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Fugu Ultra v2

Vercel AI Gateway

Official sourceAnnouncement

Sakana AI announced Fugu Ultra v2 on September 11, 2026, alongside Fugu Max, as part of its orchestration-based model family. Both share the same core orchestration architecture but are optimized for different missions: Fugu Max targets the lowest possible cost per task, while Fugu Ultra v2 targets the absolute highest The announcement explicitly states that Fugu Ultra v2 achieves its results "without Fable 5, Fable 5.1, or GPT-6-Astra in its agent pool," meaning it does not depend on those frontier models to deliver high-capability output. Sakana frames this as evidence that an orchestration layer can match or surpass closed ecosyst

Vercel AI Gateway

Coverage

BigGo Finance's September 14 coverage restates Sakana's September 11 announcement, confirming that Fugu Ultra v2's 74.3 DeepSWE score edges out GPT-6-Astra's 74.1 and Claude Fable 5.1's 67.4, and its 48.3 Chartography score substantially exceeds Claude Opus 5's 27.3. The piece notes that existing users can migrate to t API pricing is reported as $5 input / $30 output per million tokens for Ultra v2 and $2 / $6 for the cost-optimized Fugu Max, with Fugu Max claiming 40–60% lower per-output costs versus Sonnet 5, GPT 5.6 Terra, and Kimi K3. The article references Sakana's June technical report describing Fugu itself as a language model

OpenRouter

Coverage

Yahoo Finance / Forkast News reports that on September 11, 2026, Sakana AI launched Fugu Max v1.0 and Fugu Ultra v2.0, framing both as orchestration engines rather than monolithic models built on learned model coordination from ICLR 2026 papers TRINITY and Conductor. Fugu Max is priced at $2 per million input tokens an On benchmarks, the piece reports Fugu Max achieves the best overall score on six evaluations (Terminal Bench 2.1, GPQAD, AA-LCR, GDP.pdf, AutomationBench, SWEFish), while Fugu Ultra achieves best or joint-best on five of eight benchmarks including GDP.pdf, Chartography, DeepSWE, Toolathon, and SWEFish. The article argu

OpenRouter

Coverage

The Unwind AI newsletter for September 11, 2026 covers OpenAI's Codex harness becoming an Agents API, then shifts to Sakana AI's launch of Fugu Max and Fugu Ultra v2 as orchestration models that route work across a swappable pool of open and specialized models. It states Fugu Max approaches elite-model performance at 2 The newsletter also covers Edge0 running a 35B sparse mixture-of-experts model on Apple Silicon by streaming weights from SSD, achieving 14.9-17.7 tokens per second on an M4 Pro Mac mini with about 2.9 GB of active memory from a 23 GB on-disk footprint, and DeepSeek's V4.1-Flash, the smallest model in a new architectur

Vercel AI Gateway

Coverage

DataNorth's write-up details that both Fugu Max and Fugu Ultra v2 are released as hosted API only on September 11, 2026, and are currently unavailable in the EU or EEA. Fugu Max costs $2 per million input tokens and $6 per million output tokens, which Sakana claims undercuts Sonnet 5, GPT 5.6 Terra, and Kimi K3 by 40–6 For Fugu Ultra v2 specifically, DataNorth reports base pricing at $5/$30 per million input/output tokens with a long-context surcharge kicking in above 272,000 tokens (where pricing rises to a higher tier), and no published cached input rate. The piece reinforces the architectural framing that Fugu is an orchestrator m

OpenRouter

Coverage

There's An AI For That (TAAFT) documents Fugu Ultra v2 as the capability-focused member of Sakana AI's second-generation Fugu family, released alongside the cost-optimized Fugu Max and announced September 10, 2026. The directory entry describes it as a learned orchestrator built on Sakana's ICLR 2026 TRINITY and Conduc On benchmarks, the entry reports Fugu Ultra v2 achieves the best or joint-best score on five of eight benchmarks tested, including a Chartography score of 48.3 versus 27.3-29.5 for competing models and a DeepSWE score of 74.3, placing in the top two on seven of eight benchmarks overall. The model's agent pool deliberat

OpenRouter

Coverage

AI Weekly's alert reports that Sakana AI shipped Fugu Max on September 11, 2026, priced at $2 per million input tokens and $6 per million output tokens, which the company claims runs 40-60% below the output cost of Sonnet 5, GPT 5.6 Terra, and Kimi K3. The system routes each task across a swappable pool of open-weight On Fugu Ultra v2 specifically, the piece pairs it with Fugu Max as a higher-capability variant on the same routing architecture, quoting 48.3 on Chartography against Opus 5's 27.3 and 74.3 on DeepSWE. Sakana frames Ultra v2 as "What is the absolute highest capability we can achieve on complex, multi-step tasks?" versus

Vercel AI Gateway

Coverage

AI/TLDR's model profile for Fugu Ultra v2 explicitly names the model id fugu-ultra-v2.0 and characterizes it as the capability-first tier of Sakana AI's Fugu family, released on September 11, 2026 alongside the cost-first Fugu Max. It confirms Fugu is "a Multi-Agent System, Delivered as One Model" that coordinates a di The profile reports published benchmarks of 48.3 on Chartography (versus Opus 5 at 27.3 and Fable 5 at 29.5) and 74.3 on DeepSWE, plus leadership on SWEFish, all sourced from Sakana. Pricing is $5 input / $30 output / $0.50 cached input per million tokens at base tier, rising to $10/$45/$1.00 for contexts above 272K to

OpenRouter

CoverageBenchmark

OpenRouter's models catalog lists "Sakana: Fugu Ultra v2" as the higher-performance member of Sakana AI's Fugu family, described as a language model trained to route tasks across a fixed pool of open and specialized models and to recursively call instances of itself. The entry specifies a 1M-token context window, $5 pe Per the OpenRouter description, Fugu Ultra v2 prioritizes answer quality on complex multi-step reasoning, autonomous research, and full-stack software development, and explicitly does not rely on individual proprietary frontier models in its pool. It supports configurable reasoning effort (high, xhigh, max), function c

Videos about Fugu Ultra v2

More models around Fugu Ultra v2