Sulat.com
AI models
OpenRouter logo

Model details

Fugu Max

Fugu Max is positioned as an inference-time orchestrator rather than a single pretrained model. According to its Blackbox listing, it "orchestrates our largest pool of models to push the cost–performance Pareto frontier" and "dynamically finds efficient combinations of expert agents, aiming to improve model performance and cost efficiency together." This routing-agent framing implies that Fugu Max selects and composes outputs from multiple underlying specialists on a per-query basis, which is a practical approach when latency budgets, input length, and task complexity vary widely across requests.

The same listing tags Fugu Max with reasoning, tool-use, vision, and structured-output capabilities, suggesting it is intended for multimodal assistant workloads that span text and images and benefit from function calling and JSON-constrained responses. A long context window is advertised alongside these capabilities, which fits use cases involving large documents, multi-turn agentic loops, or code-and-image reasoning. Because the design emphasis is on combining expert agents dynamically, practical strengths lie in flexible cost-quality trade-offs and broad task coverage rather than a single narrow specialization.

OpenRoutersakana/fugu-maxfugu

Quick Info

Powered by
Provider
OpenRouter
Model key
sakana/fugu-max
Release date
Sep 11, 2026
Last updated
Sep 11, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$2.00
Output token cost
$6.00

Limits

Output tokens
128,000 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare Fugu Max pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Fugu Max

OpenRouter

Official sourceAnnouncement

Sakana AI announced Fugu Max on September 11, 2026, positioning it as a cost-performance orchestration model that routes each task to the leanest model in a swappable pool of open-weights and specialized models, including NVIDIA's Nemotron family. The release frames Fugu Max and Fugu Ultra v2 as the same orchestration Fugu Max is priced at $2 per million input tokens and $6 per million output tokens, with Sakana claiming this runs 40 to 60 percent below comparable output pricing for Sonnet 5, GPT 5.6 Terra, and Kimi K3. The release also reports that Fugu Max claims the best overall score on six self-reported benchmarks (Terminal Ben

OpenRouter

Coverage

DataNorth reported on September 11, 2026, that Sakana AI released Fugu Max and Fugu Ultra v2, with Fugu Max priced at $2 per million input tokens and $6 per million output tokens, which Sakana claims undercuts Sonnet 5 and Kimi K3 by 40 to 60 percent on output. Both models run as hosted APIs only and are not available The report provides a side-by-side specification table: Fugu Max has a 1M token context window, 128K maximum output, $0.25 cached input per million tokens, flat pricing regardless of context length, and no long-context surcharge. Fugu Max widens the pool with open-weight and specialist models including NVIDIA's Nemotro

OpenRouter

Coverage

The There's An AI For That directory describes Fugu Max as the cost-optimized member of Sakana AI's second-generation Fugu family of orchestration models, announced September 10, 2026. It positions Fugu Max as a learned orchestrator that routes each incoming task to the leanest model in a pool of open-weights and speci The listing reports pricing at $2 per million input tokens and $6 per million output tokens, 40-60 percent below comparable models, and claims the best overall score on six benchmarks (Terminal Bench 2.1, GPQA Diamond, AA-LCR, GDP.pdf, AutomationBench, and SWEFish). Fugu Max is accessible via an OpenAI-compatible API t

OpenRouter

Coverage

AI Weekly reported on September 11, 2026, that Sakana AI shipped Fugu Max, an orchestration engine priced at $2 per million input tokens and $6 per million output tokens that routes each request to what Sakana calls the leanest model capable of solving them. The pricing runs 40 to 60 percent below the output cost of So The report notes that Fugu Max claims the best overall score on six self-reported benchmarks including Terminal Bench 2.1, GPQAD, AA-LCR, GDP.pdf, AutomationBench, and SWEFish, with no third-party evaluation cited. Sakana paired the launch with Fugu Ultra v2, a higher-capability variant built on the same routing archit

OpenRouter

CoverageBenchmark

OpenRouter's model directory lists Sakana: Fugu Max as a distinct entry with 1M token context, $2 per million input tokens and $6 per million output tokens, dated September 11, 2026. The listing describes Fugu Max as the cost-performance model in Sakana AI's Fugu family, a learned multi-agent orchestration system that The directory entry confirms that Fugu Max supports configurable reasoning effort (high, xhigh, max), function calling, structured outputs, image and PDF input, and built-in web search and web fetch. Orchestration tokens consumed by the system are billed as standard input/output tokens. While OpenRouter is a serving ga

Videos about Fugu Max

More models around Fugu Max