Sakana AI Releases Fugu Ultra: Multi-Agent Orchestration System with 1M Context Window at $5/$30 per Million Tokens
Sakana AI has released Fugu Ultra, a multi-agent orchestration system that routes tasks across pools of underlying models rather than operating as a single monolithic model. The system supports a 1M token context window and is priced at $5 per million input tokens and $30 per million output tokens.
Fugu Ultra — Quick Specs
Sakana AI Releases Fugu Ultra: Multi-Agent Orchestration System with 1M Context Window
Sakana AI has released Fugu Ultra, the higher-performance model in its Fugu family, now available through OpenRouter. Unlike traditional language models, Fugu Ultra is a learned multi-agent orchestration system trained to route tasks across a swappable pool of underlying models and recursively call instances of itself.
Architecture and Capabilities
According to Sakana AI, Fugu Ultra prioritizes answer quality on complex, multi-step reasoning, coding, and agentic workflows. The system supports:
- 1 million token context window
- Configurable reasoning effort
- Tool calling
- Built-in web search capabilities
Orchestration tokens consumed by the system are billed as standard input/output tokens, with no separate pricing tier for the routing logic.
Pricing
Fugu Ultra is priced at:
- Input: $5 per million tokens
- Output: $30 per million tokens
The effective price can be 60-80% lower when prompt caching is applied for repeated context, according to OpenRouter's monitoring data.
Technical Approach
Rather than training a single large model, Sakana AI's approach involves training a language model to act as an orchestrator, deciding which underlying models to route tasks to and when to recursively invoke additional instances of itself. This architecture represents a departure from the monolithic model paradigm adopted by most major AI labs.
The model was released June 24, 2026, according to OpenRouter's listing. Sakana AI has not disclosed benchmark scores, parameter count, or details about the underlying model pool at this time.
Availability
Fugu Ultra is currently available exclusively through OpenRouter, which forwards requests directly to Sakana AI's infrastructure with no intermediate routing layer.
What This Means
Sakana AI's orchestration approach represents a significant architectural departure from the single-model paradigm. By routing tasks to specialized models rather than attempting to embed all capabilities in one system, Fugu Ultra could potentially offer better cost-performance ratios for complex workflows. However, without published benchmarks, it's unclear how the system compares to frontier models like Claude 3.5 Sonnet or GPT-4 on standardized reasoning and coding tasks. The success of this approach will depend on whether the orchestration overhead is offset by more efficient task routing.
Related Articles
Xiaomi Launches MiMo-V2.6-Pro, a 1T+ Parameter Model with 1M-Token Context
Xiaomi has released MiMo-V2.6-Pro, a flagship foundation model exceeding 1 trillion parameters with a 1M-token context window and native multimodal support. The model is priced at $0.435 per 1M input tokens and $0.87 per 1M output tokens, targeting agentic and long-horizon tasks.
Anthropic Ships Claude Opus 5.5, OpenAI Launches GPT-6 Sol and Luna — All Cheaper Than Predecessors
Anthropic released Claude Opus 5.5 at $4/$20 per million input/output tokens, undercutting Opus 5's $5/$25 pricing while claiming better agentic coding scores. OpenAI countered with GPT-6 Sol ($2/$10) and GPT-6 Luna ($0.10/$0.50), both up to 50% cheaper than GPT-5.6's promotional rates.
OpenAI Launches GPT-6 Luna: Fast, Low-Cost Model With 1.1M Context Window
OpenAI has released GPT-6 Luna, the fast and cost-efficient entry in its new GPT-6 model family, featuring a 1.1M token context window and pricing starting at $0.10 per 1M input tokens. The model is positioned below GPT-6 Sol and GPT-6 Astra in OpenAI's tiered lineup.
OpenAI Launches GPT-6 Sol Pro, a High-Reasoning Mode for Its Mid-Tier Model at $2/$10 per 1M Tokens
OpenAI has released GPT-6 Sol Pro, which runs the existing GPT-6 Sol model with its reasoning mode set to 'pro' for higher-quality responses on complex tasks. It carries a 1.1 million token context window and is priced at $2 per 1M input tokens and $10 per 1M output tokens.
Comments
Loading...