Sakana AI's upgraded Fugu Ultra release with stronger coding, agentic task execution, and advanced reasoning through dynamic orchestration of frontier models.
Added Jul 24, 2026
Context Window
1.0M
Max Output
16.4K
Input Price (Auto)
$5.00/1M
Output Price (Auto)
$30.00/1M
Cache Read (Auto)
$0.50/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Fugu Ultra v1.1 with similar models from the same provider or model family.
Fugu Ultra
sakana/fugu-ultraSakana AI's higher-quality Fugu model. It coordinates a deeper pool of expert agents for hard, high-stakes reasoning and coding tasks.
Fugu Max
sakana/fugu-maxSakana AI's cost-performance Fugu model uses learned multi-agent orchestration to route tasks across expert models for reasoning, coding, and tool use.
Nvidia Nemotron 3 Ultra 550B
nvidia/nemotron-3-ultra-550b-a55bNvidia's Nemotron 3 Ultra 550B A55B model from the Nemotron 3 family. It uses a hybrid Mamba-Transformer MoE architecture. Provider-specific context limits vary, with the longest current route supporting up to 1M context.
Nvidia Nemotron 3 Ultra 550B Thinking
nvidia/nemotron-3-ultra-550b-a55b:thinkingNvidia's Nemotron 3 Ultra 550B A55B model from the Nemotron 3 family. It uses a hybrid Mamba-Transformer MoE architecture. Provider-specific context limits vary, with the longest current route supporting up to 1M context. Thinking enabled.
Claude Opus 5.5
anthropic/claude-opus-5.5Claude Opus 5.5 is Anthropic's flagship model for agentic coding, computer use, complex knowledge work, and long-running tasks, with improved efficiency and more natural communication.
Nex N2.5 Mini
nex-agi/nex-n2.5-miniNex N2.5 Mini is an open-source multimodal agentic model for coding and long-horizon workflows. It supports image input, structured output, configurable reasoning, a 262K-token context window, and up to 235K output tokens.