Sakana AI's higher-quality Fugu model. It coordinates a deeper pool of expert agents for hard, high-stakes reasoning and coding tasks.
Added Jun 22, 2026
Context Window
1M
Max Output
16.4K
Input Price (Auto)
$5.00/1M
Output Price (Auto)
$30.00/1M
Cache Read (Auto)
$0.50/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Fugu Ultra with similar models from the same provider or model family.
Fugu Ultra v1.1
sakana/fugu-ultra-v1.1Sakana AI's upgraded Fugu Ultra release with stronger coding, agentic task execution, and advanced reasoning through dynamic orchestration of frontier models.
Fugu Max
sakana/fugu-maxSakana AI's cost-performance Fugu model uses learned multi-agent orchestration to route tasks across expert models for reasoning, coding, and tool use.
Nvidia Nemotron 3 Ultra 550B
nvidia/nemotron-3-ultra-550b-a55bNvidia's Nemotron 3 Ultra 550B A55B model from the Nemotron 3 family. It uses a hybrid Mamba-Transformer MoE architecture. Provider-specific context limits vary, with the longest current route supporting up to 1M context.
Nvidia Nemotron 3 Ultra 550B Thinking
nvidia/nemotron-3-ultra-550b-a55b:thinkingNvidia's Nemotron 3 Ultra 550B A55B model from the Nemotron 3 family. It uses a hybrid Mamba-Transformer MoE architecture. Provider-specific context limits vary, with the longest current route supporting up to 1M context. Thinking enabled.
Ling 3.1 Flash
inclusionai/ling-3.1-flashLing 3.1 Flash is inclusionAI's hybrid reasoning model for coding, tool use, planning, and long-context agent workflows. It has 560B total parameters with 25B active parameters per token. Thinking is enabled by default and can be turned off in settings.
GPT 6.1 Sol
openai/gpt-6.1-solGPT-6.1 Sol delivers near-Astra performance at a lower cost for complex coding, computer use, and professional work. It supports text and image input, tool calling, structured outputs, and reasoning from low through max.