Sakana AI's cost-performance Fugu model uses learned multi-agent orchestration to route tasks across expert models for reasoning, coding, and tool use.
Added Sep 11, 2026
Context Window
1.0M
Max Output
128.0K
Input Price (Auto)
$2.00/1M
Output Price (Auto)
$6.00/1M
Cache Read (Auto)
$0.25/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Fugu Max with similar models from the same provider or model family.
Fugu Ultra v1.1
sakana/fugu-ultra-v1.1Sakana AI's upgraded Fugu Ultra release with stronger coding, agentic task execution, and advanced reasoning through dynamic orchestration of frontier models.
Fugu Ultra
sakana/fugu-ultraSakana AI's higher-quality Fugu model. It coordinates a deeper pool of expert agents for hard, high-stakes reasoning and coding tasks.
Qwen3.8 Max 0902
alibaba/qwen3.8-max-0902Qwen3.8 Max 0902 is Alibaba's September 2 checkpoint of its flagship Qwen3.8 Max model for coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, selectable thinking, tool calling, structured output, and a near-million-token context window.
Qwen3.8 2.4T A95B (Max)
qwen/qwen3.8-2.4t-a95bThis is the same underlying model as Qwen3.8 Max, exposed under its architecture-based 2.4T A95B name for easier discovery. It uses the identical routing, pricing, capabilities, and non-thinking mode.
Qwen3.8 Max
qwen3.8-maxQwen3.8 Max is Qwen's 2.4T-parameter flagship model for coding, knowledge work, full-stack development, data analysis, and long-running agent workflows in non-thinking mode. It supports text, image, video, PDF input, tool calling, structured output, and a near-million-token context window.
Qwen3.8 Max Thinking
qwen3.8-max:thinkingQwen3.8 Max Thinking enables generation-time reasoning for deeper coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, tool calling, structured output, and a near-million-token context window.