Browse all Sakana text models
Provider logo

Fugu Max

sakana/fugu-max
Provider logo

Fugu Max

sakana/fugu-max

Sakana AI's cost-performance Fugu model uses learned multi-agent orchestration to route tasks across expert models for reasoning, coding, and tool use.

Added Sep 11, 2026

Context Window

1.0M

Max Output

128.0K

Input Price (Auto)

$2.00/1M

Output Price (Auto)

$6.00/1M

Cache Read (Auto)

$0.25/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

No benchmark data is available yet for this model.

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare Fugu Max with similar models from the same provider or model family.

Fugu Ultra v1.1

sakana/fugu-ultra-v1.1

Sakana AI's upgraded Fugu Ultra release with stronger coding, agentic task execution, and advanced reasoning through dynamic orchestration of frontier models.

Fugu Ultra

sakana/fugu-ultra

Sakana AI's higher-quality Fugu model. It coordinates a deeper pool of expert agents for hard, high-stakes reasoning and coding tasks.

Qwen3.8 Max 0902

alibaba/qwen3.8-max-0902

Qwen3.8 Max 0902 is Alibaba's September 2 checkpoint of its flagship Qwen3.8 Max model for coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, selectable thinking, tool calling, structured output, and a near-million-token context window.

Qwen3.8 2.4T A95B (Max)

qwen/qwen3.8-2.4t-a95b

This is the same underlying model as Qwen3.8 Max, exposed under its architecture-based 2.4T A95B name for easier discovery. It uses the identical routing, pricing, capabilities, and non-thinking mode.

Qwen3.8 Max

qwen3.8-max

Qwen3.8 Max is Qwen's 2.4T-parameter flagship model for coding, knowledge work, full-stack development, data analysis, and long-running agent workflows in non-thinking mode. It supports text, image, video, PDF input, tool calling, structured output, and a near-million-token context window.

Qwen3.8 Max Thinking

qwen3.8-max:thinking

Qwen3.8 Max Thinking enables generation-time reasoning for deeper coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, tool calling, structured output, and a near-million-token context window.