Poolside's open-weights agentic coding model with 118B total parameters and 8B activated per token, with thinking enabled for harder long-horizon software engineering, tool use, and extended reasoning. It supports a context window of up to 1M tokens.
Added Jul 21, 2026
Model weightsContext Window
1.0M
Max Output
131.1K
Input Price (Auto)
$0.10/1M
Output Price (Auto)
$0.20/1M
Cache Read (Auto)
$0.0100/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Laguna S 2.1 Thinking with similar models from the same provider or model family.
Laguna S 2.1
poolside/laguna-s-2.1Poolside's open-weights agentic coding model with 118B total parameters and 8B activated per token. It is designed for long-horizon software engineering and tool use, with a context window of up to 1M tokens. This variant keeps thinking disabled for faster direct responses.
DiffusionGemma
google/diffusiongemmaDiffusionGemma is a high-speed diffusion-based version of Gemma 4 26B A4B. It supports optional reasoning and a 262,144-token context window.
Gemma 4 26B A4B Cybersecurity
google/gemma-4-26b-a4b-it-cybersecurityGemma 4 26B A4B Cybersecurity is a cybersecurity-focused variant based on the uncensored model, with provider moderation for illegal activities. It supports optional reasoning, image understanding, tool calling, and a 262,144-token context window.
Nemotron 3.5 Content Safety
nvidia/nemotron-3.5-content-safetyNemotron 3.5 Content Safety is a content moderation classifier that labels user messages and assistant responses as safe or unsafe. It supports a 131,072-token context window and optional reasoning.
Qwen 3.8 27B Cybersecurity
qwen/qwen3.8-27b-cybersecurityQwen 3.8 27B Cybersecurity is a cybersecurity-focused variant based on the uncensored model, with provider moderation for illegal activities. It supports optional reasoning, image understanding, tool calling, and a 262,144-token context window.
GLM 5.3 Flash Cybersecurity
z-ai/glm-5.3-flash-cybersecurityGLM 5.3 Flash Cybersecurity is a cybersecurity-focused variant based on the uncensored model, with provider moderation for illegal activities. It supports always-on reasoning, image understanding, tool calling, and a 1,048,576-token context window.