Agnes 3.0 Flash

agnes-3.0-flash

Agnes 3.0 Flash

agnes-3.0-flash

Agnes 3.0 Flash is a low-cost model for coding, tool use, and multi-turn agent tasks. It supports text and image input, optional thinking, and a 512K-token context window.

Added Sep 9, 2026

Context Window

524.3K

Max Output

65.5K

Input Price (Auto)

$0.050/1M

Output Price (Auto)

$0.15/1M

Cache Read (Auto)

$0.0050/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

No benchmark data is available yet for this model.

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare Agnes 3.0 Flash with similar models from the same provider or model family.

DeepSeek V4.1 Flash

deepseek/deepseek-v4.1-flash

DeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use. Assume prompts and responses are logged by the provider and may be used for model training or service improvement. Do not send sensitive or confidential data.

DeepSeek V4.1 Flash Thinking

deepseek/deepseek-v4.1-flash:thinking

DeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use. Assume prompts and responses are logged by the provider and may be used for model training or service improvement. Do not send sensitive or confidential data.

DeepSeek V4 Flash Vision Exp Uncensored

deepseek/deepseek-v4-flash-vision-exp-uncensored

An uncensored variant of the experimental vision-enabled DeepSeek V4 Flash model for chat, image understanding, reasoning, coding, and tool use, with a 524K context window.

Synth 2.5 Flash Preview

synth-2.5-flash

Synth 2.5 Flash Preview is a low-cost text model designed for role-play, character dialogue, and collaborative storytelling.

Gemini 3.8 Flash

google/gemini-3.8-flash

Google's fast multimodal model for agentic workloads, including coding, tool use, image understanding, PDF and document extraction, audio, and video. Its capabilities, limits, reasoning behavior, and pricing currently mirror Gemini 3.7 Flash.

GLM 5.3 Flash TEE

TEE/glm-5.3-flash

GLM-5.3 Flash is Z.AI's natively multimodal 320B MoE reasoning model with 18B active parameters. This TEE deployment is verified through the selected provider: Redpill attestation with signed completion receipts or the official Tinfoil SDK's ATC/EHBP verification.