Agnes 3.0 Flash is a low-cost model for coding, tool use, and multi-turn agent tasks. It supports text and image input, optional thinking, and a 512K-token context window.
Added Sep 9, 2026
Context Window
524.3K
Max Output
65.5K
Input Price (Auto)
$0.050/1M
Output Price (Auto)
$0.15/1M
Cache Read (Auto)
$0.0050/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Agnes 3.0 Flash with similar models from the same provider or model family.
DeepSeek V4.1 Flash
deepseek/deepseek-v4.1-flashDeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use. Assume prompts and responses are logged by the provider and may be used for model training or service improvement. Do not send sensitive or confidential data.
DeepSeek V4.1 Flash Thinking
deepseek/deepseek-v4.1-flash:thinkingDeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use. Assume prompts and responses are logged by the provider and may be used for model training or service improvement. Do not send sensitive or confidential data.
DeepSeek V4 Flash Vision Exp Uncensored
deepseek/deepseek-v4-flash-vision-exp-uncensoredAn uncensored variant of the experimental vision-enabled DeepSeek V4 Flash model for chat, image understanding, reasoning, coding, and tool use, with a 524K context window.
Synth 2.5 Flash Preview
synth-2.5-flashSynth 2.5 Flash Preview is a low-cost text model designed for role-play, character dialogue, and collaborative storytelling.
Gemini 3.8 Flash
google/gemini-3.8-flashGoogle's fast multimodal model for agentic workloads, including coding, tool use, image understanding, PDF and document extraction, audio, and video. Its capabilities, limits, reasoning behavior, and pricing currently mirror Gemini 3.7 Flash.
GLM 5.3 Flash TEE
TEE/glm-5.3-flashGLM-5.3 Flash is Z.AI's natively multimodal 320B MoE reasoning model with 18B active parameters. This TEE deployment is verified through the selected provider: Redpill attestation with signed completion receipts or the official Tinfoil SDK's ATC/EHBP verification.