The flagship model for balanced reasoning and context-rich dialogue. Perfect for AI roleplay, storytelling, and assistant tasks with a 16K window.
Added Sep 20, 2025
Context Window
16.4K
Max Output
16.4K
Input Price (Auto)
$0.020/1M
Output Price (Auto)
$0.16/1M
Cache Read (Auto)
$0.010/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Provider information for this model’s automatic routing. These routes cannot be selected individually.
Loading provider options…
Related text models
Compare Manta Flash 1.0 with similar models from the same provider or model family.
Manta Mini 1.0
meganova-ai/manta-mini-1.0Lightweight tier optimized for speed and cost.
Manta Pro 1.0
meganova-ai/manta-pro-1.0Tailored for deep reasoning, long-form generation, and RAG workloads. 32K token context window.
DeepSeek V4.1 Flash TEE
TEE/deepseek-v4.1-flashDeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This route runs through Tinfoil attested inference inside a Trusted Execution Environment.
Agnes 3.0 Flash
agnes-3.0-flashAgnes 3.0 Flash is a low-cost model for coding, tool use, and multi-turn agent tasks. It supports text and image input, optional thinking, and a 512K-token context window.
Ling 3.0 Flash VL
inclusionai/ling-3.0-flash-vlLing 3.0 Flash VL is inclusionAI's native multimodal Mixture-of-Experts model with 124B total parameters and 5.5B active parameters per token. It combines image and video understanding with reasoning and tool use for document analysis, charts, visual verification, and interface-based agent tasks. Thinking is enabled by default and can be turned off in settings.
DeepSeek V4.1 Flash
deepseek/deepseek-v4.1-flashDeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use.