Fast web answers with citations, tuned for factual questions and quick summaries.
Added Dec 23, 2025
Context Window
N/A
Max Output
32.8K
Pricing
Fixed cost: $0.015
Token Pricing
N/A
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Web Answer with similar models from the same provider or model family.
Universal Summarizer
universal-summarizerSummarizes long-form text or URLs across formats with large context support.
Schematron V2 Small
inference-net/schematron-v2-smallInference.net's 3B-parameter HTML-to-JSON extraction model, focused on accuracy for complex schemas and long web pages. It turns HTML into typed, structured data for web scraping and product catalog ingestion, with a 128K-token context window. Supply HTML in the user message and extraction instructions in a JSON schema via response_format; it does not follow ordinary chat or system prompts.
Schematron V2 Turbo
inference-net/schematron-v2-turboInference.net's 3B-parameter HTML-to-JSON extraction model, optimized for throughput and low cost on high-volume workloads. It turns HTML into typed, structured data for web scraping and product catalog ingestion, with a 128K-token context window. Supply HTML in the user message and extraction instructions in a JSON schema via response_format; it does not follow ordinary chat or system prompts.
DeepSeek V4.1 Flash TEE
TEE/deepseek-v4.1-flashDeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This route runs through Tinfoil attested inference inside a Trusted Execution Environment.
GPT Astra Latest
openai/gpt-astra-latestCompatibility alias that routes to GPT 6 Astra, the latest supported GPT Astra model.
GPT Luna Latest
openai/gpt-luna-latestCompatibility alias that routes to GPT 5.6 Luna, the latest supported GPT Luna model.