Linear models offer a promising approach to significantly reduce computational costs at scale, particularly for large context lengths. Enabling a >1000x improvement in inference costs, enabling o1 inference time thinking and wider AI accessibility.
Context Window
32.0K
Max Output
8.2K
Input Price (Auto)
$0.50/1M
Output Price (Auto)
$0.50/1M
Cache Read (Auto)
$0.25/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Qwerky 72B with similar models from the same provider or model family.
Llama 3.1 70B Dracarys 2
abacusai/Dracarys-72B-InstructLlama 3.1 70b finetune that offers improvements on coding.
Muse Glimmer 30B TEE
TEE/muse-glimmer-30bMeta's Muse Glimmer 30B is a dense, open-weight multimodal model for long-horizon agentic and coding workflows. Running inside a TEE (Trusted Execution Environment), with provider attestation support.
Muse Glimmer 30B
meta/muse-glimmer-30bMeta's Muse Glimmer 30B is a dense, open-weight multimodal model distilled from Muse Spark for long-horizon agents and coding workflows. It supports multi-step reasoning, reliable tool use, failure recovery, image understanding, and more than 100 languages.
Muse Spark 1.2
meta/muse-spark-1.2Meta's Muse Spark 1.2 is a multimodal reasoning model for complex agentic and coding tasks, with tool calling, structured output, and a one-million-token context window.
Muse Spark 1.2 Contributor (Data Used for Training)
meta/muse-spark-1.2-contributorA much cheaper opt-in version of Muse Spark 1.2 with the same multimodal coding and agentic capabilities. Prompts and outputs sent to this Contributor model may be used by Meta for training and to improve its products; use the standard Muse Spark 1.2 model if you do not want your data used for training.
Muse Spark 1.1
meta/muse-spark-1.1Meta's Muse Spark 1.1 is a long-context multimodal reasoning model served by the Meta Model API.