Browse all Tencent text models
Provider logo

Tencent Hy4 Preview

tencent/hy4-preview
Provider logo

Tencent Hy4 Preview

tencent/hy4-preview

Hy4 Preview is Tencent's 770B-parameter mixture-of-experts model with 49B active parameters. It is designed for coding agents, complex tool-use workflows, and productivity tasks, with a 1M-token context window and configurable reasoning effort.

Added Aug 28, 2026

Model weights

Context Window

1.0M

Max Output

64.0K

Input Price (Auto)

$0.83/1M

Output Price (Auto)

$2.50/1M

Cache Read (Auto)

$0.042/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

No benchmark data is available yet for this model.

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…

Compare Tencent Hy4 Preview with similar models from the same provider or model family.

Tencent Hy3

tencent/hy3

Hy3 is Tencent's 295B-parameter Mixture-of-Experts model with 21B active parameters, native 256K context, and configurable reasoning modes. It is built for coding, long-context comprehension, multi-turn dialogue, and agentic task execution with a focus on high-throughput production workloads.

Schematron V2 Small

inference-net/schematron-v2-small

Inference.net's 3B-parameter HTML-to-JSON extraction model, focused on accuracy for complex schemas and long web pages. It turns HTML into typed, structured data for web scraping and product catalog ingestion, with a 128K-token context window. Supply HTML in the user message and extraction instructions in a JSON schema via response_format; it does not follow ordinary chat or system prompts.

Schematron V2 Turbo

inference-net/schematron-v2-turbo

Inference.net's 3B-parameter HTML-to-JSON extraction model, optimized for throughput and low cost on high-volume workloads. It turns HTML into typed, structured data for web scraping and product catalog ingestion, with a 128K-token context window. Supply HTML in the user message and extraction instructions in a JSON schema via response_format; it does not follow ordinary chat or system prompts.

DeepSeek V4.1 Flash TEE

TEE/deepseek-v4.1-flash

DeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This route runs through Tinfoil attested inference inside a Trusted Execution Environment.

GPT Astra Latest

openai/gpt-astra-latest

Compatibility alias that routes to GPT 6 Astra, the latest supported GPT Astra model.

GPT Luna Latest

openai/gpt-luna-latest

Compatibility alias that routes to GPT 5.6 Luna, the latest supported GPT Luna model.