Baidu

Baidu

CN
18 models available

Models served by Baidu

ModelContextInput /MOutput /MCache read /MCache hitLatencyThroughput
163k$0.29/M$0.44/M$0.03/M14.3%1.1s35 tps
163k$0.29/M$0.44/M$0.03/M0%1.1s35 tps
1.0M$0.09/M$0.18/M$0.02/M71.7%0.7s85 tps
1.0M$0.09/M$0.18/M$0.02/M59.2%0.7s85 tps
1.0M$0.05/M$0.10/M$0.01/M81.9%0.8s121 tps
1.0M$0.05/M$0.10/M$0.01/M87.8%0.8s121 tps
1.0M$1.60/M$3.19/M$0.13/M4.8%1s56 tps
1.0M$1.60/M$3.19/M$0.13/M55.5%1s56 tps
1.0M$1.18/M$3.53/M$0.12/M0%0.8s58 tps
1.0M$1.18/M$3.53/M$0.12/M29.4%0.8s58 tps
200k$0.73/M$2.35/M$0.15/MN/A1.8s52 tps
200k$0.73/M$2.35/M$0.15/MN/A1.8s52 tps
200k$0.95/M$3.00/M$0.18/M92.3%1s70 tps
200k$0.95/M$3.00/M$0.18/M6.7%1s70 tps
1.0M$0.37/M$1.18/M$0.07/M87.8%0.7s64 tps
1.0M$0.37/M$1.18/M$0.07/M87.3%0.7s64 tps
256k$0.56/M$2.37/M$0.09/MN/A1.4s45 tps
256k$0.56/M$2.37/M$0.09/MN/A1.4s45 tps

Prices are per million tokens and include the 5% provider-selection markup. Availability and pricing refresh continuously; each model page shows the live provider comparison.

Using Baidu via the API

Append :baidu to the model ID, set "provider": "baidu" in the request body, or send an X-Provider: baidu header. Explicit provider selection adds a 5% markup over that provider's base price; the prices in the table above already include it.

curl https://nano-gpt.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NANOGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek/deepseek-v3.2:baidu",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

See the API documentation for provider preferences, price-aware routing, and error behavior.