Baidu

Baidu

CN
16 models available

Models served by Baidu

ModelContextInput /MOutput /MCache read /MCache hitLatencyThroughput
163k$0.29/M$0.44/M$0.03/M45%1.1s31 tps
163k$0.29/M$0.44/M$0.03/M39.2%1.1s31 tps
1.0M$0.09/M$0.17/M$0.02/M74.2%0.7s76 tps
1.0M$0.09/M$0.17/M$0.02/M79.3%0.7s76 tps
1.0M$0.05/M$0.09/M$0.0094/M85.4%1s81 tps
1.0M$0.05/M$0.09/M$0.0094/M90.5%1s81 tps
1.0M$1.60/M$3.19/M$0.13/M5.5%0.8s38 tps
1.0M$1.60/M$3.19/M$0.13/M90.4%0.8s38 tps
200k$0.73/M$2.35/M$0.15/MN/A1.2s36 tps
200k$0.73/M$2.35/M$0.15/MN/A1.2s36 tps
200k$0.95/M$3.00/M$0.18/M0%1s65 tps
200k$0.95/M$3.00/M$0.18/M25.5%1s65 tps
1.0M$0.39/M$1.23/M$0.07/M82.7%0.8s58 tps
1.0M$0.39/M$1.23/M$0.07/M86%0.8s58 tps
256k$0.56/M$2.34/M$0.09/M64.8%1.6s38 tps
256k$0.56/M$2.34/M$0.09/M63.5%1.6s38 tps

Prices are per million tokens and include the 5% provider-selection markup. Availability and pricing refresh continuously; each model page shows the live provider comparison.

Using Baidu via the API

Append :baidu to the model ID, set "provider": "baidu" in the request body, or send an X-Provider: baidu header. Explicit provider selection adds a 5% markup over that provider's base price; the prices in the table above already include it.

curl https://nano-gpt.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NANOGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek/deepseek-v3.2:baidu",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

See the API documentation for provider preferences, price-aware routing, and error behavior.