Chutes

Chutes

US
17 models available

Models served by Chutes

ModelContextInput /MOutput /MCache read /MLatencyThroughput
Gemma 4 31B TEE
TEE verified
131k$0.16/M$0.44/M$0.08/MN/AN/A
131k$0.16/M$0.44/M$0.08/MN/AN/A
200k$1.03/M$3.23/M$0.10/M2.5s41 tps
200k$1.03/M$3.23/M$0.10/M2.5s41 tps
GLM 5.1 TEE
TEE verified
131k$1.10/M$3.68/M$0.55/MN/AN/A
131k$1.10/M$3.68/M$0.55/MN/AN/A
GLM 5.2 TEE
TEE verified
1.0M$1.47/M$4.62/M$0.73/MN/AN/A
1.0M$1.47/M$4.62/M$0.73/MN/AN/A
256k$0.61/M$3.57/M$0.06/M2s40 tps
Kimi K2.6 TEE
TEE verified
262k$1.00/M$4.20/M$0.50/MN/AN/A
256k$0.61/M$3.57/M$0.06/M2s40 tps
Kimi K3
MXFP4
1.0M$3.15/M$15.75/M$0.31/M1.8s24 tps
Kimi K3 TEE
TEE verified
1.0M$3.15/M$15.75/M$1.57/MN/AN/A
260k$0.31/M$2.10/M$0.03/M1s11 tps
260k$0.31/M$2.10/M$0.03/M1s11 tps
262k$0.37/M$2.89/M$0.04/M1.9s30 tps
262k$0.37/M$2.89/M$0.04/M1.9s30 tps

Prices are per million tokens and include the 5% provider-selection markup. Availability and pricing refresh continuously; each model page shows the live provider comparison.

Using Chutes via the API

Append :chutes to the model ID, set "provider": "chutes" in the request body, or send an X-Provider: chutes header. Explicit provider selection adds a 5% markup over that provider's base price; the prices in the table above already include it.

curl https://nano-gpt.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NANOGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "TEE/gemma4-31b:chutes",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

See the API documentation for provider preferences, price-aware routing, and error behavior.