StreamLake

StreamLake

CN
34 models available

Models served by StreamLake

ModelContextInput /MOutput /MCache read /MCache hitLatencyThroughput
128k$0.60/M$2.40/MN/AN/A3.2s47 tps
128k$0.36/M$1.08/MN/AN/A1.2s53 tps
128k$0.36/M$1.08/MN/AN/A1.2s53 tps
163k$0.23/M$0.34/M$0.02/M60.4%1.4s27 tps
163k$0.23/M$0.34/M$0.02/M32.3%1.4s27 tps
1.0M$0.05/M$0.10/M$0.01/M42%1.5s37 tps
1.0M$0.05/M$0.10/M$0.01/M79.8%1.5s37 tps
1.0M$0.06/M$0.17/M$0.0018/M75.7%1.7s31 tps
1.0M$0.06/M$0.17/M$0.0018/M42.1%1.7s31 tps
1.0M$1.83/M$3.65/M$0.02/M74.8%2.4s35 tps
1.0M$1.83/M$3.65/M$0.02/M85.6%2.4s35 tps
1.0M$0.61/M$1.83/M$0.02/MN/A2.4s46 tps
1.0M$0.61/M$1.83/M$0.02/MN/A2.4s46 tps
200k$0.63/M$2.02/M$0.13/M20%3.2s35 tps
200k$0.63/M$2.02/M$0.13/M33.4%3.2s35 tps
200k$1.01/M$3.19/M$0.19/M15.9%2s31 tps
200k$1.01/M$3.19/M$0.19/M30.2%2s31 tps
1.0M$0.58/M$1.83/M$0.11/M84.8%1.4s52 tps
1.0M$0.58/M$1.83/M$0.11/M79.2%1.4s52 tps
1.0M$0.15/M$0.49/M$0.03/M77.1%1.4s19 tps
256k$0.63/M$2.65/M$0.11/MN/A1.6s49 tps
256k$0.63/M$2.65/M$0.11/MN/A1.6s49 tps
256k$0.75/M$3.15/M$0.15/MN/A1.8s48 tps
1M$0.18/M$0.35/M$0.0035/MN/A1.4s21 tps
1.0M$0.55/M$1.10/M$0.0045/M69.4%1.3s38 tps
1.0M$0.55/M$1.10/M$0.0045/M65.6%1.3s38 tps
1M$0.18/M$0.35/M$0.0035/M49.7%1.4s21 tps
205k$0.28/M$1.13/M$0.03/MN/A1s63 tps
1M$0.31/M$1.26/M$0.06/MN/A1.1s79 tps
1M$0.31/M$1.26/M$0.06/MN/A1.1s79 tps
262k$0.22/M$0.88/MN/AN/A0.8s3 tps
262k$0.19/M$0.94/M$0.04/MN/A0.8s87 tps
258k$0.63/M$3.78/M$0.13/MN/A1s206 tps
258k$0.63/M$3.78/M$0.13/MN/A1s206 tps

Prices are per million tokens. Selectable routes include the applicable provider-selection markup; fixed routes show the standard model price. Availability and pricing refresh continuously; each model page shows the live provider comparison.

Using StreamLake via the API

For models with provider selection, append :streamlake to the model ID, set "provider": "streamlake" in the request body, or send an X-Provider: streamlake header. Explicit provider selection adds a route-specific markup over that provider's base price; the prices in the table above already include it. Models without provider selection use the displayed model ID without a provider suffix or selection surcharge.

curl https://nano-gpt.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NANOGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-ai/DeepSeek-R1-0528:streamlake",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

See the API documentation for provider preferences, price-aware routing, and error behavior.