StreamLake

StreamLake

CN
34 models available

Data handling can vary by model and deployment. The privacy badge above shows the more conservative of the provider-wide fallback and the policies for routes on this page; each model page shows the policy for that specific route.

Models served by StreamLake

34 models

ModelContextInput /MOutput /MCache read /MCache hitTTFTThroughput
128k$0.60/M$2.40/MN/A0%3.1s46 tps
128k$0.36/M$1.08/MN/A27.9%1s66 tps
128k$0.36/M$1.08/MN/A0%1s66 tps
1.0M$0.07/M$0.15/M$0.01/M23.2%1.1s79 tps
1.0M$0.07/M$0.15/M$0.01/M85.1%1.1s79 tps
1.0M$0.05/M$0.14/M$0.0015/M78.5%1.3s45 tps
1.0M$0.05/M$0.14/M$0.0015/M78.5%1.3s45 tps
1.0M$1.83/M$3.65/M$0.15/M90.1%2.2s35 tps
1.0M$1.83/M$3.65/M$0.15/M78.6%2.2s35 tps
1.0M$0.69/M$2.08/M$0.02/MN/A2.2s47 tps
1.0M$0.69/M$2.08/M$0.02/MN/A2.2s47 tps
1.0M$0.16/M$0.63/M$0.0032/M58.3%1.5s75 tps
1.0M$0.16/M$0.63/M$0.0032/M69%1.5s75 tps
198k$0.63/M$2.02/M$0.13/M66.9%3s39 tps
198k$0.63/M$2.02/M$0.13/M37%3s39 tps
200k$1.01/M$3.19/M$0.19/M47.1%1.3s48 tps
200k$1.01/M$3.19/M$0.19/M0%1.3s48 tps
1.0M$0.59/M$1.85/M$0.11/M83.3%1.1s54 tps
1.0M$0.59/M$1.85/M$0.11/M79.3%1.1s54 tps
1.0M$0.09/M$0.30/M$0.02/M0%3.9s34 tps
256k$0.63/M$2.65/M$0.11/M18.3%13.1s66 tps
256k$0.63/M$2.65/M$0.11/M68.6%13.1s66 tps
256k$0.75/M$3.15/M$0.15/MN/A1s83 tps
1M$0.18/M$0.35/M$0.0035/MN/A2.4s32.5 tps
1M$0.55/M$1.10/M$0.0045/MN/A1.2s33.5 tps
1M$0.55/M$1.10/M$0.0045/MN/A1.2s33.5 tps
1M$0.18/M$0.35/M$0.0035/MN/A2.4s32.5 tps
200k$0.28/M$1.13/M$0.03/MN/A1.1s70 tps
1M$0.31/M$1.26/M$0.06/MN/A2.6s97 tps
1M$0.31/M$1.26/M$0.06/MN/A2.6s97 tps
128k$0.22/M$0.88/MN/A<0.1%0.7s53 tps
256k$0.19/M$0.94/M$0.04/MN/A1.2s60 tps
256k$0.63/M$3.78/M$0.13/MN/A0.5s157 tps
256k$0.63/M$3.78/M$0.13/MN/A0.5s157 tps
Context
128k
Input /M
$0.60/M
Output /M
$2.40/M
Cache read /M
N/A
Cache hit
0%
TTFT
3.1s
Throughput
46 tps
Context
128k
Input /M
$0.36/M
Output /M
$1.08/M
Cache read /M
N/A
Cache hit
27.9%
TTFT
1s
Throughput
66 tps
Context
128k
Input /M
$0.36/M
Output /M
$1.08/M
Cache read /M
N/A
Cache hit
0%
TTFT
1s
Throughput
66 tps
Context
1.0M
Input /M
$0.07/M
Output /M
$0.15/M
Cache read /M
$0.01/M
Cache hit
23.2%
TTFT
1.1s
Throughput
79 tps
Context
1.0M
Input /M
$0.07/M
Output /M
$0.15/M
Cache read /M
$0.01/M
Cache hit
85.1%
TTFT
1.1s
Throughput
79 tps
Context
1.0M
Input /M
$0.05/M
Output /M
$0.14/M
Cache read /M
$0.0015/M
Cache hit
78.5%
TTFT
1.3s
Throughput
45 tps
Context
1.0M
Input /M
$0.05/M
Output /M
$0.14/M
Cache read /M
$0.0015/M
Cache hit
78.5%
TTFT
1.3s
Throughput
45 tps
Context
1.0M
Input /M
$1.83/M
Output /M
$3.65/M
Cache read /M
$0.15/M
Cache hit
90.1%
TTFT
2.2s
Throughput
35 tps
Context
1.0M
Input /M
$1.83/M
Output /M
$3.65/M
Cache read /M
$0.15/M
Cache hit
78.6%
TTFT
2.2s
Throughput
35 tps
Context
1.0M
Input /M
$0.69/M
Output /M
$2.08/M
Cache read /M
$0.02/M
Cache hit
N/A
TTFT
2.2s
Throughput
47 tps
Context
1.0M
Input /M
$0.69/M
Output /M
$2.08/M
Cache read /M
$0.02/M
Cache hit
N/A
TTFT
2.2s
Throughput
47 tps
Context
1.0M
Input /M
$0.16/M
Output /M
$0.63/M
Cache read /M
$0.0032/M
Cache hit
58.3%
TTFT
1.5s
Throughput
75 tps
Context
1.0M
Input /M
$0.16/M
Output /M
$0.63/M
Cache read /M
$0.0032/M
Cache hit
69%
TTFT
1.5s
Throughput
75 tps
Context
198k
Input /M
$0.63/M
Output /M
$2.02/M
Cache read /M
$0.13/M
Cache hit
66.9%
TTFT
3s
Throughput
39 tps
Context
198k
Input /M
$0.63/M
Output /M
$2.02/M
Cache read /M
$0.13/M
Cache hit
37%
TTFT
3s
Throughput
39 tps
Context
200k
Input /M
$1.01/M
Output /M
$3.19/M
Cache read /M
$0.19/M
Cache hit
47.1%
TTFT
1.3s
Throughput
48 tps
Context
200k
Input /M
$1.01/M
Output /M
$3.19/M
Cache read /M
$0.19/M
Cache hit
0%
TTFT
1.3s
Throughput
48 tps
Context
1.0M
Input /M
$0.59/M
Output /M
$1.85/M
Cache read /M
$0.11/M
Cache hit
83.3%
TTFT
1.1s
Throughput
54 tps
Context
1.0M
Input /M
$0.59/M
Output /M
$1.85/M
Cache read /M
$0.11/M
Cache hit
79.3%
TTFT
1.1s
Throughput
54 tps
Context
1.0M
Input /M
$0.09/M
Output /M
$0.30/M
Cache read /M
$0.02/M
Cache hit
0%
TTFT
3.9s
Throughput
34 tps
Context
256k
Input /M
$0.63/M
Output /M
$2.65/M
Cache read /M
$0.11/M
Cache hit
18.3%
TTFT
13.1s
Throughput
66 tps
Context
256k
Input /M
$0.63/M
Output /M
$2.65/M
Cache read /M
$0.11/M
Cache hit
68.6%
TTFT
13.1s
Throughput
66 tps
Context
256k
Input /M
$0.75/M
Output /M
$3.15/M
Cache read /M
$0.15/M
Cache hit
N/A
TTFT
1s
Throughput
83 tps
Context
1M
Input /M
$0.18/M
Output /M
$0.35/M
Cache read /M
$0.0035/M
Cache hit
N/A
TTFT
2.4s
Throughput
32.5 tps
Context
1M
Input /M
$0.55/M
Output /M
$1.10/M
Cache read /M
$0.0045/M
Cache hit
N/A
TTFT
1.2s
Throughput
33.5 tps
Context
1M
Input /M
$0.55/M
Output /M
$1.10/M
Cache read /M
$0.0045/M
Cache hit
N/A
TTFT
1.2s
Throughput
33.5 tps
Context
1M
Input /M
$0.18/M
Output /M
$0.35/M
Cache read /M
$0.0035/M
Cache hit
N/A
TTFT
2.4s
Throughput
32.5 tps
Context
200k
Input /M
$0.28/M
Output /M
$1.13/M
Cache read /M
$0.03/M
Cache hit
N/A
TTFT
1.1s
Throughput
70 tps
Context
1M
Input /M
$0.31/M
Output /M
$1.26/M
Cache read /M
$0.06/M
Cache hit
N/A
TTFT
2.6s
Throughput
97 tps
Context
1M
Input /M
$0.31/M
Output /M
$1.26/M
Cache read /M
$0.06/M
Cache hit
N/A
TTFT
2.6s
Throughput
97 tps
Context
128k
Input /M
$0.22/M
Output /M
$0.88/M
Cache read /M
N/A
Cache hit
<0.1%
TTFT
0.7s
Throughput
53 tps
Context
256k
Input /M
$0.19/M
Output /M
$0.94/M
Cache read /M
$0.04/M
Cache hit
N/A
TTFT
1.2s
Throughput
60 tps
Context
256k
Input /M
$0.63/M
Output /M
$3.78/M
Cache read /M
$0.13/M
Cache hit
N/A
TTFT
0.5s
Throughput
157 tps
Context
256k
Input /M
$0.63/M
Output /M
$3.78/M
Cache read /M
$0.13/M
Cache hit
N/A
TTFT
0.5s
Throughput
157 tps

Prices are per million tokens. TTFT is the observed time to first token. Selectable routes include the applicable provider-selection markup; fixed routes show the standard model price. Availability and pricing refresh continuously; each model page shows the live provider comparison.

Using StreamLake via the API

Call a model served by StreamLake using its displayed model ID, including the provider suffix when selection is supported.

cURL

curl https://nano-gpt.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NANOGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-ai/DeepSeek-R1-0528:streamlake",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

For models with provider selection, append :streamlake to the model ID, set "provider": "streamlake" in the request body, or send an X-Provider: streamlake header. Explicit provider selection adds a route-specific markup over that provider's base price; the prices in the table above already include it. Models without provider selection use the displayed model ID without a provider suffix or selection surcharge.

See the API documentation for provider preferences, price-aware routing, and error behavior.