DekaLLM

DekaLLM

ID
11 models available

Data handling can vary by model and deployment. The privacy badge above shows the more conservative of the provider-wide fallback and the policies for routes on this page; each model page shows the policy for that specific route.

Models served by DekaLLM

11 models

ModelContextInput /MOutput /MCache read /MCache hitTTFTThroughput
1.0M$0.13/M$1.26/M$0.0053/M69.8%1.5s91 tps
1.0M$0.13/M$1.26/M$0.0053/M86.6%1.5s91 tps
1.0M$0.10/M$1.05/M$0.04/MN/A2.7s41 tps
131k$0.03/M$0.19/M$0.03/M1.1%0.7s22 tps
131k$0.03/M$0.15/M$0.03/M0.4%0.7s29 tps
262k$0.08/M$0.47/MN/A0%1.8s16 tps
262k$0.08/M$0.47/MN/A0%1.8s16 tps
262k$0.10/M$1.05/M$0.05/MN/A0.4s37 tps
262k$0.10/M$1.05/M$0.05/MN/A0.4s37 tps
262k$0.05/M$3.15/M$0.02/MN/A0.7s107 tps
262k$0.05/M$3.15/M$0.02/MN/A0.7s107 tps
Context
1.0M
Input /M
$0.13/M
Output /M
$1.26/M
Cache read /M
$0.0053/M
Cache hit
69.8%
TTFT
1.5s
Throughput
91 tps
Context
1.0M
Input /M
$0.13/M
Output /M
$1.26/M
Cache read /M
$0.0053/M
Cache hit
86.6%
TTFT
1.5s
Throughput
91 tps
Context
1.0M
Input /M
$0.10/M
Output /M
$1.05/M
Cache read /M
$0.04/M
Cache hit
N/A
TTFT
2.7s
Throughput
41 tps
Context
131k
Input /M
$0.03/M
Output /M
$0.19/M
Cache read /M
$0.03/M
Cache hit
1.1%
TTFT
0.7s
Throughput
22 tps
Context
131k
Input /M
$0.03/M
Output /M
$0.15/M
Cache read /M
$0.03/M
Cache hit
0.4%
TTFT
0.7s
Throughput
29 tps
Context
262k
Input /M
$0.08/M
Output /M
$0.47/M
Cache read /M
N/A
Cache hit
0%
TTFT
1.8s
Throughput
16 tps
Context
262k
Input /M
$0.10/M
Output /M
$1.05/M
Cache read /M
$0.05/M
Cache hit
N/A
TTFT
0.4s
Throughput
37 tps
Context
262k
Input /M
$0.10/M
Output /M
$1.05/M
Cache read /M
$0.05/M
Cache hit
N/A
TTFT
0.4s
Throughput
37 tps
Context
262k
Input /M
$0.05/M
Output /M
$3.15/M
Cache read /M
$0.02/M
Cache hit
N/A
TTFT
0.7s
Throughput
107 tps
Context
262k
Input /M
$0.05/M
Output /M
$3.15/M
Cache read /M
$0.02/M
Cache hit
N/A
TTFT
0.7s
Throughput
107 tps

Prices are per million tokens. TTFT is the observed time to first token. Selectable routes include the applicable provider-selection markup; fixed routes show the standard model price. Availability and pricing refresh continuously; each model page shows the live provider comparison.

Using DekaLLM via the API

Call a model served by DekaLLM using its displayed model ID, including the provider suffix when selection is supported.

cURL

curl https://nano-gpt.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NANOGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek/deepseek-v4.1-flash:dekallm",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

For models with provider selection, append :dekallm to the model ID, set "provider": "dekallm" in the request body, or send an X-Provider: dekallm header. Explicit provider selection adds a route-specific markup over that provider's base price; the prices in the table above already include it. Models without provider selection use the displayed model ID without a provider suffix or selection surcharge.

See the API documentation for provider preferences, price-aware routing, and error behavior.