Heabsy

Heabsy

GLB
12 models available

Heabsy states that its EEA Qwen3.8 deployment processes prompts and completions in memory without durable storage, logging, or training use. Data handling and deployment region vary by model; see each model’s provider details. Its GDPR Art. 28 Data Processing Addendum is available at https://heabsy.com/dpa.

Data handling can vary by model and deployment. The privacy badge above shows the more conservative of the provider-wide fallback and the policies for routes on this page; each model page shows the policy for that specific route.

Models served by Heabsy

12 models

ModelContextInput /MOutput /MCache read /MCache hitTTFTThroughput
1.0M$1.50/M$4.50/M$0.27/M72.4%5.7s145.2 tps
1.0M$0.50/M$1.50/M$0.20/MN/AN/AN/A
1.0M$0.50/M$1.50/M$0.27/M80.4%5.4s82.5 tps
1.0M$0.46/M$1.50/M$0.15/M96.1%6s78.2 tps
524k$0.10/M$0.61/M$0.05/M45%2.4s83.2 tps
524k$0.21/M$1.47/M$0.19/M2.6%0.8s125.4 tps
524k$0.21/M$1.47/M$0.19/M86.3%8.6s145.3 tps
524k$0.10/M$0.61/M$0.05/M7.4%2s165.5 tps
524k$1.00/M$3.00/M$0.05/M96.5%1.6s237.6 tps
524k$1.00/M$3.00/M$0.05/M70.4%3.4s213.4 tps
262k$0.08/M$3.00/M$0.08/MN/A0.9s176 tps
262k$0.08/M$3.00/M$0.08/MN/A0.9s176 tps
Context
1.0M
Input /M
$1.50/M
Output /M
$4.50/M
Cache read /M
$0.27/M
Cache hit
72.4%
TTFT
5.7s
Throughput
145.2 tps
Context
1.0M
Input /M
$0.50/M
Output /M
$1.50/M
Cache read /M
$0.20/M
Cache hit
N/A
TTFT
N/A
Throughput
N/A
Context
1.0M
Input /M
$0.50/M
Output /M
$1.50/M
Cache read /M
$0.27/M
Cache hit
80.4%
TTFT
5.4s
Throughput
82.5 tps
Context
1.0M
Input /M
$0.46/M
Output /M
$1.50/M
Cache read /M
$0.15/M
Cache hit
96.1%
TTFT
6s
Throughput
78.2 tps
Context
524k
Input /M
$0.10/M
Output /M
$0.61/M
Cache read /M
$0.05/M
Cache hit
45%
TTFT
2.4s
Throughput
83.2 tps
Context
524k
Input /M
$0.21/M
Output /M
$1.47/M
Cache read /M
$0.19/M
Cache hit
2.6%
TTFT
0.8s
Throughput
125.4 tps
Context
524k
Input /M
$0.21/M
Output /M
$1.47/M
Cache read /M
$0.19/M
Cache hit
86.3%
TTFT
8.6s
Throughput
145.3 tps
Context
524k
Input /M
$0.10/M
Output /M
$0.61/M
Cache read /M
$0.05/M
Cache hit
7.4%
TTFT
2s
Throughput
165.5 tps
Context
524k
Input /M
$1.00/M
Output /M
$3.00/M
Cache read /M
$0.05/M
Cache hit
96.5%
TTFT
1.6s
Throughput
237.6 tps
Context
524k
Input /M
$1.00/M
Output /M
$3.00/M
Cache read /M
$0.05/M
Cache hit
70.4%
TTFT
3.4s
Throughput
213.4 tps
Context
262k
Input /M
$0.08/M
Output /M
$3.00/M
Cache read /M
$0.08/M
Cache hit
N/A
TTFT
0.9s
Throughput
176 tps
Context
262k
Input /M
$0.08/M
Output /M
$3.00/M
Cache read /M
$0.08/M
Cache hit
N/A
TTFT
0.9s
Throughput
176 tps

Prices are per million tokens. TTFT is the observed time to first token. Selectable routes include the applicable provider-selection markup; fixed routes show the standard model price. Availability and pricing refresh continuously; each model page shows the live provider comparison.

Using Heabsy via the API

Call a model served by Heabsy using its displayed model ID, including the provider suffix when selection is supported.

cURL

curl https://nano-gpt.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NANOGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "heabsy/cyberheabsy:heabsy",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

For models with provider selection, append :heabsy to the model ID, set "provider": "heabsy" in the request body, or send an X-Provider: heabsy header. Explicit provider selection adds a route-specific markup over that provider's base price; the prices in the table above already include it. Models without provider selection use the displayed model ID without a provider suffix or selection surcharge.

See the API documentation for provider preferences, price-aware routing, and error behavior.