- Context
- 1.0M
- Input /M
- $0.13/M
- Output /M
- $0.53/M
- Cache read /M
- $0.05/M
- Cache hit
- 96.6%
- TTFT
- 2.7s
- Throughput
- 35.1 tps
Arnict
Arnict confirms zero data retention for inference traffic: prompts and outputs are not archived or used for training. Temporary inference caches are VRAM-only and expire 15 minutes after last use. Account, usage and billing metadata may be retained.
Data handling can vary by model and deployment. The privacy badge above shows the more conservative of the provider-wide fallback and the policies for routes on this page; each model page shows the policy for that specific route.
Models served by Arnict
- Context
- 1.0M
- Input /M
- $1.31/M
- Output /M
- $2.36/M
- Cache read /M
- $0.31/M
- Cache hit
- 97.1%
- TTFT
- 1.7s
- Throughput
- 34.3 tps
Prices are per million tokens. TTFT is the observed time to first token. Selectable routes include the applicable provider-selection markup; fixed routes show the standard model price. Availability and pricing refresh continuously; each model page shows the live provider comparison.
Using Arnict via the API
Call a model served by Arnict using its displayed model ID, including the provider suffix when selection is supported.
cURL
curl https://nano-gpt.com/api/v1/chat/completions \
-H "Authorization: Bearer $NANOGPT_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "z-ai/glm-5.3-flash-uncensored:arnict",
"messages": [{"role": "user", "content": "Hello"}]
}'For models with provider selection, append :arnict to the model ID, set "provider": "arnict" in the request body, or send an X-Provider: arnict header. Explicit provider selection adds a route-specific markup over that provider's base price; the prices in the table above already include it. Models without provider selection use the displayed model ID without a provider suffix or selection surcharge.
See the API documentation for provider preferences, price-aware routing, and error behavior.