Greenference

Greenference

EU
9 models available

Data handling can vary by model and deployment. The privacy badge above shows the more conservative of the provider-wide fallback and the policies for routes on this page; each model page shows the policy for that specific route.

Greenference states that prompts and outputs are not persisted or used for training. Its in-memory prefix cache is shared across customers, is not partitioned, and has no fixed maximum lifetime. EU processing is promised, but managed datacentre hosting is not guaranteed.

Models served by Greenference

ModelContextInput /MOutput /MCache read /MCache write /MLatencyThroughput
33k$0.0095/M$0.05/M$0.0047/MFreeN/A884.7 tps
16k$0.17/M$1.42/M$0.08/MFreeN/A304.8 tps
16k$0.10/M$0.30/M$0.05/MN/AN/AN/A
8k$0.10/M$0.30/M$0.05/MFreeN/AN/A
16k$0.10/M$0.30/M$0.05/MFreeN/AN/A
33k$0.47/M$0.47/M$0.23/MFreeN/AN/A
16k$0.03/M$0.05/M$0.01/MFreeN/A189.7 tps
33k$0.05/M$0.94/M$0.03/M$0.28/MN/AN/A
33k$0.05/M$0.94/M$0.03/M$0.28/MN/AN/A

Prices are per million tokens. Selectable routes include the applicable provider-selection markup; fixed routes show the standard model price. Availability and pricing refresh continuously; each model page shows the live provider comparison.

Using Greenference via the API

For models with provider selection, append :greenference to the model ID, set "provider": "greenference" in the request body, or send an X-Provider: greenference header. Explicit provider selection adds a route-specific markup over that provider's base price; the prices in the table above already include it. Models without provider selection use the displayed model ID without a provider suffix or selection surcharge.

curl https://nano-gpt.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NANOGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "openai/gpt-oss-20b:greenference",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

See the API documentation for provider preferences, price-aware routing, and error behavior.