SiliconFlow

SiliconFlow

SG
48 models available

Models served by SiliconFlow

ModelContextInput /MOutput /MCache read /MLatencyThroughput
128k$0.53/M$2.29/MN/A1.7s17 tps
128k$0.26/M$1.05/MN/A1.6s19 tps
128k$0.28/M$1.05/MN/A1.4s16 tps
128k$0.28/M$1.05/MN/A1.9s19 tps
128k$0.28/M$1.05/MN/A1.9s19 tps
163k$0.27/M$0.44/M$0.14/M2.1s18 tps
163k$0.27/M$0.44/M$0.14/M2.1s18 tps
1.0M$0.15/M$0.29/M$0.03/M1.4s84 tps
1.0M$0.15/M$0.29/M$0.03/M1.4s84 tps
1.0M$0.23/M$0.69/M$0.03/M1.5s51 tps
1.0M$0.23/M$0.69/M$0.03/M1.5s51 tps
1.0M$0.46/M$1.39/M$0.03/M0.8s86 tps
1.0M$1.83/M$3.65/M$0.15/M2s35 tps
1.0M$1.83/M$3.65/M$0.15/M2s35 tps
1.0M$1.39/M$4.16/M$0.05/M1.3s46 tps
1.0M$1.39/M$4.16/M$0.05/M1.3s46 tps
262k$0.13/M$0.42/MN/A1.4s6 tps
262k$0.13/M$0.42/MN/A1.4s6 tps
262k$0.14/M$0.42/MN/A4.4s26 tps
262k$0.14/M$0.42/MN/A4.4s26 tps
128k$0.15/M$0.90/MN/A1.5s5 tps
128k$0.15/M$0.90/MN/A1.5s5 tps
200k$1.00/M$2.68/M$0.21/M2.3s42 tps
200k$1.00/M$2.68/M$0.21/M2.3s42 tps
200k$1.25/M$3.93/M$0.63/M1.4s28 tps
200k$1.25/M$3.93/M$0.63/M1.4s28 tps
1.0M$1.25/M$3.93/M$0.23/M1.5s40 tps
1.0M$1.25/M$3.93/M$0.23/M1.5s40 tps
1.0M$1.47/M$4.62/M$0.27/M2.2s31 tps
1.0M$1.47/M$4.62/M$0.27/M2.2s31 tps
1.0M$0.16/M$0.53/M$0.03/M4.8s25 tps
128k$0.05/M$0.47/MN/A2.2s7 tps
128k$0.04/M$0.19/MN/A1.2s43 tps
256k$0.47/M$2.36/M$0.07/M1.9s32 tps
256k$0.47/M$2.36/M$0.07/M1.9s32 tps
262k$0.90/M$3.99/M$0.19/M0.7s68.5 tps
1.0M$0.79/M$3.10/M$0.02/MN/AN/A
1.0M$0.79/M$3.10/M$0.02/MN/AN/A
205k$0.31/M$1.26/M$0.03/M1.1s45 tps
512k$0.63/M$2.52/M$0.13/MN/AN/A
512k$0.63/M$2.52/M$0.13/MN/AN/A
260k$0.31/M$3.36/MN/A1.5s19.5 tps
260k$0.31/M$3.36/MN/A1.5s19.5 tps
262k$0.21/M$1.68/MN/A2.7s36 tps
262k$0.21/M$1.68/MN/A2.7s36 tps
41k$0.15/M$0.60/MN/A1.5s35 tps
1.0M$2.10/M$6.30/M$0.26/M1.5s31 tps
1.0M$2.10/M$6.30/M$0.26/M1.5s31 tps

Prices are per million tokens and include the 5% provider-selection markup. Availability and pricing refresh continuously; each model page shows the live provider comparison.

Using SiliconFlow via the API

Append :siliconflow to the model ID, set "provider": "siliconflow" in the request body, or send an X-Provider: siliconflow header. Explicit provider selection adds a 5% markup over that provider's base price; the prices in the table above already include it.

curl https://nano-gpt.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NANOGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-ai/DeepSeek-R1-0528:siliconflow",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

See the API documentation for provider preferences, price-aware routing, and error behavior.