OpenAI compatible API · Attested · Public status

SiliconFlow

Explore SiliconFlow models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

SiliconFlowsiliconflow

No provider claim

All providers

ProviderSiliconFlow
Provider websitehttps://www.siliconflow.com/
Models28 public models
Prepaid routes28
BYOK routes28
Zero data retentionnot claimed
Confidential computenot claimed
Provider E2EEnot claimed
Policy noteNo provider-ZDR claim is tracked here. SiliconFlow's privacy policy source is linked for retention and interaction-data terms.
Policy source

Measured performance

52 samples

Continuously sampled across SiliconFlow's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT2065 ms
Effective throughput50 tok/s n=12
Uptime98.08%
Modelp50 TTFTp50 TTFBEffective throughputUptimeConfig excludedAvailability samples
moonshotai/kimi-k2.7-code 1091 ms 1091 ms 43 tok/s n=1 100.00% 3
deepseek/deepseek-v4-flash-0731 1533 ms 1533 ms 73 tok/s n=2 100.00% 2
z-ai/glm-4.5-air 1535 ms 1535 ms 100.00% 2
qwen/qwen3.5-9b 1599 ms 1599 ms 100.00% 6
moonshotai/kimi-k2.5 1637 ms 1637 ms 100.00% 1
z-ai/glm-5.2 1820 ms 1820 ms 31 tok/s n=1 100.00% 3
google/gemma-4-31b-it 1847 ms 1847 ms 12 tok/s n=1 100.00% 2
openai/gpt-oss-20b 1874 ms 1873 ms 100.00% 1
tencent/hy3 2028 ms 2028 ms 75 tok/s n=2 100.00% 3
qwen/qwen3.5-122b-a10b 2065 ms 2065 ms 100.00% 3
qwen/qwen3.6-35b-a3b 2210 ms 2210 ms 100.00% 3
deepseek/deepseek-v4-flash 2236 ms 2236 ms 47 tok/s n=1 100.00% 1
deepseek/deepseek-v3.2 2453 ms 2453 ms 100.00% 2
z-ai/glm-5v-turbo 2666 ms 2666 ms 100.00% 1
tencent/hunyuan-a13b-instruct 2977 ms 2977 ms 100.00% 2
openai/gpt-oss-120b 3285 ms 3285 ms 15 tok/s n=1 100.00% 2
moonshotai/kimi-k3 3343 ms 3343 ms 29 tok/s n=1 100.00% 1
deepseek/deepseek-v4-pro 3870 ms 3870 ms 54 tok/s n=1 100.00% 4
deepseek/deepseek-v3.1-terminus 4410 ms 4410 ms 100.00% 2
z-ai/glm-5.1 4504 ms 4504 ms 100.00% 1
deepseek/deepseek-v3.2-exp 4662 ms 4662 ms 100.00% 5
qwen/qwen3.5-27b 1675 ms 1674 ms 50.00% 2
minimax/minimax-m3 172 tok/s n=1 0

SiliconFlow performance history · Full provider & model leaderboard.

Provider models

Models served by SiliconFlow.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Endpoints Prompt Completion Routes
deepseek/deepseek-v3.1-terminus
DeepSeek: DeepSeek V3.1 Terminus
163,840 2 $0.2835/1M $1.05/1M prepaid BYOK
deepseek/deepseek-v3.2
DeepSeek: DeepSeek V3.2
IQ 103#57 163,840 2 $0.2835/1M $0.441/1M prepaid BYOK
deepseek/deepseek-v3.2-exp
DeepSeek: DeepSeek V3.2 Exp
163,840 2 $0.2835/1M $0.4305/1M prepaid BYOK
deepseek/deepseek-v4-flash
DeepSeek: DeepSeek V4 Flash 0423
IQ 108#46 1,048,576 2 $0.1365/1M $0.294/1M prepaid BYOK
deepseek/deepseek-v4-flash-0731
DeepSeek: DeepSeek V4 Flash 0731
1,048,576 2 $0.1365/1M $0.294/1M prepaid BYOK
deepseek/deepseek-v4-pro
DeepSeek: DeepSeek V4 Pro
IQ 115#30 1,048,576 2 $1.576701/1M $3.29175/1M prepaid BYOK
google/gemma-4-26b-a4b-it
Google: Gemma 4 26B A4B
IQ 96#81 262,144 2 $0.126/1M $0.42/1M prepaid BYOK
google/gemma-4-31b-it
Google: Gemma 4 31B
IQ 101#66 262,144 2 $0.1365/1M $0.42/1M prepaid BYOK
minimax/minimax-m2.5
MiniMax: MiniMax M2.5
IQ 105#56 204,800 2 $0.315/1M $1.26/1M prepaid BYOK
minimax/minimax-m3
MiniMax: MiniMax M3
IQ 114#33 1,048,576 2 $0.315/1M $1.26/1M prepaid BYOK
moonshotai/kimi-k2.5
MoonshotAI: Kimi K2.5
IQ 111#39 262,144 2 $0.4725/1M $2.3625/1M prepaid BYOK
moonshotai/kimi-k2.6
MoonshotAI: Kimi K2.6
IQ 119#19 262,144 2 $0.8085/1M $3.57/1M prepaid BYOK
moonshotai/kimi-k2.7-code
MoonshotAI: Kimi K2.7 Code
IQ 118#22 262,144 2 $0.902118/1M $3.99/1M prepaid BYOK
moonshotai/kimi-k3
MoonshotAI: Kimi K3
IQ 123#14 1,048,576 2 $3.15/1M $15.75/1M prepaid BYOK
openai/gpt-oss-120b
OpenAI: gpt-oss-120b
IQ 105#53 131,072 2 $0.0525/1M $0.4725/1M prepaid BYOK
openai/gpt-oss-20b
OpenAI: gpt-oss-20b
IQ 100#70 131,072 2 $0.042/1M $0.189/1M prepaid BYOK
qwen/qwen3.5-122b-a10b
Qwen: Qwen3.5-122B-A10B
262,144 2 $0.273/1M $2.184/1M prepaid BYOK
qwen/qwen3.5-27b
Qwen: Qwen3.5-27B
262,144 2 $0.2625/1M $2.1/1M prepaid BYOK
qwen/qwen3.5-9b
Qwen: Qwen3.5-9B
IQ 93#90 262,144 2 $0.105/1M $0.1575/1M prepaid BYOK
qwen/qwen3.6-27b
Qwen: Qwen3.6 27B
IQ 111#40 262,144 2 $0.315/1M $3.36/1M prepaid BYOK
qwen/qwen3.6-35b-a3b
Qwen: Qwen3.6 35B A3B
IQ 100#71 262,144 2 $0.21/1M $1.68/1M prepaid BYOK
tencent/hunyuan-a13b-instruct
Tencent: Hunyuan A13B Instruct
131,072 2 $0.147/1M $0.5985/1M prepaid BYOK
tencent/hy3
Tencent: Hy3
IQ 103#60 262,144 2 $0.1386/1M $0.5544/1M prepaid BYOK
z-ai/glm-4.5-air
Z.ai: GLM 4.5 Air
131,072 2 $0.147/1M $0.903/1M prepaid BYOK
z-ai/glm-5
Z.ai: GLM 5
IQ 105#52 204,800 2 $0.9975/1M $2.6775/1M prepaid BYOK
z-ai/glm-5.1
Z.ai: GLM 5.1
IQ 114#32 204,800 2 $1.2495/1M $3.927/1M prepaid BYOK
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#16 1,048,576 2 $1.47/1M $4.62/1M prepaid BYOK
z-ai/glm-5v-turbo
Z.ai: GLM 5V Turbo
202,752 2 $1.26/1M $4.2/1M prepaid BYOK
Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.