OpenAI compatible API · Attested · Public status
Chutes
Explore Chutes models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
Chuteschutes
No logs| Provider | Chutes |
|---|---|
| Provider website | https://chutes.ai/ |
| Models | 10 public models |
| Prepaid routes | 10 |
| BYOK routes | 10 |
| Zero data retention | yes |
| Confidential compute | yes |
| Provider E2EE | no |
| Policy note | Chutes documents no prompt/output storage or training and serves these routes in confidential-compute TEEs. Standard API calls are not marked provider end-to-end encrypted. Policy source |
Measured performance
56 samplesContinuously sampled across Chutes's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 2101 ms |
|---|---|
| Effective throughput | 27 tok/s n=2 |
| Uptime | 91.07% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| mistralai/mistral-nemo | 1684 ms | 1684 ms | — | 100.00% | — | 8 |
| qwen/qwen3.5-397b-a17b | 2033 ms | 2032 ms | — | 100.00% | — | 8 |
| qwen/qwen3.6-27b | 2101 ms | 2101 ms | — | 100.00% | — | 7 |
| qwen/qwen3-235b-a22b-thinking-2507 | 2408 ms | 2407 ms | — | 100.00% | — | 9 |
| google/gemma-4-31b-turbo | 2457 ms | 2457 ms | — | 100.00% | — | 5 |
| moonshotai/kimi-k2.6 | 3657 ms | 3657 ms | 27 tok/s n=2 | 100.00% | — | 4 |
| deepseek/deepseek-v3.2 | 3783 ms | 3783 ms | — | 100.00% | — | 4 |
| qwen/qwen3-32b | 17002 ms | 17002 ms | — | 100.00% | — | 1 |
| z-ai/glm-5.1 | 4464 ms | 4464 ms | — | 75.00% | — | 4 |
| z-ai/glm-5.2 | 1438 ms | 1438 ms | — | 33.33% | — | 6 |
Chutes performance history · Full provider & model leaderboard.
Provider models
Models served by Chutes.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-v3.2DeepSeek: DeepSeek V3.2 |
IQ 103#57 | 163,840 | 2 | $1.05/1M | $1.05/1M | prepaid BYOK |
google/gemma-4-31b-turbogoogle/gemma-4-31B-turbo |
— | 131,072 | 2 | $0.126/1M | $0.3885/1M | prepaid BYOK |
mistralai/mistral-nemoMistral: Mistral Nemo |
— | 131,072 | 2 | $0.025725/1M | $0.10269/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $0.609/1M | $3.57/1M | prepaid BYOK |
qwen/qwen3-235b-a22b-thinking-2507Qwen: Qwen3 235B A22B Thinking 2507 |
— | 262,144 | 2 | $0.313845/1M | $1.255485/1M | prepaid BYOK |
qwen/qwen3-32bQwen: Qwen3 32B |
— | 131,072 | 2 | $0.1092/1M | $0.4368/1M | prepaid BYOK |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | 2 | $0.4725/1M | $3.15/1M | prepaid BYOK |
qwen/qwen3.6-27bQwen: Qwen3.6 27B |
IQ 111#40 | 262,144 | 2 | $0.315/1M | $2.1/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#32 | 204,800 | 2 | $1.029/1M | $3.234/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.3125/1M | $4.1475/1M | prepaid BYOK |