OpenAI compatible API · Attested · Public status
Chutes
Chutes models on TrustedRouter with prices, routes, policy notes, and source links.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
chutes
No logs
| Provider | Chutes |
|---|---|
| Models | 10 public models |
| Prepaid routes | 10 |
| BYOK routes | 10 |
| Zero data retention | yes |
| Confidential compute | yes |
| Provider E2EE | no |
| Policy note | Chutes documents no prompt/output storage or training and serves these routes in confidential-compute TEEs. Standard API calls are not marked provider end-to-end encrypted. Policy source |
Measured performance
51 samplesContinuously sampled across Chutes's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 2838 ms |
|---|---|
| Effective throughput | 31 tok/s n=3 |
| Uptime | 94.12% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| qwen/qwen3-235b-a22b-thinking-2507 | 1379 ms | 1379 ms | — | 100.00% | — | 4 |
| mistralai/mistral-nemo | 1859 ms | 1859 ms | — | 100.00% | — | 6 |
| moonshotai/kimi-k2.6 | 2025 ms | 2025 ms | 34 tok/s n=1 | 100.00% | — | 6 |
| z-ai/glm-5.1 | 2027 ms | 2027 ms | — | 100.00% | — | 4 |
| qwen/qwen3.6-27b | 2600 ms | 2599 ms | — | 100.00% | — | 1 |
| qwen/qwen3-32b | 3575 ms | 3575 ms | — | 100.00% | — | 8 |
| z-ai/glm-5.2 | 3914 ms | 3914 ms | 31 tok/s n=2 | 100.00% | — | 4 |
| deepseek/deepseek-v3.2 | 4166 ms | 4166 ms | — | 100.00% | — | 7 |
| google/gemma-4-31b-turbo | 4308 ms | 4308 ms | — | 80.00% | — | 5 |
| qwen/qwen3.5-397b-a17b | 2838 ms | 2837 ms | — | 75.00% | — | 4 |
| z-ai/glm-5 | 1592 ms | 1592 ms | — | 50.00% | 1 unsupported_route |
2 |
Chutes performance history · Full provider & model leaderboard.
Provider models
Models served by Chutes.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-v3.2DeepSeek: DeepSeek V3.2 |
IQ 103#56 | 163,840 | 2 | $1.05/1M | $1.05/1M | prepaid BYOK |
google/gemma-4-31b-turbogoogle/gemma-4-31B-turbo |
— | 131,072 | 2 | $0.126/1M | $0.3885/1M | prepaid BYOK |
mistralai/mistral-nemoMistral: Mistral Nemo |
— | 131,072 | 2 | $0.025725/1M | $0.10269/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $0.693/1M | $3.675/1M | prepaid BYOK |
qwen/qwen3-235b-a22b-thinking-2507Qwen: Qwen3 235B A22B Thinking 2507 |
— | 262,144 | 2 | $0.313845/1M | $1.255485/1M | prepaid BYOK |
qwen/qwen3-32bQwen: Qwen3 32B |
— | 131,072 | 2 | $0.1092/1M | $0.4368/1M | prepaid BYOK |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | 2 | $0.4725/1M | $3.15/1M | prepaid BYOK |
qwen/qwen3.6-27bQwen: Qwen3.6 27B |
IQ 111#40 | 262,144 | 2 | $0.315/1M | $2.1/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#32 | 204,800 | 2 | $1.029/1M | $3.234/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.3125/1M | $4.1475/1M | prepaid BYOK |