OpenAI compatible API · Attested · Public status

Chutes

Explore Chutes models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Chuteschutes

Confidential

All providers

These privacy labels describe Chutes, the upstream model provider. ZDR is a retention policy; verified confidential inference additionally requires attested provider compute and end-to-end encryption.

ProviderChutes
Routing statusActive
Provider websitehttps://chutes.ai/
Models10 public models
Credits routes10
Zero data retentionyes All routes
Verified confidential inferenceVerified
Policy noteTrustedRouter encrypts each request to an attested Chutes workload and verifies Intel TDX plus NVIDIA GPU attestation inside the TrustedRouter enclave before sending content. Verification fails closed. Chutes also documents no prompt/output storage or training.
Policy source

Measured performance

26 samples

Continuously sampled across Chutes's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT7507 ms
Effective throughput29 tok/s n=1
Uptime76.92%
Modelp50 TTFTEffective throughputUptimeConfig excludedAvailability samples
google/gemma-4-31b-turbo 5730 ms 100.00% 2
z-ai/glm-5.1 6129 ms 50.00% 4
mistralai/mistral-nemo 6395 ms 100.00% 4
qwen/qwen3-32b 7197 ms 100.00% 1
deepseek/deepseek-v3.2 7507 ms 100.00% 4
qwen/qwen3.5-397b-a17b 9891 ms 100.00% 3
qwen/qwen3.6-27b 12543 ms 75.00% 4
moonshotai/kimi-k2.6 12901 ms 29 tok/s n=1 100.00% 1
qwen/qwen3-235b-a22b-thinking-2507 0.00% 3

Chutes performance history · Full provider & model leaderboard.

Provider models

Models served by Chutes.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Input Cached input Output
deepseek/deepseek-v3.2
DeepSeek: DeepSeek V3.2
IQ 103#77 163,840 $1.055/1M $0.1055/1M $1.055/1M
google/gemma-4-31b-turbo
google/gemma-4-31B-turbo
131,072 $0.1266/1M $0.01266/1M $0.39035/1M
mistralai/mistral-nemo
Mistral: Mistral Nemo
131,072 $0.025848/1M $0.01/1M $0.103179/1M
moonshotai/kimi-k2.6
MoonshotAI: Kimi K2.6
IQ 118#35 262,144 $0.6119/1M $0.06119/1M $3.587/1M
qwen/qwen3-235b-a22b-thinking-2507
Qwen: Qwen3 235B A22B Thinking 2507
131,072 $0.31534/1M $0.031534/1M $1.261464/1M
qwen/qwen3-32b
Qwen: Qwen3 32B
40,960 $0.10972/1M $0.010972/1M $0.43888/1M
qwen/qwen3.5-397b-a17b
Qwen: Qwen3.5 397B A17B
262,144 $0.47475/1M $0.047475/1M $3.165/1M
qwen/qwen3.6-27b
Qwen: Qwen3.6 27B
IQ 108#68 262,144 $0.3165/1M $0.03165/1M $2.11/1M
z-ai/glm-5.1
Z.ai: GLM 5.1
IQ 113#49 202,752 $1.0339/1M $0.10339/1M $3.2494/1M
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#28 1,048,576 $1.31875/1M $0.131875/1M $4.16725/1M

Questions

Does Chutes have zero data retention?

TrustedRouter records Chutes as supporting provider-level zero data retention based on the policy source linked on this page. This is a provider policy claim, separate from TrustedRouter's content-stateless real-time gateway and from end-to-end confidential compute.

Is Chutes end-to-end encrypted?

TrustedRouter records Chutes as supporting provider-side confidential compute and end-to-end encrypted inference. The route-specific model page shows whether that protection applies to a particular endpoint.

Which Chutes models are available through TrustedRouter?

This page currently lists 10 public Chutes models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.