Nscale
Explore Nscale models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.
Nscalenscale
No provider claimThese privacy labels describe Nscale, the upstream model provider. ZDR is a retention policy; verified confidential inference additionally requires attested provider compute and end-to-end encryption.
| Provider | Nscale |
|---|---|
| Routing status | Active |
| Provider website | https://www.nscale.com/ |
| Models | 14 public models |
| Credits routes | 14 |
| Zero data retention | no |
| Verified confidential inference | Not verified |
| Policy note | Nscale publishes a general nonlogging statement for serverless inference, but TrustedRouter has not verified a contractual zero-retention commitment for this account. These routes remain Standard and make no confidential-compute or E2EE claim. Policy source |
Measured performance
18 samplesContinuously sampled across Nscale's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 1860 ms |
|---|---|
| Effective throughput | — |
| Uptime | 100.00% |
| Model | p50 TTFT | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|
| qwen/qwen2.5-coder-7b-instruct | 1462 ms | — | 100.00% | — | 2 |
| qwen/qwen3-32b | 1596 ms | — | 100.00% | — | 4 |
| qwen/qwen3-4b-instruct-2507 | 1860 ms | — | 100.00% | — | 4 |
| nvidia/nemotron-3-nano-30b-a3b-bf16 | 2061 ms | — | 100.00% | — | 1 |
| openai/gpt-oss-20b | 2225 ms | — | 100.00% | — | 1 |
| qwen/qwen3-235b-a22b | 2303 ms | — | 100.00% | — | 1 |
| moonshotai/kimi-k2.5 | 3397 ms | — | 100.00% | — | 3 |
| qwen/qwen2.5-coder-3b-instruct | 3414 ms | — | 100.00% | — | 2 |
Nscale performance history · Full provider & model leaderboard.
Models served by Nscale.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Input | Cached input | Output |
|---|---|---|---|---|---|
black-forest-labs/flux.1-schnellFLUX.1 Schnell |
— | 0 | selected route | Not published | selected route |
moonshotai/kimi-k2.5MoonshotAI: Kimi K2.5 |
IQ 111#57 | 262,144 | $0.47475/1M | Not published | $2.321/1M |
nvidia/nemotron-3-nano-30b-a3b-bf16NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 |
— | 256,000 | $0.05275/1M | Not published | $0.211/1M |
openai/gpt-oss-120bOpenAI: gpt-oss-120b |
IQ 98#99 | 131,072 | $0.1055/1M | Not published | $0.422/1M |
openai/gpt-oss-20bOpenAI: gpt-oss-20b |
IQ 93#113 | 131,072 | $0.05275/1M | Not published | $0.211/1M |
qwen/qwen2.5-coder-32b-instruct@cf/qwen/qwen2.5-coder-32b-instruct |
— | 32,768 | $0.0633/1M | Not published | $0.211/1M |
qwen/qwen2.5-coder-3b-instructQwen2.5-Coder-3B-Instruct |
— | 32,768 | $0.01055/1M | Not published | $0.03165/1M |
qwen/qwen2.5-coder-7b-instructQwen2.5-Coder-7B-Instruct |
— | 131,072 | $0.01055/1M | Not published | $0.03165/1M |
qwen/qwen3-14bQwen: Qwen3 14B |
— | 40,960 | $0.07385/1M | Not published | $0.211/1M |
qwen/qwen3-235b-a22bQwen3-235B-A22B |
— | 32,768 | $0.211/1M | Not published | $0.633/1M |
qwen/qwen3-235b-a22b-instruct-2507Qwen3 235B A22B Instruct 2507 |
— | 131,072 | $0.211/1M | Not published | $0.633/1M |
qwen/qwen3-32bQwen: Qwen3 32B |
— | 40,960 | $0.0844/1M | Not published | $0.26375/1M |
qwen/qwen3-4b-instruct-2507Qwen3-4B-Instruct-2507 |
— | 262,144 | $0.01055/1M | Not published | $0.03165/1M |
qwen/qwen3-embedding-8bQwen3 Embedding 8B |
— | 32,768 | $0.0422/1M | Not published | selected route |
Questions
Does Nscale have zero data retention?
TrustedRouter does not currently mark Nscale as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.
Is Nscale end-to-end encrypted?
TrustedRouter does not currently mark Nscale as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.
Which Nscale models are available through TrustedRouter?
This page currently lists 14 public Nscale models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.