OpenAI compatible API · Attested · Public status

Nscale

Explore Nscale models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Nscalenscale

No provider claim

All providers

These privacy labels describe Nscale, the upstream model provider. ZDR is a retention policy; verified confidential inference additionally requires attested provider compute and end-to-end encryption.

ProviderNscale
Routing statusActive
Provider websitehttps://www.nscale.com/
Models14 public models
Credits routes14
Zero data retentionno
Verified confidential inferenceNot verified
Policy noteNscale publishes a general nonlogging statement for serverless inference, but TrustedRouter has not verified a contractual zero-retention commitment for this account. These routes remain Standard and make no confidential-compute or E2EE claim.
Policy source

Measured performance

18 samples

Continuously sampled across Nscale's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT1860 ms
Effective throughput
Uptime100.00%
Modelp50 TTFTEffective throughputUptimeConfig excludedAvailability samples
qwen/qwen2.5-coder-7b-instruct 1462 ms 100.00% 2
qwen/qwen3-32b 1596 ms 100.00% 4
qwen/qwen3-4b-instruct-2507 1860 ms 100.00% 4
nvidia/nemotron-3-nano-30b-a3b-bf16 2061 ms 100.00% 1
openai/gpt-oss-20b 2225 ms 100.00% 1
qwen/qwen3-235b-a22b 2303 ms 100.00% 1
moonshotai/kimi-k2.5 3397 ms 100.00% 3
qwen/qwen2.5-coder-3b-instruct 3414 ms 100.00% 2

Nscale performance history · Full provider & model leaderboard.

Provider models

Models served by Nscale.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Input Cached input Output
black-forest-labs/flux.1-schnell
FLUX.1 Schnell
0 selected route Not published selected route
moonshotai/kimi-k2.5
MoonshotAI: Kimi K2.5
IQ 111#57 262,144 $0.47475/1M Not published $2.321/1M
nvidia/nemotron-3-nano-30b-a3b-bf16
NVIDIA-Nemotron-3-Nano-30B-A3B-BF16
256,000 $0.05275/1M Not published $0.211/1M
openai/gpt-oss-120b
OpenAI: gpt-oss-120b
IQ 98#99 131,072 $0.1055/1M Not published $0.422/1M
openai/gpt-oss-20b
OpenAI: gpt-oss-20b
IQ 93#113 131,072 $0.05275/1M Not published $0.211/1M
qwen/qwen2.5-coder-32b-instruct
@cf/qwen/qwen2.5-coder-32b-instruct
32,768 $0.0633/1M Not published $0.211/1M
qwen/qwen2.5-coder-3b-instruct
Qwen2.5-Coder-3B-Instruct
32,768 $0.01055/1M Not published $0.03165/1M
qwen/qwen2.5-coder-7b-instruct
Qwen2.5-Coder-7B-Instruct
131,072 $0.01055/1M Not published $0.03165/1M
qwen/qwen3-14b
Qwen: Qwen3 14B
40,960 $0.07385/1M Not published $0.211/1M
qwen/qwen3-235b-a22b
Qwen3-235B-A22B
32,768 $0.211/1M Not published $0.633/1M
qwen/qwen3-235b-a22b-instruct-2507
Qwen3 235B A22B Instruct 2507
131,072 $0.211/1M Not published $0.633/1M
qwen/qwen3-32b
Qwen: Qwen3 32B
40,960 $0.0844/1M Not published $0.26375/1M
qwen/qwen3-4b-instruct-2507
Qwen3-4B-Instruct-2507
262,144 $0.01055/1M Not published $0.03165/1M
qwen/qwen3-embedding-8b
Qwen3 Embedding 8B
32,768 $0.0422/1M Not published selected route

Questions

Does Nscale have zero data retention?

TrustedRouter does not currently mark Nscale as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.

Is Nscale end-to-end encrypted?

TrustedRouter does not currently mark Nscale as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.

Which Nscale models are available through TrustedRouter?

This page currently lists 14 public Nscale models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.