Chutes · Confidential
Qwen/Qwen3-32B-TEE
Open weights, run inside a TEE. Inference happens sealed, so the prompt stays inside the enclave.
- Tool calling: yes
- Image input: yes
Released
Apr 2025Modalities
Text · Tool calling · Image inputBest price on gm
-55.5%In
Cached
Out
per Mtok*
Recent effective price
What buyers have actually paid on this model recently, across every provider and cache hits.
Input
$0.078
$0.104Cached input
$0.039
$0.052Output
$0.306
$0.416How the effective price is calculated
Requests are load balanced across every provider serving this model, each offering its own rate, so what you pay is a blend rather than a single number. These are the blended rates actually paid — priced at what a provider offered, capped at list.
Measured over 15 windows: 1,451 requests and 2,096,552 input-side tokens. Most recent window closed 2026-08-11 02:24 UTC.
Who's serving this model
Every provider currently serving this model, ranked by discount.
| Provider | Discount | In | Out |
|---|---|---|---|
| Provider 1 | 55.5% off | $0.046 | $0.185 |
| Provider 2 | 52.5% off | $0.049 | $0.197 |
| Provider 3 | 52.5% off | $0.049 | $0.197 |
| Provider 4 | 4.0% off | $0.100 | $0.399 |
Performance
Measured on the requests gm routed to this model.
- Throughput
- 16.2 tok/sover 6,284 requests
- Time to first token
- 4.0 sover 6,281 requests
- Request success rate
- 100.0% over 6,290 requests
Benchmarks
Fixed-seed suites gm runs against this model.
No benchmark run has been published for this model.