Qwen
Open weights, run inside a TEE. Inference happens sealed, so the prompt stays inside the enclave.
Released
Apr 2025Modalities
Text · Tool callingBest price on SayGM
-51.0%In
Cached
Out
per Mtok*
SayGM's best available rate for qwen3-235b-a22b-thinking-2507-tee next to the published list price — 51.0% off.
| Dimension | SayGM price | List price |
|---|---|---|
| Input | $0.146461per Mtok | $0.2989 |
| Output | $0.585893per Mtok | $1.1957 |
| Cache read | $0.0146461per Mtok | $0.02989 |
What buyers have actually paid on this model recently, across every provider and cache hits.
Input
$0.2028
$0.2989Cached input
$0.02088
$0.02989Output
$0.8129
$1.1957Requests are load balanced across every provider serving this model, each offering its own rate, so what you pay is a blend rather than a single number. These are the blended rates actually paid — priced at what a provider offered, capped at list.
Measured over 11 windows: 801 requests and 2,206,304 input-side tokens. Most recent window closed 2026-09-16 09:06 UTC.
Every provider currently serving this model, ranked by discount.
| Provider | Discount | In | Out |
|---|---|---|---|
| Provider 1 | 51.0% off | $0.1465 | $0.5859 |
| Provider 2 | 7.0% off | $0.2780 | $1.1120 |
Measured on the requests SayGM routed to this model.
Fixed-seed suites SayGM runs against this model.
The best available rate for qwen3-235b-a22b-thinking-2507-tee on SayGM right now is $0.1465 per million input tokens and $0.5859 per million output tokens.
SayGM's best available rate for qwen3-235b-a22b-thinking-2507-tee is 51.0% off list price.
Yes. qwen3-235b-a22b-thinking-2507-tee is served on SayGM's OpenAI-compatible chat/completions surface, so an unmodified OpenAI SDK pointed at SayGM's base URL works with it.