Gemma 4 31B turbo (TEE)
Open weights, run inside a TEE. Inference happens sealed, so the prompt stays inside the enclave.
- Tool calling: yes
- Image input: yes
Released
Apr 2026Knowledge cutoff
Jan 2025Modalities
Text · Tool calling · Image inputBest price on SayGM
-53.6%In
Cached
Out
per Mtok*
SayGM vs. list price
SayGM's best available rate for gemma-4-31b-turbo-tee next to the published list price — 53.6% off.
| Dimension | SayGM price | List price |
|---|---|---|
| Input | $0.060255per Mtok | $0.13 |
| Output | $0.1854per Mtok | $0.40 |
| Cache read | $0.02781per Mtok | $0.06 |
Recent effective price
What buyers have actually paid on this model recently, across every provider and cache hits.
Input
$0.12
$0.13Cached input
$0.05
$0.06Output
$0.36
$0.40How the effective price is calculated
Requests are load balanced across every provider serving this model, each offering its own rate, so what you pay is a blend rather than a single number. These are the blended rates actually paid — priced at what a provider offered, capped at list.
Measured over 168 windows: 851 requests and 1,591,724 input-side tokens, short of the 2,000,000 tokens the window walks for. Most recent window closed 2026-08-26 14:17 UTC.
Who's serving this model
Every provider currently serving this model, ranked by discount.
| Provider | Discount | In | Out |
|---|---|---|---|
| Provider 1 | 53.6% off | $0.06 | $0.19 |
Performance
Measured on the requests SayGM routed to this model.
- Throughput
- 26.0 tok/sover 180 requests
- Time to first token
- 3.2 sover 180 requests
- Request success rate
- 100.0% over 102 requests
Benchmarks
Fixed-seed suites SayGM runs against this model.
Frequently asked questions
How much does gemma-4-31b-turbo-tee cost on SayGM?
The best available rate for gemma-4-31b-turbo-tee on SayGM right now is $0.06 per million input tokens and $0.19 per million output tokens.
How much cheaper is SayGM than list price?
SayGM's best available rate for gemma-4-31b-turbo-tee is 53.6% off list price.
Is gemma-4-31b-turbo-tee OpenAI-compatible on SayGM?
Yes. gemma-4-31b-turbo-tee is served on SayGM's OpenAI-compatible chat/completions surface, so an unmodified OpenAI SDK pointed at SayGM's base URL works with it.