Open weights, run inside a TEE. Inference happens sealed, so the prompt stays inside the enclave.
Released
Apr 2026Knowledge cutoff
Jan 2025Modalities
Text · Tool calling · Image inputOffered by miners now
-5.0% offeredIn
Cached
Out
per Mtok*
Create an API key and set it as SAYGM_API_KEY in your terminal.
curl "https://api.saygm.com/v1/chat/completions" \
-H "Authorization: Bearer $SAYGM_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemma-4-31b-turbo-tee",
"messages": [
{
"role": "user",
"content": "Hello!"
}
]
}'SayGM's price offered by miners now for gemma-4-31b-turbo-tee next to the published list price — 5.0% off.
| Dimension | SayGM price | List price |
|---|---|---|
| Input | $0.114per Mtok | $0.12 |
| Output | $0.3515per Mtok | $0.37 |
| Cache read | $0.0114per Mtok | $0.012 |
What buyers have actually paid on this model recently, across every provider and cache hits.
Input
$0.11
$0.12Cached input
$0.011
$0.012Output
$0.34
$0.37Requests are load balanced across every provider serving this model, each offering its own rate, so what you pay is a blend rather than a single number. These are the blended rates actually paid — priced at what a provider offered, capped at list.
Measured over 77 windows: 433 requests and 2,301,002 input-side tokens. Most recent window closed 2026-10-10 16:28 UTC.
Every provider currently serving this model, ranked by discount.
No provider is currently serving this model.
Measured on the requests SayGM routed to this model.
No requests were served on this model across the last 12 windows.
Explore the Google lab and compare related models.
The best price offered by miners now for gemma-4-31b-turbo-tee on SayGM is $0.11 per million input tokens and $0.35 per million output tokens.
SayGM's best price offered by miners now for gemma-4-31b-turbo-tee is 5.0% off list price.
Yes. gemma-4-31b-turbo-tee is served on SayGM's OpenAI-compatible chat/completions surface, so an unmodified OpenAI SDK pointed at SayGM's base URL works with it.
Fixed-seed suites SayGM runs against this model.
What is private AI inference? →