Hosted by a provider reached through SayGM. The route is sealed in transit and in use; the provider receives the prompt under its API terms.
Released
May 2026Best price on SayGM
-8.0%In
Cached
Out
per Mtok*
SayGM's best available rate for gemini-3.1-flash-lite next to Google's own published list price — 8.0% off.
| Dimension | SayGM price | Google list price |
|---|---|---|
| Input | $0.23per Mtok | $0.25 |
| Output | $1.38per Mtok | $1.50 |
| Cache read | $0.023per Mtok | $0.025 |
| Audio input | $0.46per Mtok | $0.50 |
| Audio output | $1.38per Mtok | $1.50 |
| Image input | $0.23per Mtok | $0.25 |
What buyers have actually paid on this model recently, across every provider and cache hits.
Input
$0.23
$0.25Cached input
$0.023
$0.025Output
$1.38
$1.50Requests are load balanced across every provider serving this model, each offering its own rate, so what you pay is a blend rather than a single number. These are the blended rates actually paid — priced at what a provider offered, capped at list.
Measured over 29 windows: 1,091 requests and 2,141,567 input-side tokens. Most recent window closed 2026-09-16 12:44 UTC.
Every provider currently serving this model, ranked by discount.
| Provider | Discount | In | Out |
|---|---|---|---|
| Provider 1 | 8.0% off | $0.23 | $1.38 |
| Provider 2 | 8.0% off | $0.23 | $1.38 |
| Provider 3 | 8.0% off | $0.23 | $1.38 |
Measured on the requests SayGM routed to this model.
Fixed-seed suites SayGM runs against this model.
The best available rate for gemini-3.1-flash-lite on SayGM right now is $0.23 per million input tokens and $1.38 per million output tokens.
SayGM's best available rate for gemini-3.1-flash-lite is 8.0% off Google's list price.
No. gemini-3.1-flash-lite is served on SayGM's native Gemini surface, not the OpenAI-compatible endpoint.