Open weights, served by a provider reached through SayGM. The provider receives the prompt under its API terms.
Released
Aug 2026Modalities
Text · Tool calling · Image inputBest reported price on SayGM
-12.0%In
Cached
Out
per Mtok*
Create an API key and set it as SAYGM_API_KEY in your terminal.
curl "https://api.saygm.com/v1/chat/completions" \
-H "Authorization: Bearer $SAYGM_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "glm-5.3-flash",
"messages": [
{
"role": "user",
"content": "Hello!"
}
]
}'SayGM's best reported rate for glm-5.3-flash next to Z.ai's own published list price — 12.0% off.
Compare every GLM model across providers in our GLM pricing breakdown.
| Dimension | SayGM price | Z.ai list price |
|---|---|---|
| Input | $0.132per Mtok | $0.15 |
| Output | $0.44per Mtok | $0.50 |
| Cache read | $0.0264per Mtok | $0.03 |
What buyers have actually paid on this model recently, across every provider and cache hits.
Input
$0.14
$0.15Cached input
$0.03
Output
$0.45
$0.50Requests are load balanced across every provider serving this model, each offering its own rate, so what you pay is a blend rather than a single number. These are the blended rates actually paid — priced at what a provider offered, capped at list.
Measured over 1 window: 687 requests and 7,913,487 input-side tokens. Most recent window closed 2026-10-11 00:55 UTC.
Every provider currently serving this model, ranked by discount.
| Provider | Discount | In | Out |
|---|---|---|---|
| Provider 1 | 12.0% off | $0.13 | $0.44 |
| Provider 2 | 12.0% off | $0.13 | $0.44 |
| Provider 3 | 11.8% off | $0.13 | $0.44 |
| Provider 4 | 11.7% off | $0.13 | $0.44 |
| Provider 5 | 11.7% off | $0.13 | $0.44 |
| Provider 6 | 11.0% off | $0.13 | $0.45 |
| Provider 7 | 10.0% off | $0.14 | $0.45 |
| Provider 8 | 10.0% off | $0.14 | $0.45 |
| Provider 9 | 10.0% off | $0.14 | $0.45 |
| Provider 10 | 10.0% off | $0.14 | $0.45 |
| Provider 11 | 10.0% off | $0.14 | $0.45 |
| Provider 12 | 10.0% off | $0.14 | $0.45 |
| Provider 13 | 10.0% off | $0.14 | $0.45 |
| Provider 14 | — | $0.15 | $0.50 |
| Provider 15 | — | $0.15 | $0.50 |
Measured on the requests SayGM routed to this model.
Fixed-seed suites SayGM runs against this model.
Explore the Z.ai lab and compare related models.
The best reported rate for glm-5.3-flash on SayGM is $0.13 per million input tokens and $0.44 per million output tokens.
SayGM's best reported rate for glm-5.3-flash is 12.0% off Z.ai's list price.
Yes. glm-5.3-flash is served on SayGM's OpenAI-compatible chat/completions surface, so an unmodified OpenAI SDK pointed at SayGM's base URL works with it.