GLM-5.3-Flash

Z.ai

glm-5.3-flash
Open weights

Open weights, served by a provider reached through SayGM. The provider receives the prompt under its API terms.

  • Tool calling: yes
  • Image input: yes

Released

Aug 2026

Modalities

Text · Tool calling · Image input

Best price on SayGM

In

up to $0.15

Cached

up to $0.03

Out

up to $0.50

per Mtok*

SayGM vs. list price

SayGM has no current discount reading for glm-5.3-flash — check back once the next window closes.

//01

Recent effective price

What buyers have actually paid on this model recently, across every provider and cache hits.

No requests were served on this model over the last 168 windows, so there is no measured token mix to price.

How the effective price is calculated

Requests are load balanced across every provider serving this model, each offering its own rate, so what you pay is a blend rather than a single number. These are the blended rates actually paid — priced at what a provider offered, capped at list.

No requests were served on this model over the last 168 windows, so there is no measured token mix to price.

//02

Who's serving this model

Every provider currently serving this model, ranked by discount.

No provider is currently serving this model.

//03

Performance

Measured on the requests SayGM routed to this model.

Nothing has been measured for this model yet. Throughput and latency are recorded on streaming requests, and the success rate once any request completes.

No requests were served on this model across the last 12 windows.

//04

Benchmarks

Fixed-seed suites SayGM runs against this model.

No benchmark run has been published for this model.

Frequently asked questions

How much does glm-5.3-flash cost on SayGM?

glm-5.3-flash lists at $0.15 per million input tokens and $0.50 per million output tokens on SayGM.

Is glm-5.3-flash OpenAI-compatible on SayGM?

Yes. glm-5.3-flash is served on SayGM's OpenAI-compatible chat/completions surface, so an unmodified OpenAI SDK pointed at SayGM's base URL works with it.