DeepSeek
Open weights, served by a provider reached through SayGM. The provider receives the prompt under its API terms.
Released
Sep 2026Modalities
Text · Tool calling · Image inputBest price on SayGM
In
Cached
Out
per Mtok*
SayGM has no current discount reading for deepseek-v4.1-flash — check back once the next window closes.
What buyers have actually paid on this model recently, across every provider and cache hits.
No requests were served on this model over the last 168 windows, so there is no measured token mix to price.
Requests are load balanced across every provider serving this model, each offering its own rate, so what you pay is a blend rather than a single number. These are the blended rates actually paid — priced at what a provider offered, capped at list.
No requests were served on this model over the last 168 windows, so there is no measured token mix to price.
Every provider currently serving this model, ranked by discount.
No provider is currently serving this model.
Measured on the requests SayGM routed to this model.
No requests were served on this model across the last 12 windows.
Fixed-seed suites SayGM runs against this model.
No benchmark run has been published for this model.
deepseek-v4.1-flash lists at $0.30 per million input tokens and $1.20 per million output tokens on SayGM.
Yes. deepseek-v4.1-flash is served on SayGM's OpenAI-compatible chat/completions surface, so an unmodified OpenAI SDK pointed at SayGM's base URL works with it.