//01  Vercel AI Gateway alternative

Frontier models,
at or below list.

Vercel AI Gateway passes through the provider’s list price with no markup. On SayGM, providers compete to serve each request, so Claude, GPT and Gemini bill at or below list, through a gateway sealed inside a Trusted Execution Environment.

Isometric rendering of SayGM routing chips, a track carrying a signal into the lit connector
Typical saving · 8 models15%below the same model on Vercel AI Gateway, at the rate its public model list publishes.
AvailableNative OpenAI, Anthropic and Gemini APIs
[02] The price

At or below list price,
by design.

Vercel charges what the provider charges. SayGM charges at or below the model maker’s published list price, at any volume and from your first request, because independent providers compete to serve each request and that competition sets the rate.

Funding your balance costs the amount you fund it with. You pay the rate of the provider that serves each request, never above list.

Vercel AI Gateway• Provider list price, no markup, as of September 2026Promos below list on some models. Card processing fees are paid by the buyer.
SayGMAt or below list · top-ups credited in fullCapped at the model maker’s list price.
[03] The prompt

A sealed gateway is something you can check.

Vercel publishes a clear policy: it deletes prompt and response content once a request completes. SayGM’s gateway runs inside an Intel TDX trusted execution environment, and the attestation quote proving it is independently verifiable, so your prompt stays out of reach of its operators and the host machine.

For Claude, GPT and Gemini, the provider serving the model still sees your prompt, exactly as it would if you called it directly, under its own API terms. Confidential models run sealed end to end.

Verifiable by anyone

Remote attestation lets you confirm for yourself exactly what code the gateway is running.

Sealed from the routing layer too

Most privacy claims stop at the model. This one covers the gateway itself.

Built for the audit trail

Pairs with signed logs of which model handled what, for teams that need to show their work.
[04] Calculator

What you could save.

Pick a model, set your monthly volume, see the delta.
Model
50M
15M
91%

Most input tokens in real traffic are cache reads, and both gateways price them well below a fresh one — so the share matters more to the bill than the totals do. The default is what SayGM measures across its own traffic.

SayGM$286.22
Vercel AI Gateway$327.86
You keep$41.64per month · 12% lower$499.66 over a yearstart saving
[05] Model by model

Price per million tokens.

USD per 1M tokens, uncached input / output. SayGM’s side is the best rate a provider offered, as of 52 minutes ago.Vercel’s is the per-model rate its public model list publishes, as of 27 minutes ago, with no fee added because Vercel charges none.
ModelSayGMVercel AI GatewayYou save
Claude Opus 5.5Anthropic$3.49 / $17.46$4.00 / $20.0012%
Claude Sonnet 5Anthropic$1.75 / $8.73$2.00 / $10.0012%
Claude Haiku 4.5Anthropic$0.87 / $4.37$1.00 / $5.0012%
GPT-6 SolOpenAI$1.63 / $8.15$2.00 / $10.0018%
GPT-5.6 TerraOpenAI$1.63 / $9.78$2.00 / $12.0018%
GPT-6 LunaOpenAI$0.08 / $0.41$0.10 / $0.5018%
Gemini 3.1 Pro PreviewGoogle$1.70 / $10.20$2.00 / $12.0015%
Gemini 3.5 FlashGoogle$1.28 / $7.65$1.50 / $9.0015%
browse the full catalogue
[06] Side by side

SayGM vs. Vercel AI Gateway, line by line.

DimensionSayGMvs. Vercel AI Gateway
Frontier token priceAt or below the model maker’s list priceThe provider’s list price with no markup, plus promotional prices on some models
Funding your balanceEvery top-up credited in fullNo platform fee; card processing fees are yours, and purchased credits expire after a year
Prompt privacyGateway sealed in an Intel TDX TEE, with an attestation quote anyone can verifyPrompts deleted once the request completes, under Vercel’s published policy
Confidential open modelsOpen-weights models served inside a TEE end to endOpt-in routing to providers that agree to zero data retention, on Pro and Enterprise
Free usagePay as you go from the first requestA monthly free credit on a subset of models, until you buy credits
Bring your own keyCapacity comes from SayGM’s own providersYour own provider keys, with no fee, on the paid tier
Model coverageClaude, GPT, Gemini and a curated set of open and confidential modelsHundreds of models, including image, video, speech and embeddings
Spend controlsUsage and a prepaid balance in the dashboardBudgets per team, project, key or member, request logs and trace export
RoutingBuild cascade models that try the next model when one fails, or fusion models that combine several. Every request is balanced across providers by latency, reliability and price, retried on another provider if an attempt fails, and kept on one provider through a conversation where it can, so prompt caching holdsRetries a failed request on the model’s other providers, with configurable provider ordering and model fallbacks
API surfaceEach model on its maker’s native API shape: OpenAI, Anthropic Messages or GeminiThe AI SDK plus OpenAI- and Anthropic-compatible APIs across its text models
Vercel details as of September 2026, from Vercel’s own docs. SayGM lists 56 models today. For image, video or audio models, per-project budgets, or your own provider keys, Vercel AI Gateway is the better tool. You can also run both: keep Vercel for breadth and send Claude, GPT and Gemini traffic to SayGM.
[07] Questions

Common questions about switching.

Does Vercel AI Gateway mark up token prices?

No. As of September 2026 Vercel charges the provider’s list price with no markup and no platform fee on tokens, on the free and paid tiers and with your own keys, as its pricing docs set out. It also runs promotional prices below list on some models and negotiates volume discounts. SayGM prices Claude, GPT and Gemini at or below the model maker’s list price, because providers compete to serve each request and that competition sets the rate.

Is SayGM cheaper than Vercel AI Gateway?

Across the frontier models this page compares, 15% for the middle one, against the rate Vercel’s public model list publishes. It varies by model, and the calculator above prices your exact volumes. While a Vercel promo runs, a single model can cost less there, and the table shows that as it is.

Does Vercel AI Gateway have a free tier?

Yes. As of September 2026 every Vercel team gets a monthly free credit that covers a subset of models at lower per-model rate limits, and it stops applying once you buy credits, per Vercel’s FAQ. For small experiments that credit can be worth more than any per-token saving. SayGM is pay as you go from a prepaid balance, from the first request.

Does Vercel AI Gateway have rate limits?

On the free tier, yes: lower limits per model. On the paid tier Vercel applies none of its own and the upstream provider’s limits still apply, as of September 2026 per its rate limit docs. SayGM sets limits per key, and you can ask for a higher one.

Can I bring my own provider keys?

On Vercel, yes. BYOK works on the paid tier with no fee, and a request that fails on your key is retried on Vercel’s credentials and charged to your credits. SayGM supplies capacity from its own providers. If you hold a negotiated rate with a lab that beats SayGM’s price on a model, keeping that model on your own key is the cheaper choice.

Does Vercel AI Gateway store my prompts?

Vercel says it does not train on prompts and deletes prompt and response content once the request completes, keeping request metadata, per its FAQ. Routing only to providers that agree to zero data retention is an opt-in on Pro and Enterprise. SayGM works at a different layer: the gateway runs inside an Intel TDX trusted execution environment with a verifiable attestation quote, so its operators and host machines cannot read your prompt. Claude, GPT and Gemini providers process prompts under their own API terms, and confidential models run sealed end to end.

Can I use SayGM from the Vercel AI SDK?

Yes for GPT models: the AI SDK’s OpenAI provider takes SayGM’s base URL and your SayGM key, and standard text requests work as before. Claude and Gemini are served on their own native APIs, so call them with a client for those. Vercel’s gateway supports the AI SDK directly along with OpenAI- and Anthropic-compatible APIs. You can keep it for breadth while sending frontier traffic to SayGM.

Where is Vercel AI Gateway the better choice?

For breadth, spend controls and platform fit. It carries hundreds of models across text, image, video and audio, sets budgets per team, project, key or member, and signs in with OIDC on Vercel deployments, while working from any other host too. SayGM is built for teams that want frontier models at or below list through a gateway whose privacy they can verify.