Open WebUI with SayGM
The self-hosted chat platform, running SayGM's GPT and open models.
Model
GPT-6 LunaBest value
InSign in to fill these snippets with a SayGM key for Open WebUI.
Setup for GPT and open models
Use this setup with the models listed under GPT, Open models and Private (TEE) below.
- 1Admin Panel → Settings → Connections → OpenAI API → Add Connectiontext
URL https://api.saygm.com/v1 Auth Bearer API Key <YOUR_SAYGM_KEY> API Type Chat Completions Advanced → Model IDs, add each: gpt-6-luna gpt-6-astra kimi-k3-tee - 2Click Verify Connection, then Savetext
Server connection verified
Switch models from the model selector at the top left of a chat, for example to gpt-6-astra. Choose Set as default there to start new chats on it.
First request
- 1In a new chattext
Pick gpt-6-luna in the model selector at the top left, then send: say gm
Troubleshooting
- Adding another model.
- Open the SayGM connection, add the model's id under Model IDs, and save. It then shows in the model selector.
- Verify Connection fails.
- The URL must be exactly the one shown, ending in /v1 with no trailing slash. Open WebUI checks it by listing models with your key, so a wrong key fails here too.
- A Claude or Gemini model shows up, and every chat with it fails.
- Open WebUI sends every model the OpenAI way, which SayGM serves only for GPT and open models. Keep Model IDs to the ids listed on this page: left empty, it lists every SayGM model.
- Environment variables such as OPENAI_API_BASE_URL are ignored.
- Open WebUI reads them only on first start, then keeps the connections saved in Admin Panel → Settings → Connections. Change the connection there.
- The same model shows up twice.
- Another connection lists the same id. Set a Prefix ID on the SayGM connection, such as saygm, to tell them apart.
Models you can use with Open WebUI
Open WebUI runs SayGM's GPT and open models. It sends every model the way it sends OpenAI's, which SayGM serves only for those, so Claude models don't work in it.
GPT
12 models
- GPT-6 Astra
gpt-6-astraBest qualityOpenAI's top GPT-6 model, for the hardest coding and reasoning work.In$10.00$8.15Out$50.00$40.75/1M - GPT-6 Luna
gpt-6-lunaBest valueLow-cost GPT-6 for fast coding loops and high-volume agent turns.In$0.10$0.08Out$0.50$0.41/1M - GPT-6 Sol
gpt-6-solMid-priced GPT-6 for day-to-day coding with strong reasoning.In$2.00$1.63Out$10.00$8.15/1M - GPT-5.6 Sol
gpt-5.6-solThe larger GPT-5.6 model, for demanding coding and analysis.In$4.00$3.26Out$20.00$16.30/1M - GPT-5.6 Terra
gpt-5.6-terraMid-sized GPT-5.6 balancing capability and cost.In$2.00$1.63Out$12.00$9.78/1M - GPT-5.6 Luna
gpt-5.6-lunaLow-cost GPT-5.6 for quick, high-volume requests.In$0.20$0.16Out$1.20$0.98/1M - GPT-5.5
gpt-5.5GPT-5.5 reasoning model for workflows tuned to it.In$5.00$4.08Out$30.00$24.45/1M - GPT-5.4
gpt-5.4GPT-5.4 for general coding and writing at a moderate price.In$2.50$2.04Out$15.00$12.23/1M - GPT-5.4 mini
gpt-5.4-miniSmall GPT-5.4 for fast, inexpensive everyday tasks.In$0.75$0.61Out$4.50$3.67/1M - GPT-5.4 nano
gpt-5.4-nanoThe smallest GPT-5.4, for classification and simple high-volume calls.In$0.20$0.16Out$1.25$1.02/1M - o3OpenAI's o3 reasoning model for multi-step math, science and code.In
$2.00$1.63Out$8.00$6.52/1M - o4-miniCompact o-series reasoning at a low price.In
$1.10$0.90Out$4.40$3.59/1M
Open models
11 models
- gpt-oss-20bOpenAI's small open-weight model for cheap, simple tasks.In $0.07Out
$0.30$0.28/1M - Kimi K3
kimi-k3Kimi K3, an open-weight reasoning model built for coding agents.In$3.00$1.02Out$15.00$5.09/1M - GLM-5.3
glm-5.3Z.ai's GLM-5.3 open-weight model for coding and agent tasks.In$1.40$0.81Out$4.40$2.54/1M - GLM-5.3-Flash
glm-5.3-flashThe fast, low-cost GLM-5.3 for quick coding help.In$0.15$0.13Out$0.50$0.45/1M - GLM-5.2
glm-5.2GLM-5.2 open weights for coding and reasoning.In$1.40$0.44Out$4.40$1.38/1M - DeepSeek-V4.1-Flash
deepseek-v4.1-flashDeepSeek's fast V4.1 model for low-cost coding and chat.In$0.30$0.09Out$1.20$0.36/1M - DeepSeek-V4-Flash-0731
deepseek-v4-flash-0731DeepSeek V4 Flash for inexpensive, quick coding help.In$0.44$0.13Out$1.32$0.39/1M - Qwen3.8-27B
qwen3.8-27bQwen3.8 27B, a compact open-weight model for everyday coding.In$0.50$0.05Out$3.00$0.32/1M - Qwen3.6-35B-A3B
qwen3.6-35b-a3bQwen3.6 35B mixture-of-experts, fast and cheap for light tasks.In$0.248$0.052Out$1.485$0.310/1M - MiMo-V2.6-Pro-UltraSpeed
mimo-v2.6-pro-ultraspeedXiaomi MiMo V2.6 Pro tuned for speed, with a 1M-token context.In$4.35$3.92Out$8.70$7.83/1M - Ornith 1.5 397B
ornith-1.5-397bOrnith 1.5 397B, a large open-weight reasoning model.In$1.40$0.70Out$4.40$2.20/1M
Private (TEE)
13 models
- Kimi-K3 (TEE)
kimi-k3-teePrivate (TEE)Kimi K3, an open-weight coding and agent model, run inside a TEE.In$3.00$1.03Out$15.00$5.17/1M - Kimi-K2.6 (TEE)
kimi-k2.6-teeKimi K2.6 open weights inside a TEE, for private agent work.In$0.50$0.17Out$2.85$0.98/1M - GLM-5.2 (TEE)
glm-5.2-teeGLM-5.2 run inside a TEE, for coding on private code.In$1.40$0.44Out$4.40$1.38/1M - GLM-5.1 (TEE)
glm-5.1-teeGLM-5.1 run inside a TEE, for private reasoning tasks.In$1.40$0.48Out$4.40$1.52/1M - DeepSeek-V4-Flash-0731 (TEE)
deepseek-v4-flash-0731-teeDeepSeek V4 Flash inside a TEE, for low-cost private work.In$0.44$0.22Out$1.32$0.66/1M - DeepSeek-V3.2 (TEE)
deepseek-v3.2-teeDeepSeek V3.2 inside a TEE, for private general-purpose tasks.In$1.00$0.48Out$1.00$0.48/1M - Qwen3.6-27B (TEE)
qwen3.6-27b-teeQwen3.6 27B inside a TEE, for private everyday tasks.In$0.30$0.15Out$2.00$0.98/1M - Qwen3.5-397B-A17B (TEE)
qwen3.5-397b-a17b-teeQwen3.5 397B, a large open-weight model run inside a TEE.In$0.45$0.16Out$3.00$1.03/1M - Qwen3-235B-A22B-Thinking-2507 (TEE)
qwen3-235b-a22b-thinking-2507-teeQwen3 235B thinking model inside a TEE, for private step-by-step reasoning.In$0.2989$0.1449Out$1.1957$0.5797/1M - Qwen3-32B (TEE)
qwen3-32b-teeQwen3 32B inside a TEE, for low-cost private requests.In$0.104$0.051Out$0.416$0.204/1M - Gemma 4 31B turbo (TEE)
gemma-4-31b-turbo-teeGoogle's Gemma 4 31B open weights inside a TEE, low cost and private.In$0.12$0.06Out$0.37$0.18/1M - Mistral Nemo (TEE)
mistral-nemo-instruct-2407-teeMistral Nemo inside a TEE, a very low-cost model for simple private tasks.In$0.0245$0.0120Out$0.0978$0.0479/1M - Nemotron 3 Nano Omni 30B (TEE)
nemotron-3-nano-omni-30b-teeNVIDIA Nemotron 3 Nano Omni inside a TEE, very low cost and private.In$0.0245$0.0110Out$0.0978$0.0439/1M
About Open WebUI
Open WebUI is an open-source AI chat platform you host yourself, on a laptop or a shared server. It gives a team a familiar chat interface with accounts, chat history, document search and tools, on top of the model providers you connect. Teams use it to give everyone one private place to chat with the models they choose.