Run GLM 5.3 Flash Xploded in Open WebUI
Run glm-5.3-flash-xploded in Open WebUI through Wild West API: one key, a hard spend cap, and a model that answers instead of refusing.
| Model id | glm-5.3-flash-xploded |
|---|---|
| Context window | 1,048,576 tokens (1M) |
| Tool calling | Supported |
| Price per 1M tokens | $0.40 input, $0.14 on a cache hit, $1.60 output |
| Intelligence | 42 on the Artificial Analysis Intelligence Index v4.3.2, for the base model |
| Web search | Add :online to the id |
Set it up
- Settings, Admin, Connections: add an OpenAI API connection.
- URL https://wildwestapi.com/v1 (keep the /v1) and your key.
- Leave Model IDs empty to auto-detect, or allowlist glm-5.3-flash-xploded and glm-5.3-flash-xploded:online.
OPENAI_API_BASE_URL=https://wildwestapi.com/v1
OPENAI_API_KEY=sk-ww-...
# Model IDs to allow (optional)
glm-5.3-flash-xploded
glm-5.3-flash-xploded:online
What a session costs
For a 10-message chat (about 40K input tokens, no cache hits, 8K output), GLM 5.3 Flash Xploded comes to about $0.029 at today's prices.
FAQ
Can GLM 5.3 Flash Xploded be used in Open WebUI?
Yes. Open WebUI does not depend on tool calling, so any chat model works; this one also brings a 1M context.
What does a session cost?
About $0.029 for a 10-message chat (about 40K input tokens, no cache hits, 8K output), at $0.40 in, $1.60 out and $0.14 per million on cache hits. Put a spend cap on the key so a long run cannot overrun it.
Is it uncensored?
A million tokens of context, tool calling and vision, with no refusals. Point a coding agent at it.