Wild West API

Run GLM 5.3 Flash Xploded in Open WebUI

Run glm-5.3-flash-xploded in Open WebUI through Wild West API: one key, a hard spend cap, and a model that answers instead of refusing.

Model idglm-5.3-flash-xploded
Context window1,048,576 tokens (1M)
Tool callingSupported
Price per 1M tokens$0.40 input, $0.14 on a cache hit, $1.60 output
Intelligence42 on the Artificial Analysis Intelligence Index v4.3.2, for the base model
Web searchAdd :online to the id

Set it up

  1. Settings, Admin, Connections: add an OpenAI API connection.
  2. URL https://wildwestapi.com/v1 (keep the /v1) and your key.
  3. Leave Model IDs empty to auto-detect, or allowlist glm-5.3-flash-xploded and glm-5.3-flash-xploded:online.
OPENAI_API_BASE_URL=https://wildwestapi.com/v1
OPENAI_API_KEY=sk-ww-...

# Model IDs to allow (optional)
glm-5.3-flash-xploded
glm-5.3-flash-xploded:online
Open WebUI with GLM 5.3 Flash Xploded The glm-5.3-flash-xploded:online id searches the web before it answers, which turns Open WebUI into an uncensored answer engine with no plugin to install.

What a session costs

For a 10-message chat (about 40K input tokens, no cache hits, 8K output), GLM 5.3 Flash Xploded comes to about $0.029 at today's prices.

FAQ

Can GLM 5.3 Flash Xploded be used in Open WebUI?

Yes. Open WebUI does not depend on tool calling, so any chat model works; this one also brings a 1M context.

What does a session cost?

About $0.029 for a 10-message chat (about 40K input tokens, no cache hits, 8K output), at $0.40 in, $1.60 out and $0.14 per million on cache hits. Put a spend cap on the key so a long run cannot overrun it.

Is it uncensored?

A million tokens of context, tool calling and vision, with no refusals. Point a coding agent at it.

The other model in Open WebUI

Uncensored AI models on one key

OpenAI and Anthropic compatible, pay as you go. New to it? Start with uncensored AI, explained.