Wild West API

Run GLM 5.3 Xploded in Continue

Run glm-5.3-xploded in Continue through Wild West API: one key, a hard spend cap, and a model that answers instead of refusing.

Model idglm-5.3-xploded
Context window1,048,576 tokens (1M)
Tool callingSupported
Price per 1M tokens$2.50 input, $0.60 on a cache hit, $4.50 output
Intelligence45 on the Artificial Analysis Intelligence Index v4.3.2, for the base model
Web searchAdd :online to the id

Set it up

  1. Add a model block to ~/.continue/config.yaml with model: glm-5.3-xploded and the Wild West API apiBase.
  2. Put WILDWEST_API_KEY=sk-ww-... in ~/.continue/.env; the editor extensions cannot read your shell.
  3. Keep capabilities: [tool_use], or Continue will not let it act in agent mode.
name: Wild West API
version: 0.0.1
schema: v1
models:
  - name: GLM 5.3 Xploded (Wild West API)
    provider: openai
    model: glm-5.3-xploded
    apiBase: https://wildwestapi.com/v1
    apiKey: ${{ secrets.WILDWEST_API_KEY }}
    defaultCompletionOptions:
      contextLength: 1048576
      maxTokens: 32768
    capabilities:
      - tool_use
    roles:
      - chat
      - edit
      - apply
Continue with GLM 5.3 Xploded Continue only sends tools to models it believes can use them. The capabilities line is that belief; without it agent mode quietly falls back to chat.

What a session costs

For a 25-turn agent task (about 750K input tokens re-read across turns, 85% of them cache hits, 25K output), GLM 5.3 Xploded comes to about $0.78 at today's prices; the cache price does most of the work, since re-read context is billed at $0.60 rather than $2.50 per million.

FAQ

Can GLM 5.3 Xploded edit files and run commands in Continue?

Yes. Continue acts through tool calls and this model supports tool calling, so it can read, edit and run, not just chat.

What does a session cost?

About $0.78 for a 25-turn agent task (about 750K input tokens re-read across turns, 85% of them cache hits, 25K output), at $2.50 in, $4.50 out and $0.60 per million on cache hits. Put a spend cap on the key so a long run cannot overrun it.

Is it uncensored?

The full GLM 5.3, uncensored: frontier reasoning, a million tokens of context and tool calling, with no refusals. Text only: it does not take images.

The other model in Continue

Uncensored AI models on one key

OpenAI and Anthropic compatible, pay as you go. New to it? Start with uncensored AI, explained.