Wild West API

Leaving Baseten? Uncensored AI models instead

Baseten is a fast host for mainstream open models, per token or on dedicated GPUs. The alternatives below are the hosts it is usually compared with, and Wild West API for the case Baseten's per-token list does not cover: uncensored builds, per token, with a hard cap on every key.

What Baseten is, and where it stops

Baseten sells per-token Model APIs over a short list of mainstream open models (GLM-5.3 at $1.40 in and $4.40 out per million tokens, GLM-5.3-Flash, DeepSeek V4.1 Flash, Kimi K3, Nemotron 3 Ultra, GPT OSS 120B), and dedicated deployments billed by the minute for anything else (checked 2 October 2026). The per-token models are the standard aligned releases, so an uncensored build means a dedicated deployment you run yourself.

Other options people compare with Baseten

OptionWhat it is
Fireworks AIFast serverless inference for mainstream open models, plus dedicated deployments and fine-tuning.
Together AILarge catalogue of mainstream open-weight models, per token on serverless or on dedicated endpoints, plus fine-tuning.
DeepInfraOne of the cheaper per-token hosts for open-weight models, with a wide catalogue that changes as models are added and retired.
CerebrasVery fast inference on wafer-scale hardware, over two shared models (GPT OSS 120B and Qwen 3.8 27B) plus dedicated capacity.
GroqVery fast inference on its own LPU hardware, over a short list of models compiled for it.
OpenRouterA router over many providers: one key and one bill for hundreds of models, each served under its provider's policy.
Worth being straight about Nothing here is a jailbreak, and none of it makes an illegal act legal. A loosely aligned model answers more questions; it does not change what you are allowed to do with the answer, and it is frequently more confidently wrong than the model you left.

What Wild West API carries instead

Wild West API sells 5 uncensored models on one OpenAI-compatible endpoint, every one able to call tools. Ordered for this particular job rather than by a single ranking: Outlaw 1 Xploded leads because it is the cheapest of the set on output at $1.00 per million tokens.

ModelContextInOutTools
Outlaw 1 Xploded
outlaw-1-xploded
1.05M $0.300 $1.00 Yes
GLM 5.3 Flash Xploded
glm-5.3-flash-xploded
1.05M $0.400 $1.60 Yes
Qwen3.8 27B Xploded
qwen3.8-27b-xploded
524k $0.300 $2.40 Yes
MiMo V2.6 Flash Xploded
mimo-v2.6-flash-xploded
1.05M $1.00 $3.00 Yes

Switching takes one base URL

Anything that speaks the OpenAI chat completions format works unchanged. Point it here, use a Wild West API key, and name the model from the table.

curl https://wildwestapi.com/v1/chat/completions \
  -H "Authorization: Bearer sk-ww-..." \
  -H "Content-Type: application/json" \
  -d '{
    "model": "outlaw-1-xploded",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

FAQ

Is Wild West API actually uncensored, or does it just say so?

Wild West API does not moderate requests. It prices the call, holds the money against your key cap and passes the request to the model. What comes back is whatever that model does: an uncensored build refuses far less than a frontier model but can still decline, and only an abliterated build has the refusal direction removed outright. Neither tag is a promise that any given answer will arrive.

What does it cost compared with Baseten?

You pay per token at the rates in the table above, with no subscription. $0.300 in and $1.00 out per million tokens for Outlaw 1 Xploded.

Can I cap what a key is allowed to spend?

Yes, and the cap is enforced before the request goes upstream rather than reconciled afterwards. The reply cannot physically cost more than the hold placed before it was sent.

Will it work with my existing client?

If the client takes an OpenAI-compatible base URL, yes: point it at https://wildwestapi.com/v1 with a Wild West API key. Claude Code and the Anthropic SDKs use the Anthropic format, which is served too.

Related

Uncensored AI models on one key

OpenAI and Anthropic compatible, pay as you go. New to it? Start with uncensored AI, explained.