Uncensored AI for content moderation datasets
The classes your moderators see, not the classes a public API will emit
Moderation models trained on what a public generator will say are trained on the wrong distribution. Production is the content that generator refused to write.
Why the mainstream APIs fail here
You cannot close the gap between eval and production if the eval was filtered. The model under-detects the worst class because it never saw it in training.
How Wild West API is used for this
Generate the class under a T&S-only key. Store and label in your environment. Train there. Wild West API's log is the bill.
A working loop
- Sample the real policy classes.
- Cap the generator so a sweep is a known cost.
Details are on the privacy page. Per-model prices are on /models and the pricing model is on /pricing.