Uncensored pricing & benchmarks
Released Jul 9, 2025 · served by 1 provider
What Uncensored is
Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...
- Model ID:
cognitivecomputations/dolphin-mistral-24b-venice-edition - Modality: text->text
- Knowledge cutoff: 2024-04-30
- Tool calling: not supported · Structured output: not supported
- Weights:
cognitivecomputations/Dolphin-Mistral-24B-Venice-Editionon Hugging Face
Providers and prices for Uncensored
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| Venicecheapest | $0.200 | $0.900 | 128K | fp16 | 100.0% | — |
FAQ
How much does the Uncensored API cost?
$0.200 per 1M input tokens and $0.900 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.47.
Is Uncensored free?
No. It is a paid model starting at $0.200 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for Uncensored?
Venice at $0.200 input / $0.900 output per 1M tokens (fp16 quantization).
Can I self-host Uncensored?
Yes — weights are published as cognitivecomputations/Dolphin-Mistral-24B-Venice-Edition on Hugging Face, so you can run it on your own hardware.