Venice · live price
Uncensored API pricing
Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...
- Input
- $0.2
- per 1M tokens
- Output
- $0.9
- per 1M tokens
- Cached input
- –
- per 1M tokens
- Context
- 128K
- 8K max output
List price from OpenRouter, which passes Venice's price through without markup. Refreshed hourly. Released July 9, 2025. Model id cognitivecomputations/dolphin-mistral-24b-venice-edition.
What Uncensored costs in practice
1,000 chat messages
$0.65
1,000 input + 500 output tokens each
1,000 RAG questions
$2.05
8,000 input + 500 output tokens each
100 coding-agent steps
$0.78
30,000 input + 2,000 output tokens each, no caching
Uncensored price history
Uncensored price by host
- Venice$0.2 in · $0.9 out · 128K ctx · 100.0% uptime 24h
Where to buy Venice credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Uncensored cost calculator
Cheaper alternatives to Uncensored
| Model | Input / 1M | Output / 1M |
|---|---|---|
| OpenAI: GPT-6 Luna Pro | $0.1 | $0.5 |
| OpenAI: GPT-6 Luna | $0.1 | $0.5 |
| Qwen: Qwen3.8 Omni Flash | $0.15 | $0.47 |
| DeepSeek: DeepSeek V4.1 Flash | $0.03 | $0.5 |
| Qwen: Qwen3.8 Flash | $0.15 | $0.47 |
| Z.ai: GLM 5.3 Flash | $0.15 | $0.5 |