AiCostCompare

Uncensored

by Cognitivecomputations · cognitivecomputations/dolphin-mistral-24b-venice-edition · #98 cheapest of 325 paid models

Prices updated Jul 22, 2026, 2:26 AM UTC · refreshed hourly

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5GPT-5.6 Sol

Pricing

Input

$0.20

per 1M tokens

Output

$0.90

per 1M tokens

Blended (3:1)

$0.375

3 input : 1 output

Cache read

per 1M cached tokens

Cache write

per 1M tokens

Cache & batch economics

Effective prices for Uncensored (no cache-read rate published — hits billed as normal input).

Cache hit rateEffective blended $/1MExample request*vs no cache
0%$0.375$0.0213
50%$0.375$0.0213−0%
90%$0.375$0.0213−0%

*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates (available for OpenAI / Anthropic / Google — not flagged for this creator). Confirm on the vendor’s pricing page.

What a request costs

WorkloadInput tokensOutput tokensCost / requestCost / 1K requests
Short chat message500300$0.00037$0.37
Document summary8,0001,000$0.0025$2.50
Codebase question (RAG)30,0002,000$0.0078$7.80
Long-context analysis150,0005,000$0.0345$34.50

Specs

Context window

128K

tokens

Max output

8K

tokens

Input modalities

text

Tokenizer

Other