Cogito v2.1 671B
by Deepcogito · deepcogito/cogito-v2.1-671b · #198 cheapest of 325 paid models
Prices updated Jul 22, 2026, 2:50 AM UTC · refreshed hourly
Cogito v2.1 671B MoE represents one of the strongest open models globally, matching performance of frontier closed and open models. This model is trained using self play with reinforcement learning...
Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5GPT-5.6 Sol
Pricing
Input
$1.25
per 1M tokens
Output
$1.25
per 1M tokens
Blended (3:1)
$1.25
3 input : 1 output
Cache read
—
per 1M cached tokens
Cache write
—
per 1M tokens
Cache & batch economics
Effective prices for Cogito v2.1 671B (no cache-read rate published — hits billed as normal input).
| Cache hit rate | Effective blended $/1M | Example request* | vs no cache |
|---|---|---|---|
| 0% | $1.25 | $0.1288 | — |
| 50% | $1.25 | $0.1288 | −0% |
| 90% | $1.25 | $0.1288 | −0% |
*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates (available for OpenAI / Anthropic / Google — not flagged for this creator). Confirm on the vendor’s pricing page.
What a request costs
| Workload | Input tokens | Output tokens | Cost / request | Cost / 1K requests |
|---|---|---|---|---|
| Short chat message | 500 | 300 | $0.001 | $1.00 |
| Document summary | 8,000 | 1,000 | $0.0112 | $11.25 |
| Codebase question (RAG) | 30,000 | 2,000 | $0.04 | $40.00 |
| Long-context analysis | 150,000 | 5,000 | $0.1938 | $193.75 |
Specs
Context window
128K
tokens
Max output
—
tokens
Input modalities
text
Tokenizer
Other