AiCostCompare

Claude Sonnet 5 vs Inkling (batch)

Side-by-side API pricing and Artificial Analysis performance for Claude Sonnet 5 (Anthropic) and Inkling (batch) (Thinkingmachines). Green cells mark the better value in each row.

Prices updated Sep 17, 2026, 3:29 PM UTC · refreshed hourly

Claude Sonnet 5

Anthropic

Inkling (batch)

Thinkingmachines

Pricing
Input $/1M$2.00$1.00
Output $/1M$10.00$4.05
Blended $/1M (3:1)$4.00$1.76
RAG example (30K in / 2K out)$0.08$0.0381
Quality & speed
Intelligence Index38.425.5
Coding Index71.552.1
Agentic Index44.324.3
Output speed (tok/s)81
Time to first token102.09s
Specs
Context window1M524K
Input modalitiestext, image, filetext, image, audio

FAQ

Which is cheaper, Claude Sonnet 5 or Inkling (batch)?

Inkling (batch) has the lower blended API price at $1.76 per 1M tokens (3:1 input:output mix), versus $4.00 for the other model. Prices are live from OpenRouter and refresh about hourly.

Which is better for RAG workloads?

For a typical RAG request (30K input / 2K output tokens), Inkling (batch) costs about $0.0381 per request versus $0.08. Also compare context windows: Claude Sonnet 5 offers 1M and Inkling (batch) offers 524K.

Which scores higher on Artificial Analysis benchmarks?

Claude Sonnet 5 leads on the Artificial Analysis Intelligence Index (38.4 vs 25.5). Check coding and agentic indices on this page for workload-specific tradeoffs.