Use-case calculators
Start with the workload, then compare models.
Each guide uses a concrete token scenario, shows low-cost candidates from the current model database, and links into compare presets for the next decision.
Short turns
Chatbot
Estimate support or product chat costs from short input, short output, and repeated monthly messages.
Context-heavy answers
RAG
Check whether low-cost models can handle retrieved document context before you compare candidates.
Long inputs
Summarization
Compare input-heavy costs for reports, articles, notes, and other long-document workloads.
Repeated repo work
Coding agent
Price longer prompts, more context, larger outputs, and multi-step coding-agent iterations.
Vector index
Embedding
Compare input-only pricing for vector-index workloads: tokens per document and monthly indexing volume.
Need a custom token count?
Use the calculator after you pick a starting scenario, then adjust input, output, volume, and cache assumptions.
Pricing data from catalog last generated Jul 13, 2026. Verify before production decisions. Data sources.