LLM API pricing comparison
Compare 2,994 LLM API models across 140 providers — free.
Check token price, context size, and cost estimates with the calculator. Compare models side-by-side with the compare tool. Browse chatbots, RAG, summarization, and coding agent use cases.
Catalog last generated: Aug 10, 2026. Check data sources before production decisions.
Database paths
Start from the evidence layer
Model database
Browse normalized pricing, token limits, capability flags, and source links by provider route.
Providers
Browse model pricing by provider. Compare model counts, price ranges, and review route-level pricing per platform.
Local LLMs
Explore open-weight models with VRAM requirements, parameter counts, and quantization support for local inference.
Exact model pricing
Jump from exact model and route pricing searches into current catalog rows and detail pages.
Benchmark explorer
Compare source-backed benchmark rows within one benchmark instead of mixing unrelated scores.
Data sources
Check how the catalog is generated, what needs provider verification, and where corrections belong.
How it works
Understand how the pricing comparison works, from data ingestion to cost calculation and model ranking.
Use-case guides
Pick a model by the job you need to do
Chatbot
Fast replies, low cost per message, and easy scaling for support or product chat.
RAG
Large context and stable token cost for answers built from your own docs.
Summarization
Low-cost reading of long files, articles, or notes with short output.
Coding agent
Compare models for long prompts, repeated turns, and higher output volume.
Embedding
Convert text to vectors for search, retrieval, and semantic similarity at low cost.
Agentic AI
Estimate multi-step agent costs with context accumulation, tool calls, and cache optimization.
Featured articles
Read the data before you compare
Cheapest LLM API chat models for a 500-token chatbot workload
Start with a 100k-message chatbot cost screen, then jump into model pages and compare links from current pricing data.
LLM API Provider Pricing Comparison
Compare OpenAI, Anthropic, Google, Mistral, and DeepSeek across cheapest chat, mid-range, and reasoning model pricing.
LLM Fine-Tuning Cost Comparison
Compare training and inference costs across OpenAI, Together AI, Fireworks AI, and AWS Bedrock. When fine-tuning saves money.
Model Routing Cascade
A four-tier Budget, Mid, Premium, and Reasoning model selection framework for routing queries to the lowest-cost suitable tier.
LLM API Cost Optimization Checklist
Seven proven strategies to reduce LLM API costs — from prompt engineering to model routing, caching, batch processing, and output length control.
Context window pricing guide
How context window size affects LLM API pricing. Compare models across 8k, 32k, 128k, and 1M context windows.
Prompt caching cost savings guide
How prompt caching reduces LLM API costs. Compare cache hit rates, savings percentages, and model-specific caching strategies.
Agentic AI cost estimation guide
Multi-step agent costs across budget, cascade, and premium strategies. Context accumulation, tool calls, reasoning multipliers, and cache optimization.
Price structure
Input vs output price
Linear axes show absolute per-token price differences for chat models. Use the calculator for workload-specific estimates.
Capability metadata
Capability coverage
Based on LiteLLM metadata flags, not independent feature verification.
Benchmarks
Collected benchmark snapshots
Benchmark coverage is intentionally explicit and source-linked.
Need exact costs for your workload?
Use the calculator for custom token counts, request volume, and cache assumptions.