Exact model pricing
Find pricing pages for exact model and route IDs.
Use this page when you have an exact model ID, provider route, or preview model name and need the matching pricing row in the database.
exact model and provider route pricing
Exact route shortlist
Exact model and provider-route pages for quick access. Rows stay route-specific, and missing provider source links remain clearly marked.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| MAI-DS-R1 azure_ai/azure_ai/MAI-DS-R1 | azure_ai | Chat | $1.3500 / 1M tokens | $5.4000 / 1M tokens | 8.2K | Source URL |
| amazon.nova-pro-v1:0 bedrock/bedrock/us-gov-west-1/amazon.nova-pro-v1:0 | Bedrock | Chat | $0.9600 / 1M tokens | $3.8400 / 1M tokens | 10.0K | Source URL |
| qwen3-coder-480b-a35b-instruct novita/novita/qwen/qwen3-coder-480b-a35b-instruct | novita | Chat | $0.3800 / 1M tokens | $1.5500 / 1M tokens | 65.5K | Source URL |
Scope
This finder matches stored pricing rows, not launch dates.
A row appears here only when the database already contains a priced model whose id, name, provider route, or LiteLLM key matches the listed model family. The page does not infer launch dates, lifecycle status, model quality, or provider availability from the model name.
OpenRouter rows are OpenRouter route prices. Provider routing and BYOK settings can change which upstream provider serves a request and whether upstream usage is paid through OpenRouter credits or a provider account, so compare those rows separately from direct-provider API prices.
gemini live pricing, gemini-3.1-flash-live-preview pricing
Gemini live pricing
Gemini live and live-preview pricing rows matched from model IDs, route IDs, and LiteLLM keys in the current database.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| gemini-live-2.5-flash-preview-native-audio-09-2025 vertex_ai-language-models/gemini-live-2.5-flash-preview-native-audio-09-2025 | vertex_ai-language-models | Realtime | $0.3000 / 1M tokens | $2.0000 / 1M tokens | 65.5K | Source URL |
| gemini-live-2.5-flash-preview-native-audio-09-2025 gemini/gemini/gemini-live-2.5-flash-preview-native-audio-09-2025 | Realtime | $0.3000 / 1M tokens | $2.0000 / 1M tokens | 65.5K | Source URL | |
| gemini-3.1-flash-live-preview gemini/gemini-3.1-flash-live-preview | Chat | $0.7500 / 1M tokens | $4.5000 / 1M tokens | 65.5K | Source URL | |
| gemini-3.1-flash-live-preview gemini/gemini/gemini-3.1-flash-live-preview | Chat | $0.7500 / 1M tokens | $4.5000 / 1M tokens | 65.5K | Source URL |
gpt realtime pricing, gpt-4o realtime preview pricing
GPT realtime pricing
Realtime and realtime-preview GPT routes currently present in the catalog. Azure route variants stay separate because their prices can differ.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| gpt-realtime-mini openai/gpt-realtime-mini | OpenAI | Realtime | $0.6000 / 1M tokens | $2.4000 / 1M tokens | 32.0K | Source URL |
| gpt-4o-mini-realtime-preview openai/gpt-4o-mini-realtime-preview | OpenAI | Realtime | $0.6000 / 1M tokens | $2.4000 / 1M tokens | 4.1K | Source pending |
| gpt-4o-mini-realtime-preview-2024-12-17 azure/azure/gpt-4o-mini-realtime-preview-2024-12-17 | Azure | Realtime | $0.6000 / 1M tokens | $2.4000 / 1M tokens | 4.1K | Source pending |
| gpt-4o-mini-realtime-preview-2024-12-17 openai/gpt-4o-mini-realtime-preview-2024-12-17 | OpenAI | Realtime | $0.6000 / 1M tokens | $2.4000 / 1M tokens | 4.1K | Source pending |
| gpt-realtime-2.1-mini openai/gpt-realtime-2.1-mini | OpenAI | Realtime | $0.6000 / 1M tokens | $2.4000 / 1M tokens | 4.1K | Source pending |
| gpt-realtime-mini-2025-10-06 azure/azure/gpt-realtime-mini-2025-10-06 | Azure | Realtime | $0.6000 / 1M tokens | $2.4000 / 1M tokens | 4.1K | Source pending |
| gpt-realtime-mini-2025-10-06 openai/gpt-realtime-mini-2025-10-06 | OpenAI | Realtime | $0.6000 / 1M tokens | $2.4000 / 1M tokens | 4.1K | Source pending |
| gpt-realtime-mini-2025-12-15 openai/gpt-realtime-mini-2025-12-15 | OpenAI | Realtime | $0.6000 / 1M tokens | $2.4000 / 1M tokens | 4.1K | Source pending |
us.anthropic.claude-3-5-sonnet-20240620-v1:0 pricing
Bedrock Claude 3.5 Sonnet exact ID
Exact Bedrock route ids matching the searched Claude Sonnet snapshot. Region and invoke routes are kept as separate detail pages.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| anthropic.claude-3-5-sonnet-20240620-v1:0 bedrock/anthropic.claude-3-5-sonnet-20240620-v1:0 | Bedrock | Chat | $3.0000 / 1M tokens | $15.0000 / 1M tokens | 4.1K | Source pending |
| anthropic.claude-3-5-sonnet-20240620-v1:0 bedrock/bedrock/invoke/anthropic.claude-3-5-sonnet-20240620-v1:0 | Bedrock | Chat | $3.0000 / 1M tokens | $15.0000 / 1M tokens | 4.1K | Source pending |
| apac.anthropic.claude-3-5-sonnet-20240620-v1:0 bedrock/apac.anthropic.claude-3-5-sonnet-20240620-v1:0 | Bedrock | Chat | $3.0000 / 1M tokens | $15.0000 / 1M tokens | 4.1K | Source pending |
| eu.anthropic.claude-3-5-sonnet-20240620-v1:0 bedrock/eu.anthropic.claude-3-5-sonnet-20240620-v1:0 | Bedrock | Chat | $3.0000 / 1M tokens | $15.0000 / 1M tokens | 4.1K | Source pending |
| us.anthropic.claude-3-5-sonnet-20240620-v1:0 bedrock/us.anthropic.claude-3-5-sonnet-20240620-v1:0 | Bedrock | Chat | $3.0000 / 1M tokens | $15.0000 / 1M tokens | 4.1K | Source pending |
| anthropic.claude-3-5-sonnet-20240620-v1:0 bedrock/bedrock/us-gov-east-1/anthropic.claude-3-5-sonnet-20240620-v1:0 | Bedrock | Chat | $3.6000 / 1M tokens | $18.0000 / 1M tokens | 8.2K | Source pending |
| anthropic.claude-3-5-sonnet-20240620-v1:0 bedrock/bedrock/us-gov-west-1/anthropic.claude-3-5-sonnet-20240620-v1:0 | Bedrock | Chat | $3.6000 / 1M tokens | $18.0000 / 1M tokens | 8.2K | Source pending |
qwen3 coder pricing, qwen3-coder-480b-a35b-instruct pricing
Qwen3 Coder pricing
Qwen3 Coder rows that already have pricing in the catalog, including provider and regional routes where present.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| qwen3-coder openrouter/openrouter/qwen/qwen3-coder | OpenRouter | Chat | $0.2200 / 1M tokens | $0.9500 / 1M tokens | 262.1K | Source URL |
| qwen3-coder-30b-a3b pinstripes/pinstripes/ps/qwen3-coder-30b-a3b | Pinstripes | Chat | $0.3000 / 1M tokens | $0.6000 / 1M tokens | 131.1K | Source URL |
| qwen3-coder-480b-a35b-instruct novita/novita/qwen/qwen3-coder-480b-a35b-instruct | novita | Chat | $0.3800 / 1M tokens | $1.5500 / 1M tokens | 65.5K | Source URL |
| Qwen3-Coder-480B-A35B-Instruct-FP8 tensormesh/tensormesh/Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8 | Tensormesh | Chat | $0.4500 / 1M tokens | $1.8000 / 1M tokens | N/A | Source URL |
| qwen.qwen3-coder-next bedrock/bedrock/us-east-1/qwen.qwen3-coder-next | Bedrock | Chat | $0.5000 / 1M tokens | $1.2000 / 1M tokens | 8.2K | Source URL |
| qwen.qwen3-coder-next bedrock/bedrock/us-east-2/qwen.qwen3-coder-next | Bedrock | Chat | $0.5000 / 1M tokens | $1.2000 / 1M tokens | 8.2K | Source URL |
| qwen.qwen3-coder-next bedrock/bedrock/us-west-2/qwen.qwen3-coder-next | Bedrock | Chat | $0.5000 / 1M tokens | $1.2000 / 1M tokens | 8.2K | Source URL |
| qwen.qwen3-coder-next bedrock_converse/qwen.qwen3-coder-next | bedrock_converse | Chat | $0.5000 / 1M tokens | $1.2000 / 1M tokens | 8.2K | Source URL |
openrouter gemini pricing, openrouter gemini 3.1 pricing
OpenRouter Gemini pricing
OpenRouter Gemini rows with current prices. These remain provider-route pages, not Google direct API rows.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| gemini-3.1-flash-lite openrouter/openrouter/google/gemini-3.1-flash-lite | OpenRouter | Chat | $0.2500 / 1M tokens | $1.5000 / 1M tokens | 65.5K | Source URL |
| gemini-3.1-flash-lite-preview openrouter/openrouter/google/gemini-3.1-flash-lite-preview | OpenRouter | Chat | $0.2500 / 1M tokens | $1.5000 / 1M tokens | 65.5K | Source URL |
| gemini-3-flash-preview openrouter/openrouter/google/gemini-3-flash-preview | OpenRouter | Chat | $0.5000 / 1M tokens | $3.0000 / 1M tokens | 65.5K | Source URL |
| gemini-3.1-pro-preview openrouter/openrouter/google/gemini-3.1-pro-preview | OpenRouter | Chat | $2.0000 / 1M tokens | $12.0000 / 1M tokens | 65.5K | Source URL |
| gemini-2.0-flash-001 openrouter/openrouter/google/gemini-2.0-flash-001 | OpenRouter | Chat | $0.1000 / 1M tokens | $0.4000 / 1M tokens | 8.2K | Source pending |
| gemini-2.5-flash openrouter/openrouter/google/gemini-2.5-flash | OpenRouter | Chat | $0.3000 / 1M tokens | $2.5000 / 1M tokens | 8.2K | Source pending |
| gemini-2.5-pro openrouter/openrouter/google/gemini-2.5-pro | OpenRouter | Chat | $1.2500 / 1M tokens | $10.0000 / 1M tokens | 8.2K | Source pending |
| gemini-3-pro-preview openrouter/openrouter/google/gemini-3-pro-preview | OpenRouter | Chat | $2.0000 / 1M tokens | $12.0000 / 1M tokens | 65.5K | Source pending |
deepseek r1 pricing, deepseek-r1 cost
DeepSeek R1 pricing
DeepSeek R1 reasoning model routes with current prices across providers.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| DeepSeek-R1-Distill-Llama-8B nscale/nscale/deepseek-ai/DeepSeek-R1-Distill-Llama-8B | nscale | Chat | $0.0250 / 1M tokens | $0.0250 / 1M tokens | N/A | Source URL |
| DeepSeek-R1-Distill-Qwen-14B nscale/nscale/deepseek-ai/DeepSeek-R1-Distill-Qwen-14B | nscale | Chat | $0.0700 / 1M tokens | $0.0700 / 1M tokens | N/A | Source URL |
| DeepSeek-R1-Distill-Qwen-1.5B nscale/nscale/deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B | nscale | Chat | $0.0900 / 1M tokens | $0.0900 / 1M tokens | N/A | Source URL |
| DeepSeek-R1-Distill-Qwen-32B nscale/nscale/deepseek-ai/DeepSeek-R1-Distill-Qwen-32B | nscale | Chat | $0.1500 / 1M tokens | $0.1500 / 1M tokens | N/A | Source URL |
| DeepSeek-R1-Distill-Qwen-7B nscale/nscale/deepseek-ai/DeepSeek-R1-Distill-Qwen-7B | nscale | Chat | $0.2000 / 1M tokens | $0.2000 / 1M tokens | N/A | Source URL |
| DeepSeek-R1-Distill-Llama-70B nebius/nebius/deepseek-ai/DeepSeek-R1-Distill-Llama-70B | nebius | Chat | $0.2500 / 1M tokens | $0.7500 / 1M tokens | 128.0K | Source URL |
| DeepSeek-R1-Distill-Llama-70B nscale/nscale/deepseek-ai/DeepSeek-R1-Distill-Llama-70B | nscale | Chat | $0.3750 / 1M tokens | $0.3750 / 1M tokens | N/A | Source URL |
| DeepSeek-R1-0528-tput together_ai/together_ai/deepseek-ai/DeepSeek-R1-0528-tput | Together AI | Chat | $0.5500 / 1M tokens | $2.1900 / 1M tokens | N/A | Source URL |
qwen3 pricing, qwen3 api cost
Qwen3 pricing
Qwen3 family routes with current prices, including Coder and embedding variants.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| qwen3-235b-a22b-2507 openrouter/openrouter/qwen/qwen3-235b-a22b-2507 | OpenRouter | Chat | $0.0710 / 1M tokens | $0.1000 / 1M tokens | 262.1K | Source URL |
| Qwen3-14B nebius/nebius/Qwen/Qwen3-14B | nebius | Chat | $0.0800 / 1M tokens | $0.2400 / 1M tokens | 32.8K | Source URL |
| Qwen3-32B ovhcloud/ovhcloud/Qwen3-32B | ovhcloud | Chat | $0.0800 / 1M tokens | $0.2300 / 1M tokens | 32.0K | Source URL |
| Qwen3-4B nebius/nebius/Qwen/Qwen3-4B | nebius | Chat | $0.0800 / 1M tokens | $0.2400 / 1M tokens | 32.8K | Source URL |
| qwen3-30b-a3b pinstripes/pinstripes/ps/qwen3-30b-a3b | Pinstripes | Chat | $0.0900 / 1M tokens | $0.2000 / 1M tokens | 131.1K | Source URL |
| Qwen3-30B-A3B nebius/nebius/Qwen/Qwen3-30B-A3B | nebius | Chat | $0.1000 / 1M tokens | $0.3000 / 1M tokens | 32.8K | Source URL |
| Qwen3-32B nebius/nebius/Qwen/Qwen3-32B | nebius | Chat | $0.1000 / 1M tokens | $0.3000 / 1M tokens | 32.8K | Source URL |
| qwen3.5-flash-02-23 openrouter/openrouter/qwen/qwen3.5-flash-02-23 | OpenRouter | Chat | $0.1000 / 1M tokens | $0.4000 / 1M tokens | 65.5K | Source URL |
gpt-5 pricing, gpt-5 api cost
GPT-5 pricing
GPT-5 family routes including Nano, Mini, and standard variants across providers.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| databricks-gpt-5-nano databricks/databricks/databricks-gpt-5-nano | databricks | Chat | $0.0500 / 1M tokens | $0.4000 / 1M tokens | 128.0K | Source URL |
| openai.gpt-5-nano oci/oci/openai.gpt-5-nano | OCI | Chat | $0.0500 / 1M tokens | $0.4000 / 1M tokens | 128.0K | Source URL |
| gpt-5.4-nano azure_ai/azure_ai/gpt-5.4-nano | azure_ai | Chat | $0.2000 / 1M tokens | $1.2500 / 1M tokens | 128.0K | Source URL |
| gpt-5.4-nano-2026-03-17 azure_ai/azure_ai/gpt-5.4-nano-2026-03-17 | azure_ai | Chat | $0.2000 / 1M tokens | $1.2500 / 1M tokens | 128.0K | Source URL |
| databricks-gpt-5-mini databricks/databricks/databricks-gpt-5-mini | databricks | Chat | $0.2500 / 1M tokens | $2.0000 / 1M tokens | 128.0K | Source URL |
| openai.gpt-5-mini oci/oci/openai.gpt-5-mini | OCI | Chat | $0.2500 / 1M tokens | $2.0000 / 1M tokens | 128.0K | Source URL |
| gpt-5.4-mini azure_ai/azure_ai/gpt-5.4-mini | azure_ai | Chat | $0.7500 / 1M tokens | $4.5000 / 1M tokens | 128.0K | Source URL |
| gpt-5.4-mini-2026-03-17 azure_ai/azure_ai/gpt-5.4-mini-2026-03-17 | azure_ai | Chat | $0.7500 / 1M tokens | $4.5000 / 1M tokens | 128.0K | Source URL |
gemini 3 pricing, gemini 3.1 flash pricing
Gemini 3 pricing
Gemini 3 family routes with current prices, including Flash, Pro, and Lite variants.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| gemini-3.1-flash-image gemini/gemini/gemini-3.1-flash-image | Image | $0.2500 / 1M tokens | $1.5000 / 1M tokens | 32.8K | Source URL | |
| gemini-3.1-flash-image-preview gemini/gemini/gemini-3.1-flash-image-preview | Image | $0.2500 / 1M tokens | $1.5000 / 1M tokens | 32.8K | Source URL | |
| gemini-3.1-flash-lite vertex_ai-language-models/gemini-3.1-flash-lite | vertex_ai-language-models | Chat | $0.2500 / 1M tokens | $1.5000 / 1M tokens | 65.5K | Source URL |
| gemini-3.1-flash-lite gemini/gemini/gemini-3.1-flash-lite | Chat | $0.2500 / 1M tokens | $1.5000 / 1M tokens | 65.5K | Source URL | |
| gemini-3.1-flash-lite openrouter/openrouter/google/gemini-3.1-flash-lite | OpenRouter | Chat | $0.2500 / 1M tokens | $1.5000 / 1M tokens | 65.5K | Source URL |
| gemini-3.1-flash-lite vertex_ai-language-models/vertex_ai/gemini-3.1-flash-lite | vertex_ai-language-models | Chat | $0.2500 / 1M tokens | $1.5000 / 1M tokens | 65.5K | Source URL |
| gemini-3.1-flash-lite-preview vertex_ai-language-models/gemini-3.1-flash-lite-preview | vertex_ai-language-models | Chat | $0.2500 / 1M tokens | $1.5000 / 1M tokens | 65.5K | Source URL |
| gemini-3.1-flash-lite-preview gemini/gemini/gemini-3.1-flash-lite-preview | Chat | $0.2500 / 1M tokens | $1.5000 / 1M tokens | 65.5K | Source URL |
llama 4 pricing, meta llama 4 cost
Llama 4 pricing
Meta Llama 4 family routes with current prices across inference providers.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| Llama-4-Scout-17B-16E-Instruct nscale/nscale/meta-llama/Llama-4-Scout-17B-16E-Instruct | nscale | Chat | $0.0900 / 1M tokens | $0.2900 / 1M tokens | N/A | Source URL |
| Llama-3.3-Nemotron-Super-49B-v1 nebius/nebius/nvidia/Llama-3.3-Nemotron-Super-49B-v1 | nebius | Chat | $0.1000 / 1M tokens | $0.4000 / 1M tokens | 131.1K | Source URL |
| llama4-scout-instruct-basic fireworks_ai/fireworks_ai/accounts/fireworks/models/llama4-scout-instruct-basic | fireworks_ai | Chat | $0.1500 / 1M tokens | $0.6000 / 1M tokens | 131.1K | Source URL |
| Llama-4-Scout-17B-16E-Instruct azure_ai/azure_ai/Llama-4-Scout-17B-16E-Instruct | azure_ai | Chat | $0.2000 / 1M tokens | $0.7800 / 1M tokens | 16.4K | Source URL |
| llama4-maverick-instruct-basic fireworks_ai/fireworks_ai/accounts/fireworks/models/llama4-maverick-instruct-basic | fireworks_ai | Chat | $0.2200 / 1M tokens | $0.8800 / 1M tokens | 131.1K | Source URL |
| llama-4-scout-17b-128e-instruct-maas vertex_ai-llama_models/vertex_ai/meta/llama-4-scout-17b-128e-instruct-maas | vertex_ai-llama_models | Chat | $0.2500 / 1M tokens | $0.7000 / 1M tokens | 10.0M | Source URL |
| llama-4-scout-17b-16e-instruct-maas vertex_ai-llama_models/vertex_ai/meta/llama-4-scout-17b-16e-instruct-maas | vertex_ai-llama_models | Chat | $0.2500 / 1M tokens | $0.7000 / 1M tokens | 10.0M | Source URL |
| Llama-4-Maverick-17B-128E-Instruct-FP8 azure_ai/azure_ai/Llama-4-Maverick-17B-128E-Instruct-FP8 | azure_ai | Chat | $1.4100 / 1M tokens | $0.3500 / 1M tokens | 16.4K | Source URL |
claude 4 pricing, claude 4 sonnet cost, claude 4 opus pricing
Claude 4 pricing
Claude 4 family routes including Opus, Sonnet, and 4.5 variants across providers.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| anthropic.claude-haiku-4-5-20251001-v1:0 bedrock_converse/anthropic.claude-haiku-4-5-20251001-v1:0 | bedrock_converse | Chat | $1.0000 / 1M tokens | $5.0000 / 1M tokens | 64.0K | Source URL |
| anthropic.claude-haiku-4-5@20251001 bedrock_converse/anthropic.claude-haiku-4-5@20251001 | bedrock_converse | Chat | $1.0000 / 1M tokens | $5.0000 / 1M tokens | 64.0K | Source URL |
| claude-haiku-4-5 vertex_ai-anthropic_models/vertex_ai/claude-haiku-4-5 | vertex_ai-anthropic_models | Chat | $1.0000 / 1M tokens | $5.0000 / 1M tokens | 200.0K | Source URL |
| claude-haiku-4-5@20251001 vertex_ai-anthropic_models/vertex_ai/claude-haiku-4-5@20251001 | vertex_ai-anthropic_models | Chat | $1.0000 / 1M tokens | $5.0000 / 1M tokens | 200.0K | Source URL |
| global.anthropic.claude-haiku-4-5-20251001-v1:0 bedrock_converse/global.anthropic.claude-haiku-4-5-20251001-v1:0 | bedrock_converse | Chat | $1.0000 / 1M tokens | $5.0000 / 1M tokens | 64.0K | Source URL |
| databricks-claude-haiku-4-5 databricks/databricks/databricks-claude-haiku-4-5 | databricks | Chat | $1.0000 / 1M tokens | $5.0000 / 1M tokens | 64.0K | Source URL |
| apac.anthropic.claude-haiku-4-5-20251001-v1:0 bedrock_converse/apac.anthropic.claude-haiku-4-5-20251001-v1:0 | bedrock_converse | Chat | $1.1000 / 1M tokens | $5.5000 / 1M tokens | 64.0K | Source URL |
| eu.anthropic.claude-haiku-4-5-20251001-v1:0 bedrock_converse/eu.anthropic.claude-haiku-4-5-20251001-v1:0 | bedrock_converse | Chat | $1.1000 / 1M tokens | $5.5000 / 1M tokens | 64.0K | Source URL |
llama 3 pricing, meta llama 3 cost, llama 3.1 pricing
Llama 3 pricing
Meta Llama 3 family routes including 3.1, 3.2, and 3.3 variants across providers.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| Llama-Guard-3-8B nebius/nebius/meta-llama/Llama-Guard-3-8B | nebius | Chat | $0.0200 / 1M tokens | $0.0600 / 1M tokens | 128.0K | Source URL |
| Meta-Llama-3.1-8B-Instruct nebius/nebius/meta-llama/Meta-Llama-3.1-8B-Instruct | nebius | Chat | $0.0200 / 1M tokens | $0.0600 / 1M tokens | 128.0K | Source URL |
| Llama-3.1-8B-Instruct nscale/nscale/meta-llama/Llama-3.1-8B-Instruct | nscale | Chat | $0.0300 / 1M tokens | $0.0300 / 1M tokens | N/A | Source URL |
| Meta-Llama-3.2-1B-Instruct sambanova/sambanova/Meta-Llama-3.2-1B-Instruct | sambanova | Chat | $0.0400 / 1M tokens | $0.0800 / 1M tokens | 16.4K | Source URL |
| Meta-Llama-3.2-3B-Instruct sambanova/sambanova/Meta-Llama-3.2-3B-Instruct | sambanova | Chat | $0.0800 / 1M tokens | $0.1600 / 1M tokens | 4.1K | Source URL |
| Llama-3.1-8B-Instruct ovhcloud/ovhcloud/Llama-3.1-8B-Instruct | ovhcloud | Chat | $0.1000 / 1M tokens | $0.1000 / 1M tokens | 131.0K | Source URL |
| Llama-3.3-Nemotron-Super-49B-v1 nebius/nebius/nvidia/Llama-3.3-Nemotron-Super-49B-v1 | nebius | Chat | $0.1000 / 1M tokens | $0.4000 / 1M tokens | 131.1K | Source URL |
| llama-v3p1-8b-instruct fireworks_ai/fireworks_ai/accounts/fireworks/models/llama-v3p1-8b-instruct | fireworks_ai | Chat | $0.1000 / 1M tokens | $0.1000 / 1M tokens | 16.4K | Source URL |
mistral large pricing, mistral large api cost, codestral pricing
Mistral Large pricing
Mistral Large and Codestral routes with current prices across providers.
| Model route | Provider | Mode | Input | Output | Context | Source |
|---|---|---|---|---|---|---|
| mistral-large-2512 mistral/mistral/mistral-large-2512 | Mistral | Chat | $0.5000 / 1M tokens | $1.5000 / 1M tokens | 262.1K | Source URL |
| mistral-large-3 azure_ai/azure_ai/mistral-large-3 | azure_ai | Chat | $0.5000 / 1M tokens | $1.5000 / 1M tokens | 8.2K | Source URL |
| mistral-large-3 mistral/mistral/mistral-large-3 | Mistral | Chat | $0.5000 / 1M tokens | $1.5000 / 1M tokens | 262.1K | Source URL |
| mistral-large-latest mistral/mistral/mistral-large-latest | Mistral | Chat | $0.5000 / 1M tokens | $1.5000 / 1M tokens | 262.1K | Source URL |
| mistral-large-2407 azure_ai/azure_ai/mistral-large-2407 | azure_ai | Chat | $2.0000 / 1M tokens | $6.0000 / 1M tokens | 4.1K | Source URL |
| mistral-large-latest azure_ai/azure_ai/mistral-large-latest | azure_ai | Chat | $2.0000 / 1M tokens | $6.0000 / 1M tokens | 4.1K | Source URL |
| mistral-large-2512 openrouter/openrouter/mistralai/mistral-large-2512 | OpenRouter | Chat | $0.5000 / 1M tokens | $1.5000 / 1M tokens | 262.1K | Source pending |
| mistral.mistral-large-3-675b-instruct bedrock_converse/mistral.mistral-large-3-675b-instruct | bedrock_converse | Chat | $0.5000 / 1M tokens | $1.5000 / 1M tokens | 8.2K | Source pending |