gpt-4o-2024-08-06
Azure · Chat model
gpt-4o-2024-08-06 is listed here as a chat model from Azure. This page shows simple API pricing, token limits, and capability flags so you can compare it with similar options.
Provider and model identifiers are kept in their original form for accuracy.
azure-azure-gpt-4o-2024-08-06
Catalog generated: Aug 10, 2026
Quick read
Best for
Use this page when you need a fast view of cost, context size, and supported features before testing the model in your own workload.
Things to verify
Always check the provider page for discounts, cache pricing, region rules, and any model limits that may not appear in public metadata.
Pricing
| Item | Price |
|---|---|
| Input | $2.5000 / 1M tokens |
| Output | $10.0000 / 1M tokens |
| Cached input | $1.2500 / 1M tokens |
| Cache write | $2.5000 / 1M tokens |
| Embedding | $2.5000 / 1M tokens |
Limits
Capabilities
| Capability | Supported |
|---|---|
| Vision | Supported |
| Function calling | Supported |
| Parallel function calling | Supported |
| Tool choice | Supported |
| Prompt caching | Supported |
| Reasoning | - |
| Response schema | Supported |
| System messages | - |
| Audio input | - |
| Audio output | - |
| Web search | - |
| PDF input | - |
| Video input | - |
| Native streaming | - |
| Computer use | - |
| Assistant prefill | - |
| Structured output | - |
| Output config | - |
| URL context | - |
Benchmarks
Most benchmark rows are attached to the base model family rather than this provider route. Open benchmark explorer
| Benchmark | Score | Metric | Scope | Checked | Source |
|---|---|---|---|---|---|
| SWE-bench Verified | 33.2% | accuracy (%) | Base model: gpt-4o (GPT-4o (2024-11-20)) | 2026-05-31 | Link |
| Aider Polyglot | 30.7% | accuracy (%) | Base model: gpt-4o (GPT-4o (2024-11-20)) | 2026-05-31 | Link |
| Aider Polyglot | 18.2% | accuracy (%) | Base model: gpt-4o (GPT-4o (2024-11-20)) | 2026-05-31 | Link |
| IFEval | 81.0% | accuracy (%) | Base model: gpt-4o (GPT-4o (2024-11-20)) | 2026-05-31 | Link |
| Design Arena - 3D | 958 Elo | Elo | Base model: GPT-4o (openai/gpt-4o) | 2026-05-31 | Link |
| Design Arena - Code Categories | 918 Elo | Elo | Base model: GPT-4o (openai/gpt-4o) | 2026-05-31 | Link |
| Design Arena - Game Development | 983 Elo | Elo | Base model: GPT-4o (openai/gpt-4o) | 2026-05-31 | Link |
| Design Arena - Website | 882 Elo | Elo | Base model: GPT-4o (openai/gpt-4o) | 2026-05-31 | Link |
Related articles
Articles relevant to this model's provider, capabilities, and use cases.
Model Routing Cascade
A four-tier Budget, Mid, Premium, and Reasoning model selection framework for routing queries to the lowest-cost suitable tier.
Hidden Costs of LLM APIs
Fine-tuning, rate limits, latency tradeoffs, evaluation, integration, and vendor risk — the LLM API costs that don't show up in per-token comparisons.
Output vs Input Pricing Multiplier
Output tokens cost 3.6× more than input on average across 2,076 chat models. Provider breakdowns, reasoning vs non-reasoning, and real workload estimates.
Cache and Batch Pricing Guide
How cached input and batch pricing change cost estimates across providers.
Sources
| Source links | |
| Pricing data | LiteLLM model cost map |
| Synced at | 2026-05-28 |
| Catalog generated | 2026-08-10T21:27:29.020Z |