AWS Bedrock API Pricing and Model Catalog: Use-Case Guide
AWS Bedrock offers access to multiple foundation models through a single API. Here's how to use it cost-effectively.
Pricing data sourced from our catalog. Check data sources for provenance and freshness.
What is AWS Bedrock?
AWS Bedrock is a fully managed service offering:
- Multi-model access: Access to AI21, Anthropic, Cohere, Meta, Mistral, Stability, and Amazon models
- Single API: Unified API for all models
- Enterprise security: AWS IAM, VPC, KMS integration
- Fine-tuning: Model customization capabilities
AWS Bedrock model catalog
Chat models
| Model | Input | Output | Context |
|---|---|---|---|
| meta.llama3-2-1b-instruct-v1:0 | $0.1000 | $0.1000 | 4K |
| us.meta.llama3-2-1b-instruct-v1:0 | $0.1000 | $0.1000 | 4K |
| eu.meta.llama3-2-1b-instruct-v1:0 | $0.1300 | $0.1300 | 4K |
| meta.llama3-2-3b-instruct-v1:0 | $0.1500 | $0.1500 | 4K |
| us.meta.llama3-2-3b-instruct-v1:0 | $0.1500 | $0.1500 | 4K |
Cost estimation example
Let's estimate costs for a typical AWS Bedrock chat workload:
AWS Bedrock strengths
- Multi-model access: Access to multiple providers through a single API
- Enterprise security: AWS IAM, VPC, KMS integration
- Fine-tuning: Model customization capabilities
- Compliance: SOC 2, HIPAA, GDPR, and more
- AWS integration: Native integration with AWS services
When to use AWS Bedrock
- Multi-model workloads: When you need to switch between models
- Enterprise compliance: When you need AWS security and compliance
- Fine-tuning: When you need model customization
- AWS ecosystem: When you're already using AWS services
Cost optimization tips
- Use Provisioned Throughput: For predictable workloads
- Batch processing: Process multiple requests together
- Caching: Cache responses for repeated queries
- Right-size your model: Match model size to task complexity
- Monitor usage: Track token usage to optimize costs
Compare AWS Bedrock models
Ready to compare AWS Bedrock models side by side? Use our tools:
Related guides
Azure OpenAI API pricing and model catalog
A use-case guide to Azure OpenAI's API model catalog and pricing.
Cross-provider pricing comparison
How pricing compares across OpenAI, Anthropic, Google, Mistral, and DeepSeek.
Hidden costs of LLM APIs
Rate limits, latency, evaluation overhead, and vendor risk beyond per-token pricing.
Local LLM vs cloud API cost comparison
When self-hosting saves money and when it doesn't.
Frequently asked questions
What is AWS Bedrock's cheapest model?
AWS Bedrock's Titan Text Lite and Cohere Command R models are typically the most cost-effective.
How does AWS Bedrock compare to direct provider APIs?
AWS Bedrock offers a unified API for multiple providers but may have higher per-token costs than direct provider APIs.
Can I fine-tune models on AWS Bedrock?
Yes, AWS Bedrock supports fine-tuning for select models including Titan, Llama, and Cohere.