Estimate AWS Bedrock costs for Claude, Llama 3, Titan, and other AI models. Calculate input/output token pricing for your workload before you deploy.
Amazon Bedrock is AWS's fully managed service for accessing foundation models from leading AI providers — including Anthropic (Claude), Meta (Llama), Amazon (Titan), Mistral, Cohere, and Stability AI. Bedrock pricing varies significantly by model, input/output token counts, and whether you use on-demand or provisioned throughput, making cost estimation essential before deploying AI workloads.
This calculator helps you estimate Bedrock costs based on your expected usage patterns, model selection, and throughput requirements — enabling informed decisions about model selection and deployment strategy.
| Pricing Model | How It Works | Best For |
|---|---|---|
| On-Demand | Pay per input/output token with no commitment | Development, testing, variable workloads |
| Batch Inference | Up to 50% discount for async processing | Large-volume offline processing |
| Provisioned Throughput | Reserved model units for guaranteed performance | Production workloads needing consistent latency |
| Model Customization | Training costs + storage + inference | Fine-tuned models for specific use cases |
| Factor | Impact on Cost |
|---|---|
| Model selection | Claude Opus vs Haiku can differ by 30-60x per token |
| Input vs output tokens | Output tokens are typically 3-5x more expensive than input |
| Context window usage | Longer prompts = more input tokens = higher cost |
| Response length | Longer outputs significantly increase per-request cost |
| Throughput needs | Provisioned throughput has a monthly minimum commitment |
| Region | Pricing varies by AWS region |