Amazon Bedrock pricing
-
Model Pricing
-
Knowledge Bases
-
Guardrails
-
Model Evaluation
-
Data Automation
-
Intelligent Prompt Routing
-
Prompt Optimization
-
Web Search
-
Model Pricing
-
Model Pricing
Pricing is dependent on the modality, provider, and model. Please select the model provider to see detailed pricing.
Amazon Bedrock supports a variety of tiers including Standard, Flex, Priority, and Reserved tiers. Click to learn more about service tiers.
Amazon Bedrock offers select foundation models (FMs) from leading AI providers like Anthropic, Meta, Mistral AI, and Amazon for batch inference at a 50% lower price compared to on-demand inference pricing. To learn more about Batch, click here. Please refer to model list here.
-
AI21 Labs
-
Amazon
-
Anthropic
-
Cohere
-
DeepSeek
-
Google
-
Luma AI
-
Meta
-
MiniMax AI
-
Mistral AI
-
Moonshot AI
-
NVIDIA
-
OpenAI
-
Qwen
-
Stability AI
-
TwelveLabs
-
Writer
-
xAI
-
Z AI
-
Custom Model Import
-
AI21 Labs
-
AI21 Labs
On-Demand pricing
-
Amazon
-
-
Amazon Nova
-
Amazon Titan
-
Other Amazon
-
Amazon Nova
-
Amazon Nova
Pricing for Understanding Models
Global Cross-region Inference
Geo Cross-region inference and in-region
Built-In-Tools
Pricing for Creative Content Generation models
Pricing for Speech Understanding and Generation Models
On-Demand pricing for speech to speech foundation models
Note: *The text tokens input and output pricing applies to specific use cases such as speech-to-text transcription, tool calls for task completion or knowledge grounding, adding conversation history to the session etc.
On-demand inference for custom Nova models is priced the same as base Nova inference.
Pricing for Embedding models
-
Amazon Titan
-
Amazon Titan
-
Other Amazon
-
-
-
Anthropic
-
**Non-GA Models: Access to Claude Mythos 5 and Claude Mythos Preview is gated and requires approval. Contact your Anthropic account team to request access on Bedrock.
Anthropic
On-Demand and Batch pricing
Models with extended access
Provider Model Name Regions Price per 1M input tokens Price per 1M output tokens Price per 1M input tokens (batch) Price per 1M output tokens (batch) Price per 1M input tokens (cache write) Price per 1M input tokens (cache read) Anthropic Claude 3.5 Sonnet (Public Extended Access, Effective 1 Dec 2025) US East (N. Virginia), US East (Ohio), US West (Oregon), Europe (Frankfurt), Europe (Ireland), Europe (Zurich), Europe (Paris) $6.00 $30.00 $3.00 $15.00 N/A N/A Anthropic Claude 3.5 Sonnet v2 (Public Extended Access, Effective 1 Dec 2025) US East (N. Virginia), US East (Ohio), US West (Oregon) $6.00 $30.00 $3.00 $15.00 $7.50 $0.60 Reserved Tier Pricing
Latency Optimized Inference
Provisioned Throughput Pricing
For Provisioned Throughput pricing, please reach out to your account team.
-
Cohere
-
Cohere
On-Demand pricing
Cohere models Price per 1,000 queries** Rerank 3.5 $2.00 **You are charged for number of queries where a query can contain up to 100 document chunks. If the query contains more than 100 document chunks, it is counted as multiple queries. For example, if a request contains 350 documents, it will be treated as 4 queries. Please note that each document can only contain upto 500 tokens (inclusive of the query and document’s total tokens), and if the token length is higher than 512 tokens, it is broken down into multiple documents. *Total tokens trained = number of tokens in training data corpus x number of epochs
Provisioned Throughput pricing
Cohere models Price per hour per model
with no commitmentPrice per hour per model unit for 1-month commitment Price per hour per model unit for 6-month commitment
Cohere Command
$49.50 $39.60
$23.77
Cohere Command - Light $8.56 $6.85
$4.11 Embed 3 English $7.12 $6.76
$6.41 Embed 3 Multilingual $7.12 $6.76
$6.41 Please reach out to your AWS account or sales team for more details on model units.
-
DeepSeek
-
DeepSeek
On-Demand pricing
-
Standard
-
Priority
-
Flex
-
Standard
-
Regions: US East (N. Virginia), US East (Ohio) and US West (Oregon)
DeepSeek models Price per 1M input tokens Price per 1M output tokens DeepSeek v3.2 $ 0.62 $ 1.85 Regions: Asia Pacific (Mumbai), South America (São Paulo), Asia Pacific (Jakarta), Asia Pacific (Tokyo) and Europe (Stockholm)
DeepSeek models Price per 1M input tokens Price per 1M output tokens DeepSeek v3.2 $ 0.74 $ 2.22 Region: Asia Pacific (Sydney)
DeepSeek models Price per 1M input tokens Price per 1M output tokens DeepSeek v3.1 $ 0.5974 $ 1.7304 DeepSeek v3.2 $ 0.6386 $ 1.9055 -
Priority
-
Region: Asia Pacific (Sydney)
DeepSeek models Price per 1M input tokens Price per 1M output tokens DeepSeek v3.1 $ 1.0455 $ 3.0282 -
Flex
-
Region: Asia Pacific (Sydney)
DeepSeek models Price per 1M input tokens Price per 1M output tokens DeepSeek v3.1 $ 0.2987 $ 0.8652
-
-
Google
-
Google
On-Demand pricing
Regions: US East (N. Virginia), US East (Ohio) and US West (Oregon)
Google models Price per 1M input tokens Price per 1M output tokens Gemma 4 31B $0.14 $0.40 Gemma 4 26B A4B $0.13 $0.40 Gemma 4 E2B $0.04 $ 0.08 Gemma 3 4B $ 0.04 $ 0.08 Gemma 3 12B $ 0.09 $ 0.29 Gemma 3 27B $ 0.23 $ 0.38 Regions: Europe (Frankfurt)
Google models Price per 1M input tokens Price per 1M output tokens Gemma 4 31B $ 0.17 $ 0.48 Gemma 4 26B A4B $ 0.16 $ 0.48 Gemma 4 E2B $ 0.05 $ 0.10 Regions: Asia Pacific (Mumbai), Europe (Ireland) and Europe (Milan)
Google models Price per 1M input tokens Price per 1M output tokens Gemma 3 4B $ 0.05 $ 0.09 Gemma 3 12B $ 0.11 $ 0.34 Gemma 3 27B $ 0.27 $ 0.45 Regions: South America (Sao Paulo) and Asia Pacific (Tokyo)
Google models Price per 1M input tokens Price per 1M output tokens Gemma 3 4B $ 0.05 $ 0.10 Gemma 3 12B $ 0.11 $ 0.35 Gemma 3 27B $ 0.28 $ 0.46 Region: Europe (London)
Google models Price per 1M input tokens Price per 1M output tokens Gemma 3 4B $ 0.06 $ 0.12 Gemma 3 12B $ 0.14 $ 0.45 Gemma 3 27B $ 0.36 $ 0.59 Region: Asia Pacific (Sydney)
Google models Price per 1M input tokens Price per 1M output tokens Gemma 3 4B $ 0.0412 $ 0.0824 Gemma 3 12B $ 0.0927 $ 0.2987 Gemma 3 27B $ 0.2369 $ 0.3914 * Priority tier pricing is at 75% premium to Standard tier pricing
* Flex tier pricing is at 50% discount to Standard tier pricing -
Luma AI
-
-