Back to Blog
EnglishGuide

AWS Bedrock Anthropic Claude Pricing per 1M Tokens

Read Claude pricing on Bedrock by model and token type, calculate a workload estimate, and avoid confusing standard rates with caching or provisioned capacity.

C
Crazyrouter Team
October 10, 2026 / 0 views
Share:
AWS Bedrock Anthropic Claude Pricing per 1M Tokens

For the query “aws.amazon.com bedrock anthropic claude pricing per 1m tokens,” there is no single Claude rate: choose the exact model, region, and inference option first. As of 2026-10, the Claude Sonnet 4.6 provider comparison lists Bedrock’s base rates at 3permillioninputtokensand3 per million input tokens and 15 per million output tokens; confirm the applicable row on AWS Bedrock pricing before budgeting.

What does per 1M tokens mean?#

Input tokens cover material sent to the model, including your prompt and conversation history. Output tokens cover the generated response. They have different prices, so multiplying the combined token total by a single rate usually gives the wrong answer.

A million tokens is a billing unit, not a requirement to send that many tokens at once. For an ordinary on-demand estimate, multiply each token count by its corresponding rate and divide by one million. Then add any separately priced services or features that your application uses.

The figures here come from the linked October provider comparison. They are not a newly measured AWS invoice, and they should not be generalized to every Claude model or regional deployment.

AWS Bedrock Claude pricing separates input and output token costs

How much would a Sonnet 4.6 workload cost?#

Consider a workload with ten million input tokens and two million output tokens. At the listed standard rates, input costs 30andoutputcostsanother30 and output costs another 30, giving an estimated $60 in model charges.

ComponentVolumeListed Bedrock rate per 1MEstimated cost
Standard input10M tokens$3.00$30.00
Standard output2M tokens$15.00$30.00
TotalSeparate input and outputNot a blended rate$60.00

The same source lists Crazyrouter Sonnet 4.6 at 1.95inputand1.95 input and 9.75 output per million tokens. That example would total $39 at those stated rates. This is a comparison of model token charges, not a claim that the two services have identical contracts, cloud integration, or operational features.

Which settings can change the bill?#

Check whether your quote is for standard on-demand inference, batch processing, or provisioned capacity. Those purchasing choices are not interchangeable. A capacity commitment cannot be evaluated from a token-price table alone, because utilization affects its effective cost.

Caching also needs its own calculation. Distinguish uncached input, cache writes, and cache reads rather than assigning the ordinary input rate to all three. The useful question is how often your actual prompt prefix is reused under the selected model’s caching rules.

Region and inference routing matter as well. Select the intended deployment on AWS’s page and record that choice alongside the model version. If an application adds retrieval, storage, monitoring, or guardrails, account for those costs separately from the model estimate.

Where should developers verify a quote?#

Start at the official AWS pricing page, then consult the Bedrock documentation for the selected inference workflow. Save the model identifier, region, rate, and date with your estimate so a later pricing change is easy to trace.

For cross-provider budgeting, use the model pricing overview and the model-specific table above. Match output length and caching assumptions before declaring one route cheaper. A service with a lower input rate can still cost more for an output-heavy workload.

This is a billing explanation, so it does not include an inference code sample. When implementing requests, follow AWS authentication and model identifiers for Bedrock; a gateway model alias is not automatically a valid Bedrock identifier.

FAQ#

Does every Anthropic Claude model on Bedrock cost 3and3 and 15?#

No. Those figures describe the Sonnet 4.6 base row in the cited October comparison. Other model versions and inference options require their own rates.

Is the per-million rate charged for every request?#

It is the unit used to express token pricing, not a fixed charge for each call. Estimate cost from the actual input and output counts, applying each relevant rate separately.

Can I compare Bedrock directly with a model gateway?#

Yes, if you match the model, token mix, and billing assumptions. Also compare cloud integration, support, and data-handling requirements, because token price alone does not decide suitability.

Plan a measurable trial#

Estimate a representative workload, run a limited trial, and compare the resulting usage with your calculation. If you also want to evaluate the gateway route, create a Crazyrouter account and check its current model-specific price before sending requests.

<!-- crazyrouter-related-links -->

Implementation Guides

Related Articles

Claude Code Pricing Guide 2026 for Teams, Startups, and Power UsersGuide

Claude Code Pricing Guide 2026 for Teams, Startups, and Power Users

A practical Claude Code pricing guide for developers who want to understand subscription trade-offs, usage patterns, and when a unified API layer makes more sense.

Mar 19
Seedance 2.0 Pricing: Convert 46 CNY per Million Tokens to Cost per SecondGuide

Seedance 2.0 Pricing: Convert 46 CNY per Million Tokens to Cost per Second

Seedance 2.0 uses token-based video pricing. This guide converts 46 CNY per million tokens into per-second and per-video costs for pure generation and video editing.

May 25
Claude API Pricing 2026: Every Model's Price per 1M Tokens (Opus 5, Fable 5, Sonnet 5, Haiku 4.5)Guide

Claude API Pricing 2026: Every Model's Price per 1M Tokens (Opus 5, Fable 5, Sonnet 5, Haiku 4.5)

Claude API pricing for every current model in one table: Anthropic list price, cached-input price and the Crazyrouter price per 1M tokens, plus monthly cost examples, how billing works, and how to get started.

Sep 30
AI API Cost Optimization: Complete Guide to Reducing Your AI Spending in 2026Guide

AI API Cost Optimization: Complete Guide to Reducing Your AI Spending in 2026

"Learn proven strategies to cut your AI API costs by 40-70%. From model selection and caching to API routing and prompt optimization, this guide covers everything developers need to reduce AI spending."

Mar 4
Claude Code Pricing Guide for Teams in 2026: Costs, Limits, and Cheaper API WorkflowsGuide

Claude Code Pricing Guide for Teams in 2026: Costs, Limits, and Cheaper API Workflows

A developer-first Claude Code pricing guide covering subscription tiers, API costs, team budgeting, alternatives, and how to reduce spend with Crazyrouter.

Mar 15
How to Get Claude API Key in 2026: Secure Setup for Teams and CIGuide

How to Get Claude API Key in 2026: Secure Setup for Teams and CI

how to get Claude API key explained for developers with setup steps, code examples, pricing trade-offs, and a Crazyrouter-based production path.

Jun 13