Claude Haiku 5.5 is now available on Amazon Bedrock

Claude Haiku 5.5 is now available on Amazon Bedrock

I tried Claude Haiku 5.5 via AWS CLI with Bedrock Runtime. For prompts under 100K tokens, the price matches GPT-6 Luna, and max context expanded from Haiku 4.5's 200K to 1M.
2026.10.08

This page has been translated by machine translation. View original

Introduction

On October 7, 2026, Claude Haiku 5.5 became available on Amazon Bedrock.

https://aws.amazon.com/about-aws/whats-new/2026/10/claude-haiku-5-5-aws/

https://docs.aws.amazon.com/bedrock/latest/userguide/model-card-anthropic-claude-haiku-5-5.html

In this article, I called Haiku 5.5 via AWS CLI converse using IAM authentication. I also checked how it appears in CloudTrail event history and confirmed the unit prices retrieved from the Price API. The unit prices are compared with Sonnet 5.5, Opus 5.5, and others.

Haiku 5.5 Context Length and Inference Profiles

According to the model card, the Context window is 200K tokens for Haiku 4.5 and 1M tokens for Haiku 5.5. Sonnet 5.5 and Opus 5.5 are also 1M tokens.

The model ID for Haiku 5.5 is anthropic.claude-haiku-5-5, and inference profiles with the prefixes us, eu, au, jp, and global are available.

Calling converse with IAM Authentication

The sample code in the model card describes a procedure for issuing a long-term API key from the Bedrock console and calling it from Python. In this article, I called it using IAM role session credentials without using an API key.

JP Profile

This is the command specifying the JP inference profile.

aws bedrock-runtime converse --region ap-northeast-1 \
  --model-id jp.anthropic.claude-haiku-5-5 \
  --messages '[{"role":"user","content":[{"text":"Bedrock の Claude Haiku 5.5 について、一文で自己紹介してください。"}]}]' \
  --inference-config '{"maxTokens":1000}'

Here is an excerpt of the response. The reasoningContent block and metrics are omitted.

{
    "output": {
        "message": {
            "role": "assistant",
            "content": [
                {
                    "text": "私はAnthropicが開発したAIアシスタントのClaudeで、Amazon Bedrockから利用されているClaude Haiku 5.5として、質問への回答、文章作成、要約、翻訳、プログラミングなどのお手伝いをします。"
                }
            ]
        }
    },
    "stopReason": "end_turn",
    "usage": {
        "inputTokens": 39,
        "outputTokens": 469,
        "totalTokens": 508,
        "cacheReadInputTokens": 0,
        "cacheWriteInputTokens": 0
    }
}

The text in the omitted reasoningContent was an empty string. However, outputTokens is a larger value than the character count of the response body. The model card states that adaptive thinking is enabled by default. Output costs are estimated based on usage outputTokens, not the length of the response body.

Global Profile

I ran the same prompt with --model-id changed to global.anthropic.claude-haiku-5-5. The exit code was 0, stopReason was end_turn, and usage was inputTokens 39, outputTokens 335, totalTokens 374. As with JP, the text in reasoningContent was an empty string.

How It Appears in CloudTrail Event History

I checked the event history using lookup-events without creating additional trails or S3 buckets.

aws cloudtrail lookup-events --region ap-northeast-1 \
  --start-time 2026-10-08T05:44:00+09:00 --end-time 2026-10-08T05:56:00+09:00 \
  --lookup-attributes AttributeKey=EventSource,AttributeValue=bedrock.amazonaws.com

When I checked approximately 10 minutes after the calls, both the JP and Global entries were displayed. The event source is bedrock.amazonaws.com, not bedrock-runtime. A search in a different time range (05:25–05:40) also included ListGuardrails and ListTagsForResource from the AWS Config service-linked role, so I narrowed the results to events where the event name is Converse.

CloudTrailEvent is returned as a JSON string. The following is an excerpt from the expanded event for the Global call. Identifiers such as access keys and role names have been removed.

{
  "EventName": "Converse",
  "ReadOnly": "true",
  "EventTime": "2026-10-08T05:45:27+09:00",
  "EventSource": "bedrock.amazonaws.com",
  "CloudTrailEvent": {
    "eventTime": "2026-10-07T20:45:27Z",
    "eventSource": "bedrock.amazonaws.com",
    "eventName": "Converse",
    "awsRegion": "ap-northeast-1",
    "requestParameters": {
      "modelId": "global.anthropic.claude-haiku-5-5",
      "inferenceConfig": {
        "maxTokens": 1000
      }
    },
    "responseElements": null,
    "additionalEventData": {
      "inferenceRegion": "us-east-1",
      "inputTokens": 39,
      "outputTokens": 335
    },
    "readOnly": true,
    "resources": [
      {
        "accountId": "<ACCOUNT_ID>",
        "type": "AWS::Bedrock::Model",
        "ARN": "arn:aws:bedrock:ap-northeast-1::foundation-model/anthropic.claude-haiku-5-5"
      }
    ],
    "eventType": "AwsApiCall",
    "managementEvent": true,
    "eventCategory": "Management"
  }
}

The model called can be identified from modelId in requestParameters. The prompt body (messages) is not included in requestParameters.

The processing region can be identified from inferenceRegion in additionalEventData. In this case, the JP call processed in ap-northeast-1 and the Global call in us-east-1. The ARN in resources pointed to the foundation model in the calling region for both JP and Global calls, so it cannot be used to determine the processing region. The inputTokens and outputTokens in additionalEventData matched the values in the response usage.

Pricing

Checking Unit Prices via the Price API

aws pricing get-products --region us-east-1 --service-code AmazonBedrockFoundationModels \
  --filters 'Type=TERM_MATCH,Field=servicename,Value=Claude Haiku 5.5 (Amazon Bedrock Edition)' \
            'Type=TERM_MATCH,Field=regionCode,Value=us-east-1'
Execution result (excerpt)
{
    "PriceList": [
        "{\"product\":{\"attributes\":{\"regionCode\":\"us-east-1\",\"usagetype\":\"USE1-MP:USE1_output_tokens_global_standard-Units\",\"locationType\":\"AWS Region\",\"location\":\"US East (N. Virginia)\",\"servicename\":\"Claude Haiku 5.5 (Amazon Bedrock Edition)\",\"operation\":\"\"},\"sku\":\"9C6E7ATQE23Q4W66\"},\"serviceCode\":\"AmazonBedrockFoundationModels\",\"terms\":{\"OnDemand\":{\"9C6E7ATQE23Q4W66.4799GE89SK\":{\"priceDimensions\":{\"9C6E7ATQE23Q4W66.4799GE89SK.6YS6EN2CT7\":{\"unit\":\"1M tokens\",\"endRange\":\"Inf\",\"description\":\"AWS Marketplace software usage|us-east-1|Output Tokens - Standard, Global\",\"appliesTo\":[],\"rateCode\":\"9C6E7ATQE23Q4W66.4799GE89SK.6YS6EN2CT7\",\"beginRange\":\"0\",\"pricePerUnit\":{\"USD\":\"0.5000000000\"}}},\"sku\":\"9C6E7ATQE23Q4W66\",\"effectiveDate\":\"2026-10-01T00:00:00Z\",\"offerTermCode\":\"4799GE89SK\",\"termAttributes\":{}}}},\"version\":\"20261007183733\",\"publicationDate\":\"2026-10-07T18:37:33Z\"}",
        "{\"product\":{\"attributes\":{\"regionCode\":\"us-east-1\",\"usagetype\":\"USE1-MP:USE1_input_tokens_global_standard-Units\",\"locationType\":\"AWS Region\",\"location\":\"US East (N. Virginia)\",\"servicename\":\"Claude Haiku 5.5 (Amazon Bedrock Edition)\",\"operation\":\"\"},\"sku\":\"CT2JY6DA5N94ZAXN\"},\"serviceCode\":\"AmazonBedrockFoundationModels\",\"terms\":{\"OnDemand\":{\"CT2JY6DA5N94ZAXN.4799GE89SK\":{\"priceDimensions\":{\"CT2JY6DA5N94ZAXN.4799GE89SK.6YS6EN2CT7\":{\"unit\":\"1M tokens\",\"endRange\":\"Inf\",\"description\":\"AWS Marketplace software usage|us-east-1|Input Tokens - Standard, Global\",\"appliesTo\":[],\"rateCode\":\"CT2JY6DA5N94ZAXN.4799GE89SK.6YS6EN2CT7\",\"beginRange\":\"0\",\"pricePerUnit\":{\"USD\":\"0.1000000000\"}}},\"sku\":\"CT2JY6DA5N94ZAXN\",\"effectiveDate\":\"2026-10-01T00:00:00Z\",\"offerTermCode\":\"4799GE89SK\",\"termAttributes\":{}}}},\"version\":\"20261007183733\",\"publicationDate\":\"2026-10-07T18:37:33Z\"}",
        "{\"product\":{\"attributes\":{\"regionCode\":\"us-east-1\",\"usagetype\":\"USE1-MP:USE1_input_tokens_long_ctx_global_standard-Units\",\"locationType\":\"AWS Region\",\"location\":\"US East (N. Virginia)\",\"servicename\":\"Claude Haiku 5.5 (Amazon Bedrock Edition)\",\"operation\":\"\"},\"sku\":\"E7TVM9F3MW6M99X9\"},\"serviceCode\":\"AmazonBedrockFoundationModels\",\"terms\":{\"OnDemand\":{\"E7TVM9F3MW6M99X9.4799GE89SK\":{\"priceDimensions\":{\"E7TVM9F3MW6M99X9.4799GE89SK.6YS6EN2CT7\":{\"unit\":\"1M tokens\",\"endRange\":\"Inf\",\"description\":\"AWS Marketplace software usage|us-east-1|Input Tokens - Standard, Long Context, Global\",\"appliesTo\":[],\"rateCode\":\"E7TVM9F3MW6M99X9.4799GE89SK.6YS6EN2CT7\",\"beginRange\":\"0\",\"pricePerUnit\":{\"USD\":\"0.5000000000\"}}},\"sku\":\"E7TVM9F3MW6M99X9\",\"effectiveDate\":\"2026-10-01T00:00:00Z\",\"offerTermCode\":\"4799GE89SK\",\"termAttributes\":{}}}},\"version\":\"20261007183733\",\"publicationDate\":\"2026-10-07T18:37:33Z\"}",
        "{\"product\":{\"attributes\":{\"regionCode\":\"us-east-1\",\"usagetype\":\"USE1-MP:USE1_cache_read_tokens_long_ctx_standard-Units\",\"locationType\":\"AWS Region\",\"location\":\"US East (N. Virginia)\",\"servicename\":\"Claude Haiku 5.5 (Amazon Bedrock Edition)\",\"operation\":\"\"},\"sku\":\"XCA8KP4W9DAXTNDB\"},\"serviceCode\":\"AmazonBedrockFoundationModels\",\"terms\":{\"OnDemand\":{\"XCA8KP4W9DAXTNDB.4799GE89SK\":{\"priceDimensions\":{\"XCA8KP4W9DAXTNDB.4799GE89SK.6YS6EN2CT7\":{\"unit\":\"1M tokens\",\"endRange\":\"Inf\",\"description\":\"AWS Marketplace software usage|us-east-1|Cache Read Tokens - Standard, Long Context\",\"appliesTo\":[],\"rateCode\":\"XCA8KP4W9DAXTNDB.4799GE89SK.6YS6EN2CT7\",\"beginRange\":\"0\",\"pricePerUnit\":{\"USD\":\"0.0550000000\"}}},\"sku\":\"XCA8KP4W9DAXTNDB\",\"effectiveDate\":\"2026-10-01T00:00:00Z\",\"offerTermCode\":\"4799GE89SK\",\"termAttributes\":{}}}},\"version\":\"20261007183733\",\"publicationDate\":\"2026-10-07T18:37:33Z\"}"
    ],
    "FormatVersion": "aws_v1"
}

Haiku 5.5 Unit Prices (Global / US / JP)

The unit is USD / 1M tokens, based on the Price List published at 2026-10-07T18:37:33Z. The US column is for us-east-1 Standard, and the JP column is for ap-northeast-1 Standard. The Global column had the same values for both us-east-1 and ap-northeast-1. The "Prompt over 100K" rows show the unit prices when the prompt exceeds 100K tokens.

Item Global US JP
Input 0.10 0.11 0.11
Output 0.50 0.55 0.55
Cache write 5 min 0.125 0.1375 0.1375
Cache write 1 hour 0.20 0.22 0.22
Cache read 0.01 0.011 0.011
Input (prompt over 100K) 0.50 0.55 0.55
Output (prompt over 100K) 2.50 2.75 2.75
Cache write 5 min (prompt over 100K) 0.625 0.6875 0.6875
Cache write 1 hour (prompt over 100K) 1.00 1.10 1.10
Cache read (prompt over 100K) 0.05 0.055 0.055

Comparing the us-east-1 and ap-northeast-1 values, US and JP are identical across all items, and both are 1.1 times the Global price.

Comparison with Haiku 4.5, Sonnet 5.5, and Opus 5.5

According to Anthropic's pricing page, models from Claude 4.6 onward (except Haiku 5.5) can use 1M tokens of context at the standard price. For Haiku 5.5, pricing is determined by prompt length, and exceeding 100K tokens results in a higher price. Sonnet 5.5 and Opus 5.5 have a flat rate up to 1M tokens.

The following is a comparison of key items for Global (USD / 1M tokens, us-east-1).

Model Context Input Output Cache write 5 min Cache read
Haiku 4.5 200K 1.00 5.00 1.25 0.10
Haiku 5.5 (prompt 100K or less) 1M 0.10 0.50 0.125 0.01
Haiku 5.5 (prompt over 100K) 1M 0.50 2.50 0.625 0.05
Sonnet 5.5 1M 2.00 10.00 2.50 0.10
Opus 5.5 1M 4.00 20.00 5.00 0.20

For Haiku 5.5 with prompts exceeding 100K tokens, the input unit price of 0.50 is 1/4 of Sonnet 5.5's 2.00, 1/8 of Opus 5.5's 4.00, and half of Haiku 4.5's 1.00. The output unit price follows the same ratio: 2.50 is 1/4 of Sonnet 5.5's 10.00, 1/8 of Opus 5.5's 20.00, and half of Haiku 4.5's 5.00. For prompts of 100K tokens or less, the input unit price of 0.10 is 1/20 of Sonnet 5.5's price. Tasks that previously required Sonnet or Opus for 1M context become candidates for replacing with Haiku 5.5 from a pricing perspective.

Comparison with Non-Claude Models

GPT-6 Luna

The GPT-6 Luna Context window is 1,050,000 tokens. The GPT-6 Luna unit prices are from the Pricing section (Standard tier) of the Bedrock model card. The Haiku 5.5 unit prices are from the Global column in the previous section. The unit is USD / 1M tokens, and each cell shows "input / output" in that order.

Model (Global) Prompt 100K or less Prompt over 100K up to 272K Prompt over 272K
Haiku 5.5 0.10 / 0.50 0.50 / 2.50 0.50 / 2.50
GPT-6 Luna (Global CRIS) 0.10 / 0.50 0.10 / 0.50 0.20 / 0.75

The prices are equal for prompts of 100K tokens or less, and Luna has lower unit prices above 100K. The GPT-6 Luna model card states that exceeding 272K tokens applies the long-context price to the entire request. The unit prices for GPT-6 Luna Geo (US) and Mantle in-Region are 10% higher than Global.

Open Models

Unit prices are from the Bedrock Price List (us-east-1, Standard, published 2026-10-06T14:47:26Z), in USD / 1M tokens. Context length and availability start date are from each model card.

Model Input Output Context Availability on Bedrock
gpt-oss-20b 0.07 0.30 128K August 2025
gpt-oss-120b 0.15 0.60 128K August 2025
Gemma 4 26B A4B 0.13 0.40 256K March 2026
Gemma 4 31B 0.14 0.40 256K March 2026
DeepSeek V3.2 0.62 1.85 164K December 2025
Kimi K3 3.30 16.50 1M September 2026

Among the models in this table, only Kimi K3 supports a 1M context. Kimi K3's unit price is approximately 6.6 times that of Haiku 5.5 (Global) for prompts over 100K, making it more expensive than Sonnet 5.5.

Summary

Following Opus 5.5 and Sonnet 5.5 in September 2026, Claude Haiku 5.5 was released in October.

https://dev.classmethod.jp/articles/bedrock-claude-opus-5-5/

https://dev.classmethod.jp/articles/bedrock-claude-sonnet-5-5/

Claude Haiku 5.5 arrived 11 months after the release of Haiku 4.5. For prompts of 100K tokens or less, the unit price is 1/10 of Haiku 4.5, making it comparable to open models.

When prompts exceed 100K tokens the price increases fivefold, but even then, context up to 1M tokens can be handled at 1/4 the unit price of Sonnet 5.5.

Claude Haiku 5.5 comes with a JP inference profile available from the time of GA. It is easy to evaluate by calling Bedrock Runtime from the AWS CLI with IAM authentication, so please give it a try.


Claudeならクラスメソッドにお任せください

クラスメソッドは、Anthropic社とリセラー契約を締結しています。各種製品ガイドから、業種別の活用法、フェーズごとのお悩み解決などサービス支援ページにまとめております。まずはご覧いただき、お気軽にご相談ください。

サービス詳細を見る

Share this article

AWSのお困り事はクラスメソッドへ