Skip to content
important New models

Claude Haiku 5.5 released with a context window of one million tokens

clearly official source

Claude Haiku 5.5 is available through Claude API and four cloud offerings. It has a context window of one million tokens, a maximum output of 128 thousand tokens and support for adaptive thinking with the effort parameter.

Claude Haiku 5.5 has been released with the identifier claude-haiku-5-5. Its context window holds one million tokens, and its maximum output is 128 thousand tokens. The model supports adaptive thinking, or adaptive reasoning, with the effort parameter.

The model is available through Claude API and the offerings Claude in Amazon Bedrock, Claude Platform on AWS, Claude on Google Cloud and Claude in Microsoft Foundry. According to the publisher, it is designed for large volumes of requests and latency-sensitive tasks.

What changed

Why it matters

The large context window allows extensive source material to be included in a single request, while the output limit provides room for long responses. Developers can integrate the model into applications through the API or one of the listed cloud offerings.

Relevant practical impact

What this means

01

For a business

Teams developing applications that handle large volumes of requests now have another model option available. The stated focus on low latency is a claim by the publisher, so it does not in itself establish how fast the model will be in business use.

Development
What to decide Before deployment, test the model’s latency on requests representative of your application.
More business impacts →
Amazon Bedrock Claude API Claude Haiku 5.5 effort Google Cloud Microsoft Foundry

Check the original

Event sources

clearly official source · 1 publisher, 0 independent. We count feeds from the same owner only once.

2
Claude Platform Release Notes primary source · first detected We've launched Claude Haiku 5.5 ( claude-haiku-5-5 ), our most capable model tuned for high-volume and latency-sensitive work. It has a 1M token context window , 128k max output… Claude Platform Release Notes primary source We've launched Claude Haiku 5.5 ( claude-haiku-5-5 ), our most capable model tuned for high-volume and latency-sensitive work. It has a 1M token context window , 128k max output…