Claude Haiku 5.5 released with a context window of one million tokens
Claude Haiku 5.5 is available through Claude API and four cloud offerings. It has a context window of one million tokens, a maximum output of 128 thousand tokens and support for adaptive thinking with the effort parameter.
Claude Haiku 5.5 has been released with the identifier claude-haiku-5-5. Its context window holds one million tokens, and its maximum output is 128 thousand tokens. The model supports adaptive thinking, or adaptive reasoning, with the effort parameter.
The model is available through Claude API and the offerings Claude in Amazon Bedrock, Claude Platform on AWS, Claude on Google Cloud and Claude in Microsoft Foundry. According to the publisher, it is designed for large volumes of requests and latency-sensitive tasks.
Why it matters
The large context window allows extensive source material to be included in a single request, while the output limit provides room for long responses. Developers can integrate the model into applications through the API or one of the listed cloud offerings.
Relevant practical impact
What this means
For a business
Teams developing applications that handle large volumes of requests now have another model option available. The stated focus on low latency is a claim by the publisher, so it does not in itself establish how fast the model will be in business use.
DevelopmentCheck the original
Event sources
clearly official source · 1 publisher, 0 independent. We count feeds from the same owner only once.