Claude Haiku 5.5 released with a context window of one million tokens and access via API and cloud
Claude Haiku 5.5 is available through Claude API and cloud platforms. It offers a context of one million tokens, output of up to 128 thousand tokens and adaptive thinking with the effort parameter. According to the company, it is designed for a high volume of requests and latency-sensitive tasks.
The announcement reports the release of Claude Haiku 5.5 with the identifier claude-haiku-5-5. The context window has a capacity of one million tokens, and the maximum output is 128 thousand tokens. According to the company, the model is tailored to handling a high volume of requests and latency-sensitive tasks.
The model supports adaptive thinking, with the effort parameter used to adjust the level of reasoning. It is available through Claude API and the offerings Claude in Amazon Bedrock, Claude Platform on AWS, Claude on Google Cloud and Claude in Microsoft Foundry.
Why it matters
The context capacity and maximum output determine how much source material can be included in a single request and how long a response can be. Availability through an API and multiple cloud platforms expands the options for integrating the model into applications.
Two audiences, two different impacts
What this means
For individuals
Developers can use Claude API to test working with extensive source material in a context of up to one million tokens and adjust the level of reasoning with the effort parameter.
For a business
Businesses have another model option for applications that handle a high volume of requests, which the company says is tailored to latency-sensitive tasks. Availability on several cloud platforms provides more options for integrating it into business software development.
DevelopmentCheck the original
Event sources
clearly official source · 1 publisher, 0 independent. We count feeds from the same owner only once.