Anthropic released Claude Haiku 5.5, with operating costs reduced by about 75% compared to the previous generation
Anthropic released Claude Haiku 5.5 on Wednesday, claiming it to be the company's cheapest, fastest, and most capable small model to date. This model is designed for high-volume tasks such as document summarization and database queries, as well as speed-sensitive scenarios like real-time customer service and browser automation. When developers integrate it into their own products, prompts of up to 100,000 tokens are charged at $0.1 per million input tokens and $0.5 per million output tokens, consistent with the pricing of OpenAI's competing small model GPT-6 Luna released on September 22.
The corresponding prices for Haiku 4.5 are $1 and $5, with the new prices being 90% lower, and prompts exceeding 100,000 tokens receiving a 50% discount. Since about 90% of requests for the old model fall below this threshold, and Haiku 5.5 has a slightly higher token count, Anthropic estimates an average savings of about 75%. Benchmark tests show that Haiku 5.5 achieved a success rate of 72.4% on OSWorld 2.1, higher than GPT Luna's 48.9%; on Terminal-Bench 4, it scored 39.2%, surpassing Luna's 16.4% and Haiku 4.5's 0%, but lower than Sonnet 5.5's 70.6%.
This is the first Haiku model with adjustable effort settings, allowing users to balance cost and answer quality.






