The new model, described by Anthropic as its fastest and most efficient, is priced at $0.10 per million input tokens and $0.50 per million output tokens for requests under 100,000 tokens, down from $1 and $5 respectively for Haiku 4.5. For larger requests, the cost is $0.50 input and $2.50 output, though Anthropic notes that roughly 90% of Haiku 4.5 requests fell into the lower-priced tier, putting average savings at about 75% when accounting for a mix of request sizes and an updated tokenizer.
Haiku 5.5 is the first Haiku model to offer effort controls, with a default setting of medium, giving developers more control over token usage per task. The model is intended for high-volume tasks such as summarisation, classification and routing, but Anthropic also highlights its suitability for agentic workloads where speed matters, including live customer support, browser use, compaction and database queries.
Benchmarks published by Anthropic show substantial improvements over Haiku 4.5. On the OSWorld 2.1 offline subset, which tests computer use, Haiku 5.5 scored 72.4%, up from 15.7% for its predecessor and ahead of OpenAI's GPT-6 Luna at 48.9%. On agentic coding benchmarks, Terminal-Bench 4.0 returned 39.2% (up from 0.0%) and FrontierCode 1.1 gave 46.4%. On visual reasoning (Chartography, no tools), the score was 46.4% versus 6.4% for Haiku 4.5. On the Knowledge work GDPval-AA v2.1 benchmark, Haiku 5.5 scored 1620, compared with 735 for Haiku 4.5 and 1437 for GPT-6 Luna.
The launch of Haiku 5.5 follows Sonnet 5.5 and another 5.5 series model released in the past month. Anthropic has not yet released Fable 5.5, which is expected to undergo a longer review process.