Anthropic releases Claude Haiku 5.5 at $0.10 input and $0.50 output, which it says runs about 75% cheaper than Haiku 4.5 on average
Its cheapest and fastest small model, and Sonnet 5.5 cache reads are halved on the same day.
Released 7 October. Per million tokens, for prompts up to 100K tokens, it costs $0.10 input and $0.50 output ($0.50 / $2.50 above 100K), against $1 / $5 for Haiku 4.5 and $2 / $10 for Sonnet 5.5. Anthropic says it costs “around 75% less to run” on average; a footnote says the figure already counts that the new tokenizer uses slightly more tokens per task. It is the first Haiku with an adjustable effort setting, and the Claude Code release notes give a 1M-token context. On Anthropic’s own evaluations, Terminal-Bench 4.0 is 39.2% (Haiku 4.5: 0.0%, Sonnet 5.5: 70.6%) and OSWorld 2.1, offline subset, is 72.4% (Sonnet 5.5: 83.9%). The same day Anthropic cut Sonnet 5.5 cache reads from $0.20 to $0.10 (about 20% cheaper on most agentic work, it says) and said Max and Team subscribers get a monthly API credit for the Claude Platform this week: $100 for Max 5x, $200 for Max 20x, and up to $500 pooled for Team.
If you run lots of small jobs such as summaries, classification, compaction or subagent work, it is worth re-checking your unit costs. Anthropic itself says Sonnet 5.5 and Opus 5.5 remain the better choice for complex agentic coding. The numbers are Anthropic’s own evaluations, and its cyber safeguards are looser than Sonnet 5.5’s but still block penetration testing.
