Claude Haiku 5.5

Anthropic's new small model, released 7 October 2026. The headline is efficiency: about 75% cheaper to run than Haiku 4.5, and far more capable.

Claude Haiku 5.5

Key points

  • Price: $0.10 per million input tokens and $0.50 per million output tokens (prompts up to 100K), against $1.00 and $5.00 for Haiku 4.5. Anthropic says about 75% cheaper to run on average.
  • Speed: Anthropic's fastest model at standard speed. One customer reports up to 2.5x faster inference per agent turn.
  • Capability jump over Haiku 4.5: OSWorld 2.1 (offline subset) 72.4% vs 15.7%; Terminal-Bench 4.0 39.2% vs 0.0%; Humanity's Last Exam (no tools) 45.9% vs 10.2%; GDPval-AA v2.1 1620 vs 735 Elo.
  • Effort setting: adjustable, so users can trade cost against capability.
  • Availability: all platforms including AWS, Google Cloud and Azure. Model ID claude-haiku-5-5.

Testimony: Theo

Developer and YouTuber Theo (t3.gg) had written in September that "Anthropic has no small models that are worth using right now." On 10 October he replied: "This is no longer true. Haiku has placed Anthropic in a position of complete and utter domination of the market." (143K views at time of capture.) One developer's reaction, not a benchmark, but a useful signal that the cheap tier is now competitive.

Theo's tweet on Haiku

Why it matters

Cheaper, capable small models make agent workloads that were too expensive to run at volume practical. The gain here is from efficiency, not a bigger frontier model.