ChengRang

Claude Haiku 5.5

AI Chatbots Freemium
This page covers a version or sub-product of Claude. View Claude overview →

Anthropic released this small model on October 7, 2026, calling it the cheapest and fastest model it has shipped. Prompts up to 100K tokens cost $0.10 per million input and $0.50 per million output, rising to $0.50 and $2.50 beyond that, up to 90 percent below Haiku 4.5. It offers a 1M context window, up to 128K output tokens, and the first adjustable effort setting on a Haiku model, and targets high-volume work such as summarization, compaction, classification and database queries as well as subagent duty alongside Opus 5.5 and Sonnet 5.5

AnthropicSmall ModelHigh ThroughputSubagentAdjustable Effort
Visit Claude Haiku 5.5

Disclaimer: Review content represents our editorial team's views and experience, not commercial recommendation or investment advice. Product info and pricing may change; refer to official sources.

Overview

Claude Haiku 5.5 is the small model Anthropic released on October 7, 2026, billed as the cheapest and fastest model it has ever shipped. The headline is price: prompts up to 100K tokens cost $0.10 per million input tokens and $0.50 per million output tokens, one tenth of Haiku 4.5 and exactly what OpenAI charges for GPT-6 Luna. Past 100K tokens the rate goes up five times to $0.50 and $2.50, which is still half of what the previous generation charged.

Price is only half of this release. The other half is capability. Haiku 4.5 was close to blank on coding and computer use, scoring 0 on Terminal-Bench 4.0 and 15.7 percent on OSWorld, while Haiku 5.5 brings those to 39.2 percent and 72.4 percent. It is also the first Haiku with an adjustable effort setting, five notches from Low to Max with Medium as the default, so you can spend more compute when the job deserves it. For teams running tens of thousands of calls a day, this is the release that moves small models from acceptable to actually competent.

Key Features

Use Cases

Pros

Pricing

Billing is per token. Prompts up to 100K tokens cost $0.10 per million input and $0.50 per million output, with cache reads at $0.01 and cache writes at $0.125. Above 100K tokens it is $0.50 and $2.50, with cache reads at $0.05 and cache writes at $0.625. The batch API halves standard rates. Every Claude plan from Free to Enterprise can use it, and Max 5x and Max 20x subscribers receive $100 and $200 of monthly Platform API credits respectively, while Team plans get up to $500 shared across the team. Credits apply to any model on the platform, expire at the end of the month and do not roll over.

Summary

What makes Haiku 5.5 interesting is not how strong it is but how much it can do at this price. Small models used to be a trade where you bought cheap and accepted weak. This time Anthropic matched the competitor on price while fixing the parts that held the line back, so for the first time the small model qualifies for a class of work on its own.

Two things are worth deciding up front. The first is the 100K line. Most traffic stays below it, but if a retrieval flow reliably lands between 120K and 180K tokens the bill will look heavier than the headline suggests, so routing by prompt length pays off. The second is the effort dial. Medium is the default and the value starting point, classification and routing gain little from going higher, while extraction from messy documents and multi-step tool calls are where High and xhigh earn their cost.

The ceiling is real. Factual knowledge is the thin spot, so anything needing encyclopedic accuracy still goes up a tier, and automation benchmarks came in low with evaluators suspecting over-refusal from safety policy that Anthropic says it is fixing. Taken as a primary reasoning model it will disappoint. Taken as the worker model on a pipeline, it is hard to beat on value right now.

Version History

Category
AI Chatbots
Pricing
Freemium
Tags
Anthropic · Small Model · High Throughput

Related Tools