TRENDING
Claude Haiku 5.5 Explained: 75% Cheaper, 1M Token ContextClaude in Google Docs, Sheets and SlidesSiri AI Adds Five New Languages in iOS 27.2: What ChangesOpenAI Cancelled GPT-6.1 Astra's Release. Here's WhyFTC Probes OpenAI and Anthropic Over Rogue AgentsOpenAI Parts Ways With 3 Safety ResearchersMicrosoft Digital Defense Report 2026 ExplainedMeta Muse Gadgets: Open-Source SDK ExplainedGoogle Project Suncatcher: TPUs in OrbitApple Tightens macOS Disk Access for AI AgentsMicrosoft MAI Transcribe 2, Voice 2.1 ExplainedGemini 4 vs GPT-6.1 vs Claude 5.5 ComparedGoogle Gemini 4 Argon Is Here: Everything You Need to KnowNano Banana Prompts: How to Edit Photos with Gemini AITrump Meets AI CEOs at White House, Pushes 'Super Intelligence'

Claude Haiku 5.5 Explained: 75% Cheaper, 1M Token Context

Anthropic released Claude Haiku 5.5 on October 7, 2026, calling it “the cheapest, fastest, and most capable small model we’ve ever released.” The model is built for high-volume, cost-sensitive work: summaries, document compaction, database queries, classification requests, live customer support, browser use, and as a subagent running alongside Opus 5.5 and Sonnet 5.5 on coding tasks.

The headline number is cost. Anthropic says Haiku 5.5 is roughly 75% cheaper to run on average than Haiku 4.5, while also extending the context window to 1 million tokens and adding an adjustable effort setting that previous Haiku models did not have.

Here is what Anthropic has confirmed, what it costs, and what changes for anyone building on the Claude API.

⚡ Quick facts

  • Product: Claude Haiku 5.5
  • Developer: Anthropic
  • Launched: October 7, 2026
  • Model ID: claude-haiku-5-5
  • Cost: ~75% cheaper to run on average than Haiku 4.5
  • Context window: 1,000,000 tokens (up from 200K)
  • Pricing (≤100K tokens): $0.10/M input, $0.50/M output
  • Availability: Claude Platform, AWS, Google Cloud, Microsoft Azure
Advertisement

What Anthropic Announced

On October 7, 2026, Anthropic released Claude Haiku 5.5, the newest entry in its small-model line. In its own announcement, the company described it as “the cheapest, fastest, and most capable small model we’ve ever released.” The model is positioned for high-volume, cost-sensitive work rather than frontier reasoning: summaries, context compaction, database queries, classification, live customer support, browser use, and running as a subagent alongside Opus 5.5 and Sonnet 5.5 on coding tasks.

Anthropic also used the same announcement to cut the cache-read price of Claude Sonnet 5.5 in half, from $0.20 to $0.10 per million tokens, which the company says reduces Sonnet 5.5’s cost on most agentic work by about 20%. A new monthly API credit for Claude Max and Team subscribers was announced alongside both changes.

What's Confirmed

Every figure in this article comes from Anthropic's own announcement at anthropic.com/news and the model card at anthropic.com/claude-haiku-5-5, cross-checked against the Claude Platform documentation at platform.claude.com/docs/en/models/overview. Nothing here is a rumor or a leak; it is Anthropic describing its own release.

Why the Price Cut Matters

The practical effect of a 75% average cost reduction is that workloads which were previously too expensive to run on every request, high-frequency classification, large-scale document summarization, or an always-on customer support bot, become cheap enough to run constantly. That shift also makes the subagent pattern more affordable: a company can let Opus 5.5 or Sonnet 5.5 act as the lead model on a coding task while delegating narrow lookups to Haiku 5.5 without the subagent calls adding meaningfully to the bill.

Cognition's Walden Yan described exactly this pattern in Anthropic's announcement, using Haiku 5.5 as a sidekick model in Devin Fusion alongside Opus 5.5 as the lead, a combination that held a top-tier FrontierCode score of 66.2. Rogo's Alex Wang said the company uses Haiku 5.5 as a subagent for pulling specific data points, such as 10-K segment revenue, out of longer documents.

Who's Affected

The announcement is aimed at developers building on the Claude API and enterprises already running Claude in production, not at the average person using claude.ai through the chat interface. Haiku-class models already sit behind parts of the free tier of claude.ai, so most consumer users will not notice an immediate change. The companies that do notice are the ones with high request volumes: customer support platforms, CRM tools, document-analysis products, and coding agents that lean on a cheap subagent for routine lookups.

Benchmark Performance

Anthropic published benchmark comparisons against Haiku 4.5, GPT-6 Luna, and Sonnet 5.5 (shown for reference as the tier above Haiku):

Benchmark Haiku 5.5 Haiku 4.5 GPT-6 Luna Sonnet 5.5
GDPval-AA v2.1 (knowledge work) 1620 735 1437 1840
OSWorld 2.1 (computer use, offline) 72.4% 15.7% 48.9% 83.9%
Humanity's Last Exam (no tools) 45.9% 10.2% — 56.9%
Terminal-Bench 4.0 (agentic coding) 39.2% 0.0% 16.4% 70.6%

The gap over Haiku 4.5 is largest on agentic and computer-use tasks, which lines up with Anthropic's positioning of Haiku 5.5 as a model meant to handle multi-step, tool-using work rather than just one-shot text generation.

What Changes for Developers

Switching to Haiku 5.5 means updating the model ID in API calls to claude-haiku-5-5. The effort parameter is new: Medium is the default, and raising it to High, Xhigh, or Max trades latency and cost for more thorough reasoning on harder subtasks, while Low favors speed. Anthropic notes that prompts up to 100,000 tokens cover roughly 90% of requests that previously went to Haiku 4.5, which is why the pricing below is split at that threshold.

What It Costs

Pricing is per 1 million tokens, split by prompts up to 100K tokens and prompts over 100K tokens, compared with Haiku 4.5 and Sonnet 5.5:

Token type Haiku 5.5 (≤100K) Haiku 5.5 (>100K) Haiku 4.5 Sonnet 5.5
Cache reads $0.01 $0.05 $0.10 $0.10
Cache writes $0.125 $0.625 $1.25 $2.50
Input tokens $0.10 $0.50 $1.00 $2.00
Output tokens $0.50 $2.50 $5.00 $10.00

Sonnet 5.5's cache-read price was also cut in half as part of the same announcement, from $0.20 to $0.10 per million tokens, which Anthropic says lowers Sonnet 5.5's cost on most agentic work by around 20%.

Where It's Available

Claude Haiku 5.5 is available now through the Claude Platform, Amazon Web Services, Google Cloud, and Microsoft Azure, in the same regions Anthropic already supports for its other current models. Anthropic has not announced India-specific pricing or availability differences, this is a global API and cloud-platform release rather than a region-specific one.

Early Customer Results

Anthropic's announcement included results from several companies already running Haiku 5.5 in production. Asana's Aaron Vinh, a staff software engineer, reported a 30%-plus latency reduction and up to 2.5x faster inference per agent turn. HubSpot's Ze'ev Klapow said the model posted the company's best-ever score on its CRM evaluation suite, at 92.8%. AlphaSense's Daniel Campos measured 0.84 versus 0.76 over Haiku 4.5 on document question-answering. Box's Yashodha Bhavnani saw an 11-point higher score at roughly half the latency compared with Haiku 4.5.

Safety Changes

Anthropic reports major alignment improvements over Haiku 4.5 in its own evaluations, including fewer instances of misaligned behavior and lower willingness to cooperate with misuse. Cybersecurity safeguards are more restrictive than Haiku 4.5's but less restrictive than Sonnet 5.5 or Opus-class models; the model still blocks penetration testing and attacker-style techniques. Biology safeguards sit at the same tier as Sonnet 5, Sonnet 5.5, and Opus 5.

What's Next

Anthropic has not announced a specific roadmap item tied to Haiku 5.5 beyond this release. The company has not said whether a dedicated claude.ai consumer update is planned, so for now the model's reach is primarily through the API and the three cloud platforms listed above.

Frequently asked questions

What is Claude Haiku 5.5?

Claude Haiku 5.5 is Anthropic's newest small model, announced on October 7, 2026. Anthropic describes it as the cheapest, fastest, and most capable small model the company has ever released, built for high-volume tasks like summaries, classification, customer support, browser use, and as a subagent alongside Opus 5.5 and Sonnet 5.5 on coding work.

How much does Claude Haiku 5.5 cost compared to Haiku 4.5?

Haiku 5.5 is roughly 75% cheaper to run on average than Haiku 4.5. For prompts up to 100K tokens, input is $0.10 per million tokens and output is $0.50 per million tokens, versus $1.00 and $5.00 per million on Haiku 4.5.

What is Claude Haiku 5.5's context window?

Claude Haiku 5.5 supports a 1,000,000 token context window, up from 200,000 tokens on Haiku 4.5, and a maximum output of 128,000 tokens, up from 64,000. Its reliable knowledge cutoff is June 2026.

Does Claude Haiku 5.5 have adjustable effort levels?

Yes. Haiku 5.5 is the first Haiku-class model with an adjustable effort setting, with levels of Low, Medium, High, Xhigh, and Max. The default is Medium. Thinking is adaptive and on by default, and can be disabled at the Low, Medium, or High effort levels.

Is Claude Haiku 5.5 available on the free tier of claude.ai?

Anthropic has not announced a separate claude.ai consumer rollout for Haiku 5.5. Haiku-class models already power parts of the free tier of claude.ai, but the October 7 announcement is focused on the API and cloud-platform release rather than a chat-interface update, so most claude.ai users will not notice an immediate difference.

What is the API model ID for Claude Haiku 5.5?

The API model ID is claude-haiku-5-5. It is available now on the Claude Platform, Amazon Web Services, Google Cloud, and Microsoft Azure.

Related articles

Comments