Quick Facts

ProductClaude Haiku 5.5 (Anthropic)
CategorySmall frontier AI model
Best forHigh-volume, repetitive work: classification, summaries, extraction, support, voice agents, subagents
AccessClaude Platform API, Amazon Web Services, Google Cloud, Microsoft Foundry; Claude.ai and Claude Code
Model IDclaude-haiku-5-5
Context window1 million tokens (up from 200K on Haiku 4.5), per Anthropic
Price$0.10/$0.50 per 1M tokens under 100K-token prompts; $0.50/$2.50 above; cache reads $0.01
Main limitationLaunch-day model: no track record, and the tokenizer change inflates token counts ~30%

What Claude Haiku 5.5 is, and why Anthropic launched it now

Claude Haiku 5.5 arrived on October 7, 2026, completing the Claude 5.5 family after Opus 5.5 and Sonnet 5.5. Anthropic calls it the cheapest, fastest, and most capable small model it has ever released, built for high-volume, cost-sensitive work: summaries, classification, database queries, context compaction, live support, voice agents, and browser use.

The timing is not subtle. Reuters reports the launch comes as Anthropic expands its lineup ahead of a planned IPO, and it lands two weeks after OpenAI's GPT-6 Luna claimed the same $0.10/$0.50 price tier. The small-model segment has become the industry's price war, and Haiku 5.5 is Anthropic's counterpunch.

The Artificial Analysis numbers: intelligence per dollar

Artificial Analysis, the independent benchmark outfit, had Haiku 5.5 indexed on launch day across its effort settings. The Intelligence Index runs from 29 at low effort to 43 at max effort, against a median of 13 for reasoning models in the same price tier. Even the cheapest setting scores more than double the tier median.

Speed is where the effort dial shows its cost. Low effort generates 181 tokens a second with a 6.3-second time to first token; max effort reaches 243 tokens a second once it starts, but takes over 400 seconds to produce the first token. Max effort is a batch-compute mode, not a chat mode.

Effort settingIntelligence IndexOutput speedTime to first token
Low29181.2 t/s6.28 s
Medium34136.8 t/s12.63 s
Xhigh41188.3 t/s87.15 s
Max43243.4 t/s415.30 s
  • The high-effort setting scores 38 on the Index; all five settings sit well above the tier median of 13.
  • At a blended rate (7:2:1 cache-hit/input/output), Artificial Analysis puts the cost at $0.08 per million tokens.

Anthropic's own benchmarks: the honest read

Anthropic's launch figures show the biggest gains where Haiku used to be weakest. On the OSWorld 2.1 computer-use test it scores 72.4%, up from 15.7% for Haiku 4.5. On Terminal-Bench 4.0, an agentic coding test, it reaches 39.2% where its predecessor scored zero. FrontierCode 1.1 comes in at 46.4%, and Humanity's Last Exam at 45.9% without tools, rising to 57.4% with them.

Vendor benchmarks deserve the standard discount: they run in the vendor's own harness. The useful signal is the independent one, and Artificial Analysis confirmed the direction on launch day. Read the exact percentages as evidence, not promises.

The price cut, with the fine print

For prompts under 100,000 tokens, Haiku 5.5 costs $0.10 per million input tokens and $0.50 per million output tokens, with cache reads at $0.01. Anthropic says about 90% of Haiku 4.5 requests fell under that threshold, which is why it can claim roughly 90% lower costs for typical use and about 75% lower on average.

The fine print matters. Above 100,000 tokens the price quintuples to $0.50/$2.50. And the new tokenizer counts roughly 30% more tokens for the same text, according to third-party analysis of the launch, so a 90%-cheaper headline shrinks once you re-estimate with the new token counts. Prompt caching (up to 90% off cached input, per Anthropic) and batch processing (50% off) are where the real savings compound.

Haiku 5.5 vs GPT-6 Luna vs the rest of the small-model tier

The head-to-head that matters is GPT-6 Luna, OpenAI's small model launched September 22, 2026 at the identical $0.10/$0.50 tier. On Artificial Analysis's Intelligence Index, Haiku 5.5 leads at both ends: 29 vs 22 at low effort, 43 vs 38 at max effort. On the OSWorld computer-use test, Anthropic reports 72.4% for Haiku 5.5 against 48.9% reported for GPT-6 Luna.

Luna's advantage is distribution: it already reaches free-tier users through ChatGPT's desktop app, while Haiku 5.5 meets consumers through Claude.ai subscriptions. For developers the APIs are priced identically, so the choice comes down to benchmark profile and ecosystem, not price.

ModelAPI price (in / out)AA Index (low / max)OSWorld 2.1
Claude Haiku 5.5$0.10 / $0.5029 / 4372.4% (Anthropic)
GPT-6 Luna$0.10 / $0.5022 / 3848.9% (reported)
Haiku 4.5 (legacy)$1.00 / $5.00Not listed15.7% (Anthropic)
  • Mistral's ML4 sits at $0.68/$2.09 per million tokens, per industry reporting: roughly seven times Haiku 5.5's input price.
  • Open-weights models average $0.83 per million tokens on OpenRouter, per the same analysis. Haiku 5.5 undercuts the open average while beating its tier's intelligence median.

Who Haiku 5.5 is for

  • Teams running classification, extraction, summarization, or support at high volume, where per-call cost dominates.
  • Agent builders who want a cheap, fast worker model under a stronger planner like Sonnet 5.5 or Opus 5.5.
  • Anyone who burned money on Haiku 4.5: the same workloads cost roughly three-quarters less, per Anthropic's math.
  • Not for you if your prompts routinely exceed 100K tokens, if you need the strongest reasoning available, or if you want a consumer product with a signup bonus.

What we still do not know

  • There is no independent track record: launch-day benchmarks are the entire evidence base so far.
  • Real-world latency and reliability at scale are unproven; Artificial Analysis's figures are single-endpoint measurements.
  • The ~30% tokenizer inflation figure comes from third-party analysis, not Anthropic. Verify it against your own workloads.
  • Anthropic's 75%-cheaper average is a vendor estimate built on its own usage mix.
  • Long-term pricing is not promised; the 100K-token tier boundary could move.

Why We Like Claude Haiku 5.5

  • The cheapest serious intelligence per dollar right now: Artificial Analysis scores Haiku 5.5 from 29 (low effort) to 43 (max effort) on its Intelligence Index, against a tier median of 13. Nothing in its price band scores meaningfully higher.
  • Genuinely fast at default settings: 181 tokens a second at low effort with a 6.3-second time to first token, per Artificial Analysis. Higher effort settings trade that speed for reasoning; pick per task.
  • A real step up where it used to be weak: Anthropic reports 72.4% on the OSWorld computer-use test (Haiku 4.5 managed 15.7%) and 39.2% on Terminal-Bench agentic coding, where its predecessor scored zero.
  • Built to be a subagent: Anthropic positions Haiku as the worker bee under Sonnet 5.5 and Opus 5.5: the big model plans, Haiku fetches, classifies, and summarizes at scale.

What to Watch Out For

  • Prices quintuple past 100K tokens.
  • The new tokenizer counts about 30% more tokens for the same text, so re-estimate budgets (third-party analysis).
  • Max-effort mode is slow: over 400 seconds to first token in Artificial Analysis testing.
  • No public referral code or signup bonus exists.
  • Launch-day benchmarks are vendor-reported except where Artificial Analysis independently tested.

Official access

Claude Haiku 5.5 access

There is no public Claude Haiku 5.5 referral code or signup bonus. Anthropic publishes no referral program or public signup offer for its API models. Start here: https://financeappradar.com/go/claude-haiku-5-5.

Last checked: October 7, 2026. Terms change; the company's page wins.

Current Claude Haiku 5.5 Access

Official model pagehttps://financeappradar.com/go/claude-haiku-5-5
Offer statusNo public referral code or signup bonus. Anthropic announced monthly API credits for Claude Max and Team subscribers alongside the launch.
RequirementSign up on the Claude Platform or use an existing Claude subscription; there is no code to enter.
Payout timelineAvailable now; Haiku 4.5 remains listed as a legacy model.
Important limitationAPI pricing is tiered by prompt length and may change; confirm current rates on Anthropic's pricing page.

Read Anthropic's announcement

How Claude Haiku 5.5 Works

  1. Haiku 5.5 is a model, not a chatbot you download. Developers call it through the Claude Platform API with the model ID claude-haiku-5-5, or through Amazon Web Services, Google Cloud, or Microsoft Foundry.
  2. Everyday users meet it inside Claude.ai: Anthropic says Free, Pro, Max, Team, and Enterprise users can select Haiku 5.5 on web, iOS, and Android. It is also available in Claude Code.
  3. The new control is the effort setting: low, medium, high, xhigh, or max. Low effort answers fast and cheap; max effort thinks longer before answering. Same model, different trade of cost against intelligence.
  4. Pricing is tiered by prompt length. Prompts up to 100,000 tokens cost $0.10 per million input tokens and $0.50 per million output tokens. Longer prompts cost five times that. Prompt caching can cut cached input costs by up to 90%, per Anthropic.
  5. Anthropic positions Haiku as a subagent: pair it with Sonnet 5.5 or Opus 5.5, let the bigger model do the hard reasoning, and let Haiku handle the repetitive steps at scale.

Is Claude Haiku 5.5 Safe?

Anthropic says Haiku 5.5 is its first small model with built-in safeguards for a narrow set of high-risk cybersecurity requests, while its biology protections match Sonnet 5, Sonnet 5.5, and Opus 5. Those are provider-described controls, not a guarantee. The practical risks for a builder are more mundane: API keys that leak, prompts that include customer data without a retention policy, and agents given tools they do not need. Treat Haiku like any production dependency: least-privilege keys, spending limits, logging, and a human in the loop before it sends, spends, or publishes anything.

Who Should Use Claude Haiku 5.5

Claude Haiku 5.5 may be a good fit if you:

  • Run the same AI task at high volume and need the per-call cost near zero.
  • Build agents where a bigger model plans and a small model does the repetitive steps.
  • Want independent benchmark data on day one instead of marketing claims alone.

Who Should Skip Claude Haiku 5.5

You may want to skip it if you:

  • Need one model for the hardest reasoning work; Sonnet 5.5 and Opus 5.5 are Anthropic's picks for that.
  • Run mostly long prompts over 100K tokens, where the cheaper tier does not apply.
  • Want a consumer chatbot with a signup bonus; this is a developer API product.

Final Verdict: Is Claude Haiku 5.5 Worth It?

If you run AI at volume, Haiku 5.5 is the new default to beat: independent scores well above its price tier, real speed at low effort, and a price cut that holds up as long as you re-estimate for the new tokenizer and the 100K-token tier. Test it on your own workload before migrating anything that matters, keep spending limits on, and let the benchmarks inform the decision instead of making it.

Claude Haiku 5.5 FAQ

What is Claude Haiku 5.5?

Anthropic's small, fast, cheap AI model, launched October 7, 2026. It is designed for high-volume, repetitive work and as a subagent under bigger models. API model ID: claude-haiku-5-5.

How much does Claude Haiku 5.5 cost?

$0.10 per million input tokens and $0.50 per million output tokens for prompts under 100,000 tokens; $0.50 and $2.50 above that. Cache reads cost $0.01 per million. Anthropic says typical workloads run about 75% cheaper than on Haiku 4.5.

How smart is Claude Haiku 5.5?

Artificial Analysis scores it 29 to 43 on its Intelligence Index depending on the effort setting, against a tier median of 13. Anthropic reports 72.4% on OSWorld 2.1 and 39.2% on Terminal-Bench 4.0.

Is Claude Haiku 5.5 better than GPT-6 Luna?

On the published numbers, Haiku 5.5 leads GPT-6 Luna on Artificial Analysis's Intelligence Index (29/43 vs 22/38 at low/max effort) and on the OSWorld computer-use test. Both cost the same per token. Luna reaches more consumers through ChatGPT's desktop app.

What is the effort setting?

A new control for Haiku: low, medium, high, xhigh, or max. Higher effort means more reasoning before answering, at the cost of latency. Max effort is a batch mode, not for interactive chat.

Is there a Claude Haiku 5.5 referral code?

No. Anthropic publishes no referral program or signup bonus for its API models.

What is the current Claude Haiku 5.5 access status?

No public referral code or signup bonus. Anthropic announced monthly API credits for Claude Max and Team subscribers alongside the launch. Sign up on the Claude Platform or use an existing Claude subscription; there is no code to enter.

Is Claude Haiku 5.5 legit?

Yes. Anthropic announced Claude Haiku 5.5 on October 7, 2026 as the third model in its Claude 5.5 family. It is a real, publicly available model on the Claude Platform API, AWS, Google Cloud, and Microsoft Foundry, and selectable on Claude.ai.

Keep researching Claude Haiku 5.5

More ai agents coverage, plus the hubs and guides that put Claude Haiku 5.5 in context.

Sources checked

Checked October 7, 2026.