Skip to content
Merged
Show file tree
Hide file tree
Changes from all commits
Commits
File filter

Filter by extension

Filter by extension

Conversations
Failed to load comments.
Loading
Jump to
Jump to file
Failed to load files.
Loading
Diff view
Diff view
11 changes: 11 additions & 0 deletions blog/src/_data/media.json
Original file line number Diff line number Diff line change
Expand Up @@ -4610,5 +4610,16 @@
"status": "ready"
},
"inline": {}
},
"claude-haiku-5-5": {
"title": "Claude Haiku 5.5: price, benchmarks and where it fits",
"category": "concepts",
"hero": {
"file": "assets/media/claude-haiku-5-5/hero.png",
"alt": "Title card on white. Claude Haiku 5.5 in large ink type over a yellow underline, a Concepts kicker, the line Checked 8 Oct 2026, and a row of coloured dominoes part way through falling",
"prompt": "Still of the hero scene in assets/media/claude-haiku-5-5/motion.js.",
"status": "ready"
},
"inline": {}
}
}
Binary file added blog/src/assets/media/claude-haiku-5-5/cups.png
Loading
Sorry, something went wrong. Reload?
Sorry, we cannot display this file.
Sorry, this file is invalid so it cannot be displayed.
Binary file added blog/src/assets/media/claude-haiku-5-5/hero.png
Loading
Sorry, something went wrong. Reload?
Sorry, we cannot display this file.
Sorry, this file is invalid so it cannot be displayed.
13 changes: 13 additions & 0 deletions blog/src/assets/media/claude-haiku-5-5/motion.css
Original file line number Diff line number Diff line change
@@ -0,0 +1,13 @@
/* Bright coded scenes for "Claude Haiku 5.5". Pairs with motion.js. */
.prose figure.mg { margin: 42px 0; }
.prose figure.mg img, .mg .mg-svg { display: block; width: 100%; height: auto; border: 1px solid #1A1320; border-radius: 18px; background: #fff; }
.mg figcaption { font: 500 13px/1.5 "Space Grotesk", system-ui, sans-serif; color: #6B5878; padding-top: 10px; text-align: center; }
.mg .mg-svg { touch-action: manipulation; -webkit-tap-highlight-color: transparent; }
.mg .mg-svg text { -webkit-user-select: none; user-select: none; pointer-events: none; }
.mg .mg-hit { cursor: pointer; outline: none; }
.mg [role="slider"] { cursor: grab; touch-action: none; }
.mg .mg-ring { stroke-opacity: 0; transition: stroke-opacity 0.15s; }
.mg .mg-hit:focus-visible .mg-ring { stroke-opacity: 1; }
.mg [role="slider"]:focus-visible circle:nth-child(2) { stroke-width: 5; }
@media (hover: hover) { .mg .mg-hit:hover .mg-ring { stroke-opacity: 0.5; } }
@media (prefers-reduced-motion: reduce) { .mg .mg-ring { transition: none; } }
317 changes: 317 additions & 0 deletions blog/src/assets/media/claude-haiku-5-5/motion.js

Large diffs are not rendered by default.

Loading
Sorry, something went wrong. Reload?
Sorry, we cannot display this file.
Sorry, this file is invalid so it cannot be displayed.
Loading
Sorry, something went wrong. Reload?
Sorry, we cannot display this file.
Sorry, this file is invalid so it cannot be displayed.
Loading
Sorry, something went wrong. Reload?
Sorry, we cannot display this file.
Sorry, this file is invalid so it cannot be displayed.
Original file line number Diff line number Diff line change
Expand Up @@ -79,7 +79,7 @@ You can also force it: name the subagent in plain words, @-mention it with `@age

## How do you set a subagent's model and tools?

Use `model` and `tools` in the front matter. The model is resolved in this order: a model Claude passes when it spawns the subagent, then the `model` field, then the `CLAUDE_CODE_SUBAGENT_MODEL` variable, then your session's model. Since 2.1.251 that variable is only a default. To force one model on every subagent, also set `CLAUDE_CODE_SUBAGENT_MODEL_FORCE=1` (2.1.257 and later). Run `/tasks` to see which model each subagent is on.
Use `model` and `tools` in the front matter. The model is resolved in this order: a model Claude passes when it spawns the subagent, then the `model` field, then the `CLAUDE_CODE_SUBAGENT_MODEL` variable, then your session's model. Since 2.1.251 that variable is only a default. To force one model on every subagent, also set `CLAUDE_CODE_SUBAGENT_MODEL_FORCE=1` (2.1.257 and later). Run `/tasks` to see which model each subagent is on. The newest Haiku is [Claude Haiku 5.5](/blog/claude-haiku-5-5/).

`tools` is an allowlist and `disallowedTools` a denylist. A `disallowedTools` entry such as `Bash(git push *)` removes the whole Bash tool, so block single commands with a deny rule in settings instead. The Task tool was renamed Agent in 2.1.63, and old `Task(...)` rules still work.

Expand Down
112 changes: 112 additions & 0 deletions blog/src/posts/claude-haiku-5-5.md
Original file line number Diff line number Diff line change
@@ -0,0 +1,112 @@
---
title: "Claude Haiku 5.5: price, benchmarks and where it fits"
description: "Claude Haiku 5.5 is Anthropic's new small model. Its tiered price, Anthropic's own scores and how to pick it in Claude Code, checked 8 Oct 2026."
date: 2026-10-08
category: concepts
categoryLabel: Concepts
type: Non-technical
primaryKeyword: "claude haiku 5.5"
secondaryKeywords: ["claude haiku", "haiku 5.5", "claude haiku 5.5 pricing", "claude haiku 5.5 vs haiku 4.5", "claude haiku 5.5 claude code", "claude-haiku-5-5"]
tags: ["Concepts", "Claude Code", "AI Agents"]
faq:
- q: "How much does Claude Haiku 5.5 cost?"
a: "Anthropic's announcement of 7 October 2026 lists $0.10 input and $0.50 output per million tokens for prompts up to 100,000 tokens, and $0.50 input and $2.50 output for prompts over 100,000 tokens. Cache reads are $0.01 and $0.05. Checked 8 Oct 2026."
- q: "Is Claude Haiku 5.5 better than Haiku 4.5?"
a: "On Anthropic's own figures, yes on every row where both have a score. For example, the announcement lists 72.4% for Haiku 5.5 and 15.7% for Haiku 4.5 on the offline subset of OSWorld 2.1. We did not run these tests."
- q: "Is Claude Haiku 5.5 better than Sonnet 5.5?"
a: "No. Sonnet 5.5 is ahead on every row of Anthropic's own benchmark table, and the announcement says Sonnet 5.5 and Opus 5.5 remain better choices for complex agentic coding tasks. Haiku 5.5 is the cheaper and faster option."
- q: "What is the model ID for Claude Haiku 5.5?"
a: "It is claude-haiku-5-5 on the Claude API, Google Cloud, Microsoft Foundry and Claude Platform on AWS, and anthropic.claude-haiku-5-5 on Amazon Bedrock, according to Anthropic's model page read on 8 Oct 2026."
- q: "How do I use Claude Haiku 5.5 in Claude Code?"
a: "Claude Code's model docs say to run /model claude-haiku-5-5 in a session, or start with claude --model claude-haiku-5-5. The same page says to use Claude Code v2.1.293 or later with Haiku 5.5. Read on 8 Oct 2026."
---

Claude Haiku 5.5 is Anthropic's new small model: much cheaper than Haiku 4.5 on prompts up to 100,000 tokens, and behind Sonnet 5.5 on every row of Anthropic's own benchmark table. Anthropic released it on 7 October 2026. We have not tried it. Checked 8 Oct 2026.

[Munder Difflin](https://harnessmd.com/download) is free and open source: a desktop app that runs a team of coding agents such as Claude Code, Codex and Gemini CLI on your own computer. A cheaper small model matters when several agents run at once, which is why we cover it. The [install guide](/blog/how-to-install-and-use-munder-difflin/) covers setup and the [Concepts hub](/blog/topics/concepts/) explains the terms.

## What is Claude Haiku 5.5?

It is Anthropic's new small model, with the model ID `claude-haiku-5-5`. Anthropic's [announcement](https://www.anthropic.com/claude-haiku-5-5), dated October 7, 2026, calls it "the cheapest, fastest, and most capable small model we've ever released" and says it is designed for high-volume, cost-sensitive tasks.

The speed claim has a footnote. Anthropic says Haiku 5.5 is its fastest model to date at each model's standard speed, although it runs less quickly than its Opus models in Fast Mode.

Anthropic's [model page](https://platform.claude.com/docs/en/models/haiku-5-5/overview), read on 8 Oct 2026, lists a 1M token context window and a maximum output of 128K tokens. The announcement adds that this is the first Haiku-class model with an adjustable effort setting.

The [Hacker News thread](https://news.ycombinator.com/item?id=49996437) had more than 300 points and 140 comments when we checked on 8 Oct 2026.

<figure class="mg" data-scene="stopwatch"><img src="/blog/assets/media/claude-haiku-5-5/stopwatch.png" width="1600" height="1200" loading="lazy" decoding="async" alt="Animation. A stopwatch labelled Haiku 5.5 starts, its hand sweeps round and stops. Two paper labels slide in beside it: one reads fastest to date at standard speed, the other reads Opus in Fast Mode is quicker."><figcaption>Anthropic calls Haiku 5.5 its fastest model to date at standard speed, with Opus in Fast Mode quicker. Announcement of 7 October 2026.</figcaption></figure>

## What does Claude Haiku 5.5 cost?

It costs $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100,000 tokens, as of 8 Oct 2026, and five times that above. These are the prices in the table in Anthropic's [announcement](https://www.anthropic.com/claude-haiku-5-5), per 1 million tokens.

| Price per 1M tokens | Haiku 5.5, prompts up to 100k | Haiku 5.5, prompts over 100k | Haiku 4.5 | Sonnet 5.5 |
| --- | --- | --- | --- | --- |
| Input | $0.10 | $0.50 | $1.00 | $2.00 |
| Output | $0.50 | $2.50 | $5.00 | $10.00 |
| Cache writes | $0.125 | $0.625 | $1.25 | $2.50 |
| Cache reads | $0.01 | $0.05 | $0.10 | $0.10 |

Anthropic says prompts up to 100,000 tokens made up around 90% of requests to its previous Haiku model. Its headline is that Haiku 5.5 costs around 75% less to run on average. A footnote explains the gap: the price is 90% lower for requests up to 100,000 tokens, less for larger ones, and an updated tokenizer uses slightly more tokens per task. The model page says the same text counts as approximately 30% more tokens than on Haiku 4.5.

The model page also lists a 50% Batch API discount on input and output. For the wider bill, see [how much Claude Code costs](/blog/how-much-does-claude-code-cost/).

<figure class="mg" data-scene="cups"><img src="/blog/assets/media/claude-haiku-5-5/cups.png" width="1600" height="1200" loading="lazy" decoding="async" alt="Animation. A tap fills a small measuring cup, and a card under it headed prompts up to 100k tokens shows $0.10 input and $0.50 output. The water spills over the lip into a larger cup, and its card headed prompts over 100k tokens shows $0.50 input and $2.50 output."><figcaption>Anthropic's prices per 1 million tokens, in two tiers by prompt size. Checked 8 Oct 2026.</figcaption></figure>

## How does Claude Haiku 5.5 score on benchmarks?

On Anthropic's own figures it is far ahead of Haiku 4.5 and behind Sonnet 5.5. The five rows below are our selection from the eight rows in the [announcement](https://www.anthropic.com/claude-haiku-5-5) table. We did not run them. NR means not reported.

| Benchmark (as named by Anthropic) | Haiku 5.5 | Haiku 4.5 | GPT-6 Luna | Sonnet 5.5 |
| --- | --- | --- | --- | --- |
| GDPval-AA v2.1 | 1620 | 735 | 1437 | 1840 |
| OSWorld 2.1, offline subset | 72.4% | 15.7% | 48.9% | 83.9% |
| Humanity's Last Exam, no tools | 45.9% | 10.2% | NR | 56.9% |
| Terminal-Bench 4.0 | 39.2% | 0.0% | 16.4% | 70.6% |
| FrontierCode 1.1 (Main) | 46.4% | NR | 42.4% | 52.1% |

Sonnet 5.5 leads on all eight rows of the full table. Anthropic marks its FrontierCode figure for Sonnet 5.5 as Xhigh effort. Haiku 5.5 is ahead of GPT-6 Luna on all six rows where Anthropic reports both.

The coding gap is the one to notice. On Terminal-Bench 4.0, Haiku 5.5 is at 39.2% against 70.6% for Sonnet 5.5. Anthropic says so itself: "Sonnet 5.5 and Opus 5.5 remain better choices for complex agentic coding tasks like those measured by Terminal-Bench 4.0."

<figure class="mg" data-scene="planes"><img src="/blog/assets/media/claude-haiku-5-5/planes.png" width="1600" height="1200" loading="lazy" decoding="async" alt="Animation. Four paper planes for Haiku 5.5, Haiku 4.5, GPT-6 Luna and Sonnet 5.5 glide along a ruled strip and land at distances set by their scores on one benchmark at a time. Arrow buttons step through the five benchmarks."><figcaption>Anthropic's own figures, five rows we selected from its table of 7 October 2026. We did not run them.</figcaption></figure>

## Where can you use Claude Haiku 5.5?

Anthropic says it "is available now on all platforms, including Amazon Web Services, Google Cloud, and Microsoft Azure". The [model page](https://platform.claude.com/docs/en/models/haiku-5-5/overview) lists the Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. The ID is `claude-haiku-5-5`, except on Amazon Bedrock, where it is `anthropic.claude-haiku-5-5`.

Neither page says which Claude app plans include it, so we leave that out.

## How do you use Claude Haiku 5.5 in Claude Code?

Run `/model claude-haiku-5-5` in a session, or start with `claude --model claude-haiku-5-5`. That wording is from Claude Code's [model configuration docs](https://code.claude.com/docs/en/model-config), read on 8 Oct 2026. The same page says to use v2.1.293 or later with Haiku 5.5.

Three details from the docs:

* **The alias depends on your provider.** `haiku` resolves to Haiku 5.5 on the Anthropic API. On Claude Platform on AWS, Amazon Bedrock, Google Cloud's Agent Platform and Microsoft Foundry it resolves to Haiku 4.5.
* **Long prompts cost more.** The docs repeat that a Haiku 5.5 request costs more per token when its prompt is longer than 100K tokens.
* **Subagents take a model.** The [subagent docs](https://code.claude.com/docs/en/sub-agents) say the `model` field accepts an alias such as `haiku`, a full model ID or `inherit`. `CLAUDE_CODE_SUBAGENT_MODEL` sets a default.

Our post on [Claude Code subagents](/blog/claude-code-subagents-vs-multi-agent-harness/) walks through that file.

<figure class="mg" data-scene="postcards"><img src="/blog/assets/media/claude-haiku-5-5/postcards.png" width="1600" height="1200" loading="lazy" decoding="async" alt="Animation. A postcard rack turns slowly. Each card names a provider, and as it faces front it flips to show which model the haiku alias gives: Haiku 5.5 on the Anthropic API card, Haiku 4.5 on the other cards."><figcaption>In Claude Code the haiku alias resolves to Haiku 5.5 on the Anthropic API and to Haiku 4.5 on the other listed providers. Docs read 8 Oct 2026.</figcaption></figure>

## What does it mean if you run a team of coding agents?

We think it makes the split between a planner and its helpers cheaper to try. Anthropic says Haiku 5.5 "pairs well with Opus 5.5 and Sonnet 5.5 as a subagent on coding work", and that it suits narrowly scoped tasks like compaction, summarization or subagent work.

Two customer quotes from the announcement, the first on subagents and the second on speed. Rogo: "The short and high-volume work is where Claude Haiku 5.5 fits for us, like quick lookups, subagents, and summaries." Asana: "Compared with the model we use today, we saw over a 30% reduction in latency for task completions and up to 2.5x faster inference per agent turn."

Our view: put a bigger model on planning and hard debugging, and a small fast one on searching, summarising and running tests. The small one fetches, the big one decides. This is an idea, not a test. We have not tried Haiku 5.5. Our piece on [model routing](/blog/do-more-with-less-model-routing/) explains the reasoning.

## What should you do now?

Our view: try it on routine work first, and keep your bigger model where the task is hard.

1. Update Claude Code to v2.1.293 or later.
2. Set `model: haiku` on one subagent that only searches or summarises.
3. Watch prompt size. The lower price applies up to 100,000 tokens.
4. Compare results on your own repo before trusting any table.

<link rel="stylesheet" href="/blog/assets/media/claude-haiku-5-5/motion.css"><script defer src="/blog/assets/media/claude-haiku-5-5/motion.js"></script>
2 changes: 1 addition & 1 deletion docs/blog/agent-tools-today-2026-10-07/index.html
Original file line number Diff line number Diff line change
Expand Up @@ -241,7 +241,7 @@ <h2 id="how-this-list-is-built" tabindex="-1">How this list is built <a class="a

<footer class="post-foot wrap"><div class="prevnext">
<a class="pn prev" href="/blog/codex-vs-cursor/"><span class="k">← Older</span><span class="t">Codex vs Cursor: Which Should You Pick in 2026?</span></a>
<span></span>
<a class="pn next" href="/blog/claude-haiku-5-5/"><span class="k">Newer →</span><span class="t">Claude Haiku 5.5: price, benchmarks and where it fits</span></a>
</div>
<section class="related">
<h2>Related in News</h2>
Expand Down
28 changes: 14 additions & 14 deletions docs/blog/agents-rule-of-two/index.html
Original file line number Diff line number Diff line change
Expand Up @@ -331,50 +331,50 @@ <h2 id="faq">FAQ</h2>
<h2>Related in Concepts</h2>
<div class="post-grid">
<article class="card t-concepts">
<a class="thumb t-concepts" href="/blog/openai-decisions-api/" aria-hidden="true" tabindex="-1"><img src="/blog/assets/media/openai-decisions-api/hero.png" alt="" loading="lazy" decoding="async" /></a>
<a class="thumb t-concepts" href="/blog/claude-haiku-5-5/" aria-hidden="true" tabindex="-1"><img src="/blog/assets/media/claude-haiku-5-5/hero.png" alt="" loading="lazy" decoding="async" /></a>
<div class="body">
<div class="meta">
<time datetime="2026-10-07">Oct 7, 2026</time>
<time datetime="2026-10-08">Oct 8, 2026</time>
<span class="dot" aria-hidden="true"></span>
<span>5 min</span>
</div>
<h3><a href="/blog/openai-decisions-api/">OpenAI Decisions API: what it is, price and limits</a></h3>
<p class="dek">The OpenAI Decisions API is a beta endpoint that returns a probability, a pick or a score. Price, limits and how to call it, checked 7 Oct 2026.</p>
<h3><a href="/blog/claude-haiku-5-5/">Claude Haiku 5.5: price, benchmarks and where it fits</a></h3>
<p class="dek">Claude Haiku 5.5 is Anthropic&#39;s new small model. Its tiered price, Anthropic&#39;s own scores and how to pick it in Claude Code, checked 8 Oct 2026.</p>
<div class="foot">
<span class="tag kind">Concepts</span>
<a class="more" href="/blog/openai-decisions-api/" aria-label="Read: OpenAI Decisions API: what it is, price and limits">Read →</a>
<a class="more" href="/blog/claude-haiku-5-5/" aria-label="Read: Claude Haiku 5.5: price, benchmarks and where it fits">Read →</a>
</div>
</div>
</article>
<article class="card t-concepts">
<a class="thumb t-concepts" href="/blog/mistral-large-4/" aria-hidden="true" tabindex="-1"><img src="/blog/assets/media/mistral-large-4/hero.png" alt="" loading="lazy" decoding="async" /></a>
<a class="thumb t-concepts" href="/blog/openai-decisions-api/" aria-hidden="true" tabindex="-1"><img src="/blog/assets/media/openai-decisions-api/hero.png" alt="" loading="lazy" decoding="async" /></a>
<div class="body">
<div class="meta">
<time datetime="2026-10-06">Oct 6, 2026</time>
<time datetime="2026-10-07">Oct 7, 2026</time>
<span class="dot" aria-hidden="true"></span>
<span>5 min</span>
</div>
<h3><a href="/blog/mistral-large-4/">Mistral Large 4 (Le Chonk): size, price, scores and weights</a></h3>
<p class="dek">Mistral Large 4 is a 1 trillion parameter open weight model in public preview. Context window, API price, Mistral&#39;s own scores and weights, checked 6 Oct 2026.</p>
<h3><a href="/blog/openai-decisions-api/">OpenAI Decisions API: what it is, price and limits</a></h3>
<p class="dek">The OpenAI Decisions API is a beta endpoint that returns a probability, a pick or a score. Price, limits and how to call it, checked 7 Oct 2026.</p>
<div class="foot">
<span class="tag kind">Concepts</span>
<a class="more" href="/blog/mistral-large-4/" aria-label="Read: Mistral Large 4 (Le Chonk): size, price, scores and weights">Read →</a>
<a class="more" href="/blog/openai-decisions-api/" aria-label="Read: OpenAI Decisions API: what it is, price and limits">Read →</a>
</div>
</div>
</article>
<article class="card t-concepts">
<a class="thumb t-concepts" href="/blog/openai-text-watermarking/" aria-hidden="true" tabindex="-1"><img src="/blog/assets/media/openai-text-watermarking/hero.png" alt="" loading="lazy" decoding="async" /></a>
<a class="thumb t-concepts" href="/blog/mistral-large-4/" aria-hidden="true" tabindex="-1"><img src="/blog/assets/media/mistral-large-4/hero.png" alt="" loading="lazy" decoding="async" /></a>
<div class="body">
<div class="meta">
<time datetime="2026-10-06">Oct 6, 2026</time>
<span class="dot" aria-hidden="true"></span>
<span>5 min</span>
</div>
<h3><a href="/blog/openai-text-watermarking/">OpenAI text watermarking: textGrain, who gets it, limits</a></h3>
<p class="dek">OpenAI text watermarking (textGrain) is opt in for the API now and is coming to eligible ChatGPT and Codex text in the EU. Limits checked 6 Oct 2026.</p>
<h3><a href="/blog/mistral-large-4/">Mistral Large 4 (Le Chonk): size, price, scores and weights</a></h3>
<p class="dek">Mistral Large 4 is a 1 trillion parameter open weight model in public preview. Context window, API price, Mistral&#39;s own scores and weights, checked 6 Oct 2026.</p>
<div class="foot">
<span class="tag kind">Concepts</span>
<a class="more" href="/blog/openai-text-watermarking/" aria-label="Read: OpenAI text watermarking: textGrain, who gets it, limits">Read →</a>
<a class="more" href="/blog/mistral-large-4/" aria-label="Read: Mistral Large 4 (Le Chonk): size, price, scores and weights">Read →</a>
</div>
</div>
</article>
Expand Down
Loading
Loading