overview changedmodels/haiku-4-5/overview
Nearest release: v2.1.293, published an hour before this site recorded the change. Shown because the two are within 24 hours of each other. Nothing here says the release caused the edit.
Recorded here
Lines+23added
Lines−14removed
From line
1
where the diff opens
First seen
24 Aug 2026
this site's first read of the page
Recorded edits12to this page, all time
The whole hunk
from line 1, old and new numbered
/
from line 1
11---
22title: Claude Haiku 4.5
33url: https://platform.claude.com/docs/en/models/haiku-4-5/overview
4description: "Claude Haiku 4.5 at a glance: what it's for, model IDs on every platform, context window, output limits, pricing, availability, and the guides and resources for building with it."
4description: "Claude Haiku 4.5 reference: lifecycle status, model IDs on every platform, context window, output limits, pricing, and migration resources. Claude Haiku 5.5 is the current Haiku model."
55---
66
7**Latest.** Released October 15, 2025.
7**Legacy.** Released October 15, 2025.
88
99The fastest model with near-frontier intelligence
1010
11Although Claude Haiku 4.5 is still available, you should consider migrating to Claude Haiku 5.5 for improved performance. [See Claude Haiku 5.5](https://platform.claude.com/docs/en/models/haiku-5-5/overview) · [Migrate to Claude Haiku 5.5](https://platform.claude.com/docs/en/models/haiku-5-5/migration-guide)
12
1113Model ID: `claude-haiku-4-5-20251001`
1214
1315Context window: 200K tokens · Max output: 64K tokens · Input pricing: $1 / MTok · Output pricing: $5 / MTok
1416
15[Announcement](https://www.anthropic.com/news/claude-haiku-4-5) · [Migration guide](https://platform.claude.com/docs/en/models/haiku-4-5/migration-guide)
17[Announcement](https://www.anthropic.com/news/claude-haiku-4-5)
1618
1719## How it compares
1820
19| Model | Context | Max output | Price / MTok | Latency | Thinking | Default effort | Knowledge cutoff |
20| :---------------------------------------------------------------------------------- | :------ | :--------- | :----------- | :------- | :------------------- | :------------- | :--------------- |
21| [Claude Fable 5.1](https://platform.claude.com/docs/en/models/fable-5-1/overview) | 1M | 128K | $10 / $50 | Slower | Adaptive (always on) | `high` | Jun 2026 |
22| [Claude Opus 5.5](https://platform.claude.com/docs/en/models/opus-5-5/overview) | 1M | 128K | $4 / $20 | Moderate | Adaptive (always on) | `medium` | Jun 2026 |
23| [Claude Sonnet 5.5](https://platform.claude.com/docs/en/models/sonnet-5-5/overview) | 1M | 128K | $2 / $10 | Fast | Adaptive | `high` | Jun 2026 |
24| **Claude Haiku 4.5** (this model) | 200K | 64K | $1 / $5 | Fastest | Extended | — | Feb 2025 |
21| Model | Context | Max output | Price / MTok | Thinking | Default effort | Knowledge cutoff |
22| :---------------------------------------------------------------------------------- | :------ | :--------- | :----------------- | :------------------- | :------------- | :--------------- |
23| [Claude Fable 5.1](https://platform.claude.com/docs/en/models/fable-5-1/overview) | 1M | 128K | $10 / $50 | Adaptive (always on) | `high` | Jun 2026 |
24| [Claude Opus 5.5](https://platform.claude.com/docs/en/models/opus-5-5/overview) | 1M | 128K | $4 / $20 | Adaptive (always on) | `medium` | Jun 2026 |
25| [Claude Sonnet 5.5](https://platform.claude.com/docs/en/models/sonnet-5-5/overview) | 1M | 128K | $2 / $10 | Adaptive | `high` | Jun 2026 |
26| [Claude Haiku 5.5](https://platform.claude.com/docs/en/models/haiku-5-5/overview) | 1M | 128K | From $0.10 / $0.50 | Adaptive | `medium` | Jun 2026 |
27| **Claude Haiku 4.5** (this model) | 200K | 64K | $1 / $5 | Extended | — | Feb 2025 |
2528
2629* **Context:** 1M tokens is roughly 555k words or 2.5M Unicode characters on the current tokenizer (introduced with Claude Opus 4.7); models before it fit about 750k words in 1M tokens. 200k tokens is roughly 150k words.
27* **Max output:** Synchronous Messages API limit. On the Message Batches API, Claude Opus 5.5, Claude Opus 5, Claude Sonnet 5.5, Claude Sonnet 5, Claude Opus 4.8, Claude Opus 4.7, Claude Opus 4.6, and Claude Sonnet 4.6 support up to 300k output tokens with the output-300k-2026-03-24 beta header.
30* **Max output:** Synchronous Messages API limit. On the Message Batches API, Claude Opus 5.5, Claude Opus 5, Claude Sonnet 5.5, Claude Sonnet 5, Claude Haiku 5.5, Claude Opus 4.8, Claude Opus 4.7, Claude Opus 4.6, and Claude Sonnet 4.6 support up to 300k output tokens with the output-300k-2026-03-24 beta header.
2831* **Price / MTok:** Input / output, base price per million tokens. Batch API requests are 50% off; prompt caching reads cost 10% of the base input price (2.5% on Claude Fable 5.1 and Claude Mythos 5.1, 5% on Claude Opus 5.5). See Pricing for the full list.
29* **Latency:** Comparative latency, relative to the current lineup, as published in the models overview. Actual latency depends on prompt length, output length, and thinking effort.
3032* **Thinking:** Adaptive thinking lets the model decide how much to think, steered by effort. Extended thinking is the manual budget\_tokens mode on earlier models.
3133* **Default effort:** The effort parameter’s default on the Claude API. Models without a value don’t support the parameter.
3234* **Knowledge cutoff:** Reliable knowledge cutoff: the date through which the model’s knowledge is most extensive and reliable.
from line 68
6668| Max output | 64K tokens |
6769| [Thinking](https://platform.claude.com/docs/en/build-with-claude/thinking) | Extended |
6870| [Default effort](https://platform.claude.com/docs/en/build-with-claude/effort) | Not supported |
69| Comparative latency | Fastest |
7071| Input → output | Text and images → text |
7172| Reliable knowledge cutoff | Feb 2025 |
7273| Training data cutoff | Jul 2025 |
from line 76
7576
7677| Feature | Value |
7778| :---------------------------------------------------------------------------- | :--------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
78| [Status](https://platform.claude.com/docs/en/about-claude/model-deprecations) | Active (latest) |
79| [Status](https://platform.claude.com/docs/en/about-claude/model-deprecations) | Active (legacy) |
7980| Released | October 15, 2025 |
8081| Retirement | Not sooner than October 15, 2026 |
8182| Platforms | Claude API, [Amazon Bedrock](https://platform.claude.com/docs/en/build-with-claude/claude-in-amazon-bedrock), [Amazon Bedrock (InvokeModel)](https://platform.claude.com/docs/en/build-with-claude/claude-on-amazon-bedrock-legacy), [Google Cloud](https://platform.claude.com/docs/en/build-with-claude/claude-on-vertex-ai), [Microsoft Foundry](https://platform.claude.com/docs/en/build-with-claude/claude-in-microsoft-foundry), [Claude Platform on AWS](https://platform.claude.com/docs/en/build-with-claude/claude-platform-on-aws) |
from line 90
8990## Resources
9091
9192<CardGroup cols={3}>
93 <Card title="Migrate to Claude Haiku 5.5" icon="arrows-left-right" href="https://platform.claude.com/docs/en/models/haiku-5-5/migration-guide">
94 What changes when moving from Claude Haiku 4.5 to Claude Haiku 5.5.
95 </Card>
96
97 <Card title="Claude Haiku 5.5" icon="arrow-right" href="https://platform.claude.com/docs/en/models/haiku-5-5/overview">
98 The current Haiku model: overview, specs, and resources.
99 </Card>
100
92101 <Card title="Extended thinking" icon="brain" href="https://platform.claude.com/docs/en/build-with-claude/extended-thinking">
93102 Claude Haiku 4.5 supports manual extended thinking with `budget_tokens`.
94103 </Card>
from line 107
98107 </Card>
99108
100109 <Card title="Reduce latency" icon="gauge" href="https://platform.claude.com/docs/en/test-and-evaluate/strengthen-guardrails/reduce-latency">
101 Techniques that pair well with the fastest model in the lineup.
110 Techniques that pair well with a fast, low-cost model.
102111 </Card>
103112</CardGroup>
104113
No line in this hunk matches that.