Quick answer: Claude 3.5 Haiku (API: claude-3-5-haiku-20241022) is Anthropic's October 2024 fast small model. It scores 88.1% on HumanEval, 65% on MMLU-Pro, 30.1% on BigCodeBench, and 69.4% on MATH. At $1/$5 per 1M tokens with a 200K context window, it is the recommended Claude small-tier model for most tasks requiring speed at moderate cost.
Where Claude 3.5 Haiku leads
Where it lags
Best for: Claude-ecosystem integrations requiring fast, cost-moderate responses; vision tasks at the Haiku tier; teams that need Anthropic's safety properties without Sonnet-tier cost.
Claude 3.5 Haiku is Anthropic's small-model upgrade within the Claude 3.5 family, released October 22, 2024 alongside Claude 3.5 Sonnet v2 (Claude Sonnet 3.6). It replaced Claude 3 Haiku as Anthropic's recommended fast model, bringing Claude 3.5-quality instruction following and coding capability to the Haiku tier.
Unlike Claude 3 Haiku ($0.25/$1.25), Claude 3.5 Haiku is priced at $1/$5 — higher cost reflecting substantially improved capability. Its 88.1% HumanEval and 69.4% MATH scores position it well above Claude 3 Haiku and competitive with GPT-4o mini on code and math tasks, despite being in the "small" model tier.
For teams embedded in the Claude ecosystem and requiring fast responses with vision at moderate cost, Claude 3.5 Haiku is the appropriate choice. For purely cost-optimised high-volume workloads, GPT-4o mini and Gemini 2.5 Flash offer better price/performance.
| Field | Value |
|---|---|
| Organization | Anthropic |
| Parameters | Undisclosed |
| Context window | 200,000 tokens |
| API model ID | claude-3-5-haiku-20241022 |
| License | Proprietary (API only) |
| Release date | October 22, 2024 |
| Knowledge cutoff | July 2024 |
| Modality | Text + Vision (multimodal) |
| Input (per 1M tokens) | Output (per 1M tokens) | |
|---|---|---|
| Anthropic API | $1.00 | $5.00 |
Prompt caching available. Pricing per Anthropic pricing page.
Claude 3.5 Haiku has a 200,000-token context window — roughly 150 pages of text in a single request.
| Benchmark | Score | Source | Date |
|---|---|---|---|
| HumanEval | 88.1% | Benchgen evaluation | 2024-10 |
| MMLU-Pro | 65% | Benchgen evaluation | 2024-10 |
| BigCodeBench | 30.1% | Benchgen evaluation | 2024-10 |
| MATH | 69.4% | Benchgen evaluation | 2024-10 |
| Model | HumanEval | MMLU-Pro | Context | Price (in/out per 1M) |
|---|---|---|---|---|
| Claude 3.5 Haiku | 88.1% | 65% | 200K | $1 / $5 |
| Claude Haiku 3.5 | — | — | 200K | $0.80 / $4 |
| GPT-4o mini | 87.2%* | — | 128K | $0.15 / $0.60 |
| Gemini 2.5 Flash | — | — | 1M | $0.15 / $0.60 |
*From OpenAI launch report. Claude 3.5 Haiku vs GPT-4o mini: similar HumanEval, 200K vs 128K context, at 6.7× higher input cost. Claude 3.5 Haiku is appropriate when Anthropic's training properties (safety, instruction following style) are specifically needed.
Specs and scores sourced from Anthropic's official Claude 3.5 Haiku announcement (October 2024) and Benchgen evaluations. Pricing cited to the Anthropic pricing page. Last updated 2026-07-24.
This model isn’t on any benchmark leaderboard yet.