Claude Haiku 4.5
Fastest and cheapest Claude, for high-volume lightweight work.
Released Oct 15, 2025
₹528.00
per 10 lakh output tokens
Input: ₹105.60 per 10 lakh tokens
Billed at provider rates converted at ₹96 per US dollar, with 0% markup.
Pricing
Benchmarks
Scores are percentages unless marked as a rating. All benchmarks are measured independently.
76.0%
MMLU-Pro
MMLU-Pro - multitask language understanding
67.2%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
10.4%
HLE
Humanity's Last Exam
61.5%
LiveCodeBench
LiveCodeBench - contamination-free coding
42.2%
SciCode
SciCode - scientific code generation
83.7%
AIME 2025
AIME 2025 - competition mathematics
54.3%
IFBench
IFBench - precise instruction following
74.3%
Long Context
Long Context Reasoning - reasoning over long inputs
27.3%
Terminal-Bench Hard
Terminal-Bench Hard - agentic terminal tasks
44.2%
Terminal-Bench 2
Terminal-Bench 2.1 - agentic terminal tasks, second edition
About Claude Haiku 4.5
The fastest Claude, delivering near-frontier coding quality with very low latency and cost. It is built for real-time work -- chat assistants, customer service agents, pair programming -- and surpasses Claude Sonnet 4 at some tasks, including computer use, at far lower cost. It has no effort control: thinking is the older extended-thinking mode, off unless a budget is requested.
Anthropic's pitch for Haiku 4.5 was a statement about how fast the frontier moves: five months earlier Claude Sonnet 4 was state of the art, and Haiku 4.5 delivers similar coding performance at a third of the cost and more than twice the speed. It was released two weeks after Sonnet 4.5, which Anthropic was careful to say remained its frontier model; Haiku 4.5 exists as the option for near-frontier work where cost and latency decide the design.
The benchmark table backs the framing rather than the headline. Anthropic reports 73.3% on SWE-bench Verified against 72.7% for Sonnet 4, 41.0% on Terminal-Bench against 36.4%, and 50.7% on OSWorld against 42.2% - computer use is where the gap over Sonnet 4 is widest. Against its own larger sibling Sonnet 4.5 it lands lower across the board, at 73.0% on GPQA Diamond against 83.4% and 83.0% on MMMLU against 89.1%, which is the trade the tier is for.
The structural idea Anthropic put forward with it is orchestration rather than substitution: a larger model breaks a complex problem into a multi-step plan, then directs a team of Haiku 4.5 instances working subtasks in parallel. That pattern is why speed matters more than peak score here, and why the lab pointed at Claude Code and browser-based work as the places the difference is felt.
Thinking on this model is the older extended-thinking mode rather than an effort dial, and the lab's own methodology shows what that means in practice. Its published Terminal-Bench figure averages eleven runs, six of them with no thinking at all scoring 40.21% and five with a 32K thinking budget scoring 41.75%; the SWE-bench Verified number was produced with a 128K budget. Thinking is off unless a call requests a budget, and on this model the budget is the only control.
Safety is where Haiku 4.5 stood out most at release. Anthropic reports it as substantially more aligned than Haiku 3.5 and, on the automated alignment assessment, showing a statistically significantly lower overall rate of misaligned behaviour than both Sonnet 4.5 and Opus 4.1 - by that one metric, the safest model the lab had shipped. Its testing also found only limited risk around chemical, biological, radiological and nuclear weapons, so it was released under the ASL-2 standard rather than the stricter ASL-3 applied to Sonnet 4.5 and Opus 4.1.
What Anthropic announced at launch
- Sonnet 4 quality, Haiku economics
- Anthropic reports similar coding performance to Claude Sonnet 4 at a third of the cost and more than twice the speed, five months after Sonnet 4 was the state of the art.
- Computer use
- The widest gap over Sonnet 4 in the published table: 50.7% on OSWorld against 42.2%, which is why the lab pointed at browser-driven work as the clearest use.
- Built to be orchestrated
- Anthropic describes a larger model planning a complex job and then directing several Haiku 4.5 instances working subtasks in parallel, which is the pattern the speed is for.
- A budget, not a dial
- Thinking here is the older extended-thinking mode, off unless a call asks for a budget. The lab's own Terminal-Bench figure averages runs with no thinking at 40.21% and runs with a 32K budget at 41.75%.
- Safest model at release
- On the automated alignment assessment Anthropic reports a statistically significantly lower rate of misaligned behaviour than Sonnet 4.5 and Opus 4.1, and released it under ASL-2 rather than the stricter ASL-3.
- What it is not for
- Anthropic kept Sonnet 4.5 as its frontier and best coding model at the time, and Haiku 4.5 trails it on every row of the launch table. Reach for it where latency and volume decide, not where peak quality does.
Indian languages
Claude Haiku 4.5 answers in 2 Indian languages. Choose a language from the menu beside the message box to receive replies in it.
Frequently Asked Questions
Frequently asked questions about Claude Haiku 4.5.
When was Claude Haiku 4.5 released?
Anthropic released Claude Haiku 4.5 on Oct 15, 2025.
Who built Claude Haiku 4.5?
Claude Haiku 4.5 is developed by Anthropic. 99Models AI connects directly to it at the provider's published rate.
How intelligent is Claude Haiku 4.5?
It is ranked 45 of 54 chat models on intelligence, ordered by independent benchmark scores. View its complete scores in the Benchmarks table above.
How much does Claude Haiku 4.5 cost?
Usage costs ₹105.60 per million input tokens and ₹528.00 per million output tokens, with 0% markup. There is no subscription; you pay only for what you use.
What is Claude Haiku 4.5 pricing in US dollars?
The provider charges $1.10 per million input tokens and $5.50 per million output tokens. Rupee rates are converted at ₹96 per US dollar.
How long a conversation can Claude Haiku 4.5 hold?
Its context window is 2 lakh tokens. That is the total volume of text and attached files it can process in a single request.
How does Claude Haiku 4.5 rank for value?
It ranks 51 out of 54 models for value. This ranking weighs benchmark intelligence against the actual token cost.
Does Claude Haiku 4.5 support reasoning?
On by default. Where reasoning is supported, you can adjust the thinking effort level directly in the message composer.
Which Indian languages does Claude Haiku 4.5 support?
It answers in 2 Indian languages. Select your preferred language from the menu beside the message box to receive replies in it.
More models from Anthropic
Catalog updated Sep 9, 2026