Claude Sonnet 4.6
Balanced Sonnet-class model for everyday coding and analysis.
Released Feb 17, 2026
₹1,584.00
per 10 lakh output tokens
Input: ₹316.80 per 10 lakh tokens
Billed at provider rates converted at ₹96 per US dollar, with 0% markup.
Pricing
Benchmarks
Scores are percentages unless marked as a rating. All benchmarks are measured independently.
87.5%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
33.6%
HLE
Humanity's Last Exam
56.6%
IFBench
IFBench - precise instruction following
80.0%
Long Context
Long Context Reasoning - reasoning over long inputs
53.0%
Terminal-Bench Hard
Terminal-Bench Hard - agentic terminal tasks
71.2%
Terminal-Bench 2
Terminal-Bench 2.1 - agentic terminal tasks, second edition
75.2%
SWE-bench Verified
SWE-bench Verified - real-world bug fixing (Epoch AI run)
24.3%
FrontierCode
FrontierCode - long-horizon production coding tasks
60.4%
ARC-AGI-2
ARC-AGI-2 - abstract reasoning on novel puzzles
41.5%
OSWorld 2
OSWorld 2 - agentic computer use (partial credit)
1522
WebDev Arena
WebDev Arena - head-to-head web-app builds, Elo rating
About Claude Sonnet 4.6
At release, the most capable Sonnet model Anthropic had shipped, upgrading coding, computer use, long-context reasoning, agent planning, knowledge work and design in one step. Its consistency and instruction-following gains led early-access developers to prefer it to its predecessor by a wide margin, and often to Opus 4.5. Thinking is off by default; effort runs low, medium, high and max, with no xhigh rung on this model.
Sonnet 4.6 was the balanced middle of Anthropic's line from February 2026 until Sonnet 5 replaced it at the end of June, and it is still in the catalogue as the previous Sonnet generation. It was the release that brought a 1M token context window to the Sonnet tier in beta, and Anthropic made it the default on the Free and Pro plans on the day it shipped.
Computer use is the capability the post is built around. Anthropic's argument is that most organisations have specialised software predating modern interfaces, which previously needed a bespoke connector for each system; a model that clicks and types the way a person does removes that step. The lab reports 72.5% on OSWorld-Verified against 61.4% for Sonnet 4.5, and describes early users seeing human-level work on tasks like navigating a complex spreadsheet or completing a multi-step web form and then pulling the result together across several browser tabs. It says plainly that the model still lags the most skilled humans, and that OSWorld measures a controlled set of tasks while real computer use is messier, more ambiguous and carries higher stakes for errors.
Preference data is the other half of the case. In Anthropic's early Claude Code testing, users chose Sonnet 4.6 over Sonnet 4.5 roughly 70% of the time and over Opus 4.5 - the frontier model from three months earlier - 59% of the time, reporting less overengineering, better instruction following, fewer false claims of success and more consistent follow-through on multi-step work. On the published table the lab measures 79.6% on SWE-bench Verified, 89.9% on GPQA Diamond, 1633 on GDPval-AA Elo and 63.3% on Finance Agent v1.1, the last two the best figures in its comparison set.
Two results say something about the model's character. On Vending-Bench Arena, which pits models against each other running a simulated business, Anthropic reports Sonnet 4.6 inventing a strategy of its own: spending heavily on capacity for the first ten simulated months, well beyond what its competitors spent, then pivoting sharply to profit in the final stretch, with the timing of the pivot carrying it clear of the field. And on safety, the lab reports a major improvement over Sonnet 4.5 in resisting prompt injection, putting it near Opus 4.6 - which matters precisely because computer use is the thing it is best at.
Anthropic says Sonnet 4.6 performs strongly at any thinking effort, extended thinking off included, and advised anyone migrating from Sonnet 4.5 to try the whole spectrum rather than assume a setting. That is worth knowing here, where thinking is off unless a call asks for it and the ladder runs low, medium, high and max with no extra rung.
What Anthropic announced at launch
- Computer use
- Anthropic reports 72.5% on OSWorld-Verified against 61.4% for Sonnet 4.5, and describes users completing multi-step web forms and spreadsheet work across several browser tabs without a purpose-built connector.
- Preferred over the older flagship
- In the lab's early Claude Code testing users picked Sonnet 4.6 over Sonnet 4.5 roughly 70% of the time, and over Opus 4.5 - the frontier model from three months earlier - 59% of the time.
- Long-horizon planning
- On Vending-Bench Arena Anthropic reports it investing heavily in capacity for ten simulated months, then pivoting to profitability late, with the timing of that pivot carrying it clear of the other models.
- Knowledge work and finance
- The lab measures 1633 on GDPval-AA Elo and 63.3% on Finance Agent v1.1, the best figures in the comparison set it published, alongside 79.6% on SWE-bench Verified.
- Resistance to prompt injection
- Anthropic reports a major improvement over Sonnet 4.5 at resisting instructions hidden in web pages, close to Opus 4.6 - the safety property that computer use depends on most.
- What it is not for
- Anthropic said Opus 4.6 remained the stronger choice for the deepest reasoning: codebase refactoring, coordinating several agents in one workflow, and problems where getting it exactly right matters more than cost.
Indian languages
Claude Sonnet 4.6 answers in 10 Indian languages. Choose a language from the menu beside the message box to receive replies in it.
Frequently Asked Questions
Frequently asked questions about Claude Sonnet 4.6.
When was Claude Sonnet 4.6 released?
Anthropic released Claude Sonnet 4.6 on Feb 17, 2026.
Who built Claude Sonnet 4.6?
Claude Sonnet 4.6 is developed by Anthropic. 99Models AI connects directly to it at the provider's published rate.
How intelligent is Claude Sonnet 4.6?
It is ranked 29 of 54 chat models on intelligence, ordered by independent benchmark scores. View its complete scores in the Benchmarks table above.
How much does Claude Sonnet 4.6 cost?
Usage costs ₹316.80 per million input tokens and ₹1,584.00 per million output tokens, with 0% markup. There is no subscription; you pay only for what you use.
What is Claude Sonnet 4.6 pricing in US dollars?
The provider charges $3.30 per million input tokens and $16.50 per million output tokens. Rupee rates are converted at ₹96 per US dollar.
How long a conversation can Claude Sonnet 4.6 hold?
Its context window is 10 lakh tokens. That is the total volume of text and attached files it can process in a single request.
How does Claude Sonnet 4.6 rank for value?
It ranks 46 out of 54 models for value. This ranking weighs benchmark intelligence against the actual token cost.
Does Claude Sonnet 4.6 support reasoning?
On by default. Where reasoning is supported, you can adjust the thinking effort level directly in the message composer.
Which Indian languages does Claude Sonnet 4.6 support?
It answers in 10 Indian languages. Select your preferred language from the menu beside the message box to receive replies in it.
More models from Anthropic
Catalog updated Sep 9, 2026