99MODELS

क्लॉड सॉनेट 4.6

Balanced Sonnet-class model for everyday coding and analysis.

रिलीज़: 17 फ़र॰ 2026

1,584.00

प्रति 10 लाख आउटपुट Tokens

इनपुट: ₹316.80 प्रति 10 लाख Tokens

प्रदाता की आधिकारिक दर पर बिलिंग, ₹96 प्रति डॉलर पर कनवर्ट, 0% मार्कअप के साथ।

स्पेसिफिकेशन्स

Context विंडो
10,00,000 Tokens
अधिकतम आउटपुट
1,28,000 Tokens
स्वीकार्य इनपुट
टेक्स्ट, इमेज, PDF फ़ाइलें
Reasoning
डिफ़ॉल्ट रूप से चालू
प्रयास स्तर
low, medium, high, max
टूल का उपयोग
हाँ
स्ट्रक्चर्ड आउटपुट
हाँ
कोड एग्जीक्यूशन
नहीं
इंटेलिजेंस रैंक
54 में से #29
वैल्यू रैंक
54 में से #46

दरें

दरें
प्रति 10 lakh TokensINRUSD
इनपुट316.80$3.30
आउटपुट1,584.00$16.50
कैश्ड इनपुट31.68$0.33

Benchmarks

स्कोर प्रतिशत में हैं जब तक कि कोई रेटिंग न दी गई हो। सभी Benchmark स्वतंत्र रूप से मापे गए हैं।

  • 87.5%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 33.6%

    HLE

    Humanity's Last Exam

  • 56.6%

    IFBench

    IFBench - precise instruction following

  • 80.0%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 53.0%

    Terminal-Bench Hard

    Terminal-Bench Hard - agentic terminal tasks

  • 71.2%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 75.2%

    SWE-bench Verified

    SWE-bench Verified - real-world bug fixing (Epoch AI run)

  • 24.3%

    FrontierCode

    FrontierCode - long-horizon production coding tasks

  • 60.4%

    ARC-AGI-2

    ARC-AGI-2 - abstract reasoning on novel puzzles

  • 41.5%

    OSWorld 2

    OSWorld 2 - agentic computer use (partial credit)

  • 1522

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

क्लॉड सॉनेट 4.6 के बारे में

At release, the most capable Sonnet model Anthropic had shipped, upgrading coding, computer use, long-context reasoning, agent planning, knowledge work and design in one step. Its consistency and instruction-following gains led early-access developers to prefer it to its predecessor by a wide margin, and often to Opus 4.5. Thinking is off by default; effort runs low, medium, high and max, with no xhigh rung on this model.

Sonnet 4.6 was the balanced middle of Anthropic's line from February 2026 until Sonnet 5 replaced it at the end of June, and it is still in the catalogue as the previous Sonnet generation. It was the release that brought a 1M token context window to the Sonnet tier in beta, and Anthropic made it the default on the Free and Pro plans on the day it shipped.

Computer use is the capability the post is built around. Anthropic's argument is that most organisations have specialised software predating modern interfaces, which previously needed a bespoke connector for each system; a model that clicks and types the way a person does removes that step. The lab reports 72.5% on OSWorld-Verified against 61.4% for Sonnet 4.5, and describes early users seeing human-level work on tasks like navigating a complex spreadsheet or completing a multi-step web form and then pulling the result together across several browser tabs. It says plainly that the model still lags the most skilled humans, and that OSWorld measures a controlled set of tasks while real computer use is messier, more ambiguous and carries higher stakes for errors.

Preference data is the other half of the case. In Anthropic's early Claude Code testing, users chose Sonnet 4.6 over Sonnet 4.5 roughly 70% of the time and over Opus 4.5 - the frontier model from three months earlier - 59% of the time, reporting less overengineering, better instruction following, fewer false claims of success and more consistent follow-through on multi-step work. On the published table the lab measures 79.6% on SWE-bench Verified, 89.9% on GPQA Diamond, 1633 on GDPval-AA Elo and 63.3% on Finance Agent v1.1, the last two the best figures in its comparison set.

Two results say something about the model's character. On Vending-Bench Arena, which pits models against each other running a simulated business, Anthropic reports Sonnet 4.6 inventing a strategy of its own: spending heavily on capacity for the first ten simulated months, well beyond what its competitors spent, then pivoting sharply to profit in the final stretch, with the timing of the pivot carrying it clear of the field. And on safety, the lab reports a major improvement over Sonnet 4.5 in resisting prompt injection, putting it near Opus 4.6 - which matters precisely because computer use is the thing it is best at.

Anthropic says Sonnet 4.6 performs strongly at any thinking effort, extended thinking off included, and advised anyone migrating from Sonnet 4.5 to try the whole spectrum rather than assume a setting. That is worth knowing here, where thinking is off unless a call asks for it and the ladder runs low, medium, high and max with no extra rung.

लॉन्च के समय Anthropic ने क्या कहा

Computer use
Anthropic reports 72.5% on OSWorld-Verified against 61.4% for Sonnet 4.5, and describes users completing multi-step web forms and spreadsheet work across several browser tabs without a purpose-built connector.
Preferred over the older flagship
In the lab's early Claude Code testing users picked Sonnet 4.6 over Sonnet 4.5 roughly 70% of the time, and over Opus 4.5 - the frontier model from three months earlier - 59% of the time.
Long-horizon planning
On Vending-Bench Arena Anthropic reports it investing heavily in capacity for ten simulated months, then pivoting to profitability late, with the timing of that pivot carrying it clear of the other models.
Knowledge work and finance
The lab measures 1633 on GDPval-AA Elo and 63.3% on Finance Agent v1.1, the best figures in the comparison set it published, alongside 79.6% on SWE-bench Verified.
Resistance to prompt injection
Anthropic reports a major improvement over Sonnet 4.5 at resisting instructions hidden in web pages, close to Opus 4.6 - the safety property that computer use depends on most.
What it is not for
Anthropic said Opus 4.6 remained the stronger choice for the deepest reasoning: codebase refactoring, coordinating several agents in one workflow, and problems where getting it exactly right matters more than cost.

भारतीय भाषाएं

क्लॉड सॉनेट 4.6 10 भारतीय भाषाओं में जवाब दे सकता है। मैसेज बॉक्स के पास वाले मेन्यू से भाषा चुनें और उसी में जवाब पाएं।

Frequently Asked Questions

क्लॉड सॉनेट 4.6 के बारे में अक्सर पूछे जाने वाले सवाल।

क्लॉड सॉनेट 4.6 कब रिलीज़ हुआ था?

Anthropic ने क्लॉड सॉनेट 4.6 को 17 फ़र॰ 2026 को रिलीज़ किया था।

क्लॉड सॉनेट 4.6 को किसने बनाया है?

क्लॉड सॉनेट 4.6 को AI लैब Anthropic ने बनाया है। 99Models इसे सीधे प्रदाता की आधिकारिक दर पर उपलब्ध कराता है।

क्लॉड सॉनेट 4.6 कितना समझदार है?

इंटेलीजेंस रैंकिंग में यह 54 चैट Models में से 29 स्थान पर है, जो बेंचमार्क स्कोर पर आधारित है। इसके पूरे स्कोर ऊपर Benchmarks पैनल में देखे जा सकते हैं।

क्लॉड सॉनेट 4.6 का उपयोग करने का क्या खर्च है?

10 लाख इनपुट Tokens के लिए ₹316.80 और 10 लाख आउटपुट Tokens के लिए ₹1,584.00, बिना किसी अतिरिक्त मार्कअप के प्रदाता की दर पर। कोई सब्सक्रिप्शन नहीं है; आप केवल अपने उपयोग का भुगतान करते हैं।

डॉलर में क्लॉड सॉनेट 4.6 API की कीमत क्या है?

प्रदाता 10 लाख इनपुट Tokens के लिए $3.30 और 10 लाख आउटपुट Tokens के लिए $16.50 चार्ज करता है। इस पेज पर रुपये की दरें ₹96 प्रति डॉलर के हिसाब से बदली गई हैं।

क्लॉड सॉनेट 4.6 कितनी लंबी बातचीत याद रख सकता है?

इसकी Context विंडो 10 lakh Tokens है। यानी यह एक रिक्वेस्ट में पिछली बातचीत और अटैच की गई फ़ाइलों को मिलाकर इतना टेक्स्ट पढ़ सकता है।

क्या क्लॉड सॉनेट 4.6 पैसे के लिहाज से किफ़ायती है?

वैल्यू रैंकिंग में यह 54 Models में से 46 स्थान पर है। यह रैंकिंग परफ़ॉर्मेंस और टोकन की वास्तविक कीमत की तुलना करके तय की जाती है।

क्या क्लॉड सॉनेट 4.6 जवाब देने से पहले सोच-विचार (Reasoning) करता है?

डिफ़ॉल्ट रूप से चालू। जहां सोचने की क्षमता उपलब्ध है, वहां आप मैसेज कंपोज़र में सीधे Reasoning का स्तर तय कर सकते हैं।

क्लॉड सॉनेट 4.6 किन भारतीय भाषाओं में जवाब दे सकता है?

यह 10 भारतीय भाषाओं में जवाब देता है। मैसेज बॉक्स के पास वाले मेन्यू से भाषा चुनें और उसी भाषा में जवाब पाएं।

Anthropic के अन्य Models

कैटलॉग अपडेट: 9 सित॰ 2026