99MODELS

کلوڈ سونیٹ 4.6

Balanced Sonnet-class model for everyday coding and analysis.

تاریخ اجراء: 17 فروری، 2026

1,584.00

فی 10 لاکھ آؤٹ پٹ Tokens

ان پٹ: ₹316.80 فی 10 لاکھ Tokens

فراہم کنندہ کے نرخ کے مطابق ₹96 فی امریکی ڈالر پر تبدیل شدہ، بغیر کسی اضافی مارک اپ کے۔

تکنیکی تفصیلات

Context ونڈو
10,00,000 Tokens
زیادہ سے زیادہ آؤٹ پٹ
1,28,000 Tokens
سپورٹ
ٹیکسٹ, تصاویر, PDF فائلیں
Reasoning
پہلے سے آن
ایفرٹ لیولز
low, medium, high, max
ٹول کا استعمال
ہاں
سٹرکچرڈ آؤٹ پٹ
ہاں
کوڈ ایگزیکیوشن
نہیں
ذہانت کا درجہ
54 میں سے #29
ویلیو کا درجہ
54 میں سے #46

قیمتیں

قیمتیں
فی 10 lakh TokensINRUSD
ان پٹ316.80$3.30
آؤٹ پٹ1,584.00$16.50
کیشڈ ان پٹ31.68$0.33

Benchmarks

تمام اسکورز فیصد میں ہیں، ماسوائے جہاں ریٹنگ درج ہو۔ یہ تمام پیمائشیں آزادانہ طور پر کی گئی ہیں۔

  • 87.5%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 33.6%

    HLE

    Humanity's Last Exam

  • 56.6%

    IFBench

    IFBench - precise instruction following

  • 80.0%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 53.0%

    Terminal-Bench Hard

    Terminal-Bench Hard - agentic terminal tasks

  • 71.2%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 75.2%

    SWE-bench Verified

    SWE-bench Verified - real-world bug fixing (Epoch AI run)

  • 24.3%

    FrontierCode

    FrontierCode - long-horizon production coding tasks

  • 60.4%

    ARC-AGI-2

    ARC-AGI-2 - abstract reasoning on novel puzzles

  • 41.5%

    OSWorld 2

    OSWorld 2 - agentic computer use (partial credit)

  • 1522

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

کلوڈ سونیٹ 4.6 کے بارے میں

At release, the most capable Sonnet model Anthropic had shipped, upgrading coding, computer use, long-context reasoning, agent planning, knowledge work and design in one step. Its consistency and instruction-following gains led early-access developers to prefer it to its predecessor by a wide margin, and often to Opus 4.5. Thinking is off by default; effort runs low, medium, high and max, with no xhigh rung on this model.

Sonnet 4.6 was the balanced middle of Anthropic's line from February 2026 until Sonnet 5 replaced it at the end of June, and it is still in the catalogue as the previous Sonnet generation. It was the release that brought a 1M token context window to the Sonnet tier in beta, and Anthropic made it the default on the Free and Pro plans on the day it shipped.

Computer use is the capability the post is built around. Anthropic's argument is that most organisations have specialised software predating modern interfaces, which previously needed a bespoke connector for each system; a model that clicks and types the way a person does removes that step. The lab reports 72.5% on OSWorld-Verified against 61.4% for Sonnet 4.5, and describes early users seeing human-level work on tasks like navigating a complex spreadsheet or completing a multi-step web form and then pulling the result together across several browser tabs. It says plainly that the model still lags the most skilled humans, and that OSWorld measures a controlled set of tasks while real computer use is messier, more ambiguous and carries higher stakes for errors.

Preference data is the other half of the case. In Anthropic's early Claude Code testing, users chose Sonnet 4.6 over Sonnet 4.5 roughly 70% of the time and over Opus 4.5 - the frontier model from three months earlier - 59% of the time, reporting less overengineering, better instruction following, fewer false claims of success and more consistent follow-through on multi-step work. On the published table the lab measures 79.6% on SWE-bench Verified, 89.9% on GPQA Diamond, 1633 on GDPval-AA Elo and 63.3% on Finance Agent v1.1, the last two the best figures in its comparison set.

Two results say something about the model's character. On Vending-Bench Arena, which pits models against each other running a simulated business, Anthropic reports Sonnet 4.6 inventing a strategy of its own: spending heavily on capacity for the first ten simulated months, well beyond what its competitors spent, then pivoting sharply to profit in the final stretch, with the timing of the pivot carrying it clear of the field. And on safety, the lab reports a major improvement over Sonnet 4.5 in resisting prompt injection, putting it near Opus 4.6 - which matters precisely because computer use is the thing it is best at.

Anthropic says Sonnet 4.6 performs strongly at any thinking effort, extended thinking off included, and advised anyone migrating from Sonnet 4.5 to try the whole spectrum rather than assume a setting. That is worth knowing here, where thinking is off unless a call asks for it and the ladder runs low, medium, high and max with no extra rung.

لانچ کے وقت Anthropic کا بیان

Computer use
Anthropic reports 72.5% on OSWorld-Verified against 61.4% for Sonnet 4.5, and describes users completing multi-step web forms and spreadsheet work across several browser tabs without a purpose-built connector.
Preferred over the older flagship
In the lab's early Claude Code testing users picked Sonnet 4.6 over Sonnet 4.5 roughly 70% of the time, and over Opus 4.5 - the frontier model from three months earlier - 59% of the time.
Long-horizon planning
On Vending-Bench Arena Anthropic reports it investing heavily in capacity for ten simulated months, then pivoting to profitability late, with the timing of that pivot carrying it clear of the other models.
Knowledge work and finance
The lab measures 1633 on GDPval-AA Elo and 63.3% on Finance Agent v1.1, the best figures in the comparison set it published, alongside 79.6% on SWE-bench Verified.
Resistance to prompt injection
Anthropic reports a major improvement over Sonnet 4.5 at resisting instructions hidden in web pages, close to Opus 4.6 - the safety property that computer use depends on most.
What it is not for
Anthropic said Opus 4.6 remained the stronger choice for the deepest reasoning: codebase refactoring, coordinating several agents in one workflow, and problems where getting it exactly right matters more than cost.

ہندوستانی زبانیں

کلوڈ سونیٹ 4.6 10 ہندوستانی زبانوں میں جواب دیتا ہے۔ میسج باکس کے ساتھ والے مینو سے زبان کا انتخاب کریں۔

Frequently Asked Questions

کلوڈ سونیٹ 4.6 کے بارے میں اکثر پوچھے جانے والے سوالات۔

کلوڈ سونیٹ 4.6 کب جاری ہوا تھا؟

Anthropic نے کلوڈ سونیٹ 4.6 کو 17 فروری، 2026 کو جاری کیا۔

کلوڈ سونیٹ 4.6 کس نے بنایا ہے؟

کلوڈ سونیٹ 4.6 کو Anthropic نے تیار کیا ہے۔ 99Models AI فراہم کنندہ کے اصل نرخ پر اس سے براہ راست جوڑتا ہے۔

کلوڈ سونیٹ 4.6 کتنا ذہین ہے؟

یہ ذہانت کی درجہ بندی میں 54 چیٹ ماڈلز میں سے 29 نمبر پر ہے۔ اس کے تمام بینچ مارک اسکور اوپر والے ٹیبل میں دیکھے جا سکتے ہیں۔

کلوڈ سونیٹ 4.6 کا کتنا خرچ آتا ہے؟

استعمال کی لاگت ₹316.80 فی 10 لاکھ ان پٹ Tokens اور ₹1,584.00 فی 10 لاکھ آؤٹ پٹ Tokens ہے، بغیر کسی اضافی فیس کے۔ کوئی سبسکرپشن نہیں؛ صرف استعمال کی ادائیگی کریں۔

امریکی ڈالر میں کلوڈ سونیٹ 4.6 کی قیمت کیا ہے؟

فراہم کنندہ $3.30 فی 10 لاکھ ان پٹ Tokens اور $16.50 فی 10 لاکھ آؤٹ پٹ Tokens لیتا ہے۔ روپے کے نرخ ₹96 فی امریکی ڈالر کے حساب سے تبدیل کیے گئے ہیں۔

کلوڈ سونیٹ 4.6 کتنی لمبی گفتگو یاد رکھ سکتا ہے؟

اس کی Context حد 10 lakh Tokens ہے، یعنی وہ تمام متن اور فائلیں جو یہ ایک ہی میسج میں پڑھ سکتا ہے۔

کیا کلوڈ سونیٹ 4.6 مناسب قیمت میں بہترین کارکردگی دیتا ہے؟

یہ بہترین قیمت کی درجہ بندی میں 54 ماڈلز میں سے 46 نمبر پر ہے، جس میں ذہانت اور ٹوکن کی لاگت کا موازنہ کیا گیا ہے۔

کیا کلوڈ سونیٹ 4.6 جواب دینے سے پہلے سوچتا ہے؟

پہلے سے آن۔ جہاں Reasoning کی سہولت موجود ہو، وہاں آپ میسج باکس میں سوچنے کی سطح خود طے کر سکتے ہیں۔

کلوڈ سونیٹ 4.6 کن ہندوستانی زبانوں میں جواب دیتا ہے؟

یہ 10 ہندوستانی زبانوں میں جواب دیتا ہے۔ میسج باکس کے ساتھ والے مینو سے اپنی زبان منتخب کریں۔

Anthropic کے مزید ماڈلز

کیٹلاگ اپ ڈیٹ: 9 ستمبر، 2026