کلوڈ سونیٹ 5
نیاFrontier coding and agentic performance at the Sonnet price point.
تاریخ اجراء: 30 جون، 2026
₹1,056.00
فی 10 لاکھ آؤٹ پٹ Tokens
ان پٹ: ₹211.20 فی 10 لاکھ Tokens
فراہم کنندہ کے نرخ کے مطابق ₹96 فی امریکی ڈالر پر تبدیل شدہ، بغیر کسی اضافی مارک اپ کے۔
تکنیکی تفصیلات
- Context ونڈو
- 10,00,000 Tokens
- زیادہ سے زیادہ آؤٹ پٹ
- 1,28,000 Tokens
- سپورٹ
- ٹیکسٹ, تصاویر, PDF فائلیں
- Reasoning
- پہلے سے آن
- ایفرٹ لیولز
- low, medium, high, xhigh, max
- ٹول کا استعمال
- ہاں
- سٹرکچرڈ آؤٹ پٹ
- ہاں
- کوڈ ایگزیکیوشن
- نہیں
- ذہانت کا درجہ
- 54 میں سے #21
- ویلیو کا درجہ
- 54 میں سے #24
قیمتیں
Benchmarks
تمام اسکورز فیصد میں ہیں، ماسوائے جہاں ریٹنگ درج ہو۔ یہ تمام پیمائشیں آزادانہ طور پر کی گئی ہیں۔
91.1%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
41.3%
HLE
Humanity's Last Exam
54.3%
SciCode
SciCode - scientific code generation
82.0%
Long Context
Long Context Reasoning - reasoning over long inputs
80.5%
Terminal-Bench 2
Terminal-Bench 2.1 - agentic terminal tasks, second edition
42.7%
FrontierCode
FrontierCode - long-horizon production coding tasks
1539
WebDev Arena
WebDev Arena - head-to-head web-app builds, Elo rating
کلوڈ سونیٹ 5 کے بارے میں
The next generation of the Sonnet family, offering what Anthropic calls the best combination of speed and intelligence. It is a capability upgrade over Sonnet 4.6 at a lower price and a drop-in replacement for it, with the largest gains in coding and agentic tasks. It is the model to reach for when you need more than Sonnet 4.6 without moving to an Opus-class price, and it supports the same low-to-max effort ladder.
Anthropic built Sonnet 5 to be the most agentic Sonnet it has shipped: a model that makes plans, drives browsers and terminals, and runs on its own at a level that a few months earlier needed a larger and dearer model. The lab tells the history deliberately. The agentic era began with Sonnet-class models, then the clearest gains moved up into the Opus class; Sonnet 5 is the release that pulls the line back down. It became the default model on Free and Pro the day it shipped and is available on every other plan.
Anthropic presents the results as cost-performance curves across the effort ladder rather than as single peak scores, which is the honest way to read a model whose thinking budget the caller sets. On the agentic search evaluation BrowseComp and the computer-use evaluation OSWorld-Verified, the lab reports Sonnet 5 as a strict improvement over Sonnet 4.6 at every rung, covering a wider range of cost-performance options than Opus 4.8 does, with the largest efficiency gain at medium effort and higher-effort runs matching Opus 4.8 on some tasks.
The table Anthropic published puts Sonnet 5 at 63.2% on SWE-bench Pro against 58.1% for Sonnet 4.6; 80.4% on Terminal-Bench 2.1 against 67.0%, within about two points of Opus 4.8; 81.2% on OSWorld-Verified against 78.5%; 43.2% on Humanity's Last Exam without tools and 57.4% with them, against 34.6% and 46.8%; and 1618 on GDPval-AA v2, just past Opus 4.8's 1615. Early-access testers described the change as follow-through rather than raw score: finishing multi-step jobs where earlier Sonnet models stopped halfway, and checking their own output without being asked to.
One practical catch is stated plainly in the post. Sonnet 5 uses an updated tokenizer, so the same input can map to roughly 1.0 to 1.35 times as many tokens depending on the kind of content, which is a real cost difference on top of the per-token rate. Anthropic also raised rate limits across its own surfaces to absorb the heavier token use of the higher effort rungs.
The lab is specific about where the model stops. It did not deliberately train Sonnet 5 on cybersecurity, and on the Firefox exploit-development evaluation the model never produced a working exploit, though its partial-success rate sits slightly above Sonnet 4.6's, which Anthropic attributes to general capability rather than targeted training. It shipped with cyber safeguards enabled by default, less strict than Fable 5's, and Anthropic recommends Opus 4.8 for cybersecurity work that needs reduced guardrails. On its automated behavioural audit Sonnet 5 is safer overall than Sonnet 4.6 but shows more misaligned behaviour than Opus 4.8 and Claude Mythos Preview.
لانچ کے وقت Anthropic کا بیان
- Agentic work at Sonnet cost
- Anthropic reports Sonnet 5 covering a wider range of cost-performance options than Opus 4.8 across the effort ladder, with the clearest efficiency gain at medium effort and higher-effort runs matching Opus 4.8 on some tasks.
- Coding and the terminal
- The lab measures 63.2% on SWE-bench Pro against Sonnet 4.6's 58.1%, and 80.4% on Terminal-Bench 2.1 against 67.0% - within about two points of Opus 4.8 on the same harness.
- Computer use and search
- Anthropic reports 81.2% on OSWorld-Verified against 78.5% for Sonnet 4.6, and presents both computer use and BrowseComp agentic search as cost curves rather than single scores.
- Knowledge work
- On GDPval-AA v2 the lab reports 1618 against 1395 for Sonnet 4.6, edging past Opus 4.8's 1615 on the same evaluation.
- A different tokenizer
- Sonnet 5 changed tokenizers, and Anthropic says the same input can map to roughly 1.0 to 1.35 times as many tokens depending on content type. Budget for that alongside the per-token rate.
- What it is not for
- Anthropic did not train Sonnet 5 for cybersecurity: it never developed a working exploit on the lab's Firefox evaluation, ships with cyber safeguards on by default, and the lab points cyber work needing reduced guardrails at Opus 4.8 instead.
ہندوستانی زبانیں
کلوڈ سونیٹ 5 14 ہندوستانی زبانوں میں جواب دیتا ہے۔ میسج باکس کے ساتھ والے مینو سے زبان کا انتخاب کریں۔
Frequently Asked Questions
کلوڈ سونیٹ 5 کے بارے میں اکثر پوچھے جانے والے سوالات۔
کلوڈ سونیٹ 5 کب جاری ہوا تھا؟
Anthropic نے کلوڈ سونیٹ 5 کو 30 جون، 2026 کو جاری کیا۔
کلوڈ سونیٹ 5 کس نے بنایا ہے؟
کلوڈ سونیٹ 5 کو Anthropic نے تیار کیا ہے۔ 99Models AI فراہم کنندہ کے اصل نرخ پر اس سے براہ راست جوڑتا ہے۔
کلوڈ سونیٹ 5 کتنا ذہین ہے؟
یہ ذہانت کی درجہ بندی میں 54 چیٹ ماڈلز میں سے 21 نمبر پر ہے۔ اس کے تمام بینچ مارک اسکور اوپر والے ٹیبل میں دیکھے جا سکتے ہیں۔
کلوڈ سونیٹ 5 کا کتنا خرچ آتا ہے؟
استعمال کی لاگت ₹211.20 فی 10 لاکھ ان پٹ Tokens اور ₹1,056.00 فی 10 لاکھ آؤٹ پٹ Tokens ہے، بغیر کسی اضافی فیس کے۔ کوئی سبسکرپشن نہیں؛ صرف استعمال کی ادائیگی کریں۔
امریکی ڈالر میں کلوڈ سونیٹ 5 کی قیمت کیا ہے؟
فراہم کنندہ $2.20 فی 10 لاکھ ان پٹ Tokens اور $11.00 فی 10 لاکھ آؤٹ پٹ Tokens لیتا ہے۔ روپے کے نرخ ₹96 فی امریکی ڈالر کے حساب سے تبدیل کیے گئے ہیں۔
کلوڈ سونیٹ 5 کتنی لمبی گفتگو یاد رکھ سکتا ہے؟
اس کی Context حد 10 lakh Tokens ہے، یعنی وہ تمام متن اور فائلیں جو یہ ایک ہی میسج میں پڑھ سکتا ہے۔
کیا کلوڈ سونیٹ 5 مناسب قیمت میں بہترین کارکردگی دیتا ہے؟
یہ بہترین قیمت کی درجہ بندی میں 54 ماڈلز میں سے 24 نمبر پر ہے، جس میں ذہانت اور ٹوکن کی لاگت کا موازنہ کیا گیا ہے۔
کیا کلوڈ سونیٹ 5 جواب دینے سے پہلے سوچتا ہے؟
پہلے سے آن۔ جہاں Reasoning کی سہولت موجود ہو، وہاں آپ میسج باکس میں سوچنے کی سطح خود طے کر سکتے ہیں۔
کلوڈ سونیٹ 5 کن ہندوستانی زبانوں میں جواب دیتا ہے؟
یہ 14 ہندوستانی زبانوں میں جواب دیتا ہے۔ میسج باکس کے ساتھ والے مینو سے اپنی زبان منتخب کریں۔
Anthropic کے مزید ماڈلز
کیٹلاگ اپ ڈیٹ: 9 ستمبر، 2026