99MODELS

ক্লড সনেট 5

নতুন

Frontier coding and agentic performance at the Sonnet price point.

প্রকাশকাল: 30 জুন, 2026

1,056.00

প্রতি 10 লাখ আউটপুট Tokens

ইনপুট: প্রতি 10 লাখ Tokens-এ ₹211.20

এটি প্রোভাইডারের নিজস্ব রেট, যা প্রতি ডলারে ₹96 হিসেবে 0% মার্কআপে কনভার্ট করা।

স্পেসিফিকেশন

Context উইন্ডো
10,00,000 Tokens
সর্বোচ্চ আউটপুট
1,28,000 Tokens
সাপোর্ট করে
টেক্সট, ছবি, PDF ফাইল
Reasoning
ডিফল্টভাবে চালু
Effort লেভেল
low, medium, high, xhigh, max
টুল ব্যবহার
হ্যাঁ
স্ট্রাকচার্ড আউটপুট
হ্যাঁ
কোড এক্সিকিউশন
না
বুদ্ধিমত্তা র‍্যাঙ্ক
54-এর মধ্যে #21
ভ্যালু র‍্যাঙ্ক
54-এর মধ্যে #24

মূল্য তালিকা

মূল্য তালিকা
প্রতি 10 lakh TokensINRUSD
ইনপুট211.20$2.20
আউটপুট1,056.00$11.00
ক্যাশড ইনপুট21.12$0.22

Benchmarks

রেটিং হিসেবে উল্লেখ না থাকলে স্কোরগুলো শতাংশে দেখানো হয়। সব Benchmark নিরপেক্ষভাবে পরিমাপ করা হয়েছে।

  • 91.1%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 41.3%

    HLE

    Humanity's Last Exam

  • 54.3%

    SciCode

    SciCode - scientific code generation

  • 82.0%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 80.5%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 42.7%

    FrontierCode

    FrontierCode - long-horizon production coding tasks

  • 1539

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

ক্লড সনেট 5 সম্পর্কে

The next generation of the Sonnet family, offering what Anthropic calls the best combination of speed and intelligence. It is a capability upgrade over Sonnet 4.6 at a lower price and a drop-in replacement for it, with the largest gains in coding and agentic tasks. It is the model to reach for when you need more than Sonnet 4.6 without moving to an Opus-class price, and it supports the same low-to-max effort ladder.

Anthropic built Sonnet 5 to be the most agentic Sonnet it has shipped: a model that makes plans, drives browsers and terminals, and runs on its own at a level that a few months earlier needed a larger and dearer model. The lab tells the history deliberately. The agentic era began with Sonnet-class models, then the clearest gains moved up into the Opus class; Sonnet 5 is the release that pulls the line back down. It became the default model on Free and Pro the day it shipped and is available on every other plan.

Anthropic presents the results as cost-performance curves across the effort ladder rather than as single peak scores, which is the honest way to read a model whose thinking budget the caller sets. On the agentic search evaluation BrowseComp and the computer-use evaluation OSWorld-Verified, the lab reports Sonnet 5 as a strict improvement over Sonnet 4.6 at every rung, covering a wider range of cost-performance options than Opus 4.8 does, with the largest efficiency gain at medium effort and higher-effort runs matching Opus 4.8 on some tasks.

The table Anthropic published puts Sonnet 5 at 63.2% on SWE-bench Pro against 58.1% for Sonnet 4.6; 80.4% on Terminal-Bench 2.1 against 67.0%, within about two points of Opus 4.8; 81.2% on OSWorld-Verified against 78.5%; 43.2% on Humanity's Last Exam without tools and 57.4% with them, against 34.6% and 46.8%; and 1618 on GDPval-AA v2, just past Opus 4.8's 1615. Early-access testers described the change as follow-through rather than raw score: finishing multi-step jobs where earlier Sonnet models stopped halfway, and checking their own output without being asked to.

One practical catch is stated plainly in the post. Sonnet 5 uses an updated tokenizer, so the same input can map to roughly 1.0 to 1.35 times as many tokens depending on the kind of content, which is a real cost difference on top of the per-token rate. Anthropic also raised rate limits across its own surfaces to absorb the heavier token use of the higher effort rungs.

The lab is specific about where the model stops. It did not deliberately train Sonnet 5 on cybersecurity, and on the Firefox exploit-development evaluation the model never produced a working exploit, though its partial-success rate sits slightly above Sonnet 4.6's, which Anthropic attributes to general capability rather than targeted training. It shipped with cyber safeguards enabled by default, less strict than Fable 5's, and Anthropic recommends Opus 4.8 for cybersecurity work that needs reduced guardrails. On its automated behavioural audit Sonnet 5 is safer overall than Sonnet 4.6 but shows more misaligned behaviour than Opus 4.8 and Claude Mythos Preview.

উদ্বোধনে Anthropic যা জানিয়েছিল

Agentic work at Sonnet cost
Anthropic reports Sonnet 5 covering a wider range of cost-performance options than Opus 4.8 across the effort ladder, with the clearest efficiency gain at medium effort and higher-effort runs matching Opus 4.8 on some tasks.
Coding and the terminal
The lab measures 63.2% on SWE-bench Pro against Sonnet 4.6's 58.1%, and 80.4% on Terminal-Bench 2.1 against 67.0% - within about two points of Opus 4.8 on the same harness.
Computer use and search
Anthropic reports 81.2% on OSWorld-Verified against 78.5% for Sonnet 4.6, and presents both computer use and BrowseComp agentic search as cost curves rather than single scores.
Knowledge work
On GDPval-AA v2 the lab reports 1618 against 1395 for Sonnet 4.6, edging past Opus 4.8's 1615 on the same evaluation.
A different tokenizer
Sonnet 5 changed tokenizers, and Anthropic says the same input can map to roughly 1.0 to 1.35 times as many tokens depending on content type. Budget for that alongside the per-token rate.
What it is not for
Anthropic did not train Sonnet 5 for cybersecurity: it never developed a working exploit on the lab's Firefox evaluation, ships with cyber safeguards on by default, and the lab points cyber work needing reduced guardrails at Opus 4.8 instead.

ভারতীয় ভাষাসমূহ

ক্লড সনেট 5 মোট 14টি ভারতীয় ভাষায় উত্তর দেয়। মেসেজ বক্সের পাশের ভাষা মেনু থেকে সিলেক্ট করলেই সেই ভাষায় উত্তর আসবে।

Frequently Asked Questions

ক্লড সনেট 5 সম্পর্কে সাধারণ কিছু প্রশ্নোত্তর।

ক্লড সনেট 5 কবে প্রকাশিত হয়?

Anthropic 30 জুন, 2026 তারিখে ক্লড সনেট 5 প্রকাশ করেছে।

ক্লড সনেট 5 কে তৈরি করেছে?

ক্লড সনেট 5 তৈরি করেছে Anthropic। 99Models সরাসরি প্রোভাইডারের নির্ধারিত রেটেই এটি ব্যবহার করতে দেয়।

ক্লড সনেট 5 কতটা বুদ্ধিমান?

স্বাধীন বেঞ্চমার্ক স্কোরের ভিত্তিতে বুদ্ধিমত্তার র‍্যাঙ্কিংয়ে 54টি চ্যাট Model-এর মধ্যে এর স্থান 21। এর বেঞ্চমার্ক স্কোরগুলো উপরের টেবিলে দেওয়া আছে।

ক্লড সনেট 5-এর জন্য কত খরচ হয়?

খরচ প্রতি 10 লাখ ইনপুট Tokens-এ ₹211.20 এবং আউটপুট Tokens-এ ₹1,056.00, কোনো বাড়তি মার্কআপ ছাড়াই প্রোভাইডারের নিজস্ব রেটে। কোনো সাবস্ক্রিপশন নেই; যতটুকু ব্যবহার করবেন শুধুই তার পেমেন্ট করবেন।

ডলারে ক্লড সনেট 5-এর API খরচ কত?

প্রোভাইডার প্রতি 10 লাখ ইনপুট Tokens-এ $2.20 এবং আউটপুট Tokens-এ $11.00 চার্জ করে। এই পেজের টাকার হিসেব প্রতি ডলারে ₹96 রেটে কনভার্ট করা।

ক্লড সনেট 5 কত দীর্ঘ কথোপকথন মনে রাখতে পারে?

এর Context উইন্ডো হলো 10 lakh Tokens। অর্থাৎ একবারে পুরো কথোপকথন এবং যুক্ত করা ফাইল মিলিয়ে এটি মোট এতটুকু পড়তে পারে।

খরচের তুলনায় ক্লড সনেট 5-এর পারফরম্যান্স বা ভ্যালু কেমন?

বুদ্ধিমত্তা এবং খরচের অনুপাত মিলিয়ে তৈরি ভ্যালু র‍্যাঙ্কিংয়ে 54টি Model-এর মধ্যে এর স্থান 24।

ক্লড সনেট 5 কি উত্তর দেওয়ার আগে চিন্তা বা Reasoning করতে পারে?

ডিফল্টভাবে চালু। যেখানে লজিক্যাল চিন্তা বা Reasoning সমর্থিত, সেখানে আপনি মেসেজ লেখার বক্সে চিন্তার মাত্রা বা Effort লেভেল সেট করতে পারেন।

ক্লড সনেট 5 কোন কোন ভারতীয় ভাষায় উত্তর দেয়?

এটি 14টি ভারতীয় ভাষায় উত্তর দিতে পারে। মেসেজ বক্সের পাশের মেনু থেকে ভাষা বেছে নিলেই সেই ভাষায় উত্তর পাওয়া যাবে।

Anthropic-এর অন্যান্য Models

ক্যাটালগ আপডেট হয়েছে: 9 সেপ, 2026