99MODELS

କ୍ଲୋଡ୍ ଓପସ୍ 5

ନୂଆ

Anthropic's flagship for demanding reasoning, coding and long-horizon agentic work.

ରିଲିଜ୍ ତାରିଖ ଜୁଲାଇ 24, 2026

2,640.00

ପ୍ରତି 10 ଲକ୍ଷ output Tokens

Input: ପ୍ରତି 10 ଲକ୍ଷ Tokens ପାଇଁ ₹528.00

0% ମାର୍କଅପ୍ ସହିତ ₹96 ପ୍ରତି US ଡଲାର ହିସାବରେ ପ୍ରୋଭାଇଡର୍ ରେଟ୍ ରେ ବିଲ୍ କରାଯାଇଛି।

ସ୍ପେସିଫିକେସନ୍

Context window
10,00,000 Tokens
ସର୍ବାଧିକ Output
1,28,000 Tokens
ଗ୍ରହଣ କରେ
ଟେକ୍ସଟ୍, ଛବି, PDF ଫାଇଲ୍
Reasoning
ଡିଫଲ୍ଟ ଭାବରେ On
Effort ସ୍ତର
low, medium, high, xhigh, max
Tool ବ୍ୟବହାର
ହଁ
Structured output
ହଁ
Code execution
ନାହିଁ
Intelligence rank
54 ମଧ୍ୟରୁ #3

ମୂଲ୍ୟ

ମୂଲ୍ୟ
ପ୍ରତି 10 ଲକ୍ଷ TokensINRUSD
Input528.00$5.50
Output2,640.00$27.50
Cached input52.80$0.55

Benchmarks

ରେଟିଂ ଭାବେ ଚିହ୍ନିତ ନ ହେଲେ ସ୍କୋରଗୁଡ଼ିକ ପ୍ରତିଶତ ଅଟେ। ସମସ୍ତ benchmark ସ୍ୱତନ୍ତ୍ର ଭାବରେ ମପାଯାଇଛି।

  • 93.2%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 54.9%

    HLE

    Humanity's Last Exam

  • 56.4%

    SciCode

    SciCode - scientific code generation

  • 79.3%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 89.1%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 53.4%

    FrontierCode

    FrontierCode - long-horizon production coding tasks

  • 90.4%

    ARC-AGI-2

    ARC-AGI-2 - abstract reasoning on novel puzzles

  • 68.3%

    OSWorld 2

    OSWorld 2 - agentic computer use (partial credit)

  • 1691

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

କ୍ଲୋଡ୍ ଓପସ୍ 5 ବିଷୟରେ

Anthropic's model for complex agentic coding and enterprise work, described by the lab as a step change over Opus 4.8 rather than an incremental one, reaching close to Fable 5's frontier intelligence at half the price. Its largest gains are in deep reasoning, long-horizon agentic tasks and test-time compute scaling, with strong results in code review and bug finding, vision, long-context work and document tasks. Thinking is adaptive and on by default, with effort selectable from low up to max.

Anthropic released Opus 5 as the model to reach for every day rather than the one held back for the hardest hour: it became the default on Claude Max and the strongest model offered on Claude Pro on the day it shipped. The lab positions it as coming close to Claude Fable 5's frontier intelligence at half the price, and it charges exactly what Opus 4.8 charged before it, so the generational gain arrives without a rate change.

Anthropic presents almost all of its results as cost-per-task curves across the effort dial rather than as single peak scores, which is the honest way to read a model whose thinking budget the caller sets. On Frontier-Bench v0.1, its software engineering evaluation, the lab reports Opus 5 ahead of every other model and more than double Opus 4.8 at a lower cost per task. On CursorBench 3.2 at max effort it lands within 0.5% of Fable 5's peak score for half the cost per task, and beats every other model at a given cost on the high, xhigh and max rungs.

The same pattern holds away from code. Anthropic reports three times the next-best model's score on ARC-AGI 3, which scores models on problems they have not seen before; a pass rate around 1.5 times the next-best model on Zapier's AutomationBench at the same cost per task, with more tasks passed than any other model even at the lowest effort setting; and better results than any model at any cost on the OSWorld 2.0 computer-use evaluation, passing Fable 5's best score at just over a third of the cost. It is also the lab's best and most cost-efficient model on GDPval-AA v2, HLE AutomationBench and DeepSearchQA.

What early-access testers described was less a score than a working habit: the model checks itself. Anthropic's own examples include a Frontier-Bench task where the model was given a mechanical drawing and deliberately no way to view it, and answered by writing a computer vision pipeline to recover the geometry from raw pixels; a real bug in a widely used package manager where it found the root cause an accepted community patch had missed; and a market data feed built in one session, with its own test harness written because no live feed existed to validate against. On science the lab reports Opus 5 ahead of Opus 4.8 on every one of its internal life-sciences evaluations, by 10.2 points on inferring molecular structures from spectroscopy and 7.7 points on predicting how protein sequence variants behave.

Anthropic is also specific about where the model stops. It says Opus 5 does not advance the frontier in dual-use capability: on its OSS-Fuzz evaluation it comes close to Mythos 5 at finding software vulnerabilities but stays far behind at turning them into working exploits, and Mythos 5 remains the stronger model for long-running autonomous biology research. Its safety classifiers allow source-code vulnerability review while blocking binary-based scanning, penetration testing and exploit generation. On the lab's automated behavioural audit it scored 2.3 for overall misaligned behaviour, the lowest of any recent Anthropic model.

ଲଞ୍ଚ ସମୟରେ Anthropic ଯାହା କହିଥିଲା

Software engineering
On Frontier-Bench v0.1 Anthropic reports Opus 5 ahead of every other model and more than double Opus 4.8's score, at a lower cost per task. On CursorBench 3.2 at max effort it lands within 0.5% of Claude Fable 5's peak for half the cost.
Problems it has not seen
On ARC-AGI 3, which tests reasoning on genuinely novel puzzles, the lab reports a score three times as high as the next-best model.
Business tasks end to end
On Zapier's AutomationBench Anthropic reports a pass rate around 1.5 times the next-best model at the same cost per task, and more tasks passed than any other model even at the lowest effort setting.
Computer use
On OSWorld 2.0 the lab reports better results than every other model at any given cost, passing Claude Fable 5's best score at just over a third of the cost.
Scientific research
Ahead of Opus 4.8 on every internal life-sciences evaluation Anthropic ran, by 10.2 points on inferring molecular structures from spectroscopy and 7.7 points on protein sequence-variant tasks.
What it is not for
Anthropic reports Opus 5 close to Mythos 5 at finding software vulnerabilities but far behind at developing exploits, and behind it on long-running autonomous biology research. Binary-based vulnerability scanning, penetration testing and exploit generation are blocked.

ଭାରତୀୟ ଭାଷା

କ୍ଲୋଡ୍ ଓପସ୍ 5 15 ଟି ଭାରତୀୟ ଭାଷାରେ ଉତ୍ତର ଦିଏ। ସେହି ଭାଷାରେ ଉତ୍ତର ପାଇବା ପାଇଁ ମେସେଜ୍ ବକ୍ସ ପାଖରେ ଥିବା ମେନୁରୁ ଭାଷା ବାଛନ୍ତୁ।

Frequently Asked Questions

କ୍ଲୋଡ୍ ଓପସ୍ 5 ବିଷୟରେ ବାରମ୍ବାର ପଚରାଯାଉଥିବା ପ୍ରଶ୍ନ।

କ୍ଲୋଡ୍ ଓପସ୍ 5 କେବେ ଲଞ୍ଚ ହୋଇଥିଲା?

Anthropic ଜୁଲାଇ 24, 2026 ରେ କ୍ଲୋଡ୍ ଓପସ୍ 5 ଲଞ୍ଚ କରିଥିଲା।

କ୍ଲୋଡ୍ ଓପସ୍ 5 କିଏ ତିଆରି କରିଛି?

କ୍ଲୋଡ୍ ଓପସ୍ 5 କୁ Anthropic ତିଆରି କରିଛି। 99Models ଏହା ସହ ସିଧାସଳଖ ପ୍ରୋଭାଇଡରଙ୍କ ନିର୍ଦ୍ଧାରିତ ରେଟ୍ ରେ ସଂଯୋଗ କରେ।

କ୍ଲୋଡ୍ ଓପସ୍ 5 କେତେ ଶକ୍ତିଶାଳୀ?

ଇଣ୍ଟେଲିଜେନ୍ସ ତାଲିକାରେ 54 ଟି chat Model ମଧ୍ୟରୁ ଏହାର ରାଙ୍କ୍ 3। ସ୍ୱାଧୀନ ବେଞ୍ଚମାର୍କ ସ୍କୋର ଆଧାରରେ ଏହା ସ୍ଥିର କରାଯାଇଛି, ଯାହା ଉପରେ ଥିବା ଟେବୁଲରେ ଉପଲବ୍ଧ।

କ୍ଲୋଡ୍ ଓପସ୍ 5 ର ମୂଲ୍ୟ କେତେ?

ବ୍ୟବହାର ଖର୍ଚ୍ଚ ପ୍ରତି 10 ଲକ୍ଷ ଇନପୁଟ୍ Tokens ପାଇଁ ₹528.00 ଏବଂ ଆଉଟପୁଟ୍ Tokens ପାଇଁ ₹2,640.00, 0% ମାର୍କଅପ୍ ସହିତ। କୌଣସି ସବସ୍କ୍ରିପସନ୍ ନାହିଁ; ଆପଣ ଯେତିକି ବ୍ୟବହାର କରିବେ ସେତିକି ପେମେଣ୍ଟ କରିବେ।

US ଡଲାରରେ କ୍ଲୋଡ୍ ଓପସ୍ 5 ର ମୂଲ୍ୟ କେତେ?

ପ୍ରୋଭାଇଡର୍ ପ୍ରତି 10 ଲକ୍ଷ ଇନପୁଟ୍ Tokens ପାଇଁ $5.50 ଏବଂ ଆଉଟପୁଟ୍ Tokens ପାଇଁ $27.50 ଚାର୍ଜ କରେ। ଏହି ପୃଷ୍ଠାର ଟଙ୍କା ମୂଲ୍ୟ ₹96 ପ୍ରତି ଡଲାର ହିସାବରେ ରୂପାନ୍ତରିତ।

କ୍ଲୋଡ୍ ଓପସ୍ 5 କେତେ ଲମ୍ବା କଥାବାର୍ତ୍ତା ମନେ ରଖିପାରିବ?

ଏହାର Context window ହେଉଛି 10 lakh Tokens। ଗୋଟିଏ request ରେ ଏହା ସମୁଦାୟ କଥାବାର୍ତ୍ତା ଏବଂ ସଂଲଗ୍ନ ଫାଇଲ୍ ପଢ଼ିପାରିବ।

ମୂଲ୍ୟ ହିସାବରେ କ୍ଲୋଡ୍ ଓପସ୍ 5 କେତେ ଭଲ?

ଭ୍ୟାଲୁ ରାଙ୍କିଙ୍ଗରେ 54 ଟି Model ମଧ୍ୟରୁ ଏହାର ସ୍ଥାନ 22। ଏହି ରାଙ୍କିଙ୍ଗ ବେଞ୍ଚମାର୍କ କ୍ଷମତା ଏବଂ Token ଖର୍ଚ୍ଚକୁ ତୁଳନା କରି ସ୍ଥିର କରାଯାଏ।

କ୍ଲୋଡ୍ ଓପସ୍ 5 କ’ଣ Reasoning ସପୋର୍ଟ କରେ?

ଡିଫଲ୍ଟ ଭାବରେ On। ଯେଉଁଠାରେ Reasoning ଉପଲବ୍ଧ, ଆପଣ ମେସେଜ୍ ବକ୍ସରେ thinking effort ସ୍ତର ସେଟ୍ କରିପାରିବେ।

କ୍ଲୋଡ୍ ଓପସ୍ 5 କେଉଁ ଭାରତୀୟ ଭାଷାରେ ଉତ୍ତର ଦେଇପାରେ?

ଏହା 15 ଟି ଭାରତୀୟ ଭାଷାରେ ଉତ୍ତର ଦିଏ। ମେସେଜ୍ ବକ୍ସ ପାଖରେ ଥିବା ମେନୁରୁ ଆପଣଙ୍କ ପସନ୍ଦର ଭାଷା ବାଛନ୍ତୁ।

Anthropic ରୁ ଅନ୍ୟାନ୍ୟ Models

କାଟାଲଗ୍ ଅପଡେଟ୍ ହୋଇଛି: ସେପ୍ଟେମ୍ବର 9, 2026