99MODELS

GLM 5.1

Previous GLM-5 generation; solid general reasoning.

تاریخ اجراء: 7 اپریل، 2026

464.64

فی 10 لاکھ آؤٹ پٹ Tokens

ان پٹ: ₹147.84 فی 10 لاکھ Tokens

فراہم کنندہ کے نرخ کے مطابق ₹96 فی امریکی ڈالر پر تبدیل شدہ، بغیر کسی اضافی مارک اپ کے۔

تکنیکی تفصیلات

Context ونڈو
2,04,800 Tokens
زیادہ سے زیادہ آؤٹ پٹ
65,536 Tokens
سپورٹ
ٹیکسٹ
Reasoning
پہلے سے آن
ٹول کا استعمال
ہاں
سٹرکچرڈ آؤٹ پٹ
ہاں
کوڈ ایگزیکیوشن
نہیں
ذہانت کا درجہ
54 میں سے #38
ویلیو کا درجہ
54 میں سے #42

قیمتیں

قیمتیں
فی 10 lakh TokensINRUSD
ان پٹ147.84$1.54
آؤٹ پٹ464.64$4.84
کیشڈ ان پٹ27.46$0.29

Benchmarks

تمام اسکورز فیصد میں ہیں، ماسوائے جہاں ریٹنگ درج ہو۔ یہ تمام پیمائشیں آزادانہ طور پر کی گئی ہیں۔

  • 86.8%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 30.1%

    HLE

    Humanity's Last Exam

  • 76.3%

    IFBench

    IFBench - precise instruction following

  • 73.7%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 43.2%

    Terminal-Bench Hard

    Terminal-Bench Hard - agentic terminal tasks

  • 61.8%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 74.2%

    SWE-bench Verified

    SWE-bench Verified - real-world bug fixing (Epoch AI run)

  • 1509

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

GLM 5.1 کے بارے میں

Z.ai's previous flagship for agentic engineering, with significantly stronger coding than its predecessor. It is designed for long-horizon autonomous work, sustaining optimisation over hundreds of rounds and thousands of tool calls rather than plateauing early, with performance that improves as the task lengthens -- Z.ai cites sustained execution for up to eight hours. Thinking is on by default and can be disabled.

The framing of the 5.1 post is a claim about time rather than about peak scores. Z.ai argues that earlier models, its own GLM-5 included, exhaust their repertoire early: they apply the familiar techniques, take a quick gain, then plateau, and giving them longer does not help. GLM-5.1 was built so that extra runtime keeps paying, and the post spends most of its length demonstrating that rather than listing results.

It uses three tasks with progressively less structured feedback. The first is an open-source vector-database challenge scored purely on queries per second at 95% recall, normally run inside a 50-turn tool budget where the best result on record was about 3,500 queries per second. Z.ai restructured it as an outer optimisation loop and reports GLM-5.1 still finding real improvements past 600 iterations and 6,000 tool calls, finishing at 21.5 thousand queries per second. The trajectory is a staircase: six structural rewrites, each chosen by the model after reading its own benchmark logs, with recall temporarily broken around each transition and then restored.

The second is more honest about the ceiling. On the hardest level of a GPU kernel benchmark, where whole architectures have to be optimised end to end, Z.ai reports GLM-5.1 reaching a 3.6 times geometric-mean speedup and still improving late in the run, while naming Claude Opus 4.6 as the stronger model in that setting at 4.2 times with headroom left. The third has no metric at all: an eight-hour run building a Linux-style desktop as a web application, wrapped in a harness that asks the model after each round what is still missing.

The benchmark numbers behind that: Z.ai reports 58.4 on SWE-bench Pro, 42.7 on natural-language-to-repository generation and 69.0 on Terminal-Bench 2.0 with its best harness, against GLM-5's 55.1, 35.9 and 56.2, plus 68.7 on the CyberGym vulnerability-reproduction set against GLM-5's 48.3.

The post ends on what is not solved, which is the part worth reading before you plan a long autonomous run on it: escaping a local optimum earlier once incremental tuning stops paying, holding coherence across execution traces of thousands of tool calls, and, the one Z.ai flags as most important, reliable self-evaluation where there is no number to optimise against. It calls this a first step in that direction. The weights are MIT-licensed and can be self-hosted; the copy served here is Z.ai's.

لانچ کے وقت Z.ai کا بیان

Complex software engineering
Z.ai reports state-of-the-art results on SWE-bench Pro at 58.4, ahead of its own GLM-5 at 55.1 and of the closed models it compares against on that particular set.
Real-world terminal tasks
On Terminal-Bench 2.0 the lab reports 69.0 with its best harness against 56.2 for GLM-5, one of the widest generational gaps in the post.
Repository generation
On long-horizon repository generation from a natural-language brief, Z.ai reports 42.7 against 35.9 for GLM-5, with Claude Opus 4.6 still ahead at 49.8.
Cybersecurity reproduction
On CyberGym, which asks a model to reproduce known software vulnerabilities, the lab reports 68.7 against 48.3 for GLM-5 and ahead of the closed models in its table.
Runtime that keeps paying
The central claim: on a vector-search optimisation task the lab reports the model still improving past 600 iterations and 6,000 tool calls, ending roughly six times better than the best single-session result.
What it is not for
Z.ai names three unsolved problems: escaping local optima earlier, holding coherence over thousands of tool calls, and self-evaluation on tasks with no numeric target. It is behind Claude Opus 4.6 on kernel optimisation.

ہندوستانی زبانیں

GLM 5.1 11 ہندوستانی زبانوں میں جواب دیتا ہے۔ میسج باکس کے ساتھ والے مینو سے زبان کا انتخاب کریں۔

Frequently Asked Questions

GLM 5.1 کے بارے میں اکثر پوچھے جانے والے سوالات۔

GLM 5.1 کب جاری ہوا تھا؟

Z.ai نے GLM 5.1 کو 7 اپریل، 2026 کو جاری کیا۔

GLM 5.1 کس نے بنایا ہے؟

GLM 5.1 کو Z.ai نے تیار کیا ہے۔ 99Models AI فراہم کنندہ کے اصل نرخ پر اس سے براہ راست جوڑتا ہے۔

GLM 5.1 کتنا ذہین ہے؟

یہ ذہانت کی درجہ بندی میں 54 چیٹ ماڈلز میں سے 38 نمبر پر ہے۔ اس کے تمام بینچ مارک اسکور اوپر والے ٹیبل میں دیکھے جا سکتے ہیں۔

GLM 5.1 کا کتنا خرچ آتا ہے؟

استعمال کی لاگت ₹147.84 فی 10 لاکھ ان پٹ Tokens اور ₹464.64 فی 10 لاکھ آؤٹ پٹ Tokens ہے، بغیر کسی اضافی فیس کے۔ کوئی سبسکرپشن نہیں؛ صرف استعمال کی ادائیگی کریں۔

امریکی ڈالر میں GLM 5.1 کی قیمت کیا ہے؟

فراہم کنندہ $1.54 فی 10 لاکھ ان پٹ Tokens اور $4.84 فی 10 لاکھ آؤٹ پٹ Tokens لیتا ہے۔ روپے کے نرخ ₹96 فی امریکی ڈالر کے حساب سے تبدیل کیے گئے ہیں۔

GLM 5.1 کتنی لمبی گفتگو یاد رکھ سکتا ہے؟

اس کی Context حد 2 lakh Tokens ہے، یعنی وہ تمام متن اور فائلیں جو یہ ایک ہی میسج میں پڑھ سکتا ہے۔

کیا GLM 5.1 مناسب قیمت میں بہترین کارکردگی دیتا ہے؟

یہ بہترین قیمت کی درجہ بندی میں 54 ماڈلز میں سے 42 نمبر پر ہے، جس میں ذہانت اور ٹوکن کی لاگت کا موازنہ کیا گیا ہے۔

کیا GLM 5.1 جواب دینے سے پہلے سوچتا ہے؟

پہلے سے آن۔ جہاں Reasoning کی سہولت موجود ہو، وہاں آپ میسج باکس میں سوچنے کی سطح خود طے کر سکتے ہیں۔

GLM 5.1 کن ہندوستانی زبانوں میں جواب دیتا ہے؟

یہ 11 ہندوستانی زبانوں میں جواب دیتا ہے۔ میسج باکس کے ساتھ والے مینو سے اپنی زبان منتخب کریں۔

Z.ai کے مزید ماڈلز

کیٹلاگ اپ ڈیٹ: 9 ستمبر، 2026