99MODELS

GLM 5.3

പുതിയത്

Flagship GLM generation for long-horizon engineering and agents.

റിലീസ് തീയതി: 2026 ഓഗ 18

633.60

10 ലക്ഷം ഔട്ട്പുട്ട് Tokens-ന്

ഇൻപുട്ട്: 10 ലക്ഷം Tokens-ന് ₹201.60

0% മാർക്ക്അപ്പിൽ, ഒരു ഡോളറിന് ₹96 എന്ന നിരക്കിലാണ് ഈടാക്കുന്നത്.

സവിശേഷതകൾ

Context window
13,10,720 Tokens
പരമാവധി Output
9,43,717 Tokens
സ്വീകരിക്കുന്നത്
ടെക്സ്റ്റ്
Reasoning
ഡിഫോൾട്ടായി ഓൺ ആണ്
Effort ലെവലുകൾ
low, high, max
ടൂൾ ഉപയോഗം
ഉണ്ട്
സ്ട്രക്ചേർഡ് Output
ഉണ്ട്
കോഡ് എക്സിക്യൂഷൻ
ഇല്ല
ഇന്റലിജൻസ് റാങ്ക്
54-ൽ #10
വാല്യു റാങ്ക്
54-ൽ #13

നിരക്കുകൾ

നിരക്കുകൾ
10 lakh Tokens-ന്INRUSD
Input201.60$2.10
Output633.60$6.60
Cached input31.20$0.33

ബെഞ്ച്മാർക്കുകൾ

റേറ്റിംഗ് എന്ന് രേഖപ്പെടുത്താത്ത സ്കോറുകൾ ശതമാനത്തിലാണ്. എല്ലാ ബെഞ്ച്മാർക്കുകളും സ്വതന്ത്രമായി പരിശോധിച്ചവയാണ്.

  • 91.7%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 42.3%

    HLE

    Humanity's Last Exam

  • 59.0%

    SciCode

    SciCode - scientific code generation

  • 79.7%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 83.9%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 1599

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

GLM 5.3-നെ കുറിച്ച്

Z.ai's flagship generation for coding and long-horizon agentic work, and an unusual release in that it shares GLM-5.2's base model entirely -- every gain comes from scaled-up post-training. Z.ai reports it as the strongest open-weights coding model it has measured, with roughly a 50 percent improvement over 5.2 on its own coding evaluation, and it reaches those scores while spending fewer output tokens per task. A capability the lab says grew faster than expected is vulnerability analysis: it reasons across multiple stages of exploitation rather than spotting isolated flaws. Thinking is mandatory, at three effort levels.

The unusual thing about GLM-5.3 is what did not change. It runs on GLM-5.2's base model, unmodified - Z.ai's own summary of the release is that scaling post-training is all it did. Everything below is therefore a claim about how much is still available after pretraining stops, which makes this release readable as an experiment as much as a product.

The coding gains are the headline. Z.ai calls it the strongest open-weights coding model it has measured, and the numbers it reports are large enough to be worth stating individually: Terminal-Bench 3.0 from 4.6 to 28.3, DeepSWE v1.1 from 46.2 to 66.9, and Agents' Last Exam from 23.8 to 28.5. It attributes the durability of those gains to reinforcement-learning strategies carried over from 5.2, including compaction, which it says is what makes the improvements hold on long-horizon tasks rather than only on short ones.

The efficiency dimension is where Z.ai draws its sharpest comparison, because it measures score against tokens spent rather than score alone. On its in-house Z.ai Code Bench - a private benchmark it built to reduce contamination from public test sets - it reports 34.5% at roughly 75K output tokens per task at Max effort, against 23.4% at 96K for GLM-5.2: better and cheaper at once. At High effort it reports 31.4% at around 50K output tokens, ahead of Claude Opus 4.8's 29.5% at 120K. Z.ai is also explicit about where it stops: it says GLM-5.3 remains behind Claude Fable 5, which reaches 39.5% at Max effort.

The part Z.ai says surprised it is security. It added vulnerability-discovery data and environments to the post-training mix expecting the model to get better at spotting flaws, and reports that the capability kept developing faster than anticipated as training scaled - the model began reasoning across multiple stages of exploitation and forming coherent plans for complete chains, rather than identifying isolated bugs. Run against real codebases with security teams, and after expert review and deduplication, it reports 2,436 vulnerabilities found across 269 projects, 1,097 of them medium-to-high severity, spanning kernels, operating systems, browser engines, infrastructure, web applications and network protocols. Many had gone unnoticed for years; the oldest dated back around four decades.

Thinking is mandatory on this model and runs at three effort levels. Z.ai says the weights would follow the API release by about two weeks, after safety evaluation and hardening.

ലോഞ്ചിംഗ് വേളയിൽ Z.ai വ്യക്തമാക്കിയത്

Post-training alone
GLM-5.3 shares GLM-5.2's base model unchanged. Z.ai's own framing is that scaling post-training is all it did, which makes the gains below a measurement of how much is left after pretraining ends.
Frontier open-weights coding
Z.ai reports Terminal-Bench 3.0 rising from 4.6 to 28.3, DeepSWE v1.1 from 46.2 to 66.9 and Agents' Last Exam from 23.8 to 28.5, and calls it the strongest open-weights coding model it has measured.
Better and cheaper per task
On its private Z.ai Code Bench the lab reports 34.5% at about 75K output tokens at Max effort against 23.4% at 96K for GLM-5.2, and at High effort 31.4% at around 50K tokens against Claude Opus 4.8 at 29.5% with 120K.
Cyber capability that outgrew training
Z.ai says vulnerability analysis developed faster than it expected as training scaled: the model reasons across multiple stages of exploitation and forms complete chains rather than spotting isolated flaws.
Findings on real code
Run against real codebases with security teams, and after expert review and deduplication, Z.ai reports 2,436 vulnerabilities across 269 projects, 1,097 of them medium-to-high severity, some decades old.
What it is not for
Z.ai names its own ceiling: on its coding benchmark GLM-5.3 remains behind Claude Fable 5, which reaches 39.5% at Max effort. Thinking is mandatory here, so there is no cheap non-reasoning mode to fall back to.

Frequently Asked Questions

GLM 5.3 സംബന്ധിച്ച പ്രധാന ചോദ്യങ്ങളും ഉത്തരങ്ങളും.

GLM 5.3 എപ്പോഴാണ് റിലീസ് ചെയ്തത്?

Z.ai 2026 ഓഗ 18-ൽ GLM 5.3 പുറത്തിറക്കി.

GLM 5.3 വികസിപ്പിച്ചത് ആരാണ്?

GLM 5.3 വികസിപ്പിച്ചത് Z.ai ആണ്. 99Models AI ഒറിജിനൽ പ്രൊവൈഡർ നിരക്കിൽ തന്നെ ഇതിലേക്ക് കണക്ട് ചെയ്യുന്നു.

GLM 5.3 എത്രത്തോളം കാര്യക്ഷമമാണ്?

ഇന്റലിജൻസ് റാങ്കിംഗിൽ 54 മോഡലുകളിൽ 10-ാം സ്ഥാനത്താണ് ഇത്. ബെഞ്ച്മാർക്ക് സ്കോറുകൾ മുകളിലുള്ള പാനലിൽ കാണാം.

GLM 5.3 ഉപയോഗിക്കാൻ എത്ര ചെലവാകും?

10 ലക്ഷം ഇൻപുട്ട് Tokens-ന് ₹201.60 രൂപയും ഔട്ട്പുട്ടിന് ₹633.60 രൂപയുമാണ് അധിക മാർക്ക്അപ്പില്ലാത്ത നിരക്ക്. സബ്‌സ്‌ക്രിപ്ഷനില്ല, ഉപയോഗിക്കുന്നതിന് മാത്രം പണമടയ്ക്കുക.

GLM 5.3 API-യുടെ ഡോളർ നിരക്ക് എത്രയാണ്?

10 ലക്ഷം ഇൻപുട്ട് Tokens-ന് $2.10 ഡോളറും 10 ലക്ഷം ഔട്ട്പുട്ട് Tokens-ന് $6.60 ഡോളറുമാണ് നിരക്ക്. ഈ പേജിലെ രൂപ നിരക്കുകൾ ഡോളറിന് ₹96 എന്ന നിരക്കിൽ മാറ്റിയതാണ്.

GLM 5.3-ന് എത്ര നീളമുള്ള സംഭാഷണം ഓർത്തുനിൽക്കാനാകും?

ഇതിന്റെ Context window 13.1 lakh Tokens ആണ്. ഒരുമിച്ച് നൽകുന്ന സംഭാഷണങ്ങളും ഫയലുകളും ഉൾപ്പെടെ ഇതിൽ വായിക്കാൻ സാധിക്കും.

GLM 5.3 മികച്ച വാല്യൂ നൽകുന്ന ഒന്നാണോ?

പെർഫോമൻസും നിരക്കും അടിസ്ഥാനമാക്കിയുള്ള വാല്യൂ റാങ്കിംഗിൽ 54 മോഡലുകളിൽ 13-ാം സ്ഥാനത്താണ് ഇത്.

മറുപടി നൽകുന്നതിന് മുൻപ് GLM 5.3 Reasoning നടത്തുമോ?

ഡിഫോൾട്ടായി ഓൺ ആണ്. Reasoning പിന്തുണയ്ക്കുന്ന മോഡലുകളിൽ ചിന്തിക്കുന്നതിന്റെ വ്യാപ്തി മെസ്സേജ് ബോക്സിൽ ക്രമീകരിക്കാം.

Z.ai-ൽ നിന്നുള്ള മറ്റ് Models

കാറ്റലോഗ് പുതുക്കിയത്: 2026 സെപ്റ്റം 9