99MODELS

GLM 4.7

Cheap GLM generation for everyday chat and tool use.

ਲਾਂਚ ਮਿਤੀ: 22 ਦਸੰ 2025

254.40

ਪ੍ਰਤੀ 10 ਲੱਖ ਆਊਟਪੁੱਟ Tokens

ਇਨਪੁੱਟ: ₹57.60 ਪ੍ਰਤੀ 10 ਲੱਖ Tokens

0% ਮਾਰਕਅੱਪ ਦੇ ਨਾਲ, ₹96 ਪ੍ਰਤੀ ਅਮਰੀਕੀ ਡਾਲਰ ਦੀ ਦਰ 'ਤੇ ਪ੍ਰੋਵਾਈਡਰ ਰੇਟ ਅਨੁਸਾਰ ਬਿਲਿੰਗ।

ਵਿਸ਼ੇਸ਼ਤਾਵਾਂ

Context ਵਿੰਡੋ
2,04,800 Tokens
ਵੱਧ ਤੋਂ ਵੱਧ ਆਊਟਪੁੱਟ
1,31,072 Tokens
ਸਵੀਕਾਰ ਕਰਦਾ ਹੈ
ਟੈਕਸਟ
Reasoning
ਡਿਫੌਲਟ ਤੌਰ 'ਤੇ ਚਾਲੂ
ਟੂਲ ਦੀ ਵਰਤੋਂ
ਹਾਂ
ਸਟ੍ਰਕਚਰਡ ਆਊਟਪੁੱਟ
ਹਾਂ
ਕੋਡ ਐਗਜ਼ੀਕਿਊਸ਼ਨ
ਨਹੀਂ
ਇੰਟੈਲੀਜੈਂਸ ਰੈਂਕ
54 ਵਿੱਚੋਂ #41
ਵੈਲਯੂ ਰੈਂਕ
54 ਵਿੱਚੋਂ #41

ਕੀਮਤ

ਕੀਮਤ
ਪ੍ਰਤੀ 10 ਲੱਖ TokensINRUSD
ਇਨਪੁਟ57.60$0.60
ਆਊਟਪੁੱਟ254.40$2.65
ਕੈਸ਼ਡ ਇਨਪੁਟ11.52$0.12

ਬੈਂਚਮਾਰਕ

ਸਾਰੇ ਸਕੋਰ ਪ੍ਰਤੀਸ਼ਤ ਵਿੱਚ ਹਨ, ਜਦੋਂ ਤੱਕ ਕੋਈ ਰੇਟਿੰਗ ਨਾ ਦਿੱਤੀ ਗਈ ਹੋਵੇ। ਬੈਂਚਮਾਰਕ ਸੁਤੰਤਰ ਤੌਰ 'ਤੇ ਮਾਪੇ ਗਏ ਹਨ।

  • 85.6%

    MMLU-Pro

    MMLU-Pro - multitask language understanding

  • 85.9%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 27.4%

    HLE

    Humanity's Last Exam

  • 89.4%

    LiveCodeBench

    LiveCodeBench - contamination-free coding

  • 95.0%

    AIME 2025

    AIME 2025 - competition mathematics

  • 67.9%

    IFBench

    IFBench - precise instruction following

  • 71.0%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 31.8%

    Terminal-Bench Hard

    Terminal-Bench Hard - agentic terminal tasks

  • 45.3%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 1434

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

GLM 4.7 ਬਾਰੇ

Z.ai's coding partner model, upgraded in two areas over its predecessor: multi-language coding and terminal-agent performance, and more stable multi-step reasoning and execution. It is a Mixture-of-Experts model of roughly 358B parameters covering core coding, tool use and complex reasoning. It introduces interleaved thinking before every response and tool call, preserved reasoning across turns, and per-turn control to switch thinking off for lightweight requests.

GLM-4.7 is the generation before Z.ai moved to the GLM-5 line, and it sits in this catalog as the cheaper, well-understood option rather than as a flagship. Z.ai pitched it as a coding partner, and the gains it claimed over GLM-4.6 are concentrated where a coding agent actually spends its time: 73.8% on SWE-bench, up 5.8 points; 66.7% on the multilingual variant, up 12.9; and 41% on Terminal-Bench 2.0, up 16.5.

The part of this release that has aged best is the thinking control, which is finer-grained than a simple on-or-off switch. Interleaved thinking means the model reasons before every response and every tool call. Preserved thinking means that in coding-agent sessions it carries its own reasoning blocks forward across turns instead of re-deriving them, which is what stops a long session drifting into inconsistency. Turn-level thinking lets a caller disable reasoning for a lightweight request and turn it back on for a hard one, inside the same conversation.

Away from code Z.ai reports gains in tool use and browsing, at 87.4 on the tau-squared agentic benchmark against GLM-4.6's 75.2 and 52 on BrowseComp against 45.1, plus a substantial mathematics and reasoning jump that it summarises as 42.8% on Humanity's Last Exam with tools, up 12.4 points. It also claims plainly better interface output than its predecessor: cleaner, more modern web pages, and slides with more accurate layout and sizing.

Z.ai's own comparison table is where the scope of the model is honest. On Terminal-Bench 2.0 its 41 sits behind Gemini 3.0 Pro's 54.2 and GPT-5.1 High's 47.6, and on the harder terminal set its 33.3 is behind GPT-5.1 High's 43. It was never presented as the strongest agentic model available, only as the strongest for what it cost, which is the reason it is still worth having on the shelf.

The weights are published on Hugging Face and ModelScope with vLLM and SGLang support, so it can be self-hosted; the copy served here is Z.ai's. Reach for it when the task is routine, the volume is high and the rate matters more than the last few points of capability, and reach for the GLM-5 line when it does not.

ਲਾਂਚ ਸਮੇਂ Z.ai ਨੇ ਕੀ ਕਿਹਾ

Core coding
Z.ai reports 73.8% on SWE-bench, 66.7% on its multilingual variant and 41% on Terminal-Bench 2.0, gains of 5.8, 12.9 and 16.5 points over GLM-4.6.
Vibe coding
The lab claims a step up in interface quality specifically: cleaner and more modern web pages, and generated slides with more accurate layout and sizing than the previous generation.
Tool use
Z.ai reports 87.4 on the tau-squared agentic tool-use benchmark against GLM-4.6's 75.2, and 52 on BrowseComp against 45.1, as the clearest non-coding gain of the release.
Complex reasoning
A substantial mathematics and reasoning improvement, summarised by the lab as 42.8% on Humanity's Last Exam with tools, up 12.4 points on GLM-4.6.
Three thinking controls
Reasoning before every response and tool call, reasoning preserved across turns in agent sessions rather than re-derived, and per-turn control to switch it off for lightweight requests.
Where it sits now
Z.ai's own table places it behind the leading closed models of its day on terminal-agent work. It is the older, cheaper option here, not the current frontier of the GLM line.

ਭਾਰਤੀ ਭਾਸ਼ਾਵਾਂ

GLM 4.7 8 ਭਾਰਤੀ ਭਾਸ਼ਾਵਾਂ ਵਿੱਚ ਜਵਾਬ ਦਿੰਦਾ ਹੈ। ਮੈਸੇਜ ਬਾਕਸ ਦੇ ਨਾਲ ਵਾਲੇ ਮੀਨੂ ਵਿੱਚੋਂ ਭਾਸ਼ਾ ਚੁਣੋ ਅਤੇ ਉਸੇ ਵਿੱਚ ਜਵਾਬ ਪ੍ਰਾਪਤ ਕਰੋ।

Frequently Asked Questions

GLM 4.7 ਬਾਰੇ ਅਕਸਰ ਪੁੱਛੇ ਜਾਣ ਵਾਲੇ ਸਵਾਲ।

GLM 4.7 ਕਦੋਂ ਲਾਂਚ ਹੋਇਆ ਸੀ?

Z.ai ਨੇ 22 ਦਸੰ 2025 ਨੂੰ GLM 4.7 ਲਾਂਚ ਕੀਤਾ।

GLM 4.7 ਕਿਸਨੇ ਬਣਾਇਆ ਹੈ?

GLM 4.7 ਨੂੰ Z.ai ਨੇ ਬਣਾਇਆ ਹੈ। 99Models AI ਸਿੱਧਾ ਇਸ ਨਾਲ ਪ੍ਰੋਵਾਈਡਰ ਦੀ ਨਿਰਧਾਰਿਤ ਦਰ 'ਤੇ ਕਨੈਕਟ ਕਰਦਾ ਹੈ।

GLM 4.7 ਕਿੰਨਾ ਸਮਝਦਾਰ ਹੈ?

ਇੰਟੈਲੀਜੈਂਸ ਰੈਂਕਿੰਗ ਵਿੱਚ ਇਹ 54 ਚੈਟ Models ਵਿੱਚੋਂ 41ਵੇਂ ਨੰਬਰ 'ਤੇ ਹੈ, ਜੋ ਸੁਤੰਤਰ ਬੈਂਚਮਾਰਕ ਸਕੋਰਾਂ 'ਤੇ ਆਧਾਰਿਤ ਹੈ। ਇਸਦੇ ਪੂਰੇ ਸਕੋਰ ਉੱਪਰ ਦਿੱਤੇ Benchmarks ਟੇਬਲ ਵਿੱਚ ਵੇਖੋ।

GLM 4.7 ਦੀ ਲਾਗਤ ਕਿੰਨੀ ਹੈ?

10 ਲੱਖ ਇਨਪੁਟ Tokens ਲਈ ₹57.60 ਅਤੇ 10 ਲੱਖ ਆਊਟਪੁਟ Tokens ਲਈ ₹254.40, ਬਿਨਾਂ ਕਿਸੇ ਮਾਰਕਅੱਪ ਦੇ। ਕੋਈ ਸਬਸਕ੍ਰਿਪਸ਼ਨ ਨਹੀਂ ਹੈ; ਤੁਸੀਂ ਸਿਰਫ਼ ਵਰਤੋਂ ਦਾ ਭੁਗਤਾਨ ਕਰਦੇ ਹੋ।

ਅਮਰੀਕੀ ਡਾਲਰਾਂ ਵਿੱਚ GLM 4.7 ਦੀ API ਕੀਮਤ ਕੀ ਹੈ?

ਪ੍ਰੋਵਾਈਡਰ 10 ਲੱਖ ਇਨਪੁਟ Tokens ਲਈ $0.60 ਅਤੇ 10 ਲੱਖ ਆਊਟਪੁਟ Tokens ਲਈ $2.65 ਚਾਰਜ ਕਰਦਾ ਹੈ। ਰੁਪਏ ਦੀਆਂ ਦਰਾਂ ₹96 ਪ੍ਰਤੀ ਅਮਰੀਕੀ ਡਾਲਰ ਦੇ ਹਿਸਾਬ ਨਾਲ ਬਦਲੀਆਂ ਗਈਆਂ ਹਨ।

GLM 4.7 ਕਿੰਨੀ ਲੰਮੀ ਗੱਲਬਾਤ ਯਾਦ ਰੱਖ ਸਕਦਾ ਹੈ?

ਇਸਦੀ Context ਵਿੰਡੋ 2 lakh Tokens ਹੈ। ਇਹ ਉਹ ਕੁੱਲ ਟੈਕਸਟ ਅਤੇ ਫ਼ਾਈਲਾਂ ਹਨ ਜਿਨ੍ਹਾਂ ਨੂੰ ਇਹ ਇੱਕੋ ਬੇਨਤੀ ਵਿੱਚ ਪੜ੍ਹ ਸਕਦਾ ਹੈ।

ਕੀਮਤ ਅਤੇ ਗੁਣਵੱਤਾ (ਵੈਲਿਊ) ਪੱਖੋਂ GLM 4.7 ਦਾ ਰੈਂਕ ਕੀ ਹੈ?

ਵੈਲਿਊ ਰੈਂਕਿੰਗ ਵਿੱਚ ਇਹ 54 Models ਵਿੱਚੋਂ 41ਵੇਂ ਨੰਬਰ 'ਤੇ ਹੈ। ਇਹ ਰੈਂਕਿੰਗ ਬੈਂਚਮਾਰਕ ਸਮਰੱਥਾ ਅਤੇ ਟੋਕਨ ਦੀ ਕੀਮਤ ਦੀ ਤੁਲਨਾ ਕਰਦੀ ਹੈ।

ਕੀ GLM 4.7 ਜਵਾਬ ਦੇਣ ਤੋਂ ਪਹਿਲਾਂ ਸੋਚਦਾ (Reasoning ਕਰਦਾ) ਹੈ?

ਡਿਫੌਲਟ ਤੌਰ 'ਤੇ ਚਾਲੂ। ਜਿੱਥੇ Reasoning ਸਮਰਥਿਤ ਹੈ, ਤੁਸੀਂ ਮੈਸੇਜ ਬਾਕਸ ਵਿੱਚ ਸਿੱਧਾ ਸੋਚਣ ਦਾ ਪੱਧਰ ਤੈਅ ਕਰ ਸਕਦੇ ਹੋ।

GLM 4.7 ਕਿਹੜੀਆਂ ਭਾਰਤੀ ਭਾਸ਼ਾਵਾਂ ਵਿੱਚ ਜਵਾਬ ਦਿੰਦਾ ਹੈ?

ਇਹ 8 ਭਾਰਤੀ ਭਾਸ਼ਾਵਾਂ ਵਿੱਚ ਜਵਾਬ ਦਿੰਦਾ ਹੈ। ਮੈਸੇਜ ਬਾਕਸ ਦੇ ਨਾਲ ਦਿੱਤੇ ਮੀਨੂ ਵਿੱਚੋਂ ਭਾਸ਼ਾ ਚੁਣੋ ਅਤੇ ਜਵਾਬ ਉਸੇ ਭਾਸ਼ਾ ਵਿੱਚ ਪ੍ਰਾਪਤ ਕਰੋ।

Z.ai ਦੇ ਹੋਰ Models

ਕੈਟਾਲਾਗ ਅੱਪਡੇਟ: 9 ਸਤੰ 2026