99MODELS

ஜெம்மா 4 31B

Google's open-weight 31B instruction model with vision input.

வெளியான தேதி: 2 ஏப்., 2026

143.04

10 லட்சம் output Tokens-க்கு

Input: 10 லட்சம் Tokens-க்கு ₹95.04

0% கூடுதல் கட்டணத்துடன், ஒரு அமெரிக்க டாலருக்கு ₹96 என்ற மாற்று விகிதத்தில் நிறுவனத்தின் நேரடி விலை வசூலிக்கப்படுகிறது.

தொழில்நுட்ப விவரங்கள்

Context window
2,62,144 Tokens
அதிகபட்ச Output
16,384 Tokens
ஏற்றுக்கொள்பவை
உரை, படங்கள், வீடியோ
Reasoning
விருப்பத்தேர்வு
Tool பயன்பாடு
ஆம்
Structured output
ஆம்
Code execution
இல்லை
Parameters
31B
நுண்ணறிவுத் தரம்
54-இல் #47
மதிப்புத் தரம்
54-இல் #48

கட்டண விவரம்

கட்டண விவரம்
10 lakh Tokens-க்குINRUSD
Input95.04$0.99
Output143.04$1.49
Cached input95.04$0.99

பெஞ்ச்மார்க் மதிப்பெண்கள்

மதிப்பீடாகக் குறிக்கப்பட்டவை தவிர மற்ற மதிப்பெண்கள் சதவீதங்கள் ஆகும். அனைத்து பெஞ்ச்மார்க்குகளும் தனிச்சார்பின்றி அளவிடப்பட்டவை.

  • 85.7%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 23.6%

    HLE

    Humanity's Last Exam

  • 45.5%

    SciCode

    SciCode - scientific code generation

  • 75.6%

    IFBench

    IFBench - precise instruction following

  • 69.7%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 36.4%

    Terminal-Bench Hard

    Terminal-Bench Hard - agentic terminal tasks

  • 43.4%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 1364

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

ஜெம்மா 4 31B பற்றி

An open model from Google DeepMind, built from Gemini 3 research to maximise intelligence per parameter. It is the dense 31B member of the Gemma 4 family, which Google describes as maximising raw quality and providing a powerful foundation for fine-tuning, and it handles text and image input across a 256K context. Google reports it outcompeting models twenty times its size.

Gemma is the open half of Google's model line, and Gemma 4 is released under the Apache 2.0 licence - the weights are published, and anyone may download, run, fine-tune and ship them commercially. In practice that means two routes to the same model: run it yourself on hardware you control, with nothing leaving your machine, or call it here and skip the hardware. Google sizes the family for the first route explicitly, and says the unquantized weights of this 31B fit on a single 80GB accelerator, with quantized builds running on consumer GPUs. That is a claim about where it fits, not a promise that it is easy to operate; serving it well is still work.

Inside the family the 31B is the dense member, the one Google positions for raw quality and as the base to fine-tune from, with the sparse 26B beside it for latency and two much smaller models for phones and edge devices. The lab's own comparison is against Gemma 3 27B, and the margins are large: 85.2% against 67.6% on the multilingual MMMLU, 76.9% against 49.7% on MMMU Pro, 89.2% against 20.8% on AIME 2026, 80.0% against 29.1% on LiveCodeBench v6, 84.3% against 42.4% on GPQA Diamond, and 86.4% against 6.6% on the retail split of tau2-bench. All of those are with thinking enabled.

The feature list is aimed at building rather than chatting. Function calling, structured JSON output and native system instructions are supported directly, so an agent can be wired to real tools without a wrapper coaxing the format out of it. Image and video input is native across the family at variable resolutions, with optical character recognition and chart reading called out specifically, and Google trained the models on more than 140 languages.

The context window is 256K on this size, which is generous for an open model and short of the million tokens the Gemini line carries. Audio input belongs to the two edge models and not to this one. And the claim Google is making throughout is intelligence per parameter, not parity with a frontier model: this is the most capable thing of its size, which is a different statement from the most capable thing.

அறிமுகத்தின் போது Google கூறியவை

Open weights, Apache 2.0
The weights are published under a permissive licence, so the model can be downloaded, run locally, fine-tuned and shipped commercially. Calling it here is the alternative to owning the hardware.
The dense, quality-first size
Google positions the 31B as the member that maximises raw quality and serves as a foundation for fine-tuning, beside a sparse 26B built for latency and two much smaller edge models.
Large margins over Gemma 3
Against Gemma 3 27B the lab reports 84.3% against 42.4% on GPQA Diamond, 80.0% against 29.1% on LiveCodeBench v6, and 89.2% against 20.8% on AIME 2026.
Runs on one accelerator
Google says the unquantized weights fit on a single 80GB GPU, with quantized builds running on consumer cards for local coding assistants and agent workflows.
Built for tools and documents
Function calling, structured JSON output and native system instructions are supported directly, alongside native image and video input at variable resolutions with OCR and chart reading.
What it is not for
Audio input belongs to the two edge sizes, not this one, and the context window is 256K rather than the million tokens Gemini carries. The claim is intelligence per parameter, not frontier parity.

Frequently Asked Questions

ஜெம்மா 4 31B பற்றிய பொதுவான கேள்விகள்.

ஜெம்மா 4 31B எப்போது வெளியானது?

Google நிறுவனம் ஜெம்மா 4 31B மாடலை 2 ஏப்., 2026 அன்று வெளியிட்டது.

ஜெம்மா 4 31B மாடலை உருவாக்கியது யார்?

ஜெம்மா 4 31B மாடலை Google நிறுவனம் உருவாக்கியது. 99Models AI இதனை நிறுவனத்தின் நேரடி கட்டணத்திலேயே வழங்குகிறது.

ஜெம்மா 4 31B எவ்வளவு அறிவார்ந்தது?

சுயாதீன Benchmark மதிப்பெண்களின்படி வரிசைப்படுத்தப்பட்ட நுண்ணறிவு தரவரிசையில், இது 54 chat Models-களில் 47-வது இடத்தில் உள்ளது. இதன் முழு மதிப்பெண்களை மேலே உள்ள Benchmarks பகுதியில் காணலாம்.

ஜெம்மா 4 31B பயன்பாட்டுக்கான கட்டணம் எவ்வளவு?

10 லட்சம் input Tokens-க்கு ₹95.04 மற்றும் 10 லட்சம் output Tokens-க்கு ₹143.04, கூடுதல் கட்டணமின்றி நிறுவனத்தின் நேரடி விலையிலேயே வசூலிக்கப்படுகிறது. சந்தா எதுவும் இல்லை; பயன்படுத்தியதற்கு மட்டுமே கட்டணம்.

டாலரில் ஜெம்மா 4 31B API கட்டணம் என்ன?

நிறுவனம் 10 லட்சம் input Tokens-க்கு $0.99 மற்றும் 10 லட்சம் output Tokens-க்கு $1.49 வசூலிக்கிறது. இப்பக்கத்தில் உள்ள ரூபாய் கட்டணம் ஒரு டாலருக்கு ₹96 என மாற்றப்பட்டதாகும்.

ஜெம்மா 4 31B எவ்வளவு நீண்ட உரையாடலைக் கையாளும்?

இதன் Context window 2.6 lakh Tokens. ஒரே கோரிக்கையில் இது படிக்கக்கூடிய முந்தைய உரையாடல் மற்றும் இணைக்கப்பட்ட கோப்புகளின் மொத்த அளவு இதுவாகும்.

மதிப்பு (Value) தரவரிசையில் ஜெம்மா 4 31B எந்த இடத்தில் உள்ளது?

திறன் மற்றும் Token கட்டணத்தை ஒப்பிடும் எங்கள் மதிப்பு (Value) தரவரிசையில், இது 54 Models-களில் 48-வது இடத்தில் உள்ளது.

ஜெம்மா 4 31B பதிலளிக்கும் முன் யோசிக்குமா (Reasoning)?

விருப்பத்தேர்வு. Reasoning வசதி உள்ள மாடல்களில், சிந்தனை அளவை (thinking effort) செய்தி உள்ளிடும் இடத்திலேயே நீங்கள் மாற்றிக்கொள்ளலாம்.

Google நிறுவனத்தின் பிற மாடல்கள்

கேட்லாக் புதுப்பிக்கப்பட்டது 9 செப்., 2026