99MODELS

جیمنی 2.5 فلیش لائٹ

Very cheap Gemini for classification and high-volume extraction.

تاریخ اجراء: 22 جولائی، 2025

38.40

فی 10 لاکھ آؤٹ پٹ Tokens

ان پٹ: ₹9.60 فی 10 لاکھ Tokens

فراہم کنندہ کے نرخ کے مطابق ₹96 فی امریکی ڈالر پر تبدیل شدہ، بغیر کسی اضافی مارک اپ کے۔

تکنیکی تفصیلات

Context ونڈو
10,48,576 Tokens
زیادہ سے زیادہ آؤٹ پٹ
65,535 Tokens
سپورٹ
ٹیکسٹ, تصاویر, PDF فائلیں, آڈیو, ویڈیو
Reasoning
پہلے سے آن
ٹول کا استعمال
ہاں
سٹرکچرڈ آؤٹ پٹ
ہاں
کوڈ ایگزیکیوشن
ہاں
نالج کٹ آف
2025-01-31
ذہانت کا درجہ
54 میں سے #52
ویلیو کا درجہ
54 میں سے #49

قیمتیں

قیمتیں
فی 10 lakh TokensINRUSD
ان پٹ9.60$0.10
آؤٹ پٹ38.40$0.40
کیشڈ ان پٹ0.96$0.01

Benchmarks

تمام اسکورز فیصد میں ہیں، ماسوائے جہاں ریٹنگ درج ہو۔ یہ تمام پیمائشیں آزادانہ طور پر کی گئی ہیں۔

  • 75.9%

    MMLU-Pro

    MMLU-Pro - multitask language understanding

  • 62.5%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 6.8%

    HLE

    Humanity's Last Exam

  • 59.3%

    LiveCodeBench

    LiveCodeBench - contamination-free coding

  • 96.9%

    MATH-500

    MATH-500 - competition mathematics, 500 problems

  • 70.3%

    AIME 2024

    AIME 2024 - competition mathematics

  • 53.3%

    AIME 2025

    AIME 2025 - competition mathematics

  • 49.9%

    IFBench

    IFBench - precise instruction following

  • 55.7%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 4.5%

    Terminal-Bench Hard

    Terminal-Bench Hard - agentic terminal tasks

جیمنی 2.5 فلیش لائٹ کے بارے میں

Google's most cost-efficient multimodal model of its generation, offering the fastest performance for high-frequency, lightweight tasks. It is aimed at high-volume classification, simple data extraction and very low-latency applications where budget and speed are the primary constraints. Thinking is off by default, which is what makes it cheap.

Google described this model as pushing the frontier of intelligence per dollar, and everything about it follows from that sentence. It was the cheapest and fastest model in the 2.5 family at general availability, and the post pairs the price with a 40% cut to audio input pricing that had applied during the preview - a detail that matters if the workload is transcript or call data rather than text.

The speed claim is made against specific predecessors: Google reports lower latency than both the previous generation's Flash-Lite and its full Flash model across a broad sample of prompts, and higher all-round quality than the previous Flash-Lite on coding, mathematics, science, reasoning and multimodal understanding. The lab is careful to say this is a balance for latency-sensitive work such as translation and classification, not a claim about hard problems.

Being the cheap tier does not mean being a stripped one. It carries the same million-token context window as the rest of the family, controllable thinking budgets, and the native tools - grounding with Google Search, code execution and URL context - so a high-volume pipeline does not have to fall back to a larger model just to fetch a page or run a snippet.

The deployments Google chose to name are a good description of the shape of work it fits: summarising satellite telemetry in orbit, where the lab reports a 45% latency reduction and 30% lower power draw against the operator's previous model; planning and translating video content across more than 180 languages; and processing long product videos into documentation by extracting thousands of frames. All of them are the same job done a very large number of times.

What it is not for is anything demanding. Reasoning is off by default here and has to be turned on deliberately, and Google's own framing puts Flash above it for everyday tasks and Pro above that for coding and complex work. Choosing it is choosing throughput and cost over headroom.

لانچ کے وقت Google کا بیان

Intelligence per dollar
Google released it as the fastest and lowest-cost model in the 2.5 family, and cut audio input pricing by 40% from the preview rate at the same time.
Measured against its predecessors
The lab reports lower latency than both the previous generation's Flash-Lite and its full Flash model on a broad sample of prompts, and higher quality than that Flash-Lite across the board.
Cheap but not stripped
It keeps the million-token context window, controllable thinking budgets, and the native tools: grounding with Google Search, code execution and URL context.
Volume work, done many times
The deployments Google names are in-orbit telemetry summarisation with a reported 45% latency reduction, video planning and translation across 180 languages, and video-to-documentation pipelines.
What it is not for
Reasoning is off unless the caller turns it on, and Google places Flash above it for everyday tasks and Pro above that for coding and complex work. It trades headroom for throughput.

ہندوستانی زبانیں

جیمنی 2.5 فلیش لائٹ 11 ہندوستانی زبانوں میں جواب دیتا ہے۔ میسج باکس کے ساتھ والے مینو سے زبان کا انتخاب کریں۔

Frequently Asked Questions

جیمنی 2.5 فلیش لائٹ کے بارے میں اکثر پوچھے جانے والے سوالات۔

جیمنی 2.5 فلیش لائٹ کب جاری ہوا تھا؟

Google نے جیمنی 2.5 فلیش لائٹ کو 22 جولائی، 2025 کو جاری کیا۔

جیمنی 2.5 فلیش لائٹ کس نے بنایا ہے؟

جیمنی 2.5 فلیش لائٹ کو Google نے تیار کیا ہے۔ 99Models AI فراہم کنندہ کے اصل نرخ پر اس سے براہ راست جوڑتا ہے۔

جیمنی 2.5 فلیش لائٹ کتنا ذہین ہے؟

یہ ذہانت کی درجہ بندی میں 54 چیٹ ماڈلز میں سے 52 نمبر پر ہے۔ اس کے تمام بینچ مارک اسکور اوپر والے ٹیبل میں دیکھے جا سکتے ہیں۔

جیمنی 2.5 فلیش لائٹ کا کتنا خرچ آتا ہے؟

استعمال کی لاگت ₹9.60 فی 10 لاکھ ان پٹ Tokens اور ₹38.40 فی 10 لاکھ آؤٹ پٹ Tokens ہے، بغیر کسی اضافی فیس کے۔ کوئی سبسکرپشن نہیں؛ صرف استعمال کی ادائیگی کریں۔

امریکی ڈالر میں جیمنی 2.5 فلیش لائٹ کی قیمت کیا ہے؟

فراہم کنندہ $0.10 فی 10 لاکھ ان پٹ Tokens اور $0.40 فی 10 لاکھ آؤٹ پٹ Tokens لیتا ہے۔ روپے کے نرخ ₹96 فی امریکی ڈالر کے حساب سے تبدیل کیے گئے ہیں۔

جیمنی 2.5 فلیش لائٹ کتنی لمبی گفتگو یاد رکھ سکتا ہے؟

اس کی Context حد 10.5 lakh Tokens ہے، یعنی وہ تمام متن اور فائلیں جو یہ ایک ہی میسج میں پڑھ سکتا ہے۔

کیا جیمنی 2.5 فلیش لائٹ مناسب قیمت میں بہترین کارکردگی دیتا ہے؟

یہ بہترین قیمت کی درجہ بندی میں 54 ماڈلز میں سے 49 نمبر پر ہے، جس میں ذہانت اور ٹوکن کی لاگت کا موازنہ کیا گیا ہے۔

کیا جیمنی 2.5 فلیش لائٹ جواب دینے سے پہلے سوچتا ہے؟

پہلے سے آن۔ جہاں Reasoning کی سہولت موجود ہو، وہاں آپ میسج باکس میں سوچنے کی سطح خود طے کر سکتے ہیں۔

جیمنی 2.5 فلیش لائٹ کن ہندوستانی زبانوں میں جواب دیتا ہے؟

یہ 11 ہندوستانی زبانوں میں جواب دیتا ہے۔ میسج باکس کے ساتھ والے مینو سے اپنی زبان منتخب کریں۔

Google کے مزید ماڈلز

کیٹلاگ اپ ڈیٹ: 9 ستمبر، 2026