99MODELS

جیمنی 3.5 فلیش لائٹ

Cheapest current Gemini; all input modalities, minimal reasoning by default.

تاریخ اجراء: 21 جولائی، 2026

264.00

فی 10 لاکھ آؤٹ پٹ Tokens

ان پٹ: ₹31.68 فی 10 لاکھ Tokens

فراہم کنندہ کے نرخ کے مطابق ₹96 فی امریکی ڈالر پر تبدیل شدہ، بغیر کسی اضافی مارک اپ کے۔

تکنیکی تفصیلات

Context ونڈو
10,48,576 Tokens
زیادہ سے زیادہ آؤٹ پٹ
65,536 Tokens
سپورٹ
ٹیکسٹ, تصاویر, PDF فائلیں, آڈیو, ویڈیو
Reasoning
پہلے سے آن
ایفرٹ لیولز
minimal, low, medium, high
ٹول کا استعمال
ہاں
سٹرکچرڈ آؤٹ پٹ
ہاں
کوڈ ایگزیکیوشن
ہاں
ذہانت کا درجہ
54 میں سے #40
ویلیو کا درجہ
54 میں سے #36

قیمتیں

قیمتیں
فی 10 lakh TokensINRUSD
ان پٹ31.68$0.33
آؤٹ پٹ264.00$2.75
کیشڈ ان پٹ3.17$0.03

Benchmarks

تمام اسکورز فیصد میں ہیں، ماسوائے جہاں ریٹنگ درج ہو۔ یہ تمام پیمائشیں آزادانہ طور پر کی گئی ہیں۔

  • 83.8%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 18.8%

    HLE

    Humanity's Last Exam

  • 41.3%

    SciCode

    SciCode - scientific code generation

  • 76.0%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 53.6%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 10.3%

    ARC-AGI-2

    ARC-AGI-2 - abstract reasoning on novel puzzles

  • 1449

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

جیمنی 3.5 فلیش لائٹ کے بارے میں

A low-latency, cost-effective multimodal Gemini optimised for high-throughput, low-cost execution -- Google aims it squarely at subagents running one focused task inside a larger workflow, and at document parsing. It is the fastest model in the 3.5 series, at roughly 350 output tokens per second, and can be pinned to minimal or low thinking for cheap work or raised for harder tasks. Full multimodal input and a million-token context are retained.

Flash-Lite is the bottom rung of the current Gemini ladder, and Google is unusually specific about the shape of work it wants there: agentic search and document processing, high-volume production traffic, and the sub-agent role inside a larger system where one bigger model plans and many small ones execute. The post's own demonstration has 3.6 Flash acting as the master agent and 3.5 Flash-Lite generating twenty-five design concepts underneath it.

Against the previous Lite generation the lab reports a large step rather than a refinement: Terminal-Bench 2.1 at 54% against 31%, the eight-needle long context set at 72.2% against 60.1%, and real-world task execution on GDPval-AA v2 at 1140 Elo against 642. The comparison that matters more for a buyer is the one against a bigger, older model: Google reports 3.5 Flash-Lite ahead of 3 Flash on SWE-Bench Pro at 54.2% against 49.6% and on OSWorld-Verified at 74.0% against 65.1%, which makes it a faster and cheaper replacement rather than a downgrade.

The thinking control is the lever that makes it two models in one. Google's guidance is to pin it to the minimal or low levels for cheap, latency-bound, high-volume execution, and to raise the level when the same model is handed a multi-step sub-agent workload. Computer use is a built-in tool here as well, which is what lets it take agentic work at all rather than only classification and extraction.

What it is not is the model for the hardest request in a system. Google positions the Flash tier above it for demanding coding and knowledge work, and the Lite tier's whole argument is throughput and price per unit of work rather than the top of any table. Read the speed claim the same way: it is a decode rate measured on a sample of prompts, not a promise about any one long generation.

لانچ کے وقت Google کا بیان

Built for sub-agents
Google aims it at agentic search, document processing and high-volume production traffic, and demonstrates it running underneath 3.6 Flash as the executor in a master-agent setup.
A real step over the last Lite
The lab reports Terminal-Bench 2.1 at 54% against 31%, the eight-needle long context set at 72.2% against 60.1%, and GDPval-AA v2 at 1140 Elo against 642 for the previous Flash-Lite.
Ahead of an older, larger model
Google reports it beating 3 Flash on SWE-Bench Pro at 54.2% against 49.6% and on OSWorld-Verified at 74.0% against 65.1%, making it a cheaper replacement rather than a step down.
Thinking as a cost dial
Pin it to minimal or low thinking for cheap, latency-bound volume, or raise the level when the same model is given a multi-step sub-agent workload. Computer use is a built-in tool.
What it is not for
Google places the Flash tier above it for demanding coding and knowledge work. This is the throughput model, and its case is price per unit of work rather than the top of any table.

ہندوستانی زبانیں

جیمنی 3.5 فلیش لائٹ 10 ہندوستانی زبانوں میں جواب دیتا ہے۔ میسج باکس کے ساتھ والے مینو سے زبان کا انتخاب کریں۔

Frequently Asked Questions

جیمنی 3.5 فلیش لائٹ کے بارے میں اکثر پوچھے جانے والے سوالات۔

جیمنی 3.5 فلیش لائٹ کب جاری ہوا تھا؟

Google نے جیمنی 3.5 فلیش لائٹ کو 21 جولائی، 2026 کو جاری کیا۔

جیمنی 3.5 فلیش لائٹ کس نے بنایا ہے؟

جیمنی 3.5 فلیش لائٹ کو Google نے تیار کیا ہے۔ 99Models AI فراہم کنندہ کے اصل نرخ پر اس سے براہ راست جوڑتا ہے۔

جیمنی 3.5 فلیش لائٹ کتنا ذہین ہے؟

یہ ذہانت کی درجہ بندی میں 54 چیٹ ماڈلز میں سے 40 نمبر پر ہے۔ اس کے تمام بینچ مارک اسکور اوپر والے ٹیبل میں دیکھے جا سکتے ہیں۔

جیمنی 3.5 فلیش لائٹ کا کتنا خرچ آتا ہے؟

استعمال کی لاگت ₹31.68 فی 10 لاکھ ان پٹ Tokens اور ₹264.00 فی 10 لاکھ آؤٹ پٹ Tokens ہے، بغیر کسی اضافی فیس کے۔ کوئی سبسکرپشن نہیں؛ صرف استعمال کی ادائیگی کریں۔

امریکی ڈالر میں جیمنی 3.5 فلیش لائٹ کی قیمت کیا ہے؟

فراہم کنندہ $0.33 فی 10 لاکھ ان پٹ Tokens اور $2.75 فی 10 لاکھ آؤٹ پٹ Tokens لیتا ہے۔ روپے کے نرخ ₹96 فی امریکی ڈالر کے حساب سے تبدیل کیے گئے ہیں۔

جیمنی 3.5 فلیش لائٹ کتنی لمبی گفتگو یاد رکھ سکتا ہے؟

اس کی Context حد 10.5 lakh Tokens ہے، یعنی وہ تمام متن اور فائلیں جو یہ ایک ہی میسج میں پڑھ سکتا ہے۔

کیا جیمنی 3.5 فلیش لائٹ مناسب قیمت میں بہترین کارکردگی دیتا ہے؟

یہ بہترین قیمت کی درجہ بندی میں 54 ماڈلز میں سے 36 نمبر پر ہے، جس میں ذہانت اور ٹوکن کی لاگت کا موازنہ کیا گیا ہے۔

کیا جیمنی 3.5 فلیش لائٹ جواب دینے سے پہلے سوچتا ہے؟

پہلے سے آن۔ جہاں Reasoning کی سہولت موجود ہو، وہاں آپ میسج باکس میں سوچنے کی سطح خود طے کر سکتے ہیں۔

جیمنی 3.5 فلیش لائٹ کن ہندوستانی زبانوں میں جواب دیتا ہے؟

یہ 10 ہندوستانی زبانوں میں جواب دیتا ہے۔ میسج باکس کے ساتھ والے مینو سے اپنی زبان منتخب کریں۔

Google کے مزید ماڈلز

کیٹلاگ اپ ڈیٹ: 9 ستمبر، 2026