99MODELS

ஜெமினி 3.5 ஃபிளாஷ் லைட்

Cheapest current Gemini; all input modalities, minimal reasoning by default.

வெளியான தேதி: 21 ஜூலை, 2026

264.00

10 லட்சம் output Tokens-க்கு

Input: 10 லட்சம் Tokens-க்கு ₹31.68

0% கூடுதல் கட்டணத்துடன், ஒரு அமெரிக்க டாலருக்கு ₹96 என்ற மாற்று விகிதத்தில் நிறுவனத்தின் நேரடி விலை வசூலிக்கப்படுகிறது.

தொழில்நுட்ப விவரங்கள்

Context window
10,48,576 Tokens
அதிகபட்ச Output
65,536 Tokens
ஏற்றுக்கொள்பவை
உரை, படங்கள், PDF கோப்புகள், ஆடியோ, வீடியோ
Reasoning
இயல்புநிலையில் ஆன்
Effort நிலைகள்
minimal, low, medium, high
Tool பயன்பாடு
ஆம்
Structured output
ஆம்
Code execution
ஆம்
நுண்ணறிவுத் தரம்
54-இல் #40
மதிப்புத் தரம்
54-இல் #36

கட்டண விவரம்

கட்டண விவரம்
10 lakh Tokens-க்குINRUSD
Input31.68$0.33
Output264.00$2.75
Cached input3.17$0.03

பெஞ்ச்மார்க் மதிப்பெண்கள்

மதிப்பீடாகக் குறிக்கப்பட்டவை தவிர மற்ற மதிப்பெண்கள் சதவீதங்கள் ஆகும். அனைத்து பெஞ்ச்மார்க்குகளும் தனிச்சார்பின்றி அளவிடப்பட்டவை.

  • 83.8%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 18.8%

    HLE

    Humanity's Last Exam

  • 41.3%

    SciCode

    SciCode - scientific code generation

  • 76.0%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 53.6%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 10.3%

    ARC-AGI-2

    ARC-AGI-2 - abstract reasoning on novel puzzles

  • 1449

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

ஜெமினி 3.5 ஃபிளாஷ் லைட் பற்றி

A low-latency, cost-effective multimodal Gemini optimised for high-throughput, low-cost execution -- Google aims it squarely at subagents running one focused task inside a larger workflow, and at document parsing. It is the fastest model in the 3.5 series, at roughly 350 output tokens per second, and can be pinned to minimal or low thinking for cheap work or raised for harder tasks. Full multimodal input and a million-token context are retained.

Flash-Lite is the bottom rung of the current Gemini ladder, and Google is unusually specific about the shape of work it wants there: agentic search and document processing, high-volume production traffic, and the sub-agent role inside a larger system where one bigger model plans and many small ones execute. The post's own demonstration has 3.6 Flash acting as the master agent and 3.5 Flash-Lite generating twenty-five design concepts underneath it.

Against the previous Lite generation the lab reports a large step rather than a refinement: Terminal-Bench 2.1 at 54% against 31%, the eight-needle long context set at 72.2% against 60.1%, and real-world task execution on GDPval-AA v2 at 1140 Elo against 642. The comparison that matters more for a buyer is the one against a bigger, older model: Google reports 3.5 Flash-Lite ahead of 3 Flash on SWE-Bench Pro at 54.2% against 49.6% and on OSWorld-Verified at 74.0% against 65.1%, which makes it a faster and cheaper replacement rather than a downgrade.

The thinking control is the lever that makes it two models in one. Google's guidance is to pin it to the minimal or low levels for cheap, latency-bound, high-volume execution, and to raise the level when the same model is handed a multi-step sub-agent workload. Computer use is a built-in tool here as well, which is what lets it take agentic work at all rather than only classification and extraction.

What it is not is the model for the hardest request in a system. Google positions the Flash tier above it for demanding coding and knowledge work, and the Lite tier's whole argument is throughput and price per unit of work rather than the top of any table. Read the speed claim the same way: it is a decode rate measured on a sample of prompts, not a promise about any one long generation.

அறிமுகத்தின் போது Google கூறியவை

Built for sub-agents
Google aims it at agentic search, document processing and high-volume production traffic, and demonstrates it running underneath 3.6 Flash as the executor in a master-agent setup.
A real step over the last Lite
The lab reports Terminal-Bench 2.1 at 54% against 31%, the eight-needle long context set at 72.2% against 60.1%, and GDPval-AA v2 at 1140 Elo against 642 for the previous Flash-Lite.
Ahead of an older, larger model
Google reports it beating 3 Flash on SWE-Bench Pro at 54.2% against 49.6% and on OSWorld-Verified at 74.0% against 65.1%, making it a cheaper replacement rather than a step down.
Thinking as a cost dial
Pin it to minimal or low thinking for cheap, latency-bound volume, or raise the level when the same model is given a multi-step sub-agent workload. Computer use is a built-in tool.
What it is not for
Google places the Flash tier above it for demanding coding and knowledge work. This is the throughput model, and its case is price per unit of work rather than the top of any table.

இந்திய மொழிகள்

ஜெமினி 3.5 ஃபிளாஷ் லைட் 10 இந்திய மொழிகளில் பதிலளிக்கும். செய்திப் பெட்டிக்கு அருகிலுள்ள மெனுவில் மொழியைத் தேர்வுசெய்து பதிலைப் பெறுங்கள்.

Frequently Asked Questions

ஜெமினி 3.5 ஃபிளாஷ் லைட் பற்றிய பொதுவான கேள்விகள்.

ஜெமினி 3.5 ஃபிளாஷ் லைட் எப்போது வெளியானது?

Google நிறுவனம் ஜெமினி 3.5 ஃபிளாஷ் லைட் மாடலை 21 ஜூலை, 2026 அன்று வெளியிட்டது.

ஜெமினி 3.5 ஃபிளாஷ் லைட் மாடலை உருவாக்கியது யார்?

ஜெமினி 3.5 ஃபிளாஷ் லைட் மாடலை Google நிறுவனம் உருவாக்கியது. 99Models AI இதனை நிறுவனத்தின் நேரடி கட்டணத்திலேயே வழங்குகிறது.

ஜெமினி 3.5 ஃபிளாஷ் லைட் எவ்வளவு அறிவார்ந்தது?

சுயாதீன Benchmark மதிப்பெண்களின்படி வரிசைப்படுத்தப்பட்ட நுண்ணறிவு தரவரிசையில், இது 54 chat Models-களில் 40-வது இடத்தில் உள்ளது. இதன் முழு மதிப்பெண்களை மேலே உள்ள Benchmarks பகுதியில் காணலாம்.

ஜெமினி 3.5 ஃபிளாஷ் லைட் பயன்பாட்டுக்கான கட்டணம் எவ்வளவு?

10 லட்சம் input Tokens-க்கு ₹31.68 மற்றும் 10 லட்சம் output Tokens-க்கு ₹264.00, கூடுதல் கட்டணமின்றி நிறுவனத்தின் நேரடி விலையிலேயே வசூலிக்கப்படுகிறது. சந்தா எதுவும் இல்லை; பயன்படுத்தியதற்கு மட்டுமே கட்டணம்.

டாலரில் ஜெமினி 3.5 ஃபிளாஷ் லைட் API கட்டணம் என்ன?

நிறுவனம் 10 லட்சம் input Tokens-க்கு $0.33 மற்றும் 10 லட்சம் output Tokens-க்கு $2.75 வசூலிக்கிறது. இப்பக்கத்தில் உள்ள ரூபாய் கட்டணம் ஒரு டாலருக்கு ₹96 என மாற்றப்பட்டதாகும்.

ஜெமினி 3.5 ஃபிளாஷ் லைட் எவ்வளவு நீண்ட உரையாடலைக் கையாளும்?

இதன் Context window 10.5 lakh Tokens. ஒரே கோரிக்கையில் இது படிக்கக்கூடிய முந்தைய உரையாடல் மற்றும் இணைக்கப்பட்ட கோப்புகளின் மொத்த அளவு இதுவாகும்.

மதிப்பு (Value) தரவரிசையில் ஜெமினி 3.5 ஃபிளாஷ் லைட் எந்த இடத்தில் உள்ளது?

திறன் மற்றும் Token கட்டணத்தை ஒப்பிடும் எங்கள் மதிப்பு (Value) தரவரிசையில், இது 54 Models-களில் 36-வது இடத்தில் உள்ளது.

ஜெமினி 3.5 ஃபிளாஷ் லைட் பதிலளிக்கும் முன் யோசிக்குமா (Reasoning)?

இயல்புநிலையில் ஆன். Reasoning வசதி உள்ள மாடல்களில், சிந்தனை அளவை (thinking effort) செய்தி உள்ளிடும் இடத்திலேயே நீங்கள் மாற்றிக்கொள்ளலாம்.

ஜெமினி 3.5 ஃபிளாஷ் லைட் எந்தெந்த இந்திய மொழிகளில் பதிலளிக்கும்?

இது 10 இந்திய மொழிகளில் பதிலளிக்கும். செய்திப் பெட்டிக்கு அருகிலுள்ள மெனுவில் உங்களுக்கு விருப்பமான மொழியைத் தேர்வுசெய்தால், பதில் அதே மொழியில் கிடைக்கும்.

Google நிறுவனத்தின் பிற மாடல்கள்

கேட்லாக் புதுப்பிக்கப்பட்டது 9 செப்., 2026