99MODELS

ஜெமினி 2.5 ஃபிளாஷ் லைட்

Very cheap Gemini for classification and high-volume extraction.

வெளியான தேதி: 22 ஜூலை, 2025

38.40

10 லட்சம் output Tokens-க்கு

Input: 10 லட்சம் Tokens-க்கு ₹9.60

0% கூடுதல் கட்டணத்துடன், ஒரு அமெரிக்க டாலருக்கு ₹96 என்ற மாற்று விகிதத்தில் நிறுவனத்தின் நேரடி விலை வசூலிக்கப்படுகிறது.

தொழில்நுட்ப விவரங்கள்

Context window
10,48,576 Tokens
அதிகபட்ச Output
65,535 Tokens
ஏற்றுக்கொள்பவை
உரை, படங்கள், PDF கோப்புகள், ஆடியோ, வீடியோ
Reasoning
இயல்புநிலையில் ஆன்
Tool பயன்பாடு
ஆம்
Structured output
ஆம்
Code execution
ஆம்
அறிவு வரம்பு
2025-01-31
நுண்ணறிவுத் தரம்
54-இல் #52
மதிப்புத் தரம்
54-இல் #49

கட்டண விவரம்

கட்டண விவரம்
10 lakh Tokens-க்குINRUSD
Input9.60$0.10
Output38.40$0.40
Cached input0.96$0.01

பெஞ்ச்மார்க் மதிப்பெண்கள்

மதிப்பீடாகக் குறிக்கப்பட்டவை தவிர மற்ற மதிப்பெண்கள் சதவீதங்கள் ஆகும். அனைத்து பெஞ்ச்மார்க்குகளும் தனிச்சார்பின்றி அளவிடப்பட்டவை.

  • 75.9%

    MMLU-Pro

    MMLU-Pro - multitask language understanding

  • 62.5%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 6.8%

    HLE

    Humanity's Last Exam

  • 59.3%

    LiveCodeBench

    LiveCodeBench - contamination-free coding

  • 96.9%

    MATH-500

    MATH-500 - competition mathematics, 500 problems

  • 70.3%

    AIME 2024

    AIME 2024 - competition mathematics

  • 53.3%

    AIME 2025

    AIME 2025 - competition mathematics

  • 49.9%

    IFBench

    IFBench - precise instruction following

  • 55.7%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 4.5%

    Terminal-Bench Hard

    Terminal-Bench Hard - agentic terminal tasks

ஜெமினி 2.5 ஃபிளாஷ் லைட் பற்றி

Google's most cost-efficient multimodal model of its generation, offering the fastest performance for high-frequency, lightweight tasks. It is aimed at high-volume classification, simple data extraction and very low-latency applications where budget and speed are the primary constraints. Thinking is off by default, which is what makes it cheap.

Google described this model as pushing the frontier of intelligence per dollar, and everything about it follows from that sentence. It was the cheapest and fastest model in the 2.5 family at general availability, and the post pairs the price with a 40% cut to audio input pricing that had applied during the preview - a detail that matters if the workload is transcript or call data rather than text.

The speed claim is made against specific predecessors: Google reports lower latency than both the previous generation's Flash-Lite and its full Flash model across a broad sample of prompts, and higher all-round quality than the previous Flash-Lite on coding, mathematics, science, reasoning and multimodal understanding. The lab is careful to say this is a balance for latency-sensitive work such as translation and classification, not a claim about hard problems.

Being the cheap tier does not mean being a stripped one. It carries the same million-token context window as the rest of the family, controllable thinking budgets, and the native tools - grounding with Google Search, code execution and URL context - so a high-volume pipeline does not have to fall back to a larger model just to fetch a page or run a snippet.

The deployments Google chose to name are a good description of the shape of work it fits: summarising satellite telemetry in orbit, where the lab reports a 45% latency reduction and 30% lower power draw against the operator's previous model; planning and translating video content across more than 180 languages; and processing long product videos into documentation by extracting thousands of frames. All of them are the same job done a very large number of times.

What it is not for is anything demanding. Reasoning is off by default here and has to be turned on deliberately, and Google's own framing puts Flash above it for everyday tasks and Pro above that for coding and complex work. Choosing it is choosing throughput and cost over headroom.

அறிமுகத்தின் போது Google கூறியவை

Intelligence per dollar
Google released it as the fastest and lowest-cost model in the 2.5 family, and cut audio input pricing by 40% from the preview rate at the same time.
Measured against its predecessors
The lab reports lower latency than both the previous generation's Flash-Lite and its full Flash model on a broad sample of prompts, and higher quality than that Flash-Lite across the board.
Cheap but not stripped
It keeps the million-token context window, controllable thinking budgets, and the native tools: grounding with Google Search, code execution and URL context.
Volume work, done many times
The deployments Google names are in-orbit telemetry summarisation with a reported 45% latency reduction, video planning and translation across 180 languages, and video-to-documentation pipelines.
What it is not for
Reasoning is off unless the caller turns it on, and Google places Flash above it for everyday tasks and Pro above that for coding and complex work. It trades headroom for throughput.

இந்திய மொழிகள்

ஜெமினி 2.5 ஃபிளாஷ் லைட் 11 இந்திய மொழிகளில் பதிலளிக்கும். செய்திப் பெட்டிக்கு அருகிலுள்ள மெனுவில் மொழியைத் தேர்வுசெய்து பதிலைப் பெறுங்கள்.

Frequently Asked Questions

ஜெமினி 2.5 ஃபிளாஷ் லைட் பற்றிய பொதுவான கேள்விகள்.

ஜெமினி 2.5 ஃபிளாஷ் லைட் எப்போது வெளியானது?

Google நிறுவனம் ஜெமினி 2.5 ஃபிளாஷ் லைட் மாடலை 22 ஜூலை, 2025 அன்று வெளியிட்டது.

ஜெமினி 2.5 ஃபிளாஷ் லைட் மாடலை உருவாக்கியது யார்?

ஜெமினி 2.5 ஃபிளாஷ் லைட் மாடலை Google நிறுவனம் உருவாக்கியது. 99Models AI இதனை நிறுவனத்தின் நேரடி கட்டணத்திலேயே வழங்குகிறது.

ஜெமினி 2.5 ஃபிளாஷ் லைட் எவ்வளவு அறிவார்ந்தது?

சுயாதீன Benchmark மதிப்பெண்களின்படி வரிசைப்படுத்தப்பட்ட நுண்ணறிவு தரவரிசையில், இது 54 chat Models-களில் 52-வது இடத்தில் உள்ளது. இதன் முழு மதிப்பெண்களை மேலே உள்ள Benchmarks பகுதியில் காணலாம்.

ஜெமினி 2.5 ஃபிளாஷ் லைட் பயன்பாட்டுக்கான கட்டணம் எவ்வளவு?

10 லட்சம் input Tokens-க்கு ₹9.60 மற்றும் 10 லட்சம் output Tokens-க்கு ₹38.40, கூடுதல் கட்டணமின்றி நிறுவனத்தின் நேரடி விலையிலேயே வசூலிக்கப்படுகிறது. சந்தா எதுவும் இல்லை; பயன்படுத்தியதற்கு மட்டுமே கட்டணம்.

டாலரில் ஜெமினி 2.5 ஃபிளாஷ் லைட் API கட்டணம் என்ன?

நிறுவனம் 10 லட்சம் input Tokens-க்கு $0.10 மற்றும் 10 லட்சம் output Tokens-க்கு $0.40 வசூலிக்கிறது. இப்பக்கத்தில் உள்ள ரூபாய் கட்டணம் ஒரு டாலருக்கு ₹96 என மாற்றப்பட்டதாகும்.

ஜெமினி 2.5 ஃபிளாஷ் லைட் எவ்வளவு நீண்ட உரையாடலைக் கையாளும்?

இதன் Context window 10.5 lakh Tokens. ஒரே கோரிக்கையில் இது படிக்கக்கூடிய முந்தைய உரையாடல் மற்றும் இணைக்கப்பட்ட கோப்புகளின் மொத்த அளவு இதுவாகும்.

மதிப்பு (Value) தரவரிசையில் ஜெமினி 2.5 ஃபிளாஷ் லைட் எந்த இடத்தில் உள்ளது?

திறன் மற்றும் Token கட்டணத்தை ஒப்பிடும் எங்கள் மதிப்பு (Value) தரவரிசையில், இது 54 Models-களில் 49-வது இடத்தில் உள்ளது.

ஜெமினி 2.5 ஃபிளாஷ் லைட் பதிலளிக்கும் முன் யோசிக்குமா (Reasoning)?

இயல்புநிலையில் ஆன். Reasoning வசதி உள்ள மாடல்களில், சிந்தனை அளவை (thinking effort) செய்தி உள்ளிடும் இடத்திலேயே நீங்கள் மாற்றிக்கொள்ளலாம்.

ஜெமினி 2.5 ஃபிளாஷ் லைட் எந்தெந்த இந்திய மொழிகளில் பதிலளிக்கும்?

இது 11 இந்திய மொழிகளில் பதிலளிக்கும். செய்திப் பெட்டிக்கு அருகிலுள்ள மெனுவில் உங்களுக்கு விருப்பமான மொழியைத் தேர்வுசெய்தால், பதில் அதே மொழியில் கிடைக்கும்.

Google நிறுவனத்தின் பிற மாடல்கள்

கேட்லாக் புதுப்பிக்கப்பட்டது 9 செப்., 2026