जेमिनी 2.5 फ्लॅश लाइट
Very cheap Gemini for classification and high-volume extraction.
लाँच दिनांक: 22 जुलै, 2025
₹38.40
प्रति 10 लाख आउटपुट Tokens
इनपुट: ₹9.60 प्रति 10 लाख Tokens
प्रोव्हायडरच्या मूळ दरानुसार 0% मार्कअपसह प्रति US डॉलर ₹96 दराने रूपांतरित।
वैशिष्ट्ये
- Context विंडो
- 10,48,576 Tokens
- कमाल आउटपुट
- 65,535 Tokens
- स्वीकारतो
- टेक्स्ट, इमेजेस, PDF फाइल्स, ऑडिओ, व्हिडिओ
- Reasoning
- बाय डीफॉल्ट चालू
- टूल वापर
- होय
- स्ट्रक्चर्ड आउटपुट
- होय
- Code एक्झिक्यूशन
- होय
- नॉलेज कटऑफ
- 2025-01-31
- इंटेलिजन्स रँक
- 54 पैकी #52
- व्हॅल्यू रँक
- 54 पैकी #49
दरपत्रक
बेंचमार्क
रेटिंग म्हणून नमूद केलेले नसल्यास स्कोअर टक्केवारीत आहेत. सर्व बेंचमार्क स्वतंत्रपणे तपासले जातात.
75.9%
MMLU-Pro
MMLU-Pro - multitask language understanding
62.5%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
6.8%
HLE
Humanity's Last Exam
59.3%
LiveCodeBench
LiveCodeBench - contamination-free coding
96.9%
MATH-500
MATH-500 - competition mathematics, 500 problems
70.3%
AIME 2024
AIME 2024 - competition mathematics
53.3%
AIME 2025
AIME 2025 - competition mathematics
49.9%
IFBench
IFBench - precise instruction following
55.7%
Long Context
Long Context Reasoning - reasoning over long inputs
4.5%
Terminal-Bench Hard
Terminal-Bench Hard - agentic terminal tasks
जेमिनी 2.5 फ्लॅश लाइट विषयी
Google's most cost-efficient multimodal model of its generation, offering the fastest performance for high-frequency, lightweight tasks. It is aimed at high-volume classification, simple data extraction and very low-latency applications where budget and speed are the primary constraints. Thinking is off by default, which is what makes it cheap.
Google described this model as pushing the frontier of intelligence per dollar, and everything about it follows from that sentence. It was the cheapest and fastest model in the 2.5 family at general availability, and the post pairs the price with a 40% cut to audio input pricing that had applied during the preview - a detail that matters if the workload is transcript or call data rather than text.
The speed claim is made against specific predecessors: Google reports lower latency than both the previous generation's Flash-Lite and its full Flash model across a broad sample of prompts, and higher all-round quality than the previous Flash-Lite on coding, mathematics, science, reasoning and multimodal understanding. The lab is careful to say this is a balance for latency-sensitive work such as translation and classification, not a claim about hard problems.
Being the cheap tier does not mean being a stripped one. It carries the same million-token context window as the rest of the family, controllable thinking budgets, and the native tools - grounding with Google Search, code execution and URL context - so a high-volume pipeline does not have to fall back to a larger model just to fetch a page or run a snippet.
The deployments Google chose to name are a good description of the shape of work it fits: summarising satellite telemetry in orbit, where the lab reports a 45% latency reduction and 30% lower power draw against the operator's previous model; planning and translating video content across more than 180 languages; and processing long product videos into documentation by extracting thousands of frames. All of them are the same job done a very large number of times.
What it is not for is anything demanding. Reasoning is off by default here and has to be turned on deliberately, and Google's own framing puts Flash above it for everyday tasks and Pro above that for coding and complex work. Choosing it is choosing throughput and cost over headroom.
लाँचवेळी Google ने काय सांगितले
- Intelligence per dollar
- Google released it as the fastest and lowest-cost model in the 2.5 family, and cut audio input pricing by 40% from the preview rate at the same time.
- Measured against its predecessors
- The lab reports lower latency than both the previous generation's Flash-Lite and its full Flash model on a broad sample of prompts, and higher quality than that Flash-Lite across the board.
- Cheap but not stripped
- It keeps the million-token context window, controllable thinking budgets, and the native tools: grounding with Google Search, code execution and URL context.
- Volume work, done many times
- The deployments Google names are in-orbit telemetry summarisation with a reported 45% latency reduction, video planning and translation across 180 languages, and video-to-documentation pipelines.
- What it is not for
- Reasoning is off unless the caller turns it on, and Google places Flash above it for everyday tasks and Pro above that for coding and complex work. It trades headroom for throughput.
भारतीय भाषा
जेमिनी 2.5 फ्लॅश लाइट हे 11 भारतीय भाषांमध्ये उत्तरे देते। मेसेज बॉक्ससमोरील मेनूमधून भाषा निवडा आणि त्याच भाषेत उत्तर मिळवा।
Frequently Asked Questions
जेमिनी 2.5 फ्लॅश लाइट बद्दल वारंवार विचारले जाणारे प्रश्न।
जेमिनी 2.5 फ्लॅश लाइट कधी लाँच झाले?
Google ने जेमिनी 2.5 फ्लॅश लाइट मॉडेल 22 जुलै, 2025 रोजी लाँच केले.
जेमिनी 2.5 फ्लॅश लाइट ची निर्मिती कोणी केली?
जेमिनी 2.5 फ्लॅश लाइट ची निर्मिती Google ने केली आहे। 99Models प्रोव्हायडरच्या मूळ दरात थेट तिच्याशी जोडते।
जेमिनी 2.5 फ्लॅश लाइट किती कार्यक्षम आहे?
स्वतंत्र बेंचमार्क गुणांवर आधारित आमच्या बुद्धिमत्ता रँकिंगमध्ये 54 पैकी या Model चा क्रमांक 52 आहे। तिचे सर्व गुण वरील Benchmarks तक्त्यामध्ये पाहू शकता।
जेमिनी 2.5 फ्लॅश लाइट चे दर किती आहेत?
0% मार्कअपसह दर प्रति 10 लाख इनपुट Tokens साठी ₹9.60 आणि प्रति 10 लाख आउटपुट Tokens साठी ₹38.40 आहे। कोणतेही सबस्क्रिप्शन नाही; तुम्ही वापरानुसार पेमेंट करता।
जेमिनी 2.5 फ्लॅश लाइट चे अमेरिकन डॉलरमधील दर काय आहेत?
प्रोव्हायडर प्रति 10 लाख इनपुट Tokens साठी $0.10 आणि प्रति 10 लाख आउटपुट Tokens साठी $0.40 आकारतो। रुपयांचे दर प्रति अमेरिकन डॉलर ₹96 या दराने रूपांतरित केले आहेत।
जेमिनी 2.5 फ्लॅश लाइट किती मोठे संभाषण लक्षात ठेवू शकते?
याची Context विंडो 10.5 lakh Tokens आहे। एकाच विनंतीमध्ये हे Model संभाषण आणि जोडलेल्या फाइल्स मिळून एवढा एकूण मजकूर वाचू शकते।
मूल्याच्या (Value) बाबतीत जेमिनी 2.5 फ्लॅश लाइट चा क्रमांक कितवा आहे?
मूल्य रँकिंगमध्ये 54 मॉडेलपैकी हिचा क्रमांक 49 आहे। हे रँकिंग मॉडेलची बुद्धिमत्ता आणि Tokens च्या किमतीची तुलना करून ठरवले जाते।
जेमिनी 2.5 फ्लॅश लाइट उत्तर देण्यापूर्वी विचार (Reasoning) करते का?
बाय डीफॉल्ट चालू। Reasoning उपलब्ध असल्यास, तुम्ही मेसेज कंपोजरमध्ये विचार करण्याची पातळी (effort level) निवडू शकता।
जेमिनी 2.5 फ्लॅश लाइट कोणत्या भारतीय भाषांमध्ये उत्तरे देते?
हे 11 भारतीय भाषांमध्ये उत्तरे देते। मेसेज बॉक्ससमोरील मेनूमधून भाषा निवडा आणि त्याच भाषेत उत्तर मिळवा।
Google कडील इतर मॉडेल्स
कॅटलॉग अपडेट: 9 सप्टें, 2026