मीमो v2.5
Cheap 1M-context multimodal model; one of the highest-volume models anywhere.
लाँच दिनांक: 22 एप्रि, 2026
₹192.00
प्रति 10 लाख आउटपुट Tokens
इनपुट: ₹38.40 प्रति 10 लाख Tokens
प्रोव्हायडरच्या मूळ दरानुसार 0% मार्कअपसह प्रति US डॉलर ₹96 दराने रूपांतरित।
वैशिष्ट्ये
- Context विंडो
- 10,50,000 Tokens
- कमाल आउटपुट
- 1,31,072 Tokens
- स्वीकारतो
- टेक्स्ट, इमेजेस, व्हिडिओ
- Reasoning
- बाय डीफॉल्ट चालू
- टूल वापर
- होय
- स्ट्रक्चर्ड आउटपुट
- होय
- Code एक्झिक्यूशन
- नाही
- इंटेलिजन्स रँक
- 54 पैकी #46
- व्हॅल्यू रँक
- 54 पैकी #44
दरपत्रक
बेंचमार्क
रेटिंग म्हणून नमूद केलेले नसल्यास स्कोअर टक्केवारीत आहेत. सर्व बेंचमार्क स्वतंत्रपणे तपासले जातात.
76.3%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
25.2%
HLE
Humanity's Last Exam
43.1%
SciCode
SciCode - scientific code generation
41.7%
Terminal-Bench Hard
Terminal-Bench Hard - agentic terminal tasks
1438
WebDev Arena
WebDev Arena - head-to-head web-app builds, Elo rating
मीमो v2.5 विषयी
Xiaomi's natively omnimodal model, handling text, image, video and audio in one unified architecture. It is a sparse Mixture-of-Experts with 310B total and 15B active parameters, pairing a vision encoder and an audio transformer with the language backbone, and it inherits a hybrid sliding-window attention design that cuts KV-cache storage nearly sixfold. It supports up to a million tokens of context at a very low price.
Xiaomi built this one to perceive and to act in the same pass. The language backbone is inherited from the team's earlier hybrid sliding-window model, and the vision and audio encoders - both pretrained in-house - hang off it through lightweight projectors rather than being bolted on as a separate pipeline. The lab's summary of the goal is a single model that sees, hears and acts on what it perceives, and the training schedule is built around that: text pre-training for the backbone, a projector warmup to align audio and vision with the language model, large-scale multimodal pre-training, supervised fine-tuning and agentic post-training during which the window is stretched from 32K to 256K to a million tokens, and finally reinforcement learning with multi-teacher on-policy distillation.
The agent claims are where Xiaomi puts its emphasis, and it argues them on efficiency as much as accuracy. On its internal coding evaluation it reports the model matching its own far larger Pro sibling at half the cost. On its daily-agent benchmark it reports 62.3 on the general subset and places the result at the frontier of score against token spend, which is the whole pitch: frontier-level agent behaviour without frontier-level token bills. On the public coding sets it reports 65.8 on Terminal-Bench 2.0 and 56.1 on SWE-Bench Pro.
The perception numbers are the other half. Xiaomi reports 81.0 on the chart-reasoning benchmark CharXiv, 77.9 on MMMU-Pro, 88.5 on high-resolution image understanding, 87.2 on document understanding, 87.7 on general video question answering and 64.0 on the harder video-reasoning set VideoHolmes. Its framing of those is comparative and specific: level with a leading closed model on video, level with another on multimodal agentic work, and competitive rather than leading on image and document understanding.
Where it stops shows up in the same table. On the multimodal agent split - real, messy interactions rather than single questions - Xiaomi reports 23.8, behind two of the closed models it compared against, and the number is low in absolute terms for every model on that row. The lab also positions the larger Pro model above this one for the hardest long-horizon software engineering. Read this as the omnimodal workhorse: cheap, very long-context, strong at everyday agent work and at reading what it is shown, with the hardest autonomous runs left to its bigger sibling.
लाँचवेळी Xiaomi ने काय सांगितले
- One model across four modalities
- Vision and audio encoders pretrained in-house attach to the language backbone through lightweight projectors, so text, image, video and audio are reasoned about in a single architecture.
- Context grown during post-training
- Xiaomi extended the window from 32K to 256K to a million tokens across supervised fine-tuning and agentic post-training, rather than bolting long context on afterwards.
- Agent quality per token
- The lab reports 62.3 on the general subset of its daily-agent benchmark and places the model at the frontier of score against token spend, matching its Pro sibling's internal coding score at half the cost.
- Perception benchmarks
- Xiaomi reports 81.0 on CharXiv chart reasoning, 77.9 on MMMU-Pro, 88.5 on high-resolution images, 87.2 on documents and 87.7 on general video question answering.
- Coding agent results
- On the public sets the lab reports 65.8 on Terminal-Bench 2.0 and 56.1 on SWE-Bench Pro, alongside its internal coding evaluation.
- What it is not for
- On Xiaomi's own multimodal agent split it scores 23.8, behind two of the closed models it compared against, and the lab points at the larger Pro model for the hardest long-horizon engineering.
भारतीय भाषा
मीमो v2.5 हे 5 भारतीय भाषांमध्ये उत्तरे देते। मेसेज बॉक्ससमोरील मेनूमधून भाषा निवडा आणि त्याच भाषेत उत्तर मिळवा।
Frequently Asked Questions
मीमो v2.5 बद्दल वारंवार विचारले जाणारे प्रश्न।
मीमो v2.5 कधी लाँच झाले?
Xiaomi ने मीमो v2.5 मॉडेल 22 एप्रि, 2026 रोजी लाँच केले.
मीमो v2.5 ची निर्मिती कोणी केली?
मीमो v2.5 ची निर्मिती Xiaomi ने केली आहे। 99Models प्रोव्हायडरच्या मूळ दरात थेट तिच्याशी जोडते।
मीमो v2.5 किती कार्यक्षम आहे?
स्वतंत्र बेंचमार्क गुणांवर आधारित आमच्या बुद्धिमत्ता रँकिंगमध्ये 54 पैकी या Model चा क्रमांक 46 आहे। तिचे सर्व गुण वरील Benchmarks तक्त्यामध्ये पाहू शकता।
मीमो v2.5 चे दर किती आहेत?
0% मार्कअपसह दर प्रति 10 लाख इनपुट Tokens साठी ₹38.40 आणि प्रति 10 लाख आउटपुट Tokens साठी ₹192.00 आहे। कोणतेही सबस्क्रिप्शन नाही; तुम्ही वापरानुसार पेमेंट करता।
मीमो v2.5 चे अमेरिकन डॉलरमधील दर काय आहेत?
प्रोव्हायडर प्रति 10 लाख इनपुट Tokens साठी $0.40 आणि प्रति 10 लाख आउटपुट Tokens साठी $2.00 आकारतो। रुपयांचे दर प्रति अमेरिकन डॉलर ₹96 या दराने रूपांतरित केले आहेत।
मीमो v2.5 किती मोठे संभाषण लक्षात ठेवू शकते?
याची Context विंडो 10.5 lakh Tokens आहे। एकाच विनंतीमध्ये हे Model संभाषण आणि जोडलेल्या फाइल्स मिळून एवढा एकूण मजकूर वाचू शकते।
मूल्याच्या (Value) बाबतीत मीमो v2.5 चा क्रमांक कितवा आहे?
मूल्य रँकिंगमध्ये 54 मॉडेलपैकी हिचा क्रमांक 44 आहे। हे रँकिंग मॉडेलची बुद्धिमत्ता आणि Tokens च्या किमतीची तुलना करून ठरवले जाते।
मीमो v2.5 उत्तर देण्यापूर्वी विचार (Reasoning) करते का?
बाय डीफॉल्ट चालू। Reasoning उपलब्ध असल्यास, तुम्ही मेसेज कंपोजरमध्ये विचार करण्याची पातळी (effort level) निवडू शकता।
मीमो v2.5 कोणत्या भारतीय भाषांमध्ये उत्तरे देते?
हे 5 भारतीय भाषांमध्ये उत्तरे देते। मेसेज बॉक्ससमोरील मेनूमधून भाषा निवडा आणि त्याच भाषेत उत्तर मिळवा।
Xiaomi कडील इतर मॉडेल्स
कॅटलॉग अपडेट: 9 सप्टें, 2026