મીમો v2.5 પ્રો
Larger MiMo with stronger reasoning at a still-low price.
રજૂઆત: 22 એપ્રિલ, 2026
₹288.00
10 લાખ આઉટપુટ Tokens દીઠ
ઇનપુટ: 10 લાખ Tokens દીઠ ₹96.00
પ્રોવાઇડરના ભાવે બિલિંગ, $1 = ₹96 મુજબ રૂપાંતરિત, 0% માર્કઅપ સાથે.
વિશિષ્ટતાઓ
- Context વિન્ડો
- 10,50,000 Tokens
- મહત્તમ આઉટપુટ
- 1,31,072 Tokens
- સ્વીકારે છે
- ટેક્સ્ટ
- Reasoning
- ડિફૉલ્ટ રૂપે ચાલુ
- ટૂલ ઉપયોગ
- હા
- સ્ટ્રક્ચર્ડ આઉટપુટ
- હા
- કોડ એક્ઝિક્યુશન
- ના
- ઇન્ટેલિજન્સ રેન્ક
- 54 માંથી #35
- Value રેન્ક
- 54 માંથી #33
કિંમત
Benchmarks
સ્કોર ટકાવારીમાં છે સિવાય કે રેટિંગ તરીકે દર્શાવેલ હોય. તમામ બેન્ચમાર્ક સ્વતંત્ર રીતે માપવામાં આવ્યા છે.
86.6%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
35.7%
HLE
Humanity's Last Exam
50.6%
SciCode
SciCode - scientific code generation
79.9%
IFBench
IFBench - precise instruction following
79.7%
Long Context
Long Context Reasoning - reasoning over long inputs
43.2%
Terminal-Bench Hard
Terminal-Bench Hard - agentic terminal tasks
65.2%
Terminal-Bench 2
Terminal-Bench 2.1 - agentic terminal tasks, second edition
1476
WebDev Arena
WebDev Arena - head-to-head web-app builds, Elo rating
મીમો v2.5 પ્રો વિશે
Xiaomi's flagship open Mixture-of-Experts model, at 1.02T total and 42B active parameters, built for the most demanding agentic, software-engineering and long-horizon work. Xiaomi describes it as a leap in agentic coherence, sustaining complex trajectories across thousands of tool calls with strong instruction following over a million-token context. Its showcase runs include writing a compiler from scratch across hundreds of tool calls.
Xiaomi does not argue this model with a benchmark table first. It argues it with three jobs it left running. The first is a compiler: a university course project asking for a complete SysY compiler in Rust from scratch - lexer, parser, syntax tree, intermediate representation, RISC-V backend and performance work - which the model finished in 4.3 hours across 672 tool calls, passing all 233 hidden tests. The lab's reading of the trace is the interesting part: the first compile already passed 137 of 233, which suggests the architecture was designed correctly before a single test ran, and when a refactor at turn 512 regressed two cases the model diagnosed and recovered rather than thrashing. The second job produced a desktop video editor with a multi-track timeline, trimming, cross-fades, audio mixing and export - 8,192 lines over 1,868 tool calls and 11.5 hours. The third wired the model into a circuit simulator to design an analog voltage regulator in a 180nm process, landing six specifications simultaneously in about an hour of closed-loop iteration.
The behaviour Xiaomi names for all three is harness awareness: the model makes full use of what its environment offers, manages its own memory, and shapes how its context gets populated toward the goal. That is a different claim from a higher score, and it is the one the lab leads with - along with reliable adherence to subtle requirements buried in context and coherence held across very long inputs.
The efficiency argument is quantified. On its daily-agent benchmark Xiaomi reports 63.8 at roughly 70,000 tokens per trajectory, which it puts at 40 to 60% fewer tokens than the frontier closed models it compares against at comparable capability. Elsewhere on its published table: 72.9 on the multi-domain agent suite, 57.2 on SWE-Bench Pro, 78.9 on SWE-bench Verified, 68.4 on Terminal-Bench 2.0, and 34.0 on Humanity's Last Exam without tools rising to 48.0 with them.
The architecture is built for that token budget. Local sliding-window and global attention interleave at six to one with a 128-token window, which Xiaomi says cuts key-value cache storage by nearly seven times at long context while a learnable attention-sink bias preserves quality; a lightweight multi-token prediction module roughly triples output throughput and speeds up reinforcement-learning rollouts. Pre-training ran on 27 trillion tokens in FP8 mixed precision at a native 32K length before the window was extended. Where it stops is stated by the lab's own charts rather than in prose: it labels its internal coding result as closing the gap to a leading closed model rather than passing it, and on the hardest-exam and implementation-ranking rows the closed frontier is still ahead.
Xiaomi એ લોન્ચ વખતે શું જણાવ્યું હતું
- A compiler in one autonomous run
- Xiaomi reports the model completing a full SysY compiler in Rust in 4.3 hours across 672 tool calls, passing all 233 hidden tests, with 137 already passing on the first compile.
- Harness awareness
- The lab describes the model making full use of its environment, managing its own memory and shaping how its context is populated toward the final objective across very long runs.
- Fewer tokens per trajectory
- On its daily-agent benchmark Xiaomi reports 63.8 at around 70,000 tokens per trajectory, which it puts at 40 to 60% fewer tokens than the closed frontier models at comparable capability.
- Attention built for long context
- Sliding-window and global attention interleave six to one with a 128-token window, cutting key-value cache storage by nearly seven times, with a learnable attention-sink bias preserving quality.
- Agent and coding scores
- The lab reports 72.9 on its multi-domain agent suite, 78.9 on SWE-bench Verified, 68.4 on Terminal-Bench 2.0, and 34.0 on Humanity's Last Exam without tools against 48.0 with them.
- What it is not for
- Xiaomi labels its own internal coding result as closing the gap to a leading closed model rather than passing it, and its table still shows the closed frontier ahead on the hardest exams.
ભારતીય ભાષાઓ
મીમો v2.5 પ્રો કુલ 9 ભારતીય ભાષાઓમાં જવાબ આપી શકે છે. મેસેજ બોક્સની બાજુમાં આપેલા મેનૂમાંથી ભાષા પસંદ કરો.
Frequently Asked Questions
મીમો v2.5 પ્રો વિશે વારંવાર પૂછાતા પ્રશ્નો.
મીમો v2.5 પ્રો ક્યારે રજૂ કરવામાં આવ્યું હતું?
Xiaomi એ 22 એપ્રિલ, 2026 ના રોજ મીમો v2.5 પ્રો રજૂ કર્યું હતું.
મીમો v2.5 પ્રો કોણે બનાવ્યું છે?
મીમો v2.5 પ્રો ને Xiaomi લેબ દ્વારા બનાવવામાં આવ્યું છે. 99Models સીધા પ્રોવાઇડરના નિર્ધારિત દરે કનેક્ટ કરે છે.
મીમો v2.5 પ્રો કેટલું સક્ષમ અને બુદ્ધિશાળી છે?
ઇન્ટેલિજન્સ રેન્કિંગમાં તે 54 ચેટ Models માંથી 35 નંબર પર છે, જે સ્વતંત્ર બેન્ચમાર્ક સ્કોર પર આધારિત છે. તેના તમામ સ્કોર ઉપર બેન્ચમાર્ક ટેબલમાં જોઈ શકો છો.
મીમો v2.5 પ્રો નો ઉપયોગ કરવાનો ખર્ચ કેટલો છે?
વપરાશ ખર્ચ ઇનપુટ માટે ₹96.00 પ્રતિ 10 લાખ Tokens અને આઉટપુટ માટે ₹288.00 પ્રતિ 10 લાખ Tokens છે, જેમાં 0% માર્કઅપ છે. કોઈ સબ્સ્ક્રિપ્શન નથી; વપરાશ મુજબ જ પેમેન્ટ કરો.
મીમો v2.5 પ્રો નું અમેરિકન ડોલરમાં API પ્રાઇસિંગ શું છે?
પ્રોવાઇડર ઇનપુટ માટે $1.00 પ્રતિ 10 લાખ Tokens અને આઉટપુટ માટે $3.00 પ્રતિ 10 લાખ Tokens ચાર્જ કરે છે. રૂપિયાના દરો $1 = ₹96 ના આધારે ગણવામાં આવ્યા છે.
મીમો v2.5 પ્રો કેટલી લાંબી વાતચીત યાદ રાખી શકે છે?
તેની Context વિન્ડો 10.5 lakh Tokens ની છે. એટલે કે તે એક જ વિનંતીમાં અગાઉની વાતચીત અને જોડેલી ફાઇલો સહિત આટલું લખાણ વાંચી શકે છે.
શું મીમો v2.5 પ્રો વાપરવું પૈસા વસૂલ છે?
વેલ્યૂ રેન્કિંગમાં તે 54 માંથી 33 નંબર પર છે. આ રેન્કિંગ બેન્ચમાર્ક ક્ષમતા અને વાસ્તવિક Token ખર્ચની તુલના કરે છે.
શું મીમો v2.5 પ્રો જવાબ આપતાં પહેલાં વિચારે છે?
ડિફૉલ્ટ રૂપે ચાલુ. જ્યાં Reasoning સપોર્ટ ઉપલબ્ધ છે, ત્યાં તમે મેસેજ કમ્પોઝરમાં વિચારવાની ક્ષમતા (Thinking effort) જાતે સેટ કરી શકો છો.
મીમો v2.5 પ્રો કઈ ભારતીય ભાષાઓમાં જવાબ આપી શકે છે?
તે 9 ભારતીય ભાષાઓમાં જવાબ આપે છે. મેસેજ બોક્સ પાસેના મેનૂમાંથી ભાષા પસંદ કરો અને તે જ ભાષામાં જવાબ મેળવો.
Xiaomi ના અન્ય Models
કેટલોગ અપડેટ: 9 સપ્ટે, 2026