99MODELS

మినిస్ట్రాల్ 8B

Tiny edge-class Mistral for classification and routing.

విడుదలైన తేదీ: 2, డిసెం 2025

15.84

10 లక్షల అవుట్‌పుట్ టోకెన్లకు

ఇన్‌పుట్: 10 లక్షల టోకెన్లకు ₹15.84

ప్రొవైడర్ అధికారిక ధరల ప్రకారం 1 డాలర్‌కు ₹96 చొప్పున 0% మార్కప్‌తో బిల్ చేయబడుతుంది.

స్పెసిఫికేషన్లు

Context విండో
2,62,144 Tokens
గరిష్ట ఔట్‌పుట్
2,09,715 Tokens
సపోర్ట్ చేస్తుంది
టెక్స్ట్, ఇమేజ్‌లు
Reasoning
లేదు
టూల్ వినియోగం
అవును
స్ట్రక్చర్డ్ ఔట్‌పుట్
అవును
కోడ్ ఎగ్జిక్యూషన్
లేదు
పారామీటర్లు
8B
ఇంటెలిజెన్స్ ర్యాంక్
54 లో #54
విలువ ర్యాంక్
54 లో #50

ధరల వివరాలు

ధరల వివరాలు
ప్రతి 10 lakh Tokens కుINRUSD
ఇన్‌పుట్15.84$0.17
ఔట్‌పుట్15.84$0.17
Cached ఇన్‌పుట్1.58$0.02

బెంచ్‌మార్క్‌లు

రేటింగ్‌గా పేర్కొన్నవి తప్ప మిగతా స్కోర్‌లన్నీ శాతాలు. ఇవన్నీ స్వతంత్రంగా లెక్కించినవి.

  • 64.2%

    MMLU-Pro

    MMLU-Pro - multitask language understanding

  • 47.1%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 4.3%

    HLE

    Humanity's Last Exam

  • 30.3%

    LiveCodeBench

    LiveCodeBench - contamination-free coding

  • 31.7%

    AIME 2025

    AIME 2025 - competition mathematics

  • 29.1%

    IFBench

    IFBench - precise instruction following

  • 25.7%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 4.5%

    Terminal-Bench Hard

    Terminal-Bench Hard - agentic terminal tasks

  • 4.1%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

మినిస్ట్రాల్ 8B గురించి

A powerful and efficient model in Mistral's Ministral 3 family, offering best-in-class text and vision capability at its size. It is built for edge deployment and performs across diverse hardware, fitting in 12GB of VRAM at FP8 and less when further quantised, while carrying a 256K context and image understanding across dozens of languages. It does not reason -- the reasoning capability ships as a separate checkpoint rather than a runtime toggle.

The 8B sits in the middle of a three-size edge family - 3B, 8B and 14B - released as the small end of the Mistral 3 generation. Mistral gave every size three separate checkpoints rather than one switchable model: a base, an instruct and a reasoning variant, each with image understanding, all under Apache 2.0. That is the structural fact worth knowing before you pick this one, because it means the reasoning capability is a different download rather than a runtime flag.

Mistral's claim for the family is a ratio rather than a peak: it says the Ministral models offer the best cost-to-performance of any open model, and it argues the case in a way most small-model launches do not. Its point is that in real deployments the tokens generated matter as much as the parameter count, and it reports the instruct variants matching or beating comparable models while often emitting an order of magnitude fewer tokens to get there. The chart it published makes that concrete for this size: the 8B instruct model lands around 51 on GPQA Diamond at roughly 1,500 output tokens, in the same accuracy band as a comparable 8B model that spends more than ten times as many.

The lab is explicit about the division of labour inside the family. For settings where accuracy is the only concern, it points at the reasoning variants, which think longer to reach what it calls state-of-the-art accuracy for their weight class - it cites 85% on AIME 2025 for the 14B. This instruct checkpoint is the opposite trade: answer fast, answer short, stay cheap.

The deployment story is where the size pays. Mistral says the whole Mistral 3 family was trained on NVIDIA Hopper hardware and that it worked with NVIDIA on optimised deployments for desktop AI machines, RTX PCs and laptops, and embedded Jetson boards - the point being that the 8B is meant to run on the device rather than in a datacentre. It carries native multilingual coverage across more than forty languages and image understanding at the same size. What it is not is a frontier model: Mistral positions it for edge inference, classification, routing and on-device assistants, and points at its far larger models for anything that needs frontier reasoning.

లాంచ్ సమయంలో Mistral AI ప్రకటించిన వివరాలు

Three sizes, three variants each
Mistral released 3B, 8B and 14B models, and for every size a base, an instruct and a reasoning checkpoint, all with image understanding and all under Apache 2.0.
Accuracy per token, not per parameter
The lab argues generated tokens matter as much as model size in production and reports the instruct models matching comparable ones while often emitting an order of magnitude fewer tokens.
Reasoning is a separate checkpoint
For accuracy-first work Mistral points at the reasoning variants rather than a runtime toggle, citing 85% on AIME 2025 for the 14B member of the family.
Built to run on the device
Mistral worked with NVIDIA on optimised deployments for desktop AI machines, RTX PCs and laptops, and Jetson boards, so the model runs locally rather than in a datacentre.
Multilingual and multimodal at 8B
Native coverage of more than forty languages plus image understanding, both at a size that fits on consumer hardware.
What it is not for
This is the edge tier: Mistral positions the Ministral line for local and cost-sensitive work and points at its far larger models for frontier reasoning.

Frequently Asked Questions

మినిస్ట్రాల్ 8B గురించి తరచుగా అడిగే ప్రశ్నలు.

మినిస్ట్రాల్ 8B ఎప్పుడు విడుదలైంది?

Mistral AI ఈ మినిస్ట్రాల్ 8B మోడల్‌ను 2, డిసెం 2025న విడుదల చేసింది.

మినిస్ట్రాల్ 8Bను ఎవరు రూపొందించారు?

మినిస్ట్రాల్ 8Bను రూపొందించింది Mistral AI. 99Models నేరుగా ప్రొవైడర్ ధరలకే దీనికి యాక్సెస్ అందిస్తుంది.

మినిస్ట్రాల్ 8B ఎంత సమర్థవంతమైనది?

ఇండిపెండెంట్ బెంచ్‌మార్క్ స్కోర్‌ల ఆధారంగా నిర్ణయించిన ఇంటెలిజెన్స్ ర్యాంకింగ్స్‌లో ఇది 54 మోడళ్లలో 54వ స్థానంలో ఉంది. పూర్తి స్కోర్లు పైన చూడవచ్చు.

మినిస్ట్రాల్ 8B ధర ఎంత?

10 లక్షల ఇన్‌పుట్ టోకెన్లకు ₹15.84, అవుట్‌పుట్ టోకెన్లకు ₹15.84 ఛార్జ్ అవుతుంది (0% మార్కప్). ఎలాంటి సబ్‌స్క్రిప్షన్ లేదు, వాడినంతకే చెల్లిస్తారు.

డాలర్లలో మినిస్ట్రాల్ 8B ధర ఎంత?

ప్రొవైడర్ 10 లక్షల ఇన్‌పుట్ టోకెన్లకు $0.17, 10 లక్షల అవుట్‌పుట్ టోకెన్లకు $0.17 ఛార్జ్ చేస్తారు. రూపాయి ధరలు 1 డాలర్‌కు ₹96 చొప్పున కన్వర్ట్ చేయబడతాయి.

మినిస్ట్రాల్ 8B ఎంత పెద్ద సంభాషణను గుర్తుంచుకోగలదు?

దీని కాంటెక్స్ట్ విండో 2.6 lakh టోకెన్లు. అంటే ఒకే సందేశంలో ఇది అంత పరిమాణంలోని సంభాషణను, ఫైళ్లను ప్రాసెస్ చేయగలదు.

ధరకు తగిన విలువను మినిస్ట్రాల్ 8B ఇస్తుందా?

ఖర్చుకు తగిన సామర్థ్యాన్ని లెక్కించే వాల్యూ ర్యాంకింగ్స్‌లో ఇది 54 మోడళ్లలో 50వ స్థానంలో ఉంది.

Mistral AI నుండి మరిన్ని మోడల్స్

కేటలాగ్ అప్‌డేట్ చేసిన తేదీ: 9, సెప్టెం 2026