99MODELS

മിസ്ട്രൽ സ്മോൾ 3

Cheap European model with vision input and optional reasoning.

റിലീസ് തീയതി: 2026 മാർ 16

72.00

10 ലക്ഷം ഔട്ട്പുട്ട് Tokens-ന്

ഇൻപുട്ട്: 10 ലക്ഷം Tokens-ന് ₹18.00

0% മാർക്ക്അപ്പിൽ, ഒരു ഡോളറിന് ₹96 എന്ന നിരക്കിലാണ് ഈടാക്കുന്നത്.

സവിശേഷതകൾ

Context window
2,62,144 Tokens
പരമാവധി Output
2,09,715 Tokens
സ്വീകരിക്കുന്നത്
ടെക്സ്റ്റ്, ചിത്രങ്ങൾ
Reasoning
ഓപ്ഷണൽ
Effort ലെവലുകൾ
none, high
ടൂൾ ഉപയോഗം
ഉണ്ട്
സ്ട്രക്ചേർഡ് Output
ഉണ്ട്
കോഡ് എക്സിക്യൂഷൻ
ഇല്ല
ഇന്റലിജൻസ് റാങ്ക്
54-ൽ #51
വാല്യു റാങ്ക്
54-ൽ #47

നിരക്കുകൾ

നിരക്കുകൾ
10 lakh Tokens-ന്INRUSD
Input18.00$0.19
Output72.00$0.75
Cached input1.58$0.02

ബെഞ്ച്മാർക്കുകൾ

റേറ്റിംഗ് എന്ന് രേഖപ്പെടുത്താത്ത സ്കോറുകൾ ശതമാനത്തിലാണ്. എല്ലാ ബെഞ്ച്മാർക്കുകളും സ്വതന്ത്രമായി പരിശോധിച്ചവയാണ്.

  • 76.9%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 9.9%

    HLE

    Humanity's Last Exam

  • 48.2%

    IFBench

    IFBench - precise instruction following

  • 49.7%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 17.4%

    Terminal-Bench Hard

    Terminal-Bench Hard - agentic terminal tasks

  • 21.0%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

മിസ്ട്രൽ സ്മോൾ 3-നെ കുറിച്ച്

Mistral Small 4, the lab's hybrid model unifying instruct, reasoning and coding in one efficient set of weights. It is the first Mistral to fold the Magistral reasoning, Pixtral multimodal and Devstral agentic-coding lines into a single versatile model, tuned for general chat, coding, agentic tasks and complex reasoning. It is a Mixture-of-Experts with 119B total and roughly 6B active parameters, taking text and image input across a 256K context.

Small 4 is the release where Mistral stopped asking people to pick a model. The lab describes it as the first Mistral to fold its reasoning, multimodal and agentic-coding lines into one set of weights, so a fast instruct reply, a step-by-step derivation and a tool-driven coding run all come from the same checkpoint. Its sparse design routes four of 128 experts per token, which is why a 119 billion parameter model costs what a much smaller one does to serve.

The control that makes the merge work is a reasoning-effort parameter with two named settings, and Mistral describes both in terms of the models they replace. Setting it to none produces fast, lightweight answers in the same chat style as the previous Small release. Setting it to high produces deep step-by-step reasoning with roughly the verbosity of the lab's dedicated reasoning models. That is the whole switch: no second endpoint, no second deployment.

Mistral's own comparison is against its earlier models rather than the field. With reasoning on it reports 71.2 on GPQA Diamond and 78 on MMLU Pro, against 59.1 and 73.5 for the same weights in instruct mode, 48 against 35.7 on the AllenAI instruction-following benchmark IFBench, and 60 against 46.3 on the vision benchmark MMMU-Pro. Alongside the accuracy the lab pushes an efficiency claim: it says Small 4 matches or beats a comparable open model on three benchmarks while generating substantially shorter answers, and that shorter answers are the point, because they are what actually shows up as latency and cost.

The engineering numbers are stated plainly. Mistral reports a 40% cut in end-to-end completion time in a latency-tuned setup and three times the requests per second in a throughput-tuned one, against the previous Small generation. It names the minimum hardware - four H100s, two H200s or a single DGX B200 - and ships under Apache 2.0, with day-one support across the common open serving stacks. The lab's own framing of who it is for is equally plain: developers doing coding automation and codebase exploration, enterprises running chat assistants and document understanding, researchers doing maths and complex reasoning. It is the efficient tier, not the frontier one, and Mistral points at its larger models for the hardest work.

ലോഞ്ചിംഗ് വേളയിൽ Mistral AI വ്യക്തമാക്കിയത്

Three model lines in one
Mistral's first model to unify its reasoning, multimodal and agentic-coding lines, so users no longer choose between a fast instruct model, a reasoning engine and a vision assistant.
Reasoning effort as a parameter
Setting effort to none gives the previous generation chat style; setting it to high gives step-by-step reasoning with the verbosity of the dedicated reasoning models it replaces.
Sparse routing at 119B
A Mixture-of-Experts with 128 experts and four active per token, around 6B active parameters per token, which is what keeps serving cost near a much smaller model.
Reasoning versus instruct scores
The lab reports 71.2 against 59.1 on GPQA Diamond, 78 against 73.5 on MMLU Pro and 60 against 46.3 on MMMU-Pro, comparing the same weights with reasoning on and off.
Shorter answers on purpose
Mistral argues efficiency per token is the real metric and reports matching or beating a comparable open model while generating substantially less text to get there.
Serving cost and hardware floor
The lab reports a 40% reduction in end-to-end completion time and three times the requests per second against the previous Small generation, with a floor of four H100s, two H200s or one DGX B200.

ഇന്ത്യൻ ഭാഷകൾ

മിസ്ട്രൽ സ്മോൾ 3 4 ഇന്ത്യൻ ഭാഷകളിൽ മറുപടി നൽകും. മെസ്സേജ് ബോക്സിന് അടുത്തുള്ള മെനുവിൽ നിന്ന് ആവശ്യമുള്ള ഭാഷ തിരഞ്ഞെടുക്കാം.

Frequently Asked Questions

മിസ്ട്രൽ സ്മോൾ 3 സംബന്ധിച്ച പ്രധാന ചോദ്യങ്ങളും ഉത്തരങ്ങളും.

മിസ്ട്രൽ സ്മോൾ 3 എപ്പോഴാണ് റിലീസ് ചെയ്തത്?

Mistral AI 2026 മാർ 16-ൽ മിസ്ട്രൽ സ്മോൾ 3 പുറത്തിറക്കി.

മിസ്ട്രൽ സ്മോൾ 3 വികസിപ്പിച്ചത് ആരാണ്?

മിസ്ട്രൽ സ്മോൾ 3 വികസിപ്പിച്ചത് Mistral AI ആണ്. 99Models AI ഒറിജിനൽ പ്രൊവൈഡർ നിരക്കിൽ തന്നെ ഇതിലേക്ക് കണക്ട് ചെയ്യുന്നു.

മിസ്ട്രൽ സ്മോൾ 3 എത്രത്തോളം കാര്യക്ഷമമാണ്?

ഇന്റലിജൻസ് റാങ്കിംഗിൽ 54 മോഡലുകളിൽ 51-ാം സ്ഥാനത്താണ് ഇത്. ബെഞ്ച്മാർക്ക് സ്കോറുകൾ മുകളിലുള്ള പാനലിൽ കാണാം.

മിസ്ട്രൽ സ്മോൾ 3 ഉപയോഗിക്കാൻ എത്ര ചെലവാകും?

10 ലക്ഷം ഇൻപുട്ട് Tokens-ന് ₹18.00 രൂപയും ഔട്ട്പുട്ടിന് ₹72.00 രൂപയുമാണ് അധിക മാർക്ക്അപ്പില്ലാത്ത നിരക്ക്. സബ്‌സ്‌ക്രിപ്ഷനില്ല, ഉപയോഗിക്കുന്നതിന് മാത്രം പണമടയ്ക്കുക.

മിസ്ട്രൽ സ്മോൾ 3 API-യുടെ ഡോളർ നിരക്ക് എത്രയാണ്?

10 ലക്ഷം ഇൻപുട്ട് Tokens-ന് $0.19 ഡോളറും 10 ലക്ഷം ഔട്ട്പുട്ട് Tokens-ന് $0.75 ഡോളറുമാണ് നിരക്ക്. ഈ പേജിലെ രൂപ നിരക്കുകൾ ഡോളറിന് ₹96 എന്ന നിരക്കിൽ മാറ്റിയതാണ്.

മിസ്ട്രൽ സ്മോൾ 3-ന് എത്ര നീളമുള്ള സംഭാഷണം ഓർത്തുനിൽക്കാനാകും?

ഇതിന്റെ Context window 2.6 lakh Tokens ആണ്. ഒരുമിച്ച് നൽകുന്ന സംഭാഷണങ്ങളും ഫയലുകളും ഉൾപ്പെടെ ഇതിൽ വായിക്കാൻ സാധിക്കും.

മിസ്ട്രൽ സ്മോൾ 3 മികച്ച വാല്യൂ നൽകുന്ന ഒന്നാണോ?

പെർഫോമൻസും നിരക്കും അടിസ്ഥാനമാക്കിയുള്ള വാല്യൂ റാങ്കിംഗിൽ 54 മോഡലുകളിൽ 47-ാം സ്ഥാനത്താണ് ഇത്.

മറുപടി നൽകുന്നതിന് മുൻപ് മിസ്ട്രൽ സ്മോൾ 3 Reasoning നടത്തുമോ?

ഓപ്ഷണൽ. Reasoning പിന്തുണയ്ക്കുന്ന മോഡലുകളിൽ ചിന്തിക്കുന്നതിന്റെ വ്യാപ്തി മെസ്സേജ് ബോക്സിൽ ക്രമീകരിക്കാം.

മിസ്ട്രൽ സ്മോൾ 3 ഏതൊക്കെ ഇന്ത്യൻ ഭാഷകളിൽ മറുപടി നൽകും?

4 ഇന്ത്യൻ ഭാഷകളിൽ ഇത് മറുപടി നൽകും. മെസ്സേജ് ബോക്സിന് അടുത്തുള്ള മെനുവിൽ നിന്ന് ഭാഷ തിരഞ്ഞെടുക്കാം.

Mistral AI-ൽ നിന്നുള്ള മറ്റ് Models

കാറ്റലോഗ് പുതുക്കിയത്: 2026 സെപ്റ്റം 9