99MODELS

મ્યુઝ ગ્લિમર 30B

નવું

Open-weight dense 30B distilled from Muse Spark; runs fast and cheap.

રજૂઆત: 9 ઑગસ્ટ, 2026

144.00

10 લાખ આઉટપુટ Tokens દીઠ

ઇનપુટ: 10 લાખ Tokens દીઠ ₹33.60

પ્રોવાઇડરના ભાવે બિલિંગ, $1 = ₹96 મુજબ રૂપાંતરિત, 0% માર્કઅપ સાથે.

વિશિષ્ટતાઓ

Context વિન્ડો
1,31,072 Tokens
મહત્તમ આઉટપુટ
1,17,964 Tokens
સ્વીકારે છે
ટેક્સ્ટ, ઇમેજ
Reasoning
ડિફૉલ્ટ રૂપે ચાલુ
એફર્ટ લેવલ
low, medium, high, xhigh
ટૂલ ઉપયોગ
હા
સ્ટ્રક્ચર્ડ આઉટપુટ
હા
કોડ એક્ઝિક્યુશન
ના
પેરામીટર્સ
30B
ઇન્ટેલિજન્સ રેન્ક
54 માંથી #42
Value રેન્ક
54 માંથી #37

કિંમત

કિંમત
પ્રતિ 10 lakh TokensINRUSD
ઇનપુટ33.60$0.35
આઉટપુટ144.00$1.50
કૅશ ઇનપુટ3.84$0.04

Benchmarks

સ્કોર ટકાવારીમાં છે સિવાય કે રેટિંગ તરીકે દર્શાવેલ હોય. તમામ બેન્ચમાર્ક સ્વતંત્ર રીતે માપવામાં આવ્યા છે.

  • 83.5%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 22.0%

    HLE

    Humanity's Last Exam

  • 44.9%

    SciCode

    SciCode - scientific code generation

  • 83.3%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 51.7%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

મ્યુઝ ગ્લિમર 30B વિશે

Meta's open-weight 30B model, optimised for local, always-on agent workflows. It is a dense transformer with a dedicated perception encoder, distilled from Muse Spark and purpose-built to run autonomous agents on consumer hardware. Meta aims it at end-to-end agentic task completion, reliable tool use, multi-step reasoning, failure recovery and use as an evaluation judge, with text and image input and selectable reasoning strength.

The argument Meta makes for this model is about where it runs. Most agent deployments assume a network and a datacentre; an agent that manages your calendar, drafts your messages and organises your files needs deep access to personal context, and Meta's answer is to put the model on the machine that already holds it. Everything about the design follows from that constraint - a compact architecture, a distillation recipe that pulls agentic reasoning out of a far larger teacher, and inference work aimed at latency rather than peak quality.

Meta breaks the recipe into three named stages. The first trains Glimmer on Muse Spark's own outputs rather than on raw text alone, so the small model inherits the large one's distribution. The second stretches the context and tilts the corpus towards agent transcripts that show their reasoning, kept diluted with ordinary text. The third layers three techniques over one another - supervised fine-tuning, on-policy distillation, then reinforcement learning - and applies them to general, reasoning, coding and agentic work alike. Meta says the result was assessed for open-weight release under its own Advanced AI Scaling Framework across every relevant risk category.

The capability list is written for agent builders rather than chat users, and one entry stands out: what happens after a tool call goes wrong. Meta says Glimmer was trained to read the failure, work out why it happened and try again, instead of halting on the error - which is the difference between an agent that finishes and one that needs a human. Alongside that sit precise function calling with real schemas across extended workflows, multi-step planning, a dedicated perception encoder that lets it read interleaved screenshots, charts and documents, selectable reasoning strength, and training data spanning more than a hundred languages.

Making it fit was its own engineering problem, and Meta shows the arithmetic. Thirty billion parameters at full precision needs over 55 GB; quantised to roughly four bits the language model drops under 20 GB, which leaves room for the working memory, the perception encoder and a speculative-decoding drafter inside a 24 GB or 32 GB budget - with, the lab says, minimal to no degradation on agentic tasks. A lightweight drafter proposes whole blocks of tokens that the main model verifies in parallel, which Meta reports as a large speed-up on a high-end consumer GPU and a meaningful one on laptop silicon.

Meta's own comparison table is unusually honest about the shape of the trade. Against the two open models near its size it leads on the agent-completion and tool-calling columns - the ones it was built for - while trailing on terminal-style coding, on computer use, and on general knowledge and hardest-exam questions. Read it as a model tuned for finishing tool-driven tasks locally, not as a small general-purpose frontier model.

Meta એ લોન્ચ વખતે શું જણાવ્યું હતું

Distilled from the flagship
Glimmer learns from Muse Spark's outputs rather than from raw text alone, then gains context length and agent behaviour in a second stage, before a third stage stacks fine-tuning, distillation and reinforcement learning.
Recovers from failed tool calls
Meta trained the model to diagnose an error and retry when a tool call fails or returns something unexpected, rather than halting the run.
Reads screens and documents
A dedicated perception encoder accepts interleaved text and images, so an agent can interpret screenshots, charts and documents alongside the conversation.
Engineered to fit a consumer GPU
At full precision the model needs over 55 GB; quantised to roughly four bits it drops under 20 GB, leaving room for working memory, the encoder and a drafter inside a 24 GB or 32 GB budget.
Speculative decoding for responsiveness
A lightweight drafter proposes blocks of tokens that the main model verifies in parallel, which Meta reports as a large generation speed-up on a high-end consumer GPU.
What it is not for
On Meta's own table it leads its size class on agent completion and tool calling but trails on terminal-style coding, computer use and general knowledge, so it is not a small general-purpose frontier model.

ભારતીય ભાષાઓ

મ્યુઝ ગ્લિમર 30B કુલ 9 ભારતીય ભાષાઓમાં જવાબ આપી શકે છે. મેસેજ બોક્સની બાજુમાં આપેલા મેનૂમાંથી ભાષા પસંદ કરો.

Frequently Asked Questions

મ્યુઝ ગ્લિમર 30B વિશે વારંવાર પૂછાતા પ્રશ્નો.

મ્યુઝ ગ્લિમર 30B ક્યારે રજૂ કરવામાં આવ્યું હતું?

Meta એ 9 ઑગસ્ટ, 2026 ના રોજ મ્યુઝ ગ્લિમર 30B રજૂ કર્યું હતું.

મ્યુઝ ગ્લિમર 30B કોણે બનાવ્યું છે?

મ્યુઝ ગ્લિમર 30B ને Meta લેબ દ્વારા બનાવવામાં આવ્યું છે. 99Models સીધા પ્રોવાઇડરના નિર્ધારિત દરે કનેક્ટ કરે છે.

મ્યુઝ ગ્લિમર 30B કેટલું સક્ષમ અને બુદ્ધિશાળી છે?

ઇન્ટેલિજન્સ રેન્કિંગમાં તે 54 ચેટ Models માંથી 42 નંબર પર છે, જે સ્વતંત્ર બેન્ચમાર્ક સ્કોર પર આધારિત છે. તેના તમામ સ્કોર ઉપર બેન્ચમાર્ક ટેબલમાં જોઈ શકો છો.

મ્યુઝ ગ્લિમર 30B નો ઉપયોગ કરવાનો ખર્ચ કેટલો છે?

વપરાશ ખર્ચ ઇનપુટ માટે ₹33.60 પ્રતિ 10 લાખ Tokens અને આઉટપુટ માટે ₹144.00 પ્રતિ 10 લાખ Tokens છે, જેમાં 0% માર્કઅપ છે. કોઈ સબ્સ્ક્રિપ્શન નથી; વપરાશ મુજબ જ પેમેન્ટ કરો.

મ્યુઝ ગ્લિમર 30B નું અમેરિકન ડોલરમાં API પ્રાઇસિંગ શું છે?

પ્રોવાઇડર ઇનપુટ માટે $0.35 પ્રતિ 10 લાખ Tokens અને આઉટપુટ માટે $1.50 પ્રતિ 10 લાખ Tokens ચાર્જ કરે છે. રૂપિયાના દરો $1 = ₹96 ના આધારે ગણવામાં આવ્યા છે.

મ્યુઝ ગ્લિમર 30B કેટલી લાંબી વાતચીત યાદ રાખી શકે છે?

તેની Context વિન્ડો 1.3 lakh Tokens ની છે. એટલે કે તે એક જ વિનંતીમાં અગાઉની વાતચીત અને જોડેલી ફાઇલો સહિત આટલું લખાણ વાંચી શકે છે.

શું મ્યુઝ ગ્લિમર 30B વાપરવું પૈસા વસૂલ છે?

વેલ્યૂ રેન્કિંગમાં તે 54 માંથી 37 નંબર પર છે. આ રેન્કિંગ બેન્ચમાર્ક ક્ષમતા અને વાસ્તવિક Token ખર્ચની તુલના કરે છે.

શું મ્યુઝ ગ્લિમર 30B જવાબ આપતાં પહેલાં વિચારે છે?

ડિફૉલ્ટ રૂપે ચાલુ. જ્યાં Reasoning સપોર્ટ ઉપલબ્ધ છે, ત્યાં તમે મેસેજ કમ્પોઝરમાં વિચારવાની ક્ષમતા (Thinking effort) જાતે સેટ કરી શકો છો.

મ્યુઝ ગ્લિમર 30B કઈ ભારતીય ભાષાઓમાં જવાબ આપી શકે છે?

તે 9 ભારતીય ભાષાઓમાં જવાબ આપે છે. મેસેજ બોક્સ પાસેના મેનૂમાંથી ભાષા પસંદ કરો અને તે જ ભાષામાં જવાબ મેળવો.

Meta ના અન્ય Models

કેટલોગ અપડેટ: 9 સપ્ટે, 2026