म्युज स्पार्क 1.2
नयाँMeta's reasoning model for complex agentic tasks; accepts every modality.
रिलिज मिति: 2026 अगस्ट 5
₹408.00
प्रति 10 लाख आउटपुट Tokens
इनपुट: ₹120.00 प्रति 10 लाख Tokens
प्रति अमेरिकी डलर ₹96 मा 0% मार्कअपका साथ प्रदायककै दरमा गणना गरिन्छ।
विवरण
- Context विन्डो
- 10,48,576 Tokens
- अधिकतम आउटपुट
- 9,43,718 Tokens
- स्वीकार गर्छ
- टेक्स्ट, तस्बिरहरू, PDF फाइलहरू, अडियो, भिडियो
- Reasoning
- सुरुमै चालु
- प्रयास स्तर
- minimal, low, medium, high, xhigh
- टुल प्रयोग
- छ
- संरचित आउटपुट
- छ
- कोड कार्यान्वयन
- छैन
- इन्टेलिजेन्स र्याङ्क
- 54 मध्ये #14
- भ्याल्यू र्याङ्क
- 54 मध्ये #9
मूल्य
बेन्चमार्क
रेटिङ बाहेकका सबै स्कोरहरू प्रतिशतमा छन्। सबै बेन्चमार्कहरू स्वतन्त्र रूपमा मापन गरिएका हुन्।
90.4%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
45.5%
HLE
Humanity's Last Exam
57.4%
SciCode
SciCode - scientific code generation
79.0%
Long Context
Long Context Reasoning - reasoning over long inputs
80.1%
Terminal-Bench 2
Terminal-Bench 2.1 - agentic terminal tasks, second edition
1534
WebDev Arena
WebDev Arena - head-to-head web-app builds, Elo rating
म्युज स्पार्क 1.2 को बारेमा
Meta Superintelligence Labs' multimodal reasoning model for agentic tasks, with major gains in tool and computer use, coding and multimodal understanding. Version 1.2 is the coding-optimised checkpoint, built for long-horizon multi-agentic workflows such as multi-file refactors and debugging sessions that run well past a single prompt, with a million-token window so an entire repository fits in one session. It was trained with asynchronous and parallel tool calls plus planning, goal conditioning and context compaction to hold focus across extended tasks.
Meta did not ship this checkpoint on its own. It arrived with Muse Code, a terminal coding agent, and the two were co-trained - the lab says it folded rejection-sampled harness trajectories into training along with recipe work for goals, context compaction and subagents, and integrated the agent's own toolset so the model and the harness would not disagree about how a task is run. That is the frame the release asks for: this is a model tuned to a specific way of working rather than a general upgrade.
The training emphasis is long-horizon work, and Meta names the material: whole-repository generation, large end-to-end projects and automated research. It credits three mechanisms for holding a run together over hours - planning to sequence the work, goal conditioning to keep direction, and context compaction to carry forward only what still matters. It also describes a self-improvement loop in which the previous Muse Spark generated hard coding environments and instruction-following templates, then graded candidate solutions against them, producing training data at a scale hand-authoring could not reach.
On its own charts Meta reports 82.9% on Terminal-Bench 2.1 run through Muse Code and 59.3% on DeepSWE 1.1, against 76.2% and 53.0% for the previous version measured in a lighter harness. The lab is not claiming the top of either chart: on both it places a frontier competitor above Muse Spark 1.2, and on the software-engineering set two competitors sit above it. The gain it is selling is against its own predecessor and against the cost of the tier.
The most telling result Meta published is not a benchmark at all. It set the model to optimise GPU kernels over more than a thousand tool calls and up to twenty-four hours, writing, compiling, profiling and improving Triton implementations against a reference, with third-party kernel libraries explicitly forbidden so the model had to implement the algorithms rather than wrap someone else's. It reports substantial and continuing improvement over the baseline across that window. The launch demo makes the multimodal half concrete in the same spirit: a walkthrough video handed to the terminal as a file, turned into a working booking page. Meta frames the release as a step rather than a destination, saying larger and much more capable models are on the way.
लन्चको समयमा Meta ले के भन्यो
- Co-trained with its harness
- Meta trained the model together with its terminal coding agent, folding in rejection-sampled harness trajectories and the agent's own toolset so model and scaffold behave consistently.
- Trained on long-horizon work
- The lab names whole-repository generation, large end-to-end projects and automated research as training material, held together by planning, goal conditioning and context compaction.
- Coding benchmarks
- Meta reports 82.9% on Terminal-Bench 2.1 through its own agent and 59.3% on DeepSWE 1.1, against 76.2% and 53.0% for the previous Muse Spark release.
- Self-improvement loop
- The previous generation generated challenging coding environments and instruction-following templates, then graded candidate solutions, producing training data for this checkpoint at scale.
- A day-long kernel optimisation run
- Meta ran the model for over a thousand tool calls and up to twenty-four hours writing, compiling and profiling Triton GPU kernels, with third-party kernel libraries forbidden.
- What it is not for
- This is a coding-focused update, and on Meta's own two charts a competing frontier model scores above it on both. The lab describes it as a step, with larger models still to come.
भारतीय भाषाहरू
म्युज स्पार्क 1.2 ले 15 भारतीय भाषाहरूमा जवाफ दिन्छ। सन्देश बाकस छेउको मेनुबाट भाषा छान्नुहोस् र सोही भाषामा जवाफ पाउनुहोस्।
Frequently Asked Questions
म्युज स्पार्क 1.2 सम्बन्धी प्रायः सोधिने प्रश्नहरू।
म्युज स्पार्क 1.2 कहिले रिलिज भएको हो?
Meta ले म्युज स्पार्क 1.2 लाई 2026 अगस्ट 5 मा सार्वजनिक गरेको हो।
म्युज स्पार्क 1.2 कसले बनाएको हो?
म्युज स्पार्क 1.2 लाई Meta ले बनाएको हो। 99Models AI ले प्रदायककै दरमा सिधै जोड्दछ।
म्युज स्पार्क 1.2 कत्तिको सक्षम र बुद्धिमानी छ?
यो हाम्रो बौद्धिकता श्रेणीकरणमा 54 च्याट Models मध्ये 14 स्थानमा छ। यसको पूर्ण अङ्क माथिको Benchmarks तालिकामा हेर्न सकिन्छ।
म्युज स्पार्क 1.2 को लागत कति पर्छ?
यसमा 0% मार्कअपका साथ प्रति 10 लाख इनपुट Tokens को ₹120.00 र आउटपुटको ₹408.00 लाग्छ। कुनै सदस्यता छैन; तपाईंले प्रयोग गरेअनुसार मात्र भुक्तानी गर्नुहुन्छ।
डलरमा म्युज स्पार्क 1.2 को API मूल्य कति हो?
प्रदायकले प्रति 10 लाख इनपुट Tokens को $1.25 र आउटपुट Tokens को $4.25 शुल्क लिन्छ। यस पृष्ठका दरहरू प्रति अमेरिकी डलर ₹96 मा रूपान्तरण गरिएका हुन्।
म्युज स्पार्क 1.2 ले कति लामो कुराकानी सम्झन सक्छ?
यसको Context विन्डो 10.5 lakh Tokens हो। यसले एकल अनुरोधमा प्रक्रिया गर्न सक्ने कुराकानी र संलग्न फाइलहरूको कुल क्षमता यही हो।
के म्युज स्पार्क 1.2 लागत अनुसार उत्कृष्ट छ?
मूल्य र गुणस्तरको आधारमा यो 54 Models मध्ये 9 स्थानमा छ। यसले बौद्धिकता र Token लागतको तुलना गर्दछ।
के म्युज स्पार्क 1.2 ले जवाफ दिनुअघि विचार गर्छ?
सुरुमै चालु। Reasoning उपलब्ध भएको ठाउँमा तपाईंले सिधै सन्देश बक्समा सोच्ने क्षमता समायोजन गर्न सक्नुहुन्छ।
म्युज स्पार्क 1.2 ले कुन-कुन भारतीय भाषाहरूमा जवाफ दिन्छ?
यसले 15 भारतीय भाषाहरूमा जवाफ दिन्छ। सन्देश बाकस छेउको मेनुबाट भाषा छान्नुहोस् र सोही भाषामा जवाफ पाउनुहोस्।
Meta का अन्य Models
क्याटलग अद्यावधिक गरिएको मिति: 2026 सेप्टेम्बर 9