99MODELS

हुनयुआन 3

नयाँ

Fast, very cheap general model; heavy real-world adoption.

रिलिज मिति: 2026 जुलाई 6

76.80

प्रति 10 लाख आउटपुट Tokens

इनपुट: ₹19.20 प्रति 10 लाख Tokens

प्रति अमेरिकी डलर ₹96 मा 0% मार्कअपका साथ प्रदायककै दरमा गणना गरिन्छ।

विवरण

Context विन्डो
2,62,144 Tokens
अधिकतम आउटपुट
1,31,072 Tokens
स्वीकार गर्छ
टेक्स्ट
Reasoning
सुरुमै चालु
प्रयास स्तर
none, low, high
टुल प्रयोग
संरचित आउटपुट
कोड कार्यान्वयन
छैन
इन्टेलिजेन्स र्‍याङ्क
54 मध्ये #36
भ्याल्यू र्‍याङ्क
54 मध्ये #11

मूल्य

मूल्य
प्रति 10 lakh TokensINRUSD
इनपुट19.20$0.20
आउटपुट76.80$0.80
क्यास गरिएको इनपुट4.80$0.05

बेन्चमार्क

रेटिङ बाहेकका सबै स्कोरहरू प्रतिशतमा छन्। सबै बेन्चमार्कहरू स्वतन्त्र रूपमा मापन गरिएका हुन्।

  • 89.7%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 33.5%

    HLE

    Humanity's Last Exam

  • 48.6%

    SciCode

    SciCode - scientific code generation

  • 79.0%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 64.4%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

हुनयुआन 3 को बारेमा

Tencent Hunyuan's third-generation model: a 295B-parameter Mixture-of-Experts with 21B active parameters routing across 192 experts, built for agent capability and real-world usability. It already powers Tencent products including CodeBuddy and Yuanbao, and Tencent states it matches the intelligence of flagship models two to five times its parameter scale at a fraction of the cost. Unusually, thinking is off by default here and opt-in, over a 256K context.

Hy3 is the general-availability build of a model whose preview had already been in production for ten weeks. Tencent rebuilt the Hunyuan infrastructure in late January 2026, shipped the Hy3 preview in April, then used feedback from more than fifty of its own products to scale up post-training - and it counts that loop, from foundational rebuild to shipped model, as under six months. The lab attributes the improvement to stronger reinforcement learning plus better data quality and diversity rather than to more parameters.

The positioning is unusually practical for a frontier-adjacent release. Tencent names the domains it improved most - software development, office productivity, financial modelling, front-end design and game production - and argues the case with usage rather than scores: average daily token consumption across its products rose twentyfold after the preview shipped, and the number of people actively choosing the model inside its workplace agent grew sixfold. It also ran a blind evaluation with 270 domain experts using tasks from their own jobs, reporting an average of 2.67 out of 4 against 2.51 for the strongest open competitor it tested, with the biggest margins in front-end development, data and storage, and continuous integration.

Much of what changed is reliability rather than capability, and Tencent is specific about it. It reports hallucination rate in its internal real-world evaluation falling from 12.5% to 5.4% and commonsense error rate from 25.4% to 12.7%, under a stated principle of answering when grounded, saying so when the evidence is missing, and never conflating sources. Multi-turn behaviour improved on the operational failures that actually break agents - pronoun resolution, recovering an omitted subject, carrying a constraint forward across turns - with its internal issue rate falling from 17.4% to 7.9%. It also reports that accuracy on SWE-bench Verified varies by less than four points across three different agent scaffolds, which is a claim about portability rather than peak score.

On its own benchmark table Tencent puts Hy3 at 71.7 on Terminal Bench 2.1, 57.9 on SWE-bench Pro, 75.8 on SWE-bench Multilingual, 84.2 on the browsing benchmark BrowseComp, 79.1 on the public MCP Atlas split, 90.4 on GPQA Diamond, and 53.2 on Humanity's Last Exam with tools against 37.0 without them. Where it stops is the frontier: on most of those same rows the lab's table shows the largest closed models ahead of it, and its argument is the ratio, not the ceiling - intelligence comparable to flagship models two to five times its parameter scale, at a fraction of the cost. It is open under Apache 2.0, so the cheapest way to be wrong about it is to not try it.

लन्चको समयमा Tencent ले के भन्यो

Shaped by real product use
Tencent scaled up post-training on feedback from more than fifty of its own products after the April preview, and reports daily token consumption across them rising twentyfold.
Judged by working experts
A blind evaluation with 270 domain experts using tasks from their own jobs scored Hy3 at 2.67 out of 4 against 2.51 for the strongest open competitor Tencent tested.
Fewer fabrications
In Tencent's internal real-world evaluation the hallucination rate fell from 12.5% to 5.4% and the commonsense error rate from 25.4% to 12.7%.
Holds a long conversation
Multi-turn work targeted pronoun resolution, recovering omitted subjects and carrying constraints across turns, with the internal issue rate falling from 17.4% to 7.9%.
Portable across agent scaffolds
Tencent reports SWE-bench Verified accuracy varying by less than four points across three different agent frameworks, so behaviour does not hinge on one harness.
What it is not for
The lab's own table places the largest closed models above Hy3 on most rows; its claim is intelligence comparable to models two to five times its scale at a fraction of the cost, not the frontier itself.

भारतीय भाषाहरू

हुनयुआन 3 ले 12 भारतीय भाषाहरूमा जवाफ दिन्छ। सन्देश बाकस छेउको मेनुबाट भाषा छान्नुहोस् र सोही भाषामा जवाफ पाउनुहोस्।

Frequently Asked Questions

हुनयुआन 3 सम्बन्धी प्रायः सोधिने प्रश्नहरू।

हुनयुआन 3 कहिले रिलिज भएको हो?

Tencent ले हुनयुआन 3 लाई 2026 जुलाई 6 मा सार्वजनिक गरेको हो।

हुनयुआन 3 कसले बनाएको हो?

हुनयुआन 3 लाई Tencent ले बनाएको हो। 99Models AI ले प्रदायककै दरमा सिधै जोड्दछ।

हुनयुआन 3 कत्तिको सक्षम र बुद्धिमानी छ?

यो हाम्रो बौद्धिकता श्रेणीकरणमा 54 च्याट Models मध्ये 36 स्थानमा छ। यसको पूर्ण अङ्क माथिको Benchmarks तालिकामा हेर्न सकिन्छ।

हुनयुआन 3 को लागत कति पर्छ?

यसमा 0% मार्कअपका साथ प्रति 10 लाख इनपुट Tokens को ₹19.20 र आउटपुटको ₹76.80 लाग्छ। कुनै सदस्यता छैन; तपाईंले प्रयोग गरेअनुसार मात्र भुक्तानी गर्नुहुन्छ।

डलरमा हुनयुआन 3 को API मूल्य कति हो?

प्रदायकले प्रति 10 लाख इनपुट Tokens को $0.20 र आउटपुट Tokens को $0.80 शुल्क लिन्छ। यस पृष्ठका दरहरू प्रति अमेरिकी डलर ₹96 मा रूपान्तरण गरिएका हुन्।

हुनयुआन 3 ले कति लामो कुराकानी सम्झन सक्छ?

यसको Context विन्डो 2.6 lakh Tokens हो। यसले एकल अनुरोधमा प्रक्रिया गर्न सक्ने कुराकानी र संलग्न फाइलहरूको कुल क्षमता यही हो।

के हुनयुआन 3 लागत अनुसार उत्कृष्ट छ?

मूल्य र गुणस्तरको आधारमा यो 54 Models मध्ये 11 स्थानमा छ। यसले बौद्धिकता र Token लागतको तुलना गर्दछ।

के हुनयुआन 3 ले जवाफ दिनुअघि विचार गर्छ?

सुरुमै चालु। Reasoning उपलब्ध भएको ठाउँमा तपाईंले सिधै सन्देश बक्समा सोच्ने क्षमता समायोजन गर्न सक्नुहुन्छ।

हुनयुआन 3 ले कुन-कुन भारतीय भाषाहरूमा जवाफ दिन्छ?

यसले 12 भारतीय भाषाहरूमा जवाफ दिन्छ। सन्देश बाकस छेउको मेनुबाट भाषा छान्नुहोस् र सोही भाषामा जवाफ पाउनुहोस्।

क्याटलग अद्यावधिक गरिएको मिति: 2026 सेप्टेम्बर 9