Ling 3.0 Flash VL
नयाँEfficient vision reasoner for multimodal agents and document understanding.
रिलिज मिति: 2026 सेप्टेम्बर 10
₹17.37
प्रति 10 लाख आउटपुट Tokens
इनपुट: ₹5.79 प्रति 10 लाख Tokens
प्रति अमेरिकी डलर ₹96.5 मा 0% मार्कअपका साथ प्रदायककै दरमा गणना गरिन्छ।
विवरण
- Context विन्डो
- 1,31,072 Tokens
- अधिकतम आउटपुट
- 32,768 Tokens
- स्वीकार गर्छ
- टेक्स्ट, तस्बिरहरू
- Reasoning
- सुरुमै चालु
- टुल प्रयोग
- छ
- संरचित आउटपुट
- छ
- कोड कार्यान्वयन
- छैन
- प्यारामिटरहरू
- 124B A5.5B MoE
- इन्टेलिजेन्स र्याङ्क
- 59 मध्ये #46
- भ्याल्यू र्याङ्क
- 59 मध्ये #4
मूल्य
बेन्चमार्क
रेटिङ बाहेकका सबै स्कोरहरू प्रतिशतमा छन्। सबै बेन्चमार्कहरू स्वतन्त्र रूपमा मापन गरिएका हुन्।
86.2%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
22.0%
HLE
Humanity's Last Exam
44.2%
SciCode
SciCode - scientific code generation
78.3%
Long Context
Long Context Reasoning - reasoning over long inputs
64.4%
Terminal-Bench 2
Terminal-Bench 2.1 - agentic terminal tasks, second edition
Ling 3.0 Flash VL को बारेमा
Ling 3.0 Flash VL is InclusionAI's multimodal reasoning model, with 124 billion total and 5.5 billion active parameters. It extends Ling 3.0 Flash with visual understanding for charts, documents and tool-assisted workflows. The hosted model accepts image inputs in a 131,072-token context and supports structured output and function calling.
Ling-3.0-flash-VL extends inclusionAI's text model into visual work. The lab describes a system that uses visual evidence throughout reasoning and verification, with examples spanning document interpretation, charts and software interfaces. Its ambition reaches beyond naming objects in an image.
The model combines a visual encoder with a sparse mixture of experts: 124 billion parameters overall, with 5.5 billion active for each token. Its hybrid attention backbone is designed to process long task histories efficiently. These are architectural choices aimed at balancing capacity with inference cost, rather than a guarantee of performance on every workload.
The launch materials distinguish understanding visual content, reasoning from it and translating interface observations into actions. They also describe video support in the underlying model. The inputs and context exposed here follow the serving provider's configuration, so the capabilities shown on this page are the practical limits for using it through 99models.
लन्चको समयमा InclusionAI ले के भन्यो
- Visual reasoning
- Uses visual evidence for calculation and verification.
- Sparse computation
- Activates 5.5 billion parameters per token.
- Interface understanding
- The lab highlights tasks involving software and web interfaces.
Frequently Asked Questions
Ling 3.0 Flash VL सम्बन्धी प्रायः सोधिने प्रश्नहरू।
Ling 3.0 Flash VL कहिले रिलिज भएको हो?
InclusionAI ले Ling 3.0 Flash VL लाई 2026 सेप्टेम्बर 10 मा सार्वजनिक गरेको हो।
Ling 3.0 Flash VL कसले बनाएको हो?
Ling 3.0 Flash VL लाई InclusionAI ले बनाएको हो। 99Models AI ले प्रदायककै दरमा सिधै जोड्दछ।
Ling 3.0 Flash VL कत्तिको सक्षम र बुद्धिमानी छ?
यो हाम्रो बौद्धिकता श्रेणीकरणमा 59 च्याट Models मध्ये 46 स्थानमा छ। यसको पूर्ण अङ्क माथिको Benchmarks तालिकामा हेर्न सकिन्छ।
Ling 3.0 Flash VL को लागत कति पर्छ?
यसमा 0% मार्कअपका साथ प्रति 10 लाख इनपुट Tokens को ₹5.79 र आउटपुटको ₹17.37 लाग्छ। कुनै सदस्यता छैन; तपाईंले प्रयोग गरेअनुसार मात्र भुक्तानी गर्नुहुन्छ।
डलरमा Ling 3.0 Flash VL को API मूल्य कति हो?
प्रदायकले प्रति 10 लाख इनपुट Tokens को $0.06 र आउटपुट Tokens को $0.18 शुल्क लिन्छ। यस पृष्ठका दरहरू प्रति अमेरिकी डलर ₹96.5 मा रूपान्तरण गरिएका हुन्।
Ling 3.0 Flash VL ले कति लामो कुराकानी सम्झन सक्छ?
यसको Context विन्डो 1.3 lakh Tokens हो। यसले एकल अनुरोधमा प्रक्रिया गर्न सक्ने कुराकानी र संलग्न फाइलहरूको कुल क्षमता यही हो।
के Ling 3.0 Flash VL लागत अनुसार उत्कृष्ट छ?
मूल्य र गुणस्तरको आधारमा यो 59 Models मध्ये 4 स्थानमा छ। यसले बौद्धिकता र Token लागतको तुलना गर्दछ।
के Ling 3.0 Flash VL ले जवाफ दिनुअघि विचार गर्छ?
सुरुमै चालु। Reasoning उपलब्ध भएको ठाउँमा तपाईंले सिधै सन्देश बक्समा सोच्ने क्षमता समायोजन गर्न सक्नुहुन्छ।
क्याटलग अद्यावधिक गरिएको मिति: 2026 सेप्टेम्बर 27