ડીપસીક V4 Flash Vision
નવુંExperimental vision build of V4 Flash; reads images at V4 Flash prices.
રજૂઆત: 21 ઑગસ્ટ, 2026
₹126.72
10 લાખ આઉટપુટ Tokens દીઠ
ઇનપુટ: 10 લાખ Tokens દીઠ ₹42.24
પ્રોવાઇડરના ભાવે બિલિંગ, $1 = ₹96 મુજબ રૂપાંતરિત, 0% માર્કઅપ સાથે.
વિશિષ્ટતાઓ
- Context વિન્ડો
- 10,48,576 Tokens
- મહત્તમ આઉટપુટ
- 2,62,144 Tokens
- સ્વીકારે છે
- ટેક્સ્ટ, ઇમેજ
- Reasoning
- ડિફૉલ્ટ રૂપે ચાલુ
- એફર્ટ લેવલ
- low, high, max
- ટૂલ ઉપયોગ
- હા
- સ્ટ્રક્ચર્ડ આઉટપુટ
- હા
- કોડ એક્ઝિક્યુશન
- ના
- ઇન્ટેલિજન્સ રેન્ક
- 54 માંથી #27
- Value રેન્ક
- 54 માંથી #6
કિંમત
Benchmarks
સ્કોર ટકાવારીમાં છે સિવાય કે રેટિંગ તરીકે દર્શાવેલ હોય. તમામ બેન્ચમાર્ક સ્વતંત્ર રીતે માપવામાં આવ્યા છે.
91.3%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
34.5%
HLE
Humanity's Last Exam
49.7%
SciCode
SciCode - scientific code generation
81.3%
Long Context
Long Context Reasoning - reasoning over long inputs
74.2%
Terminal-Bench 2
Terminal-Bench 2.1 - agentic terminal tasks, second edition
ડીપસીક V4 Flash Vision વિશે
An experimental vision build of V4 Flash: the same sparse Mixture-of-Experts design with 13B active parameters out of 284B total, extended to read images while matching the base model on text, agents, reasoning and world knowledge. DeepSeek aims it at document and chart understanding, visual question answering and multimodal agent workflows that interleave text and images, and reports its multimodal agent scores as a large jump over plain V4 Flash. Images are tokenised for billing at up to 384 tokens each, so vision costs V4 Flash rates. It is labelled experimental by DeepSeek, and the id may change.
This is V4 Flash with eyes, and DeepSeek is careful to say it is nothing more than that on the text side. It is the same sparse Mixture-of-Experts design - 13B active parameters out of 284B total - and the lab's claim is that it matches plain V4 Flash on text capabilities including agents, reasoning and world knowledge, with image understanding added rather than traded for.
Where it moves is the multimodal half. DeepSeek reports that on multimodal agent benchmarks the model makes a major leap over V4 Flash, bringing multimodal agent performance close to Claude Opus 4.8. That is the claim the release exists to make, and it is a specific one: not that the model sees well in isolation, but that it holds up when seeing is one step inside a longer agentic loop.
The uses DeepSeek names follow from that framing - document and chart understanding, visual question answering, and multimodal agent workflows that interleave text and images. It is the interleaving that matters. A model that reads a chart is useful; a model that reads a chart in the middle of a tool-calling sequence, without losing the thread of what it was doing, is a different capability, and it is the one being claimed here.
Billing is the part most likely to surprise. Images are tokenised at up to 384 tokens each and charged at ordinary V4 Flash rates, so vision costs essentially nothing beyond the tokens - a screenshot is priced like a short paragraph. Access to the Files API is free. Against a model already among the cheapest capable options available, that makes visual input unusually inexpensive to use in volume.
The caveat is in the name and DeepSeek does not soften it: this is labelled experimental, and its id carries an "exp" suffix to say so. Experimental ids get renamed, superseded or withdrawn on the lab's schedule rather than yours, so treat it as a capability to try rather than one to build a dependency on. There is no separate published architecture or context detail beyond what V4 Flash already carries.
DeepSeek એ લોન્ચ વખતે શું જણાવ્યું હતું
- V4 Flash, plus sight
- The same sparse Mixture-of-Experts model with 13B active parameters out of 284B total. DeepSeek says it matches plain V4 Flash on text, agents, reasoning and world knowledge, with vision added rather than substituted.
- Multimodal agents, not just images
- DeepSeek reports a major leap over V4 Flash on multimodal agent benchmarks, bringing that performance close to Claude Opus 4.8 - the claim is about seeing inside a longer loop, not about single-image tasks.
- What it is built to read
- Documents and charts, visual question answering, and agent workflows that interleave text and images without losing the thread of the task in between.
- Vision at text prices
- Images are tokenised at up to 384 tokens each and billed at ordinary V4 Flash rates, so a screenshot costs about what a short paragraph costs. Files API access is free.
- What it is not for
- DeepSeek labels this experimental. Experimental ids get renamed, superseded or withdrawn on the lab's timetable, so it is a capability worth trying rather than one to build a hard dependency on.
Frequently Asked Questions
ડીપસીક V4 Flash Vision વિશે વારંવાર પૂછાતા પ્રશ્નો.
ડીપસીક V4 Flash Vision ક્યારે રજૂ કરવામાં આવ્યું હતું?
DeepSeek એ 21 ઑગસ્ટ, 2026 ના રોજ ડીપસીક V4 Flash Vision રજૂ કર્યું હતું.
ડીપસીક V4 Flash Vision કોણે બનાવ્યું છે?
ડીપસીક V4 Flash Vision ને DeepSeek લેબ દ્વારા બનાવવામાં આવ્યું છે. 99Models સીધા પ્રોવાઇડરના નિર્ધારિત દરે કનેક્ટ કરે છે.
ડીપસીક V4 Flash Vision કેટલું સક્ષમ અને બુદ્ધિશાળી છે?
ઇન્ટેલિજન્સ રેન્કિંગમાં તે 54 ચેટ Models માંથી 27 નંબર પર છે, જે સ્વતંત્ર બેન્ચમાર્ક સ્કોર પર આધારિત છે. તેના તમામ સ્કોર ઉપર બેન્ચમાર્ક ટેબલમાં જોઈ શકો છો.
ડીપસીક V4 Flash Vision નો ઉપયોગ કરવાનો ખર્ચ કેટલો છે?
વપરાશ ખર્ચ ઇનપુટ માટે ₹42.24 પ્રતિ 10 લાખ Tokens અને આઉટપુટ માટે ₹126.72 પ્રતિ 10 લાખ Tokens છે, જેમાં 0% માર્કઅપ છે. કોઈ સબ્સ્ક્રિપ્શન નથી; વપરાશ મુજબ જ પેમેન્ટ કરો.
ડીપસીક V4 Flash Vision નું અમેરિકન ડોલરમાં API પ્રાઇસિંગ શું છે?
પ્રોવાઇડર ઇનપુટ માટે $0.44 પ્રતિ 10 લાખ Tokens અને આઉટપુટ માટે $1.32 પ્રતિ 10 લાખ Tokens ચાર્જ કરે છે. રૂપિયાના દરો $1 = ₹96 ના આધારે ગણવામાં આવ્યા છે.
ડીપસીક V4 Flash Vision કેટલી લાંબી વાતચીત યાદ રાખી શકે છે?
તેની Context વિન્ડો 10.5 lakh Tokens ની છે. એટલે કે તે એક જ વિનંતીમાં અગાઉની વાતચીત અને જોડેલી ફાઇલો સહિત આટલું લખાણ વાંચી શકે છે.
શું ડીપસીક V4 Flash Vision વાપરવું પૈસા વસૂલ છે?
વેલ્યૂ રેન્કિંગમાં તે 54 માંથી 6 નંબર પર છે. આ રેન્કિંગ બેન્ચમાર્ક ક્ષમતા અને વાસ્તવિક Token ખર્ચની તુલના કરે છે.
શું ડીપસીક V4 Flash Vision જવાબ આપતાં પહેલાં વિચારે છે?
ડિફૉલ્ટ રૂપે ચાલુ. જ્યાં Reasoning સપોર્ટ ઉપલબ્ધ છે, ત્યાં તમે મેસેજ કમ્પોઝરમાં વિચારવાની ક્ષમતા (Thinking effort) જાતે સેટ કરી શકો છો.
DeepSeek ના અન્ય Models
કેટલોગ અપડેટ: 9 સપ્ટે, 2026