डीपसीक V4 फ्लैश 0731
नयाExtremely cheap MoE reasoner; one of the most-used models in production.
रिलीज़: 31 जुल॰ 2026
₹126.72
प्रति 10 लाख आउटपुट Tokens
इनपुट: ₹42.24 प्रति 10 लाख Tokens
प्रदाता की आधिकारिक दर पर बिलिंग, ₹96 प्रति डॉलर पर कनवर्ट, 0% मार्कअप के साथ।
स्पेसिफिकेशन्स
- Context विंडो
- 13,10,720 Tokens
- अधिकतम आउटपुट
- 9,43,718 Tokens
- स्वीकार्य इनपुट
- टेक्स्ट
- Reasoning
- डिफ़ॉल्ट रूप से चालू
- प्रयास स्तर
- low, high, max
- टूल का उपयोग
- हाँ
- स्ट्रक्चर्ड आउटपुट
- हाँ
- कोड एग्जीक्यूशन
- नहीं
- इंटेलिजेंस रैंक
- 54 में से #34
- वैल्यू रैंक
- 54 में से #15
दरें
Benchmarks
स्कोर प्रतिशत में हैं जब तक कि कोई रेटिंग न दी गई हो। सभी Benchmark स्वतंत्र रूप से मापे गए हैं।
85.2%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
18.8%
FrontierCode
FrontierCode - long-horizon production coding tasks
61.4%
ARC-AGI-2
ARC-AGI-2 - abstract reasoning on novel puzzles
1579
WebDev Arena
WebDev Arena - head-to-head web-app builds, Elo rating
डीपसीक V4 फ्लैश 0731 के बारे में
The lighter, high-throughput tier of the DeepSeek V4 pair: a Mixture-of-Experts model with 284B total and 13B activated parameters and a one-million-token context. It uses a hybrid attention architecture combining Compressed Sparse Attention and Heavily Compressed Attention to make long context efficient. This re-post-trained revision is tuned for coding, reasoning and agent workflows, and offers non-think plus high and max thinking settings.
This is the checkpoint that took DeepSeek-V4-Flash out of preview. DeepSeek is unusually precise about what changed: the architecture and the parameter count are identical to the preview weights, and the entire difference is a fresh round of post-training aimed at agent work. Nothing about the model got bigger, and the price did not move; what moved is how reliably it finishes a task once a harness hands it tools.
The gain the lab reports is large enough that it reframes the tier. DeepSeek measures the 0731 revision at 82.7 on Terminal Bench 2.1, 54.2 on NL2Repo, 76.7 on Cybergym, 54.4 on DeepSWE, 70.3 on Toolathlon-Verified, 25.2 on Agents' Last Exam and 25.1 on the public AutomationBench split, plus 68.7 and 59.6 on DSBench-FullStack and DSBench-Hard, its own internal full-stack and hard-coding-agent sets. Its point is the comparison against its own previous flagship rather than against anyone else: on every one of those the cheap Flash checkpoint clears what V4-Pro-Preview scored. The lab is explicit that these were run through DeepSeek Harness in minimal mode at the max effort level, with top-p 0.95 and temperature 1.0, so they are peak-effort numbers, not defaults.
The release also made the model easier to drop into an existing agent. DeepSeek added native support for the OpenAI Responses API format alongside its OpenAI- and Anthropic-compatible chat endpoints, with a configuration path aimed specifically at Codex, and the calling name stayed `deepseek-v4-flash` so nothing downstream had to be rewritten. A fortnight later the lab exposed a low / high / max thinking-effort control across both V4 models, and recommends high for everyday agent loops and max only for genuinely hard scenarios.
Where it stops is a matter of tier rather than defect. DeepSeek shipped this update to the Flash API alone - the Pro API and the app and web experiences were untouched that day - and it published the Pro GA two weeks later with materially higher scores on the same evaluations. Flash is the model to reach for when throughput and cost dominate; the lab points at Pro for the hardest reasoning and the longest agent runs.
लॉन्च के समय DeepSeek ने क्या कहा
- Same weights, new post-training
- DeepSeek states that 0731 keeps the architecture and size of the preview release exactly and was only re-post-trained, so the improvement is entirely a post-training result rather than a bigger model.
- Agent scores past the old flagship
- The lab measures 82.7 on Terminal Bench 2.1, 54.2 on NL2Repo, 76.7 on Cybergym and 54.4 on DeepSWE, and notes these clear what V4-Pro-Preview scored on the same set.
- Tool-heavy workloads
- DeepSeek reports 70.3 on Toolathlon-Verified and 25.1 on the public AutomationBench split, alongside 68.7 and 59.6 on its internal DSBench full-stack and hard-problem sets.
- Drops into existing harnesses
- The release added native OpenAI Responses API support with a Codex-oriented setup path, and the API model name was left unchanged so callers pick up the revision without a code change.
- Peak-effort measurement
- Every code-agent figure the lab published was produced through DeepSeek Harness in minimal mode at max effort with top-p 0.95 and temperature 1.0, so it is the ceiling rather than the default-setting result.
- What it is not for
- This update covered the Flash API only, leaving the Pro API and the app and web models unchanged, and DeepSeek released V4-Pro two weeks later with higher scores on the same evaluations for harder work.
भारतीय भाषाएं
डीपसीक V4 फ्लैश 0731 12 भारतीय भाषाओं में जवाब दे सकता है। मैसेज बॉक्स के पास वाले मेन्यू से भाषा चुनें और उसी में जवाब पाएं।
Frequently Asked Questions
डीपसीक V4 फ्लैश 0731 के बारे में अक्सर पूछे जाने वाले सवाल।
डीपसीक V4 फ्लैश 0731 कब रिलीज़ हुआ था?
DeepSeek ने डीपसीक V4 फ्लैश 0731 को 31 जुल॰ 2026 को रिलीज़ किया था।
डीपसीक V4 फ्लैश 0731 को किसने बनाया है?
डीपसीक V4 फ्लैश 0731 को AI लैब DeepSeek ने बनाया है। 99Models इसे सीधे प्रदाता की आधिकारिक दर पर उपलब्ध कराता है।
डीपसीक V4 फ्लैश 0731 कितना समझदार है?
इंटेलीजेंस रैंकिंग में यह 54 चैट Models में से 34 स्थान पर है, जो बेंचमार्क स्कोर पर आधारित है। इसके पूरे स्कोर ऊपर Benchmarks पैनल में देखे जा सकते हैं।
डीपसीक V4 फ्लैश 0731 का उपयोग करने का क्या खर्च है?
10 लाख इनपुट Tokens के लिए ₹42.24 और 10 लाख आउटपुट Tokens के लिए ₹126.72, बिना किसी अतिरिक्त मार्कअप के प्रदाता की दर पर। कोई सब्सक्रिप्शन नहीं है; आप केवल अपने उपयोग का भुगतान करते हैं।
डॉलर में डीपसीक V4 फ्लैश 0731 API की कीमत क्या है?
प्रदाता 10 लाख इनपुट Tokens के लिए $0.44 और 10 लाख आउटपुट Tokens के लिए $1.32 चार्ज करता है। इस पेज पर रुपये की दरें ₹96 प्रति डॉलर के हिसाब से बदली गई हैं।
डीपसीक V4 फ्लैश 0731 कितनी लंबी बातचीत याद रख सकता है?
इसकी Context विंडो 13.1 lakh Tokens है। यानी यह एक रिक्वेस्ट में पिछली बातचीत और अटैच की गई फ़ाइलों को मिलाकर इतना टेक्स्ट पढ़ सकता है।
क्या डीपसीक V4 फ्लैश 0731 पैसे के लिहाज से किफ़ायती है?
वैल्यू रैंकिंग में यह 54 Models में से 15 स्थान पर है। यह रैंकिंग परफ़ॉर्मेंस और टोकन की वास्तविक कीमत की तुलना करके तय की जाती है।
क्या डीपसीक V4 फ्लैश 0731 जवाब देने से पहले सोच-विचार (Reasoning) करता है?
डिफ़ॉल्ट रूप से चालू। जहां सोचने की क्षमता उपलब्ध है, वहां आप मैसेज कंपोज़र में सीधे Reasoning का स्तर तय कर सकते हैं।
डीपसीक V4 फ्लैश 0731 किन भारतीय भाषाओं में जवाब दे सकता है?
यह 12 भारतीय भाषाओं में जवाब देता है। मैसेज बॉक्स के पास वाले मेन्यू से भाषा चुनें और उसी भाषा में जवाब पाएं।
DeepSeek के अन्य Models
कैटलॉग अपडेट: 9 सित॰ 2026