99MODELS

GPT 5.6 लूना

नया

Fast, low-cost GPT-5.6 tier for chat and high-volume tasks.

रिलीज़: 9 जुल॰ 2026

126.72

प्रति 10 लाख आउटपुट Tokens

इनपुट: ₹21.12 प्रति 10 लाख Tokens

प्रदाता की आधिकारिक दर पर बिलिंग, ₹96 प्रति डॉलर पर कनवर्ट, 0% मार्कअप के साथ।

स्पेसिफिकेशन्स

Context विंडो
10,50,000 Tokens
अधिकतम आउटपुट
1,28,000 Tokens
स्वीकार्य इनपुट
टेक्स्ट, इमेज, PDF फ़ाइलें
Reasoning
डिफ़ॉल्ट रूप से चालू
प्रयास स्तर
none, low, medium, high, xhigh, max
टूल का उपयोग
हाँ
स्ट्रक्चर्ड आउटपुट
हाँ
कोड एग्जीक्यूशन
नहीं
नॉलेज कटऑफ़
2026-02-16
इंटेलिजेंस रैंक
54 में से #22
वैल्यू रैंक
54 में से #3

दरें

दरें
प्रति 10 lakh TokensINRUSD
इनपुट21.12$0.22
आउटपुट126.72$1.32
कैश्ड इनपुट2.11$0.02

Benchmarks

स्कोर प्रतिशत में हैं जब तक कि कोई रेटिंग न दी गई हो। सभी Benchmark स्वतंत्र रूप से मापे गए हैं।

  • 91.1%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 39.5%

    HLE

    Humanity's Last Exam

  • 53.6%

    SciCode

    SciCode - scientific code generation

  • 83.7%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 80.9%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 39.8%

    FrontierCode

    FrontierCode - long-horizon production coding tasks

  • 59.5%

    ARC-AGI-2

    ARC-AGI-2 - abstract reasoning on novel puzzles

  • 1518

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

GPT 5.6 लूना के बारे में

The GPT-5.6 model optimised for cost-sensitive, high-volume work, corresponding to the nano tier of earlier generations. OpenAI describes it as the fastest and most affordable model in the series, bringing strong capability at the lowest cost. It still reasons when asked, so quality degrades gracefully rather than falling off a cliff on harder inputs.

Luna is the cheapest and fastest tier in the family, and three weeks after launch OpenAI cut its price by 80%, to twenty cents per million input tokens and $1.20 per million output. The lab's framing for that move is worth reading twice: it says Luna delivers performance comparable to models that were frontier-class a year earlier at roughly six cents on the dollar per task and nearly nine times the speed, and that on its long-horizon professional evaluation Luna beats the strongest competing model at an estimated cost per task nearly 99% lower.

That is not a stripped-down completion engine. OpenAI reports Luna at 50.3% on Agents' Last Exam, 84.7% on Terminal-Bench 2.1, 62.7% on SWE-Bench Pro and 92.3% on GPQA Diamond, and says that on its coding-agent index Luna comes out ahead of the previous generation's flagship competitor. It reasons when asked, uses tools and completes multi-step workflows, which is what makes high-volume automation practical rather than merely cheap.

The workflow OpenAI itself sketches is the one to copy. Use the flagship to resolve uncertainty and settle the plan, then hand Luna the well-specified parts: implement the change, write and run the tests, evaluate the result. Deciding what to build is where frontier intelligence earns its rate; carrying out a decision that has already been made is where Luna's rate wins by two orders of magnitude.

The limits are specific and they are about capacity, not effort. On the lab's multi-round retrieval evaluation Luna scores 41.3% at both 256K to 512K and 512K to 1M tokens, against the flagship's 91.5% and 73.8% - so the million-token window is there, but do not rely on Luna to find a needle deep inside it. Abstract reasoning is the other cliff: 0.18% on the lab's newest novel-puzzle evaluation, against 7.78% for the flagship. Cybersecurity and self-improvement results fall off in the same way. Luna is for volume and for tasks whose shape you already know.

लॉन्च के समय OpenAI ने क्या कहा

Eighty per cent cheaper
Three weeks after launch OpenAI cut Luna to twenty cents per million input tokens and $1.20 per million output, and describes it as the fastest and most affordable model in the family.
Frontier-class, a year late
OpenAI says Luna matches models that were frontier-class a year earlier at roughly six cents on the dollar per task and nearly nine times the speed.
Still a capable agent
The lab reports 50.3% on Agents' Last Exam, 84.7% on Terminal-Bench 2.1, 62.7% on SWE-Bench Pro and 92.3% on GPQA Diamond, with tool use and multi-step workflows intact.
Plan high, execute cheap
OpenAI suggests using the flagship to resolve uncertainty and define the plan, then Luna to implement well-specified changes, write and run tests, and evaluate the results.
What it is not for
Deep retrieval and genuinely novel reasoning are the cliffs. The lab reports 41.3% on multi-round retrieval at every depth past 256K against the flagship's 91.5%, and 0.18% on its newest novel-puzzle evaluation against 7.78%.

भारतीय भाषाएं

GPT 5.6 लूना 13 भारतीय भाषाओं में जवाब दे सकता है। मैसेज बॉक्स के पास वाले मेन्यू से भाषा चुनें और उसी में जवाब पाएं।

Frequently Asked Questions

GPT 5.6 लूना के बारे में अक्सर पूछे जाने वाले सवाल।

GPT 5.6 लूना कब रिलीज़ हुआ था?

OpenAI ने GPT 5.6 लूना को 9 जुल॰ 2026 को रिलीज़ किया था।

GPT 5.6 लूना को किसने बनाया है?

GPT 5.6 लूना को AI लैब OpenAI ने बनाया है। 99Models इसे सीधे प्रदाता की आधिकारिक दर पर उपलब्ध कराता है।

GPT 5.6 लूना कितना समझदार है?

इंटेलीजेंस रैंकिंग में यह 54 चैट Models में से 22 स्थान पर है, जो बेंचमार्क स्कोर पर आधारित है। इसके पूरे स्कोर ऊपर Benchmarks पैनल में देखे जा सकते हैं।

GPT 5.6 लूना का उपयोग करने का क्या खर्च है?

10 लाख इनपुट Tokens के लिए ₹21.12 और 10 लाख आउटपुट Tokens के लिए ₹126.72, बिना किसी अतिरिक्त मार्कअप के प्रदाता की दर पर। कोई सब्सक्रिप्शन नहीं है; आप केवल अपने उपयोग का भुगतान करते हैं।

डॉलर में GPT 5.6 लूना API की कीमत क्या है?

प्रदाता 10 लाख इनपुट Tokens के लिए $0.22 और 10 लाख आउटपुट Tokens के लिए $1.32 चार्ज करता है। इस पेज पर रुपये की दरें ₹96 प्रति डॉलर के हिसाब से बदली गई हैं।

GPT 5.6 लूना कितनी लंबी बातचीत याद रख सकता है?

इसकी Context विंडो 10.5 lakh Tokens है। यानी यह एक रिक्वेस्ट में पिछली बातचीत और अटैच की गई फ़ाइलों को मिलाकर इतना टेक्स्ट पढ़ सकता है।

क्या GPT 5.6 लूना पैसे के लिहाज से किफ़ायती है?

वैल्यू रैंकिंग में यह 54 Models में से 3 स्थान पर है। यह रैंकिंग परफ़ॉर्मेंस और टोकन की वास्तविक कीमत की तुलना करके तय की जाती है।

क्या GPT 5.6 लूना जवाब देने से पहले सोच-विचार (Reasoning) करता है?

डिफ़ॉल्ट रूप से चालू। जहां सोचने की क्षमता उपलब्ध है, वहां आप मैसेज कंपोज़र में सीधे Reasoning का स्तर तय कर सकते हैं।

GPT 5.6 लूना किन भारतीय भाषाओं में जवाब दे सकता है?

यह 13 भारतीय भाषाओं में जवाब देता है। मैसेज बॉक्स के पास वाले मेन्यू से भाषा चुनें और उसी भाषा में जवाब पाएं।

OpenAI के अन्य Models

कैटलॉग अपडेट: 9 सित॰ 2026