99MODELS

GPT 5.6 Luna

New

Fast, low-cost GPT-5.6 tier for chat and high-volume tasks.

Released Jul 9, 2026

126.72

per 10 lakh output tokens

Input: ₹21.12 per 10 lakh tokens

Billed at provider rates converted at ₹96 per US dollar, with 0% markup.

Specifications

Context window
10,50,000 tokens
Max output
1,28,000 tokens
Accepts
Text, Images, PDF files
Reasoning
On by default
Effort levels
none, low, medium, high, xhigh, max
Tool use
Yes
Structured output
Yes
Code execution
No
Knowledge cutoff
2026-02-16
Intelligence rank
#22 of 54
Value rank
#3 of 54

Pricing

Pricing
Per 10 lakh tokensINRUSD
Input21.12$0.22
Output126.72$1.32
Cached input2.11$0.02

Benchmarks

Scores are percentages unless marked as a rating. All benchmarks are measured independently.

  • 91.1%

    GPQA Diamond

    GPQA Diamond - graduate-level science Q&A

  • 39.5%

    HLE

    Humanity's Last Exam

  • 53.6%

    SciCode

    SciCode - scientific code generation

  • 83.7%

    Long Context

    Long Context Reasoning - reasoning over long inputs

  • 80.9%

    Terminal-Bench 2

    Terminal-Bench 2.1 - agentic terminal tasks, second edition

  • 39.8%

    FrontierCode

    FrontierCode - long-horizon production coding tasks

  • 59.5%

    ARC-AGI-2

    ARC-AGI-2 - abstract reasoning on novel puzzles

  • 1518

    WebDev Arena

    WebDev Arena - head-to-head web-app builds, Elo rating

About GPT 5.6 Luna

The GPT-5.6 model optimised for cost-sensitive, high-volume work, corresponding to the nano tier of earlier generations. OpenAI describes it as the fastest and most affordable model in the series, bringing strong capability at the lowest cost. It still reasons when asked, so quality degrades gracefully rather than falling off a cliff on harder inputs.

Luna is the cheapest and fastest tier in the family, and three weeks after launch OpenAI cut its price by 80%, to twenty cents per million input tokens and $1.20 per million output. The lab's framing for that move is worth reading twice: it says Luna delivers performance comparable to models that were frontier-class a year earlier at roughly six cents on the dollar per task and nearly nine times the speed, and that on its long-horizon professional evaluation Luna beats the strongest competing model at an estimated cost per task nearly 99% lower.

That is not a stripped-down completion engine. OpenAI reports Luna at 50.3% on Agents' Last Exam, 84.7% on Terminal-Bench 2.1, 62.7% on SWE-Bench Pro and 92.3% on GPQA Diamond, and says that on its coding-agent index Luna comes out ahead of the previous generation's flagship competitor. It reasons when asked, uses tools and completes multi-step workflows, which is what makes high-volume automation practical rather than merely cheap.

The workflow OpenAI itself sketches is the one to copy. Use the flagship to resolve uncertainty and settle the plan, then hand Luna the well-specified parts: implement the change, write and run the tests, evaluate the result. Deciding what to build is where frontier intelligence earns its rate; carrying out a decision that has already been made is where Luna's rate wins by two orders of magnitude.

The limits are specific and they are about capacity, not effort. On the lab's multi-round retrieval evaluation Luna scores 41.3% at both 256K to 512K and 512K to 1M tokens, against the flagship's 91.5% and 73.8% - so the million-token window is there, but do not rely on Luna to find a needle deep inside it. Abstract reasoning is the other cliff: 0.18% on the lab's newest novel-puzzle evaluation, against 7.78% for the flagship. Cybersecurity and self-improvement results fall off in the same way. Luna is for volume and for tasks whose shape you already know.

What OpenAI announced at launch

Eighty per cent cheaper
Three weeks after launch OpenAI cut Luna to twenty cents per million input tokens and $1.20 per million output, and describes it as the fastest and most affordable model in the family.
Frontier-class, a year late
OpenAI says Luna matches models that were frontier-class a year earlier at roughly six cents on the dollar per task and nearly nine times the speed.
Still a capable agent
The lab reports 50.3% on Agents' Last Exam, 84.7% on Terminal-Bench 2.1, 62.7% on SWE-Bench Pro and 92.3% on GPQA Diamond, with tool use and multi-step workflows intact.
Plan high, execute cheap
OpenAI suggests using the flagship to resolve uncertainty and define the plan, then Luna to implement well-specified changes, write and run tests, and evaluate the results.
What it is not for
Deep retrieval and genuinely novel reasoning are the cliffs. The lab reports 41.3% on multi-round retrieval at every depth past 256K against the flagship's 91.5%, and 0.18% on its newest novel-puzzle evaluation against 7.78%.

Indian languages

GPT 5.6 Luna answers in 13 Indian languages. Choose a language from the menu beside the message box to receive replies in it.

Frequently Asked Questions

Frequently asked questions about GPT 5.6 Luna.

When was GPT 5.6 Luna released?

OpenAI released GPT 5.6 Luna on Jul 9, 2026.

Who built GPT 5.6 Luna?

GPT 5.6 Luna is developed by OpenAI. 99Models AI connects directly to it at the provider's published rate.

How intelligent is GPT 5.6 Luna?

It is ranked 22 of 54 chat models on intelligence, ordered by independent benchmark scores. View its complete scores in the Benchmarks table above.

How much does GPT 5.6 Luna cost?

Usage costs ₹21.12 per million input tokens and ₹126.72 per million output tokens, with 0% markup. There is no subscription; you pay only for what you use.

What is GPT 5.6 Luna pricing in US dollars?

The provider charges $0.22 per million input tokens and $1.32 per million output tokens. Rupee rates are converted at ₹96 per US dollar.

How long a conversation can GPT 5.6 Luna hold?

Its context window is 10.5 lakh tokens. That is the total volume of text and attached files it can process in a single request.

How does GPT 5.6 Luna rank for value?

It ranks 3 out of 54 models for value. This ranking weighs benchmark intelligence against the actual token cost.

Does GPT 5.6 Luna support reasoning?

On by default. Where reasoning is supported, you can adjust the thinking effort level directly in the message composer.

Which Indian languages does GPT 5.6 Luna support?

It answers in 13 Indian languages. Select your preferred language from the menu beside the message box to receive replies in it.

More models from OpenAI

Catalog updated Sep 9, 2026