GPT 5.6 സോൾ
പുതിയത്Flagship of the GPT-5.6 series; strongest at multi-step coding and agents.
റിലീസ് തീയതി: 2026 ജൂലൈ 9
₹3,168.00
10 ലക്ഷം ഔട്ട്പുട്ട് Tokens-ന്
ഇൻപുട്ട്: 10 ലക്ഷം Tokens-ന് ₹528.00
0% മാർക്ക്അപ്പിൽ, ഒരു ഡോളറിന് ₹96 എന്ന നിരക്കിലാണ് ഈടാക്കുന്നത്.
സവിശേഷതകൾ
- Context window
- 10,50,000 Tokens
- പരമാവധി Output
- 1,28,000 Tokens
- സ്വീകരിക്കുന്നത്
- ടെക്സ്റ്റ്, ചിത്രങ്ങൾ, PDF ഫയലുകൾ
- Reasoning
- ഡിഫോൾട്ടായി ഓൺ ആണ്
- Effort ലെവലുകൾ
- none, low, medium, high, xhigh, max
- ടൂൾ ഉപയോഗം
- ഉണ്ട്
- സ്ട്രക്ചേർഡ് Output
- ഉണ്ട്
- കോഡ് എക്സിക്യൂഷൻ
- ഇല്ല
- Knowledge cutoff
- 2026-02-16
- ഇന്റലിജൻസ് റാങ്ക്
- 54-ൽ #6
- വാല്യു റാങ്ക്
- 54-ൽ #27
നിരക്കുകൾ
ബെഞ്ച്മാർക്കുകൾ
റേറ്റിംഗ് എന്ന് രേഖപ്പെടുത്താത്ത സ്കോറുകൾ ശതമാനത്തിലാണ്. എല്ലാ ബെഞ്ച്മാർക്കുകളും സ്വതന്ത്രമായി പരിശോധിച്ചവയാണ്.
94.1%
GPQA Diamond
GPQA Diamond - graduate-level science Q&A
49.5%
HLE
Humanity's Last Exam
57.1%
SciCode
SciCode - scientific code generation
72.7%
IFBench
IFBench - precise instruction following
84.0%
Long Context
Long Context Reasoning - reasoning over long inputs
65.9%
Terminal-Bench Hard
Terminal-Bench Hard - agentic terminal tasks
88.0%
Terminal-Bench 2
Terminal-Bench 2.1 - agentic terminal tasks, second edition
47.5%
FrontierCode
FrontierCode - long-horizon production coding tasks
92.5%
ARC-AGI-2
ARC-AGI-2 - abstract reasoning on novel puzzles
62.7%
OSWorld 2
OSWorld 2 - agentic computer use (partial credit)
1619
WebDev Arena
WebDev Arena - head-to-head web-app builds, Elo rating
GPT 5.6 സോൾ-നെ കുറിച്ച്
OpenAI's frontier model for complex professional work and the flagship of the GPT-5.6 family, corresponding to the unsuffixed tier of earlier GPT-5 generations. It is built for demanding reasoning, coding and agentic workflows, and is particularly strong at command-line and multi-step coding tasks. Reasoning effort spans none through max, so one model covers a quick lookup and a long deliberate build.
Sol is the top of a three-tier family that OpenAI shipped together after a limited preview, and the naming is deliberate: the number is the generation, while Sol, Terra and Luna are durable capability tiers that the lab says can advance on their own cadence. Sol is the tier it puts in front of work that has to be right - long professional tasks, real codebases, agents that keep going without a person re-steering them every few steps.
The claim OpenAI leads with is not a peak score, it is work per token. On Agents' Last Exam, an evaluation of long-running professional workflows across 55 fields, the lab reports Sol setting a new high of 53.6, ahead of the strongest competing model by 13.1 points, and says that even at medium reasoning effort it stays ahead by 11.4 points at roughly a quarter of the estimated cost per task. Its published table puts Sol at 52.7% there against 46.9% for GPT-5.5. The same argument runs through the coding results: OpenAI reports new state-of-the-art scores on Terminal-Bench 2.1 at 88.8% and on DeepSWE at 72.7%, with 64.6% on SWE-Bench Pro, and says it reaches them while spending less than half the output tokens and less than half the time of the model it displaces.
The effort dial is the feature that makes those numbers readable. Reasoning effort runs from none through low, medium, high and xhigh to max, and OpenAI describes max as giving the model more time than xhigh to reason, explore alternatives, run checks and revise its approach. Above that sits ultra, which coordinates four agents in parallel by default and trades higher token use for both stronger results and faster time to an answer. Programmatic Tool Calling in the Responses API lets Sol write and run small programs that coordinate tools and filter intermediate data, so tool-heavy work advances with fewer round trips through the model.
Away from code, OpenAI reports 90.4% on BrowseComp and 62.6% on OSWorld 2.0, the latter with 85% fewer output tokens than the model it beats, plus 70.6% on BenchCAD and 83.4% with a Python tool available. It also puts weight on design judgement: stronger computer use lets the model inspect the rendered result rather than only emit the code behind it, catch visual and functional problems, and finish the work before handing it back. On documents, decks and spreadsheets the lab says the biggest gain is in following a supplied template or reference file faithfully, including rules carried in a slide master.
OpenAI is specific about where the model stops. It reports GPT-5.6 as more capable than its earlier models in both biology and cybersecurity while crossing the Critical threshold in neither, and says its testing shows the model is better at finding and fixing vulnerabilities than at reliably carrying out autonomous end-to-end attacks against hardened targets, and that in biology it can support legitimate research without providing the end-to-end capability needed to create a dangerous novel threat. It also warns about the cost of that caution: Sol's cyber safeguards block roughly ten times more potentially harmful activity than previous models, which creates friction for benign work, so ChatGPT and Codex offer a one-press retry on a lower-capability model.
ലോഞ്ചിംഗ് വേളയിൽ OpenAI വ്യക്തമാക്കിയത്
- Results per dollar
- OpenAI presents Sol as an efficiency argument rather than a peak score: a new high of 53.6 on Agents' Last Exam, and 11.4 points ahead of the next model even at medium effort for roughly a quarter of the estimated cost per task.
- Coding and the terminal
- The lab reports new state-of-the-art results on Terminal-Bench 2.1 at 88.8% and DeepSWE at 72.7%, plus 64.6% on SWE-Bench Pro, reached with less than half the output tokens of the model it displaces.
- From none to ultra
- Reasoning effort spans none through max, where OpenAI says the model gets more time than xhigh to explore alternatives and check itself. Ultra goes further again, coordinating four agents in parallel by default.
- Browsing and computer use
- OpenAI reports 90.4% on BrowseComp and 62.6% on OSWorld 2.0, the latter using 85% fewer output tokens than the model it beats, and 83.4% on BenchCAD with a Python tool available.
- Work it can finish
- Stronger computer use lets Sol inspect the rendered result instead of only the code behind it. On decks and spreadsheets the lab reports the clearest gain in following a supplied template or reference file, including rules held in a slide master.
- What it is not for
- OpenAI reports GPT-5.6 as better at finding and fixing vulnerabilities than at carrying out autonomous end-to-end attacks, and as unable to provide end-to-end capability for creating a dangerous novel biological threat. Its cyber safeguards block roughly ten times more activity than earlier models, so legitimate security work can be refused.
ഇന്ത്യൻ ഭാഷകൾ
GPT 5.6 സോൾ 15 ഇന്ത്യൻ ഭാഷകളിൽ മറുപടി നൽകും. മെസ്സേജ് ബോക്സിന് അടുത്തുള്ള മെനുവിൽ നിന്ന് ആവശ്യമുള്ള ഭാഷ തിരഞ്ഞെടുക്കാം.
Frequently Asked Questions
GPT 5.6 സോൾ സംബന്ധിച്ച പ്രധാന ചോദ്യങ്ങളും ഉത്തരങ്ങളും.
GPT 5.6 സോൾ എപ്പോഴാണ് റിലീസ് ചെയ്തത്?
OpenAI 2026 ജൂലൈ 9-ൽ GPT 5.6 സോൾ പുറത്തിറക്കി.
GPT 5.6 സോൾ വികസിപ്പിച്ചത് ആരാണ്?
GPT 5.6 സോൾ വികസിപ്പിച്ചത് OpenAI ആണ്. 99Models AI ഒറിജിനൽ പ്രൊവൈഡർ നിരക്കിൽ തന്നെ ഇതിലേക്ക് കണക്ട് ചെയ്യുന്നു.
GPT 5.6 സോൾ എത്രത്തോളം കാര്യക്ഷമമാണ്?
ഇന്റലിജൻസ് റാങ്കിംഗിൽ 54 മോഡലുകളിൽ 6-ാം സ്ഥാനത്താണ് ഇത്. ബെഞ്ച്മാർക്ക് സ്കോറുകൾ മുകളിലുള്ള പാനലിൽ കാണാം.
GPT 5.6 സോൾ ഉപയോഗിക്കാൻ എത്ര ചെലവാകും?
10 ലക്ഷം ഇൻപുട്ട് Tokens-ന് ₹528.00 രൂപയും ഔട്ട്പുട്ടിന് ₹3,168.00 രൂപയുമാണ് അധിക മാർക്ക്അപ്പില്ലാത്ത നിരക്ക്. സബ്സ്ക്രിപ്ഷനില്ല, ഉപയോഗിക്കുന്നതിന് മാത്രം പണമടയ്ക്കുക.
GPT 5.6 സോൾ API-യുടെ ഡോളർ നിരക്ക് എത്രയാണ്?
10 ലക്ഷം ഇൻപുട്ട് Tokens-ന് $5.50 ഡോളറും 10 ലക്ഷം ഔട്ട്പുട്ട് Tokens-ന് $33.00 ഡോളറുമാണ് നിരക്ക്. ഈ പേജിലെ രൂപ നിരക്കുകൾ ഡോളറിന് ₹96 എന്ന നിരക്കിൽ മാറ്റിയതാണ്.
GPT 5.6 സോൾ-ന് എത്ര നീളമുള്ള സംഭാഷണം ഓർത്തുനിൽക്കാനാകും?
ഇതിന്റെ Context window 10.5 lakh Tokens ആണ്. ഒരുമിച്ച് നൽകുന്ന സംഭാഷണങ്ങളും ഫയലുകളും ഉൾപ്പെടെ ഇതിൽ വായിക്കാൻ സാധിക്കും.
GPT 5.6 സോൾ മികച്ച വാല്യൂ നൽകുന്ന ഒന്നാണോ?
പെർഫോമൻസും നിരക്കും അടിസ്ഥാനമാക്കിയുള്ള വാല്യൂ റാങ്കിംഗിൽ 54 മോഡലുകളിൽ 27-ാം സ്ഥാനത്താണ് ഇത്.
മറുപടി നൽകുന്നതിന് മുൻപ് GPT 5.6 സോൾ Reasoning നടത്തുമോ?
ഡിഫോൾട്ടായി ഓൺ ആണ്. Reasoning പിന്തുണയ്ക്കുന്ന മോഡലുകളിൽ ചിന്തിക്കുന്നതിന്റെ വ്യാപ്തി മെസ്സേജ് ബോക്സിൽ ക്രമീകരിക്കാം.
GPT 5.6 സോൾ ഏതൊക്കെ ഇന്ത്യൻ ഭാഷകളിൽ മറുപടി നൽകും?
15 ഇന്ത്യൻ ഭാഷകളിൽ ഇത് മറുപടി നൽകും. മെസ്സേജ് ബോക്സിന് അടുത്തുള്ള മെനുവിൽ നിന്ന് ഭാഷ തിരഞ്ഞെടുക്കാം.
OpenAI-ൽ നിന്നുള്ള മറ്റ് Models
കാറ്റലോഗ് പുതുക്കിയത്: 2026 സെപ്റ്റം 9