GPT-6 Luna
Released
GPT-6 Luna is an AI model released by OpenAI on Tuesday, Sep 22 2026, the same day as GPT-6 Sol. It has a 1.05M token context window. Benchmark results (shown below) cover DeepSWE 1.1.
Get OpenAI releases and AI news.
One email, every other Monday.
API pricing
- InputWhat you pay for everything you send the model — your question, plus any documents or earlier conversation you include with it.
- $0.10
- Cached inputA reduced rate for text you send over and over. If every request starts with the same instructions or the same document, the provider keeps a copy ready and charges less to read it again.
- $0.01
- Cache writeA one-off charge for storing text so later requests can reuse it at the cheaper cached rate.
- $0.125
- OutputWhat you pay for the text the model writes back. It is normally the dearer half: producing an answer costs more than reading one.
- $0.50
- InputWhat you pay for everything you send the model — your question, plus any documents or earlier conversation you include with it.
- $0.20
- Cached inputA reduced rate for text you send over and over. If every request starts with the same instructions or the same document, the provider keeps a copy ready and charges less to read it again.
- $0.02
- Cache writeA one-off charge for storing text so later requests can reuse it at the cheaper cached rate.
- $0.25
- OutputWhat you pay for the text the model writes back. It is normally the dearer half: producing an answer costs more than reading one.
- $0.75
- Reasoning or thinking is supported.
- OpenAI also publishes a Fast mode for this model at exactly twice the standard rate in every column — $0.20 per million input tokens and $1.00 per million output at short context.
Benchmarks
Compare GPT-6 Luna with
Suggested comparisons
GPT-6 LunavsGPT-6 SolGPT-6 LunavsClaude Opus 5.5GPT-6 LunavsGemini 3.8 FlashGPT-6 LunavsMuse Spark 1.3GPT-6 LunavsGrok 4.7GPT-6 LunavsDeepSeek-V4.1-FlashGPT-6 LunavsMistral Medium 3.5GPT-6 LunavsKimi K3GPT-6 LunavsGLM-5.3-FlashGPT-6 LunavsQwen3.8-Max-0902GPT-6 LunavsNemotron 3.5 LightningAbout
GPT-6 Luna, released September 22, 2026, was the small model of OpenAI's GPT-6 generation, launched alongside GPT-6 Sol nineteen days after GPT-6 Astra opened the line. OpenAI's API documentation called it "our most efficient model for focused, high-volume tasks", and the launch post placed the pair as the cost-efficiency end of a family Astra still led. It launched at $0.10 per million input tokens and $0.50 per million output — down from the $0.20 and $1.20 of GPT-5.6 Luna's promotional pricing, a cut OpenAI put at 50%, and a twentieth of Sol's rate in both columns — with cached input at $0.01, cache writes at $0.125, and the family's long-context rule past 272K input tokens: double on input and cache, 1.5x on output.
The launch figures were cost-per-task arguments. At max effort it scored 66.6% on DeepSWE v1.1, which OpenAI set beside Claude Opus 5 and Claude Fable 5 at medium effort at 93% and 96% less per task; at high effort it improved on GPT-5.6 Luna by 5.4 points on AutomationBench at 58% lower cost per task, and on the offline set of OSWorld 2.0 it passed GPT-5.6 Sol at medium effort at a tenth of the cost. On OpenAI's internal factuality evaluation it matched GPT-5.6 Sol at the higher effort levels at about a hundredth of the cost.
For a model sold on price it gave up little of the family's surface: the same 1,050,000-token context window as Sol and Astra, 128,000 max output tokens, text and image input, text output, and the full reasoning-effort range from none to max with medium as the default, on the Responses, Chat Completions and Batch endpoints. Its knowledge cutoff was May 18, 2026, four weeks later than Sol's. Luna was also the one GPT-6 model that reached free users at launch: Free and Go users got it in the desktop app, while Plus, Pro, Business, Enterprise and Edu users had both models in ChatGPT Work and Codex, and the API served it as gpt-6-luna.
Frequently asked questions
GPT-6 Luna was released by OpenAI on Tuesday, Sep 22 2026.