GPT-5.6 Luna
Budgetby OpenAI
GPT-5.6 Luna is the fastest and cheapest model in the GPT-5.6 family, previewed June 26, 2026. It launched at $1 per million input tokens and $6 per million output, and OpenAI cut it 80 percent on July 30 to $0.20 input, $0.02 cached, and $1.20 output, with a 1 million token context window. Like its siblings, Luna entered as a limited preview to roughly 20 pre-approved organizations under a US Government arrangement, and reached general availability on July 9, 2026. OpenAI is positioning it for high-volume, latency-sensitive workloads where Sol-level reasoning is not required. Its successor GPT-6 Luna shipped on September 22, 2026 at $0.10/$0.50.
Input Price
$0.20
per 1M tokens
Output Price
$1.20
per 1M tokens
Context Window
1M
tokens
Released
2026-07
API access
Capabilities
Key Strengths
- ✓Lowest pricing in the GPT-5.6 family ($0.20/$1.20 per 1M)
- ✓1M token context window
- ✓Frontier-family lineage at budget-tier economics
- ✓Reasoning, vision, tool use, and code included
Best For
- ▸High-volume chat and classification
- ▸Cost-sensitive coding assistants
- ▸Real-time tool-augmented workflows
- ▸Latency-sensitive customer-facing applications
Benchmark Scores
| Benchmark | Score | Description |
|---|---|---|
| SWE-bench | 93.0 | Real-world software engineering tasks from GitHub issues (SWE-bench Verified) |
| MMLU-Pro | 86.0 | General knowledge and reasoning across 57 subjects |
| GPQA Diamond | 92.3 | Graduate-level science questions verified by domain experts |
| OSWorld 2.0 | 45.6 | Computer use across real desktop applications and multi-step GUI tasks |
| BrowseComp | 83.3 | Agentic web search and browsing over hard-to-find facts |
| FrontierCode v1.1 | 39.8 | Agentic coding on frontier software engineering tasks (Main split) |
| Terminal-Bench 4.0 | 4.5 | Long-horizon agentic work in a terminal across software, science, ML, operations, hardware, security, and media (66 tasks, all-or-nothing verifiers) |
Scores sourced from public benchmark datasets. See full benchmark leaderboard for all models.
Pricing Details
Input tokens
$0.20
per 1M tokens
Output tokens
$1.20
per 1M tokens
Estimated cost per 1K requests
$0.80
~1K input + ~500 output tokens avg
Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.
Related Models
GPT-6 Astra
FlagshipOpenAI
$10.00 in / $50.00 out
GPT-6 Sol
Mid-tierOpenAI
$2.00 in / $10.00 out
GPT-6 Luna
BudgetOpenAI
$0.10 in / $0.50 out
GPT-5.6 Sol
FlagshipOpenAI
$4.00 in / $20.00 out
GPT-5.6 Terra
Mid-tierOpenAI
$2.00 in / $12.00 out
GPT-5.5
FlagshipOpenAI
$5.00 in / $30.00 out
GPT-4o
FlagshipOpenAI
$2.50 in / $10.00 out
GPT-4o-mini
BudgetOpenAI
$0.15 in / $0.60 out
o1
FlagshipOpenAI
$15.00 in / $60.00 out
o3-mini
Mid-tierOpenAI
$1.10 in / $4.40 out
Claude Haiku 4.5
BudgetAnthropic
$1.00 in / $5.00 out
Gemini 3.1 Flash-Lite
Budget$0.25 in / $1.50 out
Gemini 3.5 Flash-Lite
Budget$0.30 in / $2.50 out
Mistral Small
BudgetMistral
$0.15 in / $0.60 out