Back to Models
OpenAI: GPT-5.6 Luna
Openaitext9.1 / 10 Overall Rating
Developed by OpenAI, GPT-5.6 Luna is a high-speed, cost-optimized model tailored for low-latency and high-volume operations. It provides a 1,050,000-token context window and a 128,000 max output token limit at $0.20 per million input tokens. It targets high-throughput classification, routing, and lightweight agentic tasks.
Context Window
1050K
Knowledge Cutoff
2025
Max Output
16K
License
Proprietary
Where it Excels (Pros)
- Extremely low API pricing
- 1.05M token context window
- 128K max output capacity
- Fast latency for high-volume tasks
Limitations (Cons)
- Weaker reasoning than flagship models
- Text-only output generation
Benchmark Breakdown
Reasoning (MMLU)82%
Coding (HumanEval)78%
Mathematics (MATH)76%
Long-Context Retrieval95%
Pricing Matrix
Input cost / 1M tokens$0.20
Output cost / 1M tokens$1.20
Compare OpenAI: GPT-5.6 Luna with other frontier models:
Related Skills, Tools & Automations
Tool
You.com
Data & Analytics
Tool
Replit
Code Assistant
Tool
v0
Code Assistant
Tool
Lovable
Code Assistant
AutomationMake
AI Customer Support Bot
Route customer inquiries to AI for instant responses, with human escalation for complex issues.
Automationn8n
Content Repurposer
Transform long-form articles into Twitter threads, LinkedIn posts, and newsletter snippets.