Back to Models

OpenAI: GPT-5.6 Luna

Openaitext
9.1 / 10 Overall Rating

Developed by OpenAI, GPT-5.6 Luna is a high-speed, cost-optimized model tailored for low-latency and high-volume operations. It provides a 1,050,000-token context window and a 128,000 max output token limit at $0.20 per million input tokens. It targets high-throughput classification, routing, and lightweight agentic tasks.

Context Window
1050K
Knowledge Cutoff
2025
Max Output
16K
License
Proprietary

Where it Excels (Pros)

  • Extremely low API pricing
  • 1.05M token context window
  • 128K max output capacity
  • Fast latency for high-volume tasks

Limitations (Cons)

  • Weaker reasoning than flagship models
  • Text-only output generation

Benchmark Breakdown

Reasoning (MMLU)82%
Coding (HumanEval)78%
Mathematics (MATH)76%
Long-Context Retrieval95%

Pricing Matrix

Input cost / 1M tokens$0.20
Output cost / 1M tokens$1.20

Compare OpenAI: GPT-5.6 Luna with other frontier models:

Related Skills, Tools & Automations