Model Terminal

GPT-5.6 Luna

GPT-5.6 Luna is OpenAI's smallest, fastest, and most cost-efficient model in the GPT-5.6 family, designed for high-volume, cost-sensitive workloads such as classification, extraction, chat, and lightweight agent workflows. It offers a 1,050,000-token context window and 128,000 max output tokens, positioned as the throughput-optimized tier below Terra (balanced) and Sol (flagship). Source: supabase.

Value score
Not scored
Context
tokens
Max output
tokens
Price

Capability radar

Peer value bars

Identity

Developer
OpenAI
Openness
Modalities
Release
Knowledge cutoff
Deprecation
API docs

Benchmarks

No benchmark scores yet.

Compare nearby

Pricing

Input / 1M
Output / 1M
Speed