GPT-6 Luna
About this model
GPT-6 Luna is the smallest, most cost-efficient member of OpenAI's GPT-6 lineup, released in September 2026 alongside GPT-6 Sol and sitting below the flagship GPT-6 Astra. OpenAI's own model guidance positions it for cost-sensitive, high-volume workloads, while Sol balances intelligence and cost and Astra handles the most complex reasoning and coding. It accepts text and images, emits text, and exposes a 1.05 million-token context window with up to 128K output tokens.
Compared with its direct predecessor, GPT-5.6 Luna, the continuity is notable: both share the same 1.05M-token window and the same selectable reasoning-effort ladder running from none through low, medium, high and max, plus vision input, function calling and hosted tools such as web search. The GPT-6 release refreshes this efficiency tier within the newer generation, so the same integration surface — Responses and Chat Completions — carries over for existing 5.6 Luna deployments.
Practically, Luna suits classification, extraction, routing, retrieval over very long documents, and focused agent steps, with reasoning effort dialled up when a task needs more deliberation and dialled down for latency-sensitive, high-throughput traffic. Teams needing deeper multi-step reasoning or heavier coding work are pointed by OpenAI's documentation toward Sol or Astra instead, keeping Luna as the volume workhorse of the family.
This About section is AI-generated from public sources (Claude Opus 5), with no human editing. It may contain inaccuracies — verify critical details against the sources listed above.
Data sources: Venice API · HuggingFace · Wikipedia — enrichment updated 2h ago