GPT-6 Luna
A low-cost OpenAI model for focused tasks repeated at scale, such as extraction, classification and short replies.
A low-cost OpenAI model for focused tasks repeated at scale, such as extraction, classification and short replies.
Plain-English overview
GPT-6 Luna accepts text and images and produces text, with the same published context and output limits as Sol. It is a separate model from GPT-5.6 Luna. Its lower token price makes it a candidate for repetitive tasks; test accuracy before moving complex work to it.
The Responses API connects it to search, files, code execution and other tools. Reasoning ranges from none to max, with medium as the default. Choose the setting by testing your own tasks, not by assuming a higher setting is always better.
Category comparison
These are provider-published specifications, not Cody benchmark scores. Follow the linked sources for current limits and endpoint-specific exceptions.
Pricing & comparisons
Set your usage. Your estimate updates as you type.
Example: 6,000 input tokens for 10 pages, plus a 500-token summary. Page lengths vary; adjust the numbers below.
One run sends your input to the model once and receives an answer. Tokens are pieces of text: input is what you send, output is the answer you receive.
GPT-6 Luna
OpenAI
Estimated total (USD)
For the usage above · USD · API pricing, not a subscription
Estimates exclude taxes, tools, cache storage/writes, free allowances and custom discounts. Image estimates cover output only, not prompt or reference-image charges. Quality modes differ by model. Unlisted settings are not treated as free.
API and provider access
Availability
Direct API access in supported countries; account and regional restrictions apply.
EU data residency requires Standard processing. Check OpenAI's supported countries and account eligibility.
Check live availabilityData and training
OpenAI says API inputs and outputs are not used for model training by default. Default abuse-monitoring logs may be retained for up to 30 days; qualifying organizations can request other controls.
This is a concise reading of the cited provider material, not legal advice. A third-party gateway can have different storage, routing, training, and residency terms from the model maker's direct API.
Read the provider policyFrequently asked questions
A low-cost OpenAI model for focused tasks repeated at scale, such as extraction, classification and short replies. GPT-6 Luna accepts text and images and produces text, with the same published context and output limits as Sol. It is a separate model from GPT-5.6 Luna. Its lower token price makes it a candidate for repetitive tasks; test accuracy before moving complex work to it.
GPT-6 Luna was released on September 22, 2026 according to the cited provider materials.
Direct API access in supported countries; account and regional restrictions apply. The access routes listed in this guide are OpenAI.
$0.10 input · $0.50 output / 1M tokens. Above 272K input tokens, the whole request uses 2× input/cache rates and 1.5× output rates. Cache writes cost 1.25× input rates. Batch/Flex cost 50% of Standard; Fast mode costs 2×. Regional processing adds 10%; tools cost extra.
Direct API access in supported countries; account and regional restrictions apply. EU data residency requires Standard processing. Check OpenAI's supported countries and account eligibility.
OpenAI says API inputs and outputs are not used for model training by default. Default abuse-monitoring logs may be retained for up to 30 days; qualifying organizations can request other controls. The policy belongs to the provider route and account terms, so verify it again before production use.
One price per model, using the same input/output mix. We weight each API rate by its share of one million total tokens and use the lowest matching reviewed route. This is a comparison rate, not a cost per task or a quality score. Tokenizers, caching, reasoning and long context can change your actual bill.
75% input + 25% output · no cache discounts
OpenAI
This model
Mistral AI
OpenAI
Anthropic
OpenAI
Anthropic
OpenAI
OpenAI
Anthropic
OpenAI
Anthropic
Anthropic
OpenAI
Anthropic
Related comparisons
An OpenAI model for coding projects and assistants that carry out several steps with tools.
Open full comparisonA lower-cost GPT-5.6 option for coding, analysis and assistants that use tools.
Open full comparisonThe budget GPT-5.6 tier for processing many short, repeatable tasks.
Open full comparisonA compact Claude model for quick replies, classification and narrowly scoped assistants.
Open full comparisonGoogle's stable, high-efficiency multimodal model for agents, software work, and large mixed-media inputs at an introductory Flash-tier price.
Open full comparisonMistral's lower-cost open model that combines normal instruction following, reasoning, coding, vision, and agent tools in one endpoint.
Open full comparison