OpenAI: GPT-6 Luna

openai/gpt-6-luna

OpenAI GPT-6 Luna API pricing

GPT-6 Luna costs $0.040 per million input tokens and $0.200 per million output tokens on Oxyy. These are the pay-as-you-go prices after the current 60% discount (Launch Promo; standard price $0.100 input and $0.500 output), and cached input tokens are billed at $0.004 per million. It accepts text and images and returns text, with a 1.05M-token context window and up to 128K output tokens per response. It is the lowest-priced OpenAI text model billed per token on Oxyy; for every model and rate, see OpenAI API pricing.

OpenAI's most efficient model for focused, high-volume tasks.
ModalitiesTextImageText
Price$0.040 / $0.200 per 1M tokens in / outAll OpenAI API pricing
Context1.05M

Providers

Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.

ProviderInput /MOutput /MCache read /M
OpenAI$0.040$0.200$0.004

Capabilities (17)

StreamingTool callingParallel tool callsStructured outputReasoningSystem promptVisionMultilingualPrompt cachingBatchFlexFast ModeWeb searchFile SearchComputer useFunction CallingData Residency

Supported parameters (6)

reasoning_effortservice_tierstreamtoolsprompt_cache_breakpointresponse_format

Any other parameter you send is ignored rather than rejected.

Pricing

What this model costs on Oxyy, pay as you go. "Charged" is the rate a request is billed at; the standard price is shown beside it for reference.

Input price
$0.040
per 1M input tokens · standard price $0.100
Output price
$0.200
per 1M output tokens · standard price $0.500
RateStandard priceChargedUnit
Input$0.100$0.040per 1M input tokens
Output$0.500$0.200per 1M output tokens
Cached input$0.010$0.004per 1M cached input tokens
Cache write$0.125$0.050per 1M cache-write tokens

Discount. Charged rates include the current "Launch Promo" discount of 60% on every rate.

Long-context pricing. Once the prompt is larger than the threshold, the whole request is billed at these charged rates per 1M tokens; rates not listed keep the price above.Above 272,000 tokens in the prompt: input $0.080 / output $0.300 / cache read $0.008 / cache write $0.100.

UTC

Specifications

Model ID
gpt-6-luna
Developer
OpenAI
Type
Text
Input
Text and images
Output
Text
Context window
1,050,000 tokens
Max output
128,000 tokens
Knowledge cutoff
May 2026
Tokenizer
GPT
Series
GPT-6
Image input
Up to 20 MB each
Endpoint
/v1/chat/completions

Best for

ReasoningCodingAgenticLong context

Quickstart

Call gpt-6-luna with a POST to /v1/chat/completions on https://api.oxyy.ai, using your Oxyy API key from the OXYY_API_KEY environment variable. The same key works for every model in the catalog.

import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.oxyy.ai/v1",
    api_key=os.environ["OXYY_API_KEY"],
)

response = client.chat.completions.create(
    model="gpt-6-luna",
    messages=[
        {"role": "user", "content": "Explain unified AI APIs in one sentence."},
    ],
)
print(response.choices[0].message.content)

Activity

Token volume and request traffic to this model over time. Daily totals, UTC.

Prompt133.8M
Reasoning2.2M
Completion11.7M

Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length.

Frequently asked questions

What is GPT-6 Luna?
OpenAI's most efficient model for focused, high-volume tasks.
How much does GPT-6 Luna cost?
GPT-6 Luna costs $0.040 per million input tokens and $0.200 per million output tokens on Oxyy. These are the pay-as-you-go prices after the current 60% discount (Launch Promo; standard price $0.100 input and $0.500 output), and cached input tokens are billed at $0.004 per million. You pay per request, with no subscription.
What is the context length of GPT-6 Luna?
GPT-6 Luna has a context window of 1,050,000 tokens, and can return up to 128,000 tokens in one response.
Does GPT-6 Luna support tool calling and structured outputs?
Yes — GPT-6 Luna supports both tool calling and structured outputs.
What inputs and outputs does GPT-6 Luna support?
GPT-6 Luna accepts text and images as input and returns text.
Which API endpoint does GPT-6 Luna use?
Send a POST request to https://api.oxyy.ai/v1/chat/completions with "model": "gpt-6-luna", authenticated with your Oxyy API key.
What is the knowledge cutoff of GPT-6 Luna?
GPT-6 Luna's training data runs up to May 2026.
What other models does OpenAI have?
OpenAI also offers chat-latest, GPT-4o, GPT-4o mini and GPT-5.1 through Oxyy.

Alternatives to GPT-6 Luna

Models of the same kind from other vendors, closest in price first. Prices are pay-as-you-go on Oxyy, through the same API key.

Explore more models

More models from OpenAI