OpenAI: GPT-6 Luna
OpenAI GPT-6 Luna API pricing
GPT-6 Luna costs $0.040 per million input tokens and $0.200 per million output tokens on Oxyy. These are the pay-as-you-go prices after the current 60% discount (Launch Promo; standard price $0.100 input and $0.500 output), and cached input tokens are billed at $0.004 per million. It accepts text and images and returns text, with a 1.05M-token context window and up to 128K output tokens per response. It is the lowest-priced OpenAI text model billed per token on Oxyy; for every model and rate, see OpenAI API pricing.
Providers
Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.
| Provider | Input /M | Output /M | Cache read /M |
|---|---|---|---|
| OpenAI | $0.040 | $0.200 | $0.004 |
Capabilities (17)
Supported parameters (6)
Any other parameter you send is ignored rather than rejected.
Pricing
What this model costs on Oxyy, pay as you go. "Charged" is the rate a request is billed at; the standard price is shown beside it for reference.
| Rate | Standard price | Charged | Unit |
|---|---|---|---|
| Input | $0.100 | $0.040 | per 1M input tokens |
| Output | $0.500 | $0.200 | per 1M output tokens |
| Cached input | $0.010 | $0.004 | per 1M cached input tokens |
| Cache write | $0.125 | $0.050 | per 1M cache-write tokens |
Discount. Charged rates include the current "Launch Promo" discount of 60% on every rate.
Long-context pricing. Once the prompt is larger than the threshold, the whole request is billed at these charged rates per 1M tokens; rates not listed keep the price above.Above 272,000 tokens in the prompt: input $0.080 / output $0.300 / cache read $0.008 / cache write $0.100.
Specifications
- Model ID
gpt-6-luna- Developer
- OpenAI
- Type
- Text
- Input
- Text and images
- Output
- Text
- Context window
- 1,050,000 tokens
- Max output
- 128,000 tokens
- Knowledge cutoff
- May 2026
- Tokenizer
- GPT
- Series
- GPT-6
- Image input
- Up to 20 MB each
- Endpoint
/v1/chat/completions
Best for
Quickstart
Call gpt-6-luna with a POST to /v1/chat/completions on https://api.oxyy.ai, using your Oxyy API key from the OXYY_API_KEY environment variable. The same key works for every model in the catalog.
import os
from openai import OpenAI
client = OpenAI(
base_url="https://api.oxyy.ai/v1",
api_key=os.environ["OXYY_API_KEY"],
)
response = client.chat.completions.create(
model="gpt-6-luna",
messages=[
{"role": "user", "content": "Explain unified AI APIs in one sentence."},
],
)
print(response.choices[0].message.content)Activity
Token volume and request traffic to this model over time. Daily totals, UTC.
Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length.
Frequently asked questions
What is GPT-6 Luna?
How much does GPT-6 Luna cost?
What is the context length of GPT-6 Luna?
Does GPT-6 Luna support tool calling and structured outputs?
What inputs and outputs does GPT-6 Luna support?
Which API endpoint does GPT-6 Luna use?
What is the knowledge cutoff of GPT-6 Luna?
What other models does OpenAI have?
Alternatives to GPT-6 Luna
Models of the same kind from other vendors, closest in price first. Prices are pay-as-you-go on Oxyy, through the same API key.


