OpenAI: GPT-6 Sol
OpenAI GPT-6 Sol API pricing
GPT-6 Sol costs $0.800 per million input tokens and $4.00 per million output tokens on Oxyy. These are the pay-as-you-go prices after the current 60% discount (Launch Promo; standard price $2.00 input and $10.00 output), and cached input tokens are billed at $0.080 per million. It accepts text and images and returns text, with a 1.05M-token context window and up to 128K output tokens per response. A cheaper OpenAI option is GPT-5.1 at $0.500 / $4.00 per 1M tokens; for every model and rate, see OpenAI API pricing.
Providers
Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.
| Provider | Input /M | Output /M | Cache read /M | Latency | Throughput |
|---|---|---|---|---|---|
| OpenAI | $0.800 | $4.00 | $0.080 | 1.00s | 55 tps |
Capabilities (17)
Supported parameters (6)
Any other parameter you send is ignored rather than rejected.
Pricing
What this model costs on Oxyy, pay as you go. "Charged" is the rate a request is billed at; the standard price is shown beside it for reference.
| Rate | Standard price | Charged | Unit |
|---|---|---|---|
| Input | $2.00 | $0.800 | per 1M input tokens |
| Output | $10.00 | $4.00 | per 1M output tokens |
| Cached input | $0.200 | $0.080 | per 1M cached input tokens |
| Cache write | $2.50 | $1.00 | per 1M cache-write tokens |
Discount. Charged rates include the current "Launch Promo" discount of 60% on every rate.
Long-context pricing. Once the prompt is larger than the threshold, the whole request is billed at these charged rates per 1M tokens; rates not listed keep the price above.Above 272,000 tokens in the prompt: input $1.60 / output $6.00 / cache read $0.160 / cache write $2.00.
Specifications
- Model ID
gpt-6-sol- Developer
- OpenAI
- Type
- Text
- Input
- Text and images
- Output
- Text
- Context window
- 1,050,000 tokens
- Max output
- 128,000 tokens
- Knowledge cutoff
- April 2026
- Tokenizer
- GPT
- Series
- GPT-6
- Image input
- Up to 20 MB each
- Endpoint
/v1/chat/completions
Best for
Quickstart
Call gpt-6-sol with a POST to /v1/chat/completions on https://api.oxyy.ai, using your Oxyy API key from the OXYY_API_KEY environment variable. The same key works for every model in the catalog.
import os
from openai import OpenAI
client = OpenAI(
base_url="https://api.oxyy.ai/v1",
api_key=os.environ["OXYY_API_KEY"],
)
response = client.chat.completions.create(
model="gpt-6-sol",
messages=[
{"role": "user", "content": "Explain unified AI APIs in one sentence."},
],
)
print(response.choices[0].message.content)Performance
Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better). Throughput and TTFT are measured on streaming requests, which are the only ones with a first-token moment to time.
Throughput
Latency
Time to first token
Cache hit rate
Availability
The share of requests to this model that completed successfully over the last 30 days. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.
Activity
Token volume and request traffic to this model over time. Daily totals, UTC.
Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length.
Frequently asked questions
What is GPT-6 Sol?
How much does GPT-6 Sol cost?
What is the context length of GPT-6 Sol?
Does GPT-6 Sol support tool calling and structured outputs?
What inputs and outputs does GPT-6 Sol support?
Which API endpoint does GPT-6 Sol use?
What is the knowledge cutoff of GPT-6 Sol?
What other models does OpenAI have?
Alternatives to GPT-6 Sol
Models of the same kind from other vendors, closest in price first. Prices are pay-as-you-go on Oxyy, through the same API key.


