OpenAI: GPT-6 Sol

openai/gpt-6-sol

OpenAI GPT-6 Sol API pricing

GPT-6 Sol costs $0.800 per million input tokens and $4.00 per million output tokens on Oxyy. These are the pay-as-you-go prices after the current 60% discount (Launch Promo; standard price $2.00 input and $10.00 output), and cached input tokens are billed at $0.080 per million. It accepts text and images and returns text, with a 1.05M-token context window and up to 128K output tokens per response. A cheaper OpenAI option is GPT-5.1 at $0.500 / $4.00 per 1M tokens; for every model and rate, see OpenAI API pricing.

Built for complex coding and agentic workflows.
ModalitiesTextImageText
Price$0.800 / $4.00 per 1M tokens in / outAll OpenAI API pricing
Context1.05M

Providers

Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.

ProviderInput /MOutput /MCache read /MLatencyThroughput
OpenAI$0.800$4.00$0.0801.00s55 tps

Capabilities (17)

StreamingTool callingParallel tool callsStructured outputReasoningSystem promptVisionMultilingualPrompt cachingBatchFlexFast ModeWeb searchFile SearchComputer useFunction CallingData Residency

Supported parameters (6)

reasoning_effortservice_tierstreamtoolsprompt_cache_breakpointresponse_format

Any other parameter you send is ignored rather than rejected.

Pricing

What this model costs on Oxyy, pay as you go. "Charged" is the rate a request is billed at; the standard price is shown beside it for reference.

Input price
$0.800
per 1M input tokens · standard price $2.00
Output price
$4.00
per 1M output tokens · standard price $10.00
RateStandard priceChargedUnit
Input$2.00$0.800per 1M input tokens
Output$10.00$4.00per 1M output tokens
Cached input$0.200$0.080per 1M cached input tokens
Cache write$2.50$1.00per 1M cache-write tokens

Discount. Charged rates include the current "Launch Promo" discount of 60% on every rate.

Long-context pricing. Once the prompt is larger than the threshold, the whole request is billed at these charged rates per 1M tokens; rates not listed keep the price above.Above 272,000 tokens in the prompt: input $1.60 / output $6.00 / cache read $0.160 / cache write $2.00.

UTC

Specifications

Model ID
gpt-6-sol
Developer
OpenAI
Type
Text
Input
Text and images
Output
Text
Context window
1,050,000 tokens
Max output
128,000 tokens
Knowledge cutoff
April 2026
Tokenizer
GPT
Series
GPT-6
Image input
Up to 20 MB each
Endpoint
/v1/chat/completions

Best for

ReasoningCodingAgenticLong context

Quickstart

Call gpt-6-sol with a POST to /v1/chat/completions on https://api.oxyy.ai, using your Oxyy API key from the OXYY_API_KEY environment variable. The same key works for every model in the catalog.

import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.oxyy.ai/v1",
    api_key=os.environ["OXYY_API_KEY"],
)

response = client.chat.completions.create(
    model="gpt-6-sol",
    messages=[
        {"role": "user", "content": "Explain unified AI APIs in one sentence."},
    ],
)
print(response.choices[0].message.content)

Performance

Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better). Throughput and TTFT are measured on streaming requests, which are the only ones with a first-token moment to time.

Throughput
55tok/s
P50, streaming requests
Latency
1.00s
P50, end to end
UTC

Throughput

P9955 tok/s
P9555 tok/s
P5055 tok/s

Latency

P501003ms
P751003ms
P951003ms
P991003ms

Time to first token

P50894ms
P95894ms

Cache hit rate

0.00%
Share of input tokens served from the prompt cache over the window.

Availability

The share of requests to this model that completed successfully over the last 30 days. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.

Requests (30 days)
5
A success rate is shown from 100 completed requests
Days with traffic
4
of the last 30 days

Activity

Token volume and request traffic to this model over time. Daily totals, UTC.

Prompt553
Reasoning103
Completion186

Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length.

Frequently asked questions

What is GPT-6 Sol?
Built for complex coding and agentic workflows.
How much does GPT-6 Sol cost?
GPT-6 Sol costs $0.800 per million input tokens and $4.00 per million output tokens on Oxyy. These are the pay-as-you-go prices after the current 60% discount (Launch Promo; standard price $2.00 input and $10.00 output), and cached input tokens are billed at $0.080 per million. You pay per request, with no subscription.
What is the context length of GPT-6 Sol?
GPT-6 Sol has a context window of 1,050,000 tokens, and can return up to 128,000 tokens in one response.
Does GPT-6 Sol support tool calling and structured outputs?
Yes — GPT-6 Sol supports both tool calling and structured outputs.
What inputs and outputs does GPT-6 Sol support?
GPT-6 Sol accepts text and images as input and returns text.
Which API endpoint does GPT-6 Sol use?
Send a POST request to https://api.oxyy.ai/v1/chat/completions with "model": "gpt-6-sol", authenticated with your Oxyy API key.
What is the knowledge cutoff of GPT-6 Sol?
GPT-6 Sol's training data runs up to April 2026.
What other models does OpenAI have?
OpenAI also offers chat-latest, GPT-4o, GPT-4o mini and GPT-5.1 through Oxyy.

Alternatives to GPT-6 Sol

Models of the same kind from other vendors, closest in price first. Prices are pay-as-you-go on Oxyy, through the same API key.

Explore more models

More models from OpenAI