oxyy.ai
HomeModelsProvidersPricingBlogDocs
Sign inGet API key
  1. Home/
  2. Providers/
  3. OpenAI/
  4. GPT-4o

OpenAI: GPT-4o

openai/gpt-4o
Compare Playground Get API key
Versatile, high-intelligence GPT model. Accepts text and image input and produces text, including Structured Outputs.
ModalitiesTextImageText
In / out price$1.00 / $4.00 per 1M
Context128K
Released—
ProvidersPricingPerformanceAvailabilityAppsActivityFAQExplore

Providers

Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.

ProviderInput /MOutput /MCache read /MLatencyThroughput
OpenAI$1.00$4.00$1.255.25s55 tps

Capabilities (9)

StreamingTool callingStructured outputSystem promptVisionPrompt cachingBatchFine TunablePredicted Outputs

Supported parameters (2)

toolsresponse_format

Any other parameter you send is ignored rather than rejected.

Pricing

What this model costs to run, next to the rate it is posted at. Caching and discounts mean the price actually paid is often below the listed one.

Effective input price
$1.00
/M tokens · listed $2.50
Effective output price
$4.00
/M tokens · listed $10.00
RateListedChargedUnit
Input$2.50$1.00per 1M tokens
Output$10.00$4.00per 1M tokens
Cache read$1.25$1.25per 1M tokens
Cache write$0$0per 1M tokens
UTC

Performance

Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better). Throughput and TTFT are measured on streaming requests, which are the only ones with a first-token moment to time.

Throughput
55tok/s
P50, streaming requests
Latency
5.25s
P50, end to end
UTC

Throughput

P99108 tok/s
P9588 tok/s
P5058 tok/s

Latency

P505749ms
P758405ms
P9538.1s
P9955.3s

Time to first token

P501327ms
P953883ms

Cache hit rate

0.00%
Share of input tokens served from the prompt cache over the window.

Availability

The share of requests to this model that completed successfully. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.

Success rate (30d)
80.00%
over 2,280 requests
Days with traffic
31
of the last 30

Days are UTC days. Days with no requests are omitted rather than drawn at 100% — no traffic is not evidence of availability.

Apps

Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for.

No apps have identified themselves for this model yet. Clients can opt in by sending an X-Title header with their request.

Activity

Token volume and request traffic to this model over time. Daily totals, UTC.

Prompt1.5M
Completion217.8K

Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length. Reasoning is not priced separately for this model, so it is not itemised here.

Frequently asked questions

What is GPT-4o?
Versatile, high-intelligence GPT model. Accepts text and image input and produces text, including Structured Outputs.
How much does GPT-4o cost?
GPT-4o costs $2.50 per million input tokens and $10.00 per million output tokens. Cached prompt tokens are billed at $1.25 per million. You pay per request, with no subscription.
What is the context length of GPT-4o?
GPT-4o has a context window of 128,000 tokens, and can return up to 16,384 tokens in one response.
Does GPT-4o support tool calling and structured outputs?
Yes — GPT-4o supports both tool calling and structured outputs.
What inputs and outputs does GPT-4o support?
GPT-4o accepts text, image as input and returns text.
What other models does OpenAI have?
OpenAI also offers chat-latest, GPT-4o mini, GPT-4o Mini TTS, GPT-5.1 through Oxyy.
When was GPT-4o released?
No release date is published for GPT-4o.

Explore more models

Browse the full catalog ModelsCompare pricing Pricing

More models from OpenAI

chat-latestThe model currently powering ChatGPT, exposed through the API. A rolling alias whose underlying snapshot changes over t…text · — context · $5.00 / $30.00GPT-4o miniFast, affordable small model for focused tasks. Text and image input, text output.text · 128K context · $0.150 / $0.600GPT-4o Mini TTSText-to-speech model powered by GPT-4o Mini, with an instructions parameter for tone and style steering.audio tts · — contextGPT-5.1GPT-5.1 reasoning model with a no-reasoning default for fast responses.text · 400K context · $1.25 / $10.00GPT-5.2Previous flagship model for complex professional work.text · 400K context · $1.75 / $14.00GPT-5.3 CodexOpenAI's Codex model, optimized for agentic coding workflows.text · — context · $1.75 / $14.00
oxyy.ai

One endpoint for every major model. The OpenAI, Anthropic and Gemini SDKs work unchanged.

Product

  • Models
  • Providers
  • Pricing
  • Status
  • Startup program

Company

  • About
  • Blog
  • Contact
  • Terms of Service
  • Privacy Policy
  • Refund Policy
  • Cookie Policy

Developer

  • Documentation
  • Quickstart
  • SDKs
  • Model catalog
  • AI providers

Connect

  • Telegram
  • Discord
  • WhatsApp
© 2026 Oxyy.ai. All rights reserved.StatusAbout