oxyy.ai
HomeModelsProvidersPricingBlogDocs
Sign inGet API key
  1. Home/
  2. Providers/
  3. Anthropic/
  4. Claude Opus 4.8

Anthropic: Claude Opus 4.8

anthropic/claude-opus-4-8
Compare Playground Get API key
Previous-generation Opus model with adaptive thinking. Anthropic recommends migrating to Claude Opus 5.
ModalitiesTextImageText
In / out price$2.00 / $10.00 per 1M
Context1M
Released—
ProvidersPricingPerformanceAvailabilityAppsActivityFAQExplore

Providers

Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.

ProviderInput /MOutput /MCache read /MLatencyThroughput
Anthropic$2.00$10.00$0.5007.99s70 tps

Capabilities (18)

StreamingTool callingStructured outputThinkingReasoningSystem promptVisionMultilingualPrompt cachingBatchFast ModeWeb searchWeb FetchCode executionComputer useBrowser UseBash ToolText Editor Tool

Supported parameters (7)

max_tokenseffortspeedinference_geotoolsresponse_formatreasoning_effort

Any other parameter you send is ignored rather than rejected.

Pricing

What this model costs to run, next to the rate it is posted at. Caching and discounts mean the price actually paid is often below the listed one.

Effective input price
$2.00
/M tokens · listed $5.00
Effective output price
$10.00
/M tokens · listed $25.00
RateListedChargedUnit
Input$5.00$2.00per 1M tokens
Output$25.00$10.00per 1M tokens
Cache read$0.500$0.500per 1M tokens
Cache write$6.25$6.25per 1M tokens
Cache write (1h)$10.00$10.00per 1M tokens
UTC

Performance

Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better). Throughput and TTFT are measured on streaming requests, which are the only ones with a first-token moment to time.

Throughput
70tok/s
P50, streaming requests
Latency
7.99s
P50, end to end
UTC

Throughput

P99910 tok/s
P95135 tok/s
P5070 tok/s

Latency

P509810ms
P7523.4s
P9563.8s
P99116.6s

Time to first token

P502135ms
P9550.6s

Cache hit rate

10.18%
Share of input tokens served from the prompt cache over the window.

Availability

The share of requests to this model that completed successfully. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.

Success rate (30d)
95.90%
over 3,653 requests
Days with traffic
30
of the last 30

Days are UTC days. Days with no requests are omitted rather than drawn at 100% — no traffic is not evidence of availability.

Apps

Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for.

  1. 1.Oxyy Tester73.9K tokens

Activity

Token volume and request traffic to this model over time. Daily totals, UTC.

Prompt183.2M
Reasoning59
Completion5.3M

Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length.

Frequently asked questions

What is Claude Opus 4.8?
Previous-generation Opus model with adaptive thinking. Anthropic recommends migrating to Claude Opus 5.
How much does Claude Opus 4.8 cost?
Claude Opus 4.8 costs $5.00 per million input tokens and $25.00 per million output tokens. Cached prompt tokens are billed at $0.500 per million. You pay per request, with no subscription.
What is the context length of Claude Opus 4.8?
Claude Opus 4.8 has a context window of 1,000,000 tokens, and can return up to 128,000 tokens in one response.
Does Claude Opus 4.8 support tool calling and structured outputs?
Yes — Claude Opus 4.8 supports both tool calling and structured outputs.
What inputs and outputs does Claude Opus 4.8 support?
Claude Opus 4.8 accepts text, image as input and returns text.
What other models does Anthropic have?
Anthropic also offers Claude Fable 5, Claude Fable 5.1, Claude Haiku 4.5, Claude Opus 4.5 through Oxyy.
When was Claude Opus 4.8 released?
No release date is published for Claude Opus 4.8.

Explore more models

Browse the full catalog ModelsCompare pricing Pricing

More models from Anthropic

Claude Fable 5Previous-generation Fable-tier model for demanding reasoning and long-horizon agentic work. Superseded by Claude Fable…text · 1M context · $10.00 / $50.00Claude Fable 5.1Anthropic's most capable model, for demanding reasoning and long-horizon agentic work.text · 1M context · $10.00 / $50.00Claude Haiku 4.5The fastest Claude model with near-frontier intelligence.text · 200K context · $1.00 / $5.00Claude Opus 4.5Legacy Opus model with extended thinking and a 200K context window.text · 200K context · $5.00 / $25.00Claude Opus 4.6Legacy Opus model. The first Opus with the full 1M token context window at standard pricing.text · 1M context · $5.00 / $25.00Claude Opus 4.7Legacy Opus model with adaptive thinking and the xhigh effort level for long-running agentic and coding tasks.text · 1M context · $5.00 / $25.00
oxyy.ai

One endpoint for every major model. The OpenAI, Anthropic and Gemini SDKs work unchanged.

Product

  • Models
  • Providers
  • Pricing
  • Status
  • Startup program

Company

  • About
  • Blog
  • Contact
  • Terms of Service
  • Privacy Policy
  • Refund Policy
  • Cookie Policy

Developer

  • Documentation
  • Quickstart
  • SDKs
  • Model catalog
  • AI providers

Connect

  • Telegram
  • Discord
  • WhatsApp
© 2026 Oxyy.ai. All rights reserved.StatusAbout