oxyy.ai
HomeModelsProvidersPricingBlogDocs
Sign inGet API key
  1. Home/
  2. Providers/
  3. DeepSeek/
  4. DeepSeek-V4-Pro-0813

DeepSeek: DeepSeek-V4-Pro-0813

deepseek/deepseek-v4-pro
Compare Playground Get API key
DeepSeek's flagship model, with a 1M-token context window and both thinking and non-thinking modes. Text only.
ModalitiesTextText
In / out price$0.264 / $0.792 per 1M
Context1M
ReleasedAug 13, 2026
ProvidersPricingPerformanceAvailabilityAppsActivityFAQExplore

Providers

Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.

ProviderInput /MOutput /MCache read /MLatencyThroughput
DeepSeek$0.264$0.792$0.0223.12s16 tps

Capabilities (13)

StreamingTool callingStructured outputJson ModeThinkingReasoningSystem promptPrompt cachingResponses ApiAnthropic ApiChat Prefix CompletionFim CompletionFiles API

Supported parameters (2)

toolsresponse_format

Any other parameter you send is ignored rather than rejected.

Pricing

What this model costs to run, next to the rate it is posted at. Caching and discounts mean the price actually paid is often below the listed one.

Effective input price
$0.264
/M tokens · listed $0.660
Effective output price
$0.792
/M tokens · listed $1.98
RateListedChargedUnit
Input$0.660$0.264per 1M tokens
Output$1.98$0.792per 1M tokens
Cache read$0.022$0.022per 1M tokens
UTC

Performance

Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better). Throughput and TTFT are measured on streaming requests, which are the only ones with a first-token moment to time.

Throughput
16tok/s
P50, streaming requests
Latency
3.12s
P50, end to end
UTC

Throughput

P9916 tok/s
P9516 tok/s
P5016 tok/s

Latency

P503120ms
P753120ms
P953120ms
P993120ms

Time to first token

P502108ms
P952108ms

Cache hit rate

0.00%
Share of input tokens served from the prompt cache over the window.

Availability

The share of requests to this model that completed successfully. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.

Success rate (30d)
1.00%
over 97 requests
Days with traffic
3
of the last 30

Days are UTC days. Days with no requests are omitted rather than drawn at 100% — no traffic is not evidence of availability.

Apps

Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for.

  1. 1.Agnai.Chat 3.0K tokens
  2. 2.Oxyy Tester112 tokens

Activity

Token volume and request traffic to this model over time. Daily totals, UTC.

Prompt3.2K
Reasoning16
Completion16

Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length.

Frequently asked questions

What is DeepSeek-V4-Pro-0813?
DeepSeek's flagship model, with a 1M-token context window and both thinking and non-thinking modes. Text only.
How much does DeepSeek-V4-Pro-0813 cost?
DeepSeek-V4-Pro-0813 costs $0.660 per million input tokens and $1.98 per million output tokens. Cached prompt tokens are billed at $0.022 per million. You pay per request, with no subscription.
What is the context length of DeepSeek-V4-Pro-0813?
DeepSeek-V4-Pro-0813 has a context window of 1,000,000 tokens, and can return up to 384,000 tokens in one response.
Does DeepSeek-V4-Pro-0813 support tool calling and structured outputs?
Yes — DeepSeek-V4-Pro-0813 supports both tool calling and structured outputs.
What inputs and outputs does DeepSeek-V4-Pro-0813 support?
DeepSeek-V4-Pro-0813 accepts text as input and returns text.
What other models does DeepSeek have?
DeepSeek also offers DeepSeek-V4.1-Flash through Oxyy.
When was DeepSeek-V4-Pro-0813 released?
DeepSeek-V4-Pro-0813 was released on Aug 13, 2026.

Explore more models

Browse the full catalog ModelsCompare pricing Pricing

More models from DeepSeek

DeepSeek-V4.1-FlashThe smallest model in DeepSeek's new architecture family, with native multimodal visual understanding. Designed for a h…text · 1M context · $0.150 / $0.600
oxyy.ai

One endpoint for every major model. The OpenAI, Anthropic and Gemini SDKs work unchanged.

Product

  • Models
  • Providers
  • Pricing
  • Status
  • Startup program

Company

  • About
  • Blog
  • Contact
  • Terms of Service
  • Privacy Policy
  • Refund Policy
  • Cookie Policy

Developer

  • Documentation
  • Quickstart
  • SDKs
  • Model catalog
  • AI providers

Connect

  • Telegram
  • Discord
  • WhatsApp
© 2026 Oxyy.ai. All rights reserved.StatusAbout