xAI: Grok Uncensored
xAI Grok Uncensored API pricing
Grok Uncensored costs $0.800 per million input tokens and $2.40 per million output tokens on Oxyy. These are the pay-as-you-go prices after the current 60% discount (Launch Promo; standard price $2.00 input and $6.00 output), and cached input tokens are billed at $0.120 per million. It accepts text and images and returns text, with a 500K-token context window. A cheaper xAI option is Grok 4.20 (0309) Non-Reasoning at $0.500 / $1.00 per 1M tokens; for every model and rate, see Grok API pricing.
xAI frontier model for coding, knowledge work and STEM, co-trained with Cursor and tuned for agentic tool calling.
Providers
Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.
| Provider | Input /M | Output /M | Cache read /M | Latency | Throughput |
|---|---|---|---|---|---|
| xAI | $0.800 | $2.40 | $0.120 | 7.46s | 75 tps |
Capabilities (14)
Supported parameters (15)
Any other parameter you send is ignored rather than rejected.
Pricing
What this model costs on Oxyy, pay as you go. "Charged" is the rate a request is billed at; the standard price is shown beside it for reference.
| Rate | Standard price | Charged | Unit |
|---|---|---|---|
| Input | $2.00 | $0.800 | per 1M input tokens |
| Output | $6.00 | $2.40 | per 1M output tokens |
| Cached input | $0.300 | $0.120 | per 1M cached input tokens |
| Cache write | $0 | $0 | per 1M cache-write tokens |
Discount. Charged rates include the current "Launch Promo" discount of 60% on every rate.
Long-context pricing. Once the prompt is larger than the threshold, the whole request is billed at these charged rates per 1M tokens; rates not listed keep the price above.From 200,000 tokens in the prompt: input $1.60 / output $4.80 / cache read $0.240.
Specifications
- Model ID
grok-uncensored- Developer
- xAI
- Type
- Text
- Input
- Text and images
- Output
- Text
- Context window
- 500,000 tokens
- Released
- Jul 8, 2026
- Tokenizer
- Grok
- Series
- Grok 4
- Image input
- Up to 20 MB each · JPEG, PNG
- Endpoint
/v1/chat/completions
Best for
Quickstart
Call grok-uncensored with a POST to /v1/chat/completions on https://api.oxyy.ai, using your Oxyy API key from the OXYY_API_KEY environment variable. The same key works for every model in the catalog.
import os
from openai import OpenAI
client = OpenAI(
base_url="https://api.oxyy.ai/v1",
api_key=os.environ["OXYY_API_KEY"],
)
response = client.chat.completions.create(
model="grok-uncensored",
messages=[
{"role": "user", "content": "Explain unified AI APIs in one sentence."},
],
)
print(response.choices[0].message.content)Performance
Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better). Throughput and TTFT are measured on streaming requests, which are the only ones with a first-token moment to time.
Throughput
Latency
Time to first token
Cache hit rate
Availability
The share of requests to this model that completed successfully over the last 30 days. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.
Days are UTC days. Days with no completed requests are omitted rather than drawn at 100% — no traffic is not evidence of availability.
Activity
Token volume and request traffic to this model over time. Daily totals, UTC.
Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length.
Frequently asked questions
What is Grok Uncensored?
How much does Grok Uncensored cost?
What is the context length of Grok Uncensored?
Does Grok Uncensored support tool calling and structured outputs?
What inputs and outputs does Grok Uncensored support?
Which API endpoint does Grok Uncensored use?
When was Grok Uncensored released?
What other models does xAI have?
Alternatives to Grok Uncensored
Models of the same kind from other vendors, closest in price first. Prices are pay-as-you-go on Oxyy, through the same API key.


