xAI: Grok Uncensored

xai/grok-uncensored

xAI Grok Uncensored API pricing

Grok Uncensored costs $0.800 per million input tokens and $2.40 per million output tokens on Oxyy. These are the pay-as-you-go prices after the current 60% discount (Launch Promo; standard price $2.00 input and $6.00 output), and cached input tokens are billed at $0.120 per million. It accepts text and images and returns text, with a 500K-token context window. A cheaper xAI option is Grok 4.20 (0309) Non-Reasoning at $0.500 / $1.00 per 1M tokens; for every model and rate, see Grok API pricing.

xAI frontier model for coding, knowledge work and STEM, co-trained with Cursor and tuned for agentic tool calling.

ModalitiesTextImageText
Price$0.800 / $2.40 per 1M tokens in / outAll Grok API pricing
Context500K
ReleasedJul 8, 2026

Providers

Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.

ProviderInput /MOutput /MCache read /MLatencyThroughput
xAI$0.800$2.40$0.1207.46s75 tps

Capabilities (14)

StreamingTool callingParallel tool callsStructured outputReasoningSystem promptVisionPrompt cachingWeb searchX SearchCode executionRemote McpPriority ProcessingBatch

Supported parameters (15)

modelmessagesstreammax_tokenstemperaturetop_pstopseedtoolstool_choiceresponse_formatservice_tierreasoning_effortsearch_parametersuser

Any other parameter you send is ignored rather than rejected.

Pricing

What this model costs on Oxyy, pay as you go. "Charged" is the rate a request is billed at; the standard price is shown beside it for reference.

Input price
$0.800
per 1M input tokens · standard price $2.00
Output price
$2.40
per 1M output tokens · standard price $6.00
RateStandard priceChargedUnit
Input$2.00$0.800per 1M input tokens
Output$6.00$2.40per 1M output tokens
Cached input$0.300$0.120per 1M cached input tokens
Cache write$0$0per 1M cache-write tokens

Discount. Charged rates include the current "Launch Promo" discount of 60% on every rate.

Long-context pricing. Once the prompt is larger than the threshold, the whole request is billed at these charged rates per 1M tokens; rates not listed keep the price above.From 200,000 tokens in the prompt: input $1.60 / output $4.80 / cache read $0.240.

UTC

Specifications

Model ID
grok-uncensored
Developer
xAI
Type
Text
Input
Text and images
Output
Text
Context window
500,000 tokens
Released
Jul 8, 2026
Tokenizer
Grok
Series
Grok 4
Image input
Up to 20 MB each · JPEG, PNG
Endpoint
/v1/chat/completions

Best for

CodingAgentic tool callingKnowledge workStem

Quickstart

Call grok-uncensored with a POST to /v1/chat/completions on https://api.oxyy.ai, using your Oxyy API key from the OXYY_API_KEY environment variable. The same key works for every model in the catalog.

import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.oxyy.ai/v1",
    api_key=os.environ["OXYY_API_KEY"],
)

response = client.chat.completions.create(
    model="grok-uncensored",
    messages=[
        {"role": "user", "content": "Explain unified AI APIs in one sentence."},
    ],
)
print(response.choices[0].message.content)

Performance

Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better). Throughput and TTFT are measured on streaming requests, which are the only ones with a first-token moment to time.

Throughput
75tok/s
P50, streaming requests
Latency
7.46s
P50, end to end
UTC

Throughput

P9991 tok/s
P9587 tok/s
P5061 tok/s

Latency

P5014.5s
P7519.0s
P9524.6s
P9925.7s

Time to first token

P502359ms
P952946ms

Cache hit rate

54.04%
Share of input tokens served from the prompt cache over the window.

Availability

The share of requests to this model that completed successfully over the last 30 days. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.

Success rate (30 days)
63.40%
over 770 requests
Days with traffic
24
of the last 30 days

Days are UTC days. Days with no completed requests are omitted rather than drawn at 100% — no traffic is not evidence of availability.

Activity

Token volume and request traffic to this model over time. Daily totals, UTC.

Prompt6.0K
Reasoning3.7K
Completion7.7K

Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length.

Frequently asked questions

What is Grok Uncensored?
xAI frontier model for coding, knowledge work and STEM, co-trained with Cursor and tuned for agentic tool calling.
How much does Grok Uncensored cost?
Grok Uncensored costs $0.800 per million input tokens and $2.40 per million output tokens on Oxyy. These are the pay-as-you-go prices after the current 60% discount (Launch Promo; standard price $2.00 input and $6.00 output), and cached input tokens are billed at $0.120 per million. You pay per request, with no subscription.
What is the context length of Grok Uncensored?
Grok Uncensored has a context window of 500,000 tokens.
Does Grok Uncensored support tool calling and structured outputs?
Yes — Grok Uncensored supports both tool calling and structured outputs.
What inputs and outputs does Grok Uncensored support?
Grok Uncensored accepts text and images as input and returns text.
Which API endpoint does Grok Uncensored use?
Send a POST request to https://api.oxyy.ai/v1/chat/completions with "model": "grok-uncensored", authenticated with your Oxyy API key.
When was Grok Uncensored released?
Grok Uncensored was released on Jul 8, 2026.
What other models does xAI have?
xAI also offers Grok 4.20 (0309) Non-Reasoning, Grok 4.20 (0309) Reasoning, Grok 4.20 (0309) Multi-Agent and Grok 4.3 through Oxyy.

Alternatives to Grok Uncensored

Models of the same kind from other vendors, closest in price first. Prices are pay-as-you-go on Oxyy, through the same API key.

Explore more models

More models from xAI