oxyy.ai
HomeModelsProvidersPricingBlogDocs
Sign inGet API key
  1. Home/
  2. Providers/
  3. Google/
  4. Gemini 2.5 Pro TTS

Google: Gemini 2.5 Pro TTS

google/gemini-2.5-pro-preview-tts
Compare Playground Get API key
High-fidelity speech synthesis optimized for quality in structured workflows such as podcasts and audiobooks, with more natural outputs and easier-to-steer prompts.
ModalitiesTextAudio
In / out price$0.400 / — per 1M
Context8K
ReleasedMay 20, 2025
ProvidersPricingPerformanceAvailabilityAppsActivityFAQExplore

Providers

Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.

ProviderInput /MOutput /MCache read /MLatencyThroughput
Google$0.400————

Capabilities (7)

StreamingAudio GenerationVoice InstructionsMulti SpeakerLanguage DetectionSynthid WatermarkBatch

Supported parameters (5)

speech_config.voice_namespeech_config.multi_speaker_voice_configresponse_modalitiessystem_instructiontemperature

Any other parameter you send is ignored rather than rejected.

Pricing

What this model costs to run, next to the rate it is posted at. Caching and discounts mean the price actually paid is often below the listed one.

Effective input price
$0.400
/M tokens · listed $1.00
Effective output price
—
/M tokens
RateListedChargedUnit
Input$1.00$0.400per 1M tokens
UTC

Performance

Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better). Throughput and TTFT are measured on streaming requests, which are the only ones with a first-token moment to time.

Throughput
—tok/s
P50, streaming requests
Latency
—s
P50, end to end
UTC
No requests to this model in the selected window, so there is nothing to measure yet.

Availability

The share of requests to this model that completed successfully. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.

Success rate (30d)
69.60%
over 23 requests
Days with traffic
5
of the last 30

Days are UTC days. Days with no requests are omitted rather than drawn at 100% — no traffic is not evidence of availability.

Apps

Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for.

No apps have identified themselves for this model yet. Clients can opt in by sending an X-Title header with their request.

Activity

Token volume and request traffic to this model over time. Daily totals, UTC.

Prompt0
Completion0

Prompt tokens measure input size. Reasoning tokens show internal thinking before a response. Completion tokens reflect total output length. Reasoning is not priced separately for this model, so it is not itemised here.

Frequently asked questions

What is Gemini 2.5 Pro TTS?
High-fidelity speech synthesis optimized for quality in structured workflows such as podcasts and audiobooks, with more natural outputs and easier-to-steer prompts.
How much does Gemini 2.5 Pro TTS cost?
Gemini 2.5 Pro TTS costs $1.00 per million input tokens and — per million output tokens. You pay per request, with no subscription.
What is the context length of Gemini 2.5 Pro TTS?
Gemini 2.5 Pro TTS has a context window of 8,000 tokens, and can return up to 16,000 tokens in one response.
Does Gemini 2.5 Pro TTS support tool calling and structured outputs?
Neither tool calling nor structured outputs are listed among this model's capabilities. Any unsupported parameter you send is ignored rather than rejected.
What inputs and outputs does Gemini 2.5 Pro TTS support?
Gemini 2.5 Pro TTS accepts text as input and returns audio.
What other models does Google have?
Google also offers Gemini 2.5 Flash, Gemini 2.5 Flash-Lite, Gemini 2.5 Flash TTS, Gemini 2.5 Pro through Oxyy.
When was Gemini 2.5 Pro TTS released?
Gemini 2.5 Pro TTS was released on May 20, 2025.

Explore more models

Browse the full catalog ModelsCompare pricing Pricing

More models from Google

Gemini 2.5 FlashGoogle's first hybrid reasoning model with configurable thinking budgets; best price-performance for low-latency, high-…text · 1.05M context · $0.300 / $2.50Gemini 2.5 Flash-LiteSmallest and most cost-effective multimodal model in the 2.5 family, built for at-scale usage.text · 1.05M context · $0.100 / $0.400Gemini 2.5 Flash TTSFast and controllable text-to-speech for low-latency, cost-efficient applications and real-time assistants, with fine c…audio tts · 8K context · $0.500 / —Gemini 2.5 ProMost advanced model of the 2.5 family, with deep reasoning and coding capability for complex tasks.text · 1.05M context · $1.25 / $10.00Nano Banana 2High-efficiency production-scale image generation and editing, balancing speed with 4K generation, world knowledge and…image generation · — context · $0.500 / $3.00Gemini 3.1 Flash-LiteCost-efficient multimodal model for high-volume agentic tasks, translation, and simple data extraction where budget and…text · 1.05M context · $0.250 / $1.50
oxyy.ai

One endpoint for every major model. The OpenAI, Anthropic and Gemini SDKs work unchanged.

Product

  • Models
  • Providers
  • Pricing
  • Status
  • Startup program

Company

  • About
  • Blog
  • Contact
  • Terms of Service
  • Privacy Policy
  • Refund Policy
  • Cookie Policy

Developer

  • Documentation
  • Quickstart
  • SDKs
  • Model catalog
  • AI providers

Connect

  • Telegram
  • Discord
  • WhatsApp
© 2026 Oxyy.ai. All rights reserved.StatusAbout