OpenAI: Whisper large-v3

openai/whisper-large-v3

OpenAI Whisper large-v3 API

Oxyy has not published a price for Whisper large-v3 yet. It takes audio as input and returns text, from audio files of up to 50 MB. For every OpenAI model and rate, see OpenAI API pricing.

Third-generation large Whisper model for multilingual speech recognition and translation.
ModalitiesAudioText
PriceNot yet publishedAll OpenAI API pricing

Providers

Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.

ProviderPriceLatencyThroughput
OpenAINot yet published3.60s—

Capabilities (5)

Language DetectionTranslationWord TimestampsLocal InferenceFine Tunable

Supported parameters (4)

tasklanguagetemperaturereturn_timestamps

Any other parameter you send is ignored rather than rejected.

Pricing

Models on Oxyy are billed pay as you go, with no subscription. This one has no published price yet.

Price
Not yet published
Oxyy has not published a price for this model yet.
RateStandard priceChargedUnit
Price not yet published

Specifications

Model ID
whisper-large-v3
Developer
OpenAI
Type
Speech to text
Input
Audio
Output
Text
Tokenizer
GPT
Series
Whisper
Audio input
Up to 50 MB each
Endpoint
/v1/audio/transcriptions

Best for

Speech to textOpen weightsTranscriptionMultilingualLocal inference

Quickstart

Call whisper-large-v3 with a POST to /v1/audio/transcriptions on https://api.oxyy.ai, using your Oxyy API key from the OXYY_API_KEY environment variable. The same key works for every model in the catalog.

import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.oxyy.ai/v1",
    api_key=os.environ["OXYY_API_KEY"],
)

with open("interview.mp3", "rb") as audio_file:
    response = client.audio.transcriptions.create(
        model="whisper-large-v3",
        file=audio_file,
    )
print(response.text)

Availability

The share of requests to this model that completed successfully over the last 30 days. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.

Requests (30 days)
19
A success rate is shown from 100 completed requests
Days with traffic
8
of the last 30 days

Frequently asked questions

What is Whisper large-v3?
Third-generation large Whisper model for multilingual speech recognition and translation.
What audio files can Whisper large-v3 transcribe?
Whisper large-v3 accepts audio uploads up to 50 MB each.
How do I call Whisper large-v3 through the API?
Send a POST request to https://api.oxyy.ai/v1/audio/transcriptions with "model": "whisper-large-v3", authenticated with your Oxyy API key. It is also served on /v1/audio/translations. Upload the audio as multipart form data in the "file" field.
What other models does OpenAI have?
OpenAI also offers GPT-Live-Transcribe, GPT-Realtime-Translate, GPT-Realtime-Whisper and GPT-Transcribe through Oxyy.

Alternatives to Whisper large-v3

Models of the same kind from other vendors, closest in price first. Prices are pay-as-you-go on Oxyy, through the same API key.

Explore more models

More models from OpenAI