OpenAI: Whisper large-v3-turbo

openai/whisper-large-v3-turbo

OpenAI Whisper large-v3-turbo API

Oxyy has not published a price for Whisper large-v3-turbo yet. It takes audio as input and returns text, from audio files of up to 50 MB. For every OpenAI model and rate, see OpenAI API pricing.

Fine-tuned, pruned version of large-v3 with the decoder cut from 32 layers to 4, for much faster inference at a small quality cost.
ModalitiesAudioText
PriceNot yet publishedAll OpenAI API pricing

Providers

Where this model runs, and what each route costs. Requests are sent to a healthy provider automatically; if one errors, the gateway retries against another serving the same model.

ProviderPrice
OpenAINot yet published

Capabilities (5)

Language DetectionTranslationWord TimestampsLocal InferenceFine Tunable

Supported parameters (4)

tasklanguagetemperaturereturn_timestamps

Any other parameter you send is ignored rather than rejected.

Pricing

Models on Oxyy are billed pay as you go, with no subscription. This one has no published price yet.

Price
Not yet published
Oxyy has not published a price for this model yet.
RateStandard priceChargedUnit
Price not yet published

Specifications

Model ID
whisper-large-v3-turbo
Developer
OpenAI
Type
Speech to text
Input
Audio
Output
Text
Tokenizer
GPT
Series
Whisper
Audio input
Up to 50 MB each
Endpoint
/v1/audio/transcriptions

Best for

Speech to textOpen weightsTranscriptionMultilingualLocal inference

Quickstart

Call whisper-large-v3-turbo with a POST to /v1/audio/transcriptions on https://api.oxyy.ai, using your Oxyy API key from the OXYY_API_KEY environment variable. The same key works for every model in the catalog.

import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.oxyy.ai/v1",
    api_key=os.environ["OXYY_API_KEY"],
)

with open("interview.mp3", "rb") as audio_file:
    response = client.audio.transcriptions.create(
        model="whisper-large-v3-turbo",
        file=audio_file,
    )
print(response.text)

Availability

The share of requests to this model that completed successfully over the last 30 days. When an upstream provider errors, the gateway retries against another provider serving the same model, so a single provider incident does not necessarily show up here.

Success rate (30 days)
0.60%
over 181 requests
Days with traffic
12
of the last 30 days

Days are UTC days. Days with no completed requests are omitted rather than drawn at 100% — no traffic is not evidence of availability.

Frequently asked questions

What is Whisper large-v3-turbo?
Fine-tuned, pruned version of large-v3 with the decoder cut from 32 layers to 4, for much faster inference at a small quality cost.
What audio files can Whisper large-v3-turbo transcribe?
Whisper large-v3-turbo accepts audio uploads up to 50 MB each.
How do I call Whisper large-v3-turbo through the API?
Send a POST request to https://api.oxyy.ai/v1/audio/transcriptions with "model": "whisper-large-v3-turbo", authenticated with your Oxyy API key. It is also served on /v1/audio/translations. Upload the audio as multipart form data in the "file" field.
What other models does OpenAI have?
OpenAI also offers GPT-Live-Transcribe, GPT-Realtime-Translate, GPT-Realtime-Whisper and GPT-Transcribe through Oxyy.

Alternatives to Whisper large-v3-turbo

Models of the same kind from other vendors, closest in price first. Prices are pay-as-you-go on Oxyy, through the same API key.

Explore more models

More models from OpenAI