Skip to main content
POST
The request and response follow OpenAI’s audio/translations API, so OpenAI SDKs work unchanged. The output is always English. The model is the name of a speech-to-text deployment that supports translation.

Headers

Form Data

There is no language field: the source language is detected, and the target is always English.

Which deployments can translate

Translation is a capability of the deployment, not of the endpoint. A deployment of a provider that cannot translate returns 404 with the message Model '…' not found or does not support audio_translation, even though the same deployment works on Transcribe Audio. Providers that can translate an uploaded file to English are OpenAI (whisper-1) and Groq (with a model that supports translation, such as whisper-large-v3). Other providers, including Deepgram, ElevenLabs, AssemblyAI, Gladia and Speechmatics, do not translate uploads; use Transcribe Audio for them.

Response headers

Errors

Errors use OpenAI’s error format, and the status codes are the same as for Transcribe Audio.
404