Gateway APIs
Translate Audio
Convert an audio file to English text, whatever language it is spoken in.
POST
audio/translations API, so OpenAI SDKs work unchanged.
The output is always English. The model is the name of a speech-to-text deployment that supports
translation.
Headers
Form Data
There is no
language field: the source language is detected, and the target is always English.
Which deployments can translate
Translation is a capability of the deployment, not of the endpoint. A deployment of a provider that cannot translate returns404 with the message Model '…' not found or does not support audio_translation, even though the same deployment works on
Transcribe Audio.
Providers that can translate an uploaded file to English are OpenAI (whisper-1) and Groq (with a
model that supports translation, such as whisper-large-v3). Other providers, including Deepgram,
ElevenLabs, AssemblyAI, Gladia and Speechmatics, do not translate uploads; use
Transcribe Audio for them.
Response headers
Errors
Errors use OpenAI’s error format, and the status codes are the same as for Transcribe Audio.404