Voice
The audio endpoints transcribe speech to text. Transcribe a complete recording in one request, or stream live audio over a WebSocket and receive transcripts as the audio arrives.
For modes, speaker labels, endpointing, and worked examples, see the Speech to text feature page.