Leading streaming word error rate, and it holds that accuracy on your vocabulary through contextual and keyword biasing, with no fine-tuning.
Live speaker and turn awareness
Speaker attribution for 20+ speakers and end-of-speech detection happen inside the recognition model, not in a batch pass at the end of the audio stream.
Production pricing
Ship with leading transcription accuracy and streaming capabilities without compromising on costs.
Price / 1,000 min
Price / hour
muse-voice-transcribe-1.0
$3.00
$0.18
Quick startYour first working request in under five minutes. Point your existing OpenAI SDK compatible client at Meta Model API.