🔊Introducing Voxtral TTS: our new frontier open-weight model for...

🎭Realistic, emotionally expressive speech.
🌍Supports 9 languages and accurately captures diverse dialects.
⚡Very low latency for time-to-first-audio.
🔄Easily adaptable to new voices
✅ Full audio intelligence: Works with Voxtral Transcribe for end-to-end speech-to-speech, or plugs into any STT + LLM stack.
✅ Built for business: From customer support to real-time translation, it’s the output layer that passes the human test.
🎥 See it in action:
