1/ Audio is now first-class on OpenRouter. Two new endpoints live...

Specialized models are dramatically faster and cheaper than general audio LLMs for the narrow jobs of "read this aloud" and "transcribe this file." Both paths are available so you can pick the right one per use case.
Transcription launches with OpenAI Whisper, GPT-4o Transcribe, Google Chirp 3, and Groq's fast Whisper inference. openrouter.ai/models?output_β¦
OpenAI speech models accept an instructions field for tone control ("speak in a warm, friendly tone").
Transcription accepts a language hint that meaningfully improves non-English accuracy. Full provider surface area, no per-vendor rewrites.
One API, one bill, automatic fallbacks, easy observability. Announcement: openrouter.ai/announcements/β¦




