OpenAI
OpenAI: Whisper Large V3
openai/whisper-large-v3
Whisper Large V3 is OpenAI's open-source automatic speech recognition model offering both audio transcription and translation. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm, flac, and ogg. With 1,550M parameters, it achieves a 10.3% word error rate and is well-suited for noise-robust, multilingual transcription in demanding conditions. Supports timestamp granularities at word and segment levels.
- Input: audio
- Output: transcription
View on OpenRouter. Model data sourced from OpenRouter.