Whisper
Open-source multilingual speech recognition.
Open-weight speech recognition model family supporting ~100 languages; widely embedded in local apps.
- Vendor
- OpenAI
- Category
- Voice & audio
- Pricing
- Open source
- Open source
- Yes
- License
- MIT
- Platforms
- Python, Local
- Launched
- 2022
- Website
- github.com/openai/whisper
Features
- Multilingual STT
- Translation
- Runs locally
Best for
- Local & on-device
- Multimodal
More voice & audio
- ElevenLabs — ElevenLabs. Lifelike AI voice, dubbing and voice agents.
- Suno — Suno. Make any song you can imagine.
- Deepgram — Deepgram. Speech-to-text and voice AI APIs.
- LiveKit Agents — LiveKit. Framework for real-time voice AI agents.
- AssemblyAI — AssemblyAI. Speech AI models for transcription and understanding.
- Vapi — Vapi. Voice AI agents for developers.
- Pipecat — Daily. Open-source framework for real-time voice and multimodal agents.
- Cartesia — Cartesia. Ultra-low-latency voice API for real-time agents.