Technology
OpenAI Whisper
Whisper is OpenAI's robust, open-source Automatic Speech Recognition (ASR) system, trained on 680,000 hours of diverse audio.
This is Whisper: a high-performance, general-purpose ASR model from OpenAI. It was trained on a massive 680,000 hours of multilingual, multitask data, resulting in exceptional robustness against accents, background noise, and technical language. The model is a Transformer sequence-to-sequence architecture, engineered for multiple tasks: multilingual transcription, speech-to-English translation, and language identification. Developers leverage the open-source code and various model sizes (tiny, base, small, medium, large) to balance transcription speed with near human-level accuracy for diverse applications.
What builders pair with OpenAI Whisper
Projects using both technologies. Select a pairing to see a project.
12 more pairings
Pairing: Llama 3
Aura: A Locally Hosted AI Gaming Companion
Pairing: PyTorch
AI and music - entertaining people's ears with AI
Pairing: React
AI powered agent to improve public speaking sklils
Pairing: Amazon Polly
UzbekVoice
Pairing: Amazon Rekognition
Easy indexing of NASCAR archived footage
Pairing: Amazon S3
Easy indexing of NASCAR archived footage
Recent Talks & Demos
Showing 1-10 of 10