OpenAI Releases 1.6 Billion Parameter Multilingual Speech Recognition AI Whisper

Опубликовано: 24 Июнь 2026
на канале: Tech Flash News
22
0

OpenAI recently released Whisper, a 1.6 billion parameter AI model that can transcribe and translate speech audio from 97 different languages. Whisper was trained on 680,000 hours of audio data collected from the web and shows robust zero-shot performance on a wide range of automated speech recognition tasks. Unlike most state-of-the-art ASR models, Whisper is not fine-tuned on any benchmark dataset; instead, it is trained using "weak" supervision on a large-scale, noisy dataset of speech audio ...

#decoder #automated #436k #wav2vec #internet #nvidia #hacker

https://techflash.news/en/TFN-376-ai-en