
Intermediate
Speech & Audio AI: Speech & Audio Processing
Comprehensive AI for Speech & Audio course — from audio signal processing, Speech Recognition (ASR) with Whisper, Text-to-Speech (TTS) with VITS, Voice Cloning, Speaker Verification, to Music AI. Practice with Python, PyTorch, Hugging Face, librosa, and state-of-the-art models.
15 lessons45h