Course Details
Contents
Feature Extraction: Acoustic theory of speech production and parametric representation of speech signal
Automatic Speech Recognition: Template matching approaches, hidden Markov models, deep acoustic modeling, language modeling
Speaker Recognition: Gaussian mixture modeling, universal background models, minimum divergence criteria, probabilistic LDA, system building
Speech Synthesis: Text analysis, Pronunciation, prosody, waveform generation using unit selection, HTS and wavenets, voice building and modification.