
Chapter 3
Speech Processing
Thierry Dutoit and Stéphane Dupont
Faculté Polytechnique de Mons, Belgium
3.1. Introduction 26
3.2. Speech Recognition 28
3.2.1. Feature Extraction 28
3.2.2. Acoustic Modelling 30
3.2.3. Language Modelling 33
3.2.4. Decoding 34
3.2.5. Multiple Sensors 35
3.2.6. Confidence Measures 37
3.2.7. Robustness 38
3.3. Speaker Recognition 40
3.3.1. Overview 40
3.3.2. Robustness 43
3.4. Text-to-Speech Synthesis 44
3.4.1. Natural Language Processing for Speech
Synthesis 44
3.4.2. Concatenative Synthesis with a Fixed Inventory 46
3.4.3. Unit Selection-Based Synthesis 50
3.4.4. Statistical Parametric Synthesis 53
3.5. Conclusions 56
References 57
Multimodal ...