
156
PART|II Multimodal Signal Processing and Modelling
chapter focuses on issues raised by combining multiple modalities
in HCI systems. Combining several systems has been investi-
gated in pattern recognition [2] in general; in applications related
to audio-visual speech processing [3–5]; in speech recognition –
examples of methods are multiband [6], multistream [7, 8], front-
end multi-feature [9] approaches and the union model [10]; in the
form of ensemble [11]; in audio-visual person authentication [12];
and in multibiometrics [1,13–16] (and references herein), among
others. In fact, for audio-visual person authentication, one of the ear-
liest works ...