
Chapter | 3 Speech Processing
57
however. State-of-the art systems are very sensitive to both intrinsic
(for instance, non-native speech) and extrinsic (background noise)
sources of variation. For instance, word error rates are still as high
as 10% for digit strings recognition in a car driven on a highway and
using a microphone positioned in the dashboard, while it is below
1% when the noise level is very low. Current research is nevertheless
attempting to tackle these vulnerabilities, as well as the assumptions
and approximations inherent in current technology, with a long-
term goal of reaching human-like capabilities. This will hopefully
be evidenced ...