Crossroads
Task Integration in Multimodal Speech Recognition Environments
April 1, 1997
The benefits of speech-driven user interfaces have been advocated for several years. Speech is a natural form of communication that is pervasive, efficient, and can be used at a distance. However, widespread acceptance of speech as a human-computer interface has yet to occur. Taking this into account, several research efforts have begun to focus on speech as an ancillary input channel in multimodal environments. A model of complementary behavior has been proposed based on arguments that direct manipulation and speech recognition interfaces have reciprocal strengths and weaknesses. This suggests that user interface performance and acceptance may increase by adopting a multimodal approach that combines speech and direct manipulation. More theoretical work is needed in order to understand how to leverage this advantage. In this paper, a framework is presented to empirically evaluate the types of tasks that might benefit from such a multimodal interface.
Article
ACM
3
3
DOI: 10.1109/ICMI.2002.1166981
Downloads: 3065 downloads
Google Scholar Citations: 10 citations