| 2003 | Combining speech and haptics for intuitive and efficient navigation through image databases. | Thomas Kster, Michael Pfeiffer, Christian Bauckhage |
| 2003 | Mutual disambiguation of 3D multimodal interaction in augmented and virtual reality. | Edward C. Kaiser, Alex Olwal, David McGee, Hrvoje Benko, Andrea Corradini, Xiaoguang Li, Philip R. Cohen, Steven Feiner |
| 2003 | Multimodal user interfaces: who's the user? | Anil K. Jain |
| 2003 | Learning and reasoning about interruption. | Eric Horvitz, Johnson Apacible |
| 2003 | Distributed and local sensing techniques for face-to-face collaboration. | Ken Hinckley |
| 2003 | Towards robust person recognition on handheld devices using face and speaker identification technologies. | Timothy J. Hazen, Eugene Weinstein, Alex Park |
| 2003 | A visually grounded natural language interface for reference to spatial scenes. | Peter Gorniak, Deb Roy |
| 2003 | Augmenting user interfaces with adaptive speech commands. | Peter Gorniak, Deb Roy |
| 2003 | A framework for rapid development of multimodal interfaces. | Frans Flippo, Allan Meng Krebs, Ivan Marsic |
| 2003 | Large vocabulary sign language recognition based on hierarchical decision trees. | Gaolin Fang, Wen Gao, Debin Zhao |
| 2003 | Architecture and implementation of multimodal plug and play. | Christian Elting, Stefan Rapp, Gregor Mhler, Michael Strube |
| 2003 | Mouthbrush: drawing and painting by hand and mouth. | Chi-Ho Chan, Michael J. Lyons, Nobuji Tetsutani |
| 2003 | A system for fast, full-text entry for small electronic devices. | Saied Bozorgui-Nesbat |
| 2003 | IRYS: a visualization tool for temporal analysis of multimodal interaction. | Daniel Bauer, James D. Hollan |
| 2003 | Interactive skills using active gaze tracking. | Rowel Atienza, Alexander Zelinsky |
| 2003 | Sensitivity to haptic-audio asynchrony. | Bernard D. Adelstein, Durand R. Begault, Mark R. Anderson, Elizabeth M. Wenzel |
| 2003 | Algorithms for controlling cooperation between output modalities in 2D embodied conversational agents. | Sarkis Abrilian, Jean-Claude Martin, Stphanie Buisine |
| 2002 | A Multi-Modal Interface for an Interactive Simulated Vascular Reconstruction System. | Elena V. Zudilova, Peter M. A. Sloot, Robert G. Belleman |
| 2002 | Research of Machine Learning Method for Specific Information Recognition on the Internet. | Dequan Zheng, Yi Hu, Tiejun Zhao, Hao Yu, Sheng Li |
| 2002 | A PDA-Based Sign Translator. | Jing Zhang, Xilin Chen, Jie Yang, Alex Waibel |
| 2002 | A Video Based Interface to Textual Information for the Visually Impaired. | Ali Zandifar, Ramani Duraiswami, Antoine Chahine, Larry S. Davis |
| 2002 | Attentional Object Spotting by Integrating Multimodal Input. | Chen Yu, Dana H. Ballard, Shenghuo Zhu |
| 2002 | Musically Expressive Doll in Face-to-Face Communication. | Tomoko Yonezawa, Kenji Mase |
| 2002 | Hand Gesture Symmetric Behavior Detection and Analysis in Natural Conversation. | Yingen Xiong, Francis K. H. Quek, David McNeill |
| 2002 | Improved Information Maximization based Face and Facial Feature Detection from Real-time Video and Application in a Multi-Modal Person Identification System. | Ziyou Xiong, Yunqiang Chen, Roy Wang, Thomas S. Huang |