| 2008 | ICASSP | MAAI: Media analytics for actionable intelligence. | Upendra V. Chaudhari, Sarah Conrod, Alexander Faisman, Giridharan Iyengar, Dimitri Kanevsky, Mark Kogan, Ganesh N. Ramaswamy, Paola Virga |
| 2007 | ICASSP | Unsupervised Audio Segmentation using Extended Baum-Welch Transformations. | Tara N. Sainath, Dimitri Kanevsky, Giridharan Iyengar |
| 2005 | ICASSP | Semantic Annotation of Multimedia using Maximum Entropy Models. | Janne Argillander, Giridharan Iyengar, Harriet J. Nock |
| 2004 | ICASSP | Multimodal video search techniques: late fusion of speech-based retrieval and visual content-based retrieval. | Arnon Amir, Giridharan Iyengar, Ching-Yung Lin, Milind R. Naphade, Apostol Natsev, Chalapathy Neti, Harriet J. Nock, John R. Smith, Belle L. Tseng |
| 2004 | ICASSP | News video story segmentation using fusion of multi-level multi-modal features in TRECVID 2003. | Winston H. Hsu, Lyndon S. Kennedy, Chih-Wei Huang, Shih-Fu Chang, Ching-Yung Lin, Giridharan Iyengar |
| 2004 | ICASSP | Improved face and feature finding for audio-visual speech recognition in visually challenging environments. | Jintao Jiang, Gerasimos Potamianos, Harriet J. Nock, Giridharan Iyengar, Chalapathy Neti |
| 2003 | ICASSP | Audio-visual synchrony for detection of monologues in video archives. | Giridharan Iyengar, Harriet J. Nock, Chalapathy Neti |
| 2003 | Interspeech | Impact of audio segmentation and segment clustering on automated transcription accuracy of large spoken archives. | Bhuvana Ramabhadran, Jing Huang, Upendra V. Chaudhari, Giridharan Iyengar, Harriet J. Nock |
| 2001 | Interspeech | Large-vocabulary audio-visual speech recognition by machines and humans. | Gerasimos Potamianos, Chalapathy Neti, Giridharan Iyengar, Eric Helmuth |
| 2001 | MMSP | Detection of faces under shadows and lighting variations. | Giridharan Iyengar, Chalapathy Neti |
| 2001 | MMSP | Robust detection of visual ROI for automatic speechreading. | Giridharan Iyengar, Gerasimos Potamianos, Chalapathy Neti, Tanveer A. Faruquie, Ashish Verma |
| 2000 | ICIP | Distributional Clustering for Efficient Content-Based Retrieval of Images and Video. | Giridharan Iyengar, Andrew Lippman |
| 2000 | Interspeech | Perceptual interfaces for information interaction: joint processing of audio and visual information for human-computer interaction. | Chalapathy Neti, Giridharan Iyengar, Gerasimos Potamianos, Andrew W. Senior, Benot Maison |
| 1996 | ICIP | VideoBook: an experiment in characterization of video. | Giridharan Iyengar, Andrew Lippman |