| 2023 | ICASSP | Multi-Scale Compositional Constraints for Representation Learning on Videos. | Georgios Paraskevopoulos, Chandrashekhar Lavania, Lovish Chum, Shiva Sundaram |
| 2022 | ICASSP | Enhancing Contrastive Learning with Temporal Cognizance for Audio-Visual Representation Generation. | Chandrashekhar Lavania, Shiva Sundaram, Sundararajan Srinivasan, Katrin Kirchhoff |
| 2022 | ICIP | Scene Representation Learning from Videos Using Self-Supervised and Weakly-Supervised Techniques. | Raghuveer Peri, Srinivas Parthasarathy, Shiva Sundaram |
| 2021 | ICASSP | Audiovisual Highlight Detection in Videos. | Karel Mundnich, Alexandra Fenster, Aparna Khare, Shiva Sundaram |
| 2021 | ICASSP | Disentanglement for Audio-Visual Emotion Recognition Using Multitask Setup. | Raghuveer Peri, Srinivas Parthasarathy, Charles Bradshaw, Shiva Sundaram |
| 2020 | ACL | Multimodal and Multiresolution Speech Recognition with Transformers. | Georgios Paraskevopoulos, Srinivas Parthasarathy, Aparna Khare, Shiva Sundaram |
| 2020 | ICASSP | Robust Multi-Channel Speech Recognition Using Frequency Aligned Network. | Taejin Park, Ken'ichi Kumatani, Minhua Wu, Shiva Sundaram |
| 2020 | ICASSP | Fully Learnable Front-End for Multi-Channel Acoustic Modeling Using Semi-Supervised Learning. | Sanna Wager, Aparna Khare, Minhua Wu, Ken'ichi Kumatani, Shiva Sundaram |
| 2020 | ICMI | Training Strategies to Handle Missing Modalities for Audio-Visual Expression Recognition. | Srinivas Parthasarathy, Shiva Sundaram |
| 2020 | Interspeech | Multi-Modal Embeddings Using Multi-Task Learning for Emotion Recognition. | Aparna Khare, Srinivas Parthasarathy, Shiva Sundaram |
| 2019 | ICASSP | Multi-geometry Spatial Acoustic Modeling for Distant Speech Recognition. | Ken'ichi Kumatani, Minhua Wu, Shiva Sundaram, Nikko Strm, Bjrn Hoffmeister |
| 2019 | ICASSP | Improving Noise Robustness of Automatic Speech Recognition via Parallel Data and Teacher-student Learning. | Ladislav Mosner, Minhua Wu, Anirudh Raju, Sree Hari Krishnan Parthasarathi, Ken'ichi Kumatani, Shiva Sundaram, Roland Maas, Bjrn Hoffmeister |
| 2019 | ICASSP | Frequency Domain Multi-channel Acoustic Modeling for Distant Speech Recognition. | Minhua Wu, Ken'ichi Kumatani, Shiva Sundaram, Nikko Strm, Bjrn Hoffmeister |
| 2018 | Interspeech | Detecting Media Sound Presence in Acoustic Scenes. | Constantinos Papayiannis, Justice Amoh, Viktor Rozgic, Shiva Sundaram, Chao Wang |
| 2013 | Interspeech | Affective classification of generic audio clips using regression models. | Nikos Malandrakis, Shiva Sundaram, Alexandros Potamianos |
| 2012 | ICASSP | Latent perceptual mapping with data-driven variable-length acoustic units for template-based speech recognition. | Shiva Sundaram, Jerome R. Bellegarda |
| 2011 | ICASSP | Experiments in context-independent recognition of non-lexical 'yes' or 'no' responses. | Shiva Sundaram, Robert Schleicher, Nathalie Diehl |
| 2010 | ICASSP | Using nave text queries for robust audio information retrieval. | Samuel Kim, Panayiotis G. Georgiou, Shrikanth S. Narayanan, Shiva Sundaram |
| 2010 | Interspeech | Latent perceptual mapping: a new acoustic modeling framework for speech recognition. | Shiva Sundaram, Jerome R. Bellegarda |
| 2010 | MMSP | An N-gram model for unstructured audio signals toward information retrieval. | Samuel Kim, Shiva Sundaram, Panayiotis G. Georgiou, Shrikanth S. Narayanan |
| 2009 | Interspeech | Emotion classification in children's speech using fusion of acoustic and linguistic features. | Tim Polzehl, Shiva Sundaram, Hamed Ketabdar, Michael Wagner, Florian Metze |
| 2009 | MMSP | Saliency-driven unstructured acoustic scene classification using latent perceptual indexing. | Ozlem Kalinli, Shiva Sundaram, Shrikanth S. Narayanan |
| 2008 | ICASSP | Audio retrieval by latent perceptual indexing. | Shiva Sundaram, Shrikanth S. Narayanan |
| 2007 | ICASSP | Discriminating Two Types of Noise Sources using Cortical Representation and Dimension Reduction Technique. | Shiva Sundaram, Shrikanth S. Narayanan |
| 2007 | ICASSP | Analysis of Audio Clustering using Word Descriptions. | Shiva Sundaram, Shrikanth S. Narayanan |
| 2007 | MMSP | Experiments in Automatic Genre Classification of Full-length Music Tracks using Audio Activity Rate. | Shiva Sundaram, Shrikanth S. Narayanan |
| 2006 | ICASSP | Speech Recognition Engineering Issues in Speech to Speech Translation System Design for Low Resource Languages and Domains. | Shrikanth S. Narayanan, Panayiotis G. Georgiou, Abhinav Sethy, Dagen Wang, Murtaza Bulut, Shiva Sundaram, Emil Ettelaie, Sankaranarayanan Ananthakrishnan, Horacio Franco, Kristin Precoda, Dimitra Vergyri, Jing Zheng, Wen Wang, Venkata Ramana Rao Gadde, Martin Graciarena, Victor Abrash, Michael W. Frandsen, Colleen Richey |
| 2006 | MMSP | An attribute-based approach to audio description applied to segmenting vocal sections in popular music songs. | Shiva Sundaram, Shrikanth S. Narayanan |
| 2003 | Interspeech | An empirical text transformation method for spontaneous speech synthesizers. | Shiva Sundaram, Shrikanth S. Narayanan |