| 2025 | ICCV | Local2Global Query Alignment for Video Instance Segmentation. | Rajat Koner, Zhipeng Wang, Srinivas Parthasarathy, Chinghang Chen |
| 2022 | ICIP | Scene Representation Learning from Videos Using Self-Supervised and Weakly-Supervised Techniques. | Raghuveer Peri, Srinivas Parthasarathy, Shiva Sundaram |
| 2021 | ICASSP | Disentanglement for Audio-Visual Emotion Recognition Using Multitask Setup. | Raghuveer Peri, Srinivas Parthasarathy, Charles Bradshaw, Shiva Sundaram |
| 2020 | ACL | Multimodal and Multiresolution Speech Recognition with Transformers. | Georgios Paraskevopoulos, Srinivas Parthasarathy, Aparna Khare, Shiva Sundaram |
| 2020 | ICMI | Training Strategies to Handle Missing Modalities for Audio-Visual Expression Recognition. | Srinivas Parthasarathy, Shiva Sundaram |
| 2020 | Interspeech | Multi-Modal Embeddings Using Multi-Task Learning for Emotion Recognition. | Aparna Khare, Srinivas Parthasarathy, Shiva Sundaram |
| 2019 | ICASSP | Improving Emotion Classification through Variational Inference of Latent Variables. | Srinivas Parthasarathy, Viktor Rozgic, Ming Sun, Chao Wang |
| 2018 | Interspeech | Preference-Learning with Qualitative Agreement for Sentence Level Emotional Annotations. | Srinivas Parthasarathy, Carlos Busso |
| 2018 | Interspeech | Ladder Networks for Emotion Recognition: Using Unsupervised Auxiliary Tasks to Improve Predictions of Emotional Attributes. | Srinivas Parthasarathy, Carlos Busso |
| 2018 | Interspeech | Role of Regularization in the Prediction of Valence from Speech. | Kusha Sridhar, Srinivas Parthasarathy, Carlos Busso |
| 2017 | ACII | Predicting speaker recognition reliability by considering emotional content. | Srinivas Parthasarathy, Carlos Busso |
| 2017 | ICASSP | Ranking emotional attributes with deep neural networks. | Srinivas Parthasarathy, Reza Lotfian, Carlos Busso |
| 2017 | ICASSP | A study of speaker verification performance with expressive speech. | Srinivas Parthasarathy, Chunlei Zhang, John H. L. Hansen, Carlos Busso |
| 2017 | Interspeech | Jointly Predicting Arousal, Valence and Dominance with Multi-Task Learning. | Srinivas Parthasarathy, Carlos Busso |
| 2016 | ICASSP | Automatic composition of broadcast news summaries using rank classifiers trained with acoustic and lexical features. | Taufiq Hasan, Mohammed Abdel-Wahab, Srinivas Parthasarathy, Carlos Busso, Yang Liu |
| 2016 | Interspeech | Defining Emotionally Salient Regions Using Qualitative Agreement Method. | Srinivas Parthasarathy, Carlos Busso |
| 2015 | ICASSP | Automatic broadcast news summarization via rank classifiers and crowdsourced annotation. | Srinivas Parthasarathy, Taufiq Hasan |