| 2026 | WACV | Face-LLaVA: Facial Expression and Attribute Understanding through Instruction Tuning. | Ashutosh Chaubey, Xulang Guan, Mohammad Soleymani |
| 2025 | EMNLP | Can VLMs Recall Factual Associations From Visual References? | Dhananjay Ashok, Ashutosh Chaubey, Hirona Jacqueline Arai, Jonathan May, Jesse Thomason |
| 2025 | ICCV | Ditailistener: Controllable High Fidelity Listener Video Generation with Diffusion. | Maksim Siniukov, Di Chang, Minh Tran, Hongkun Gong, Ashutosh Chaubey, Mohammad Soleymani |
| 2025 | WACV | ContextIQ: A Multimodal Expert-Based Video Retrieval System for Contextual Advertising. | Ashutosh Chaubey, Anoubhav Agrawal, Sartaki Sinha Roy, Aayush Agrawal, Susmita Ghose |
| 2023 | ASRU | Meta-Learning Framework for End-to-End Imposter Identification in Unseen Speaker Recognition. | Ashutosh Chaubey, Sparsh Sinha, Susmita Ghose |
| 2022 | CVPR | OPAD: An Optimized Policy-based Active Learning Framework for Document Content Analysis. | Sumit Shekhar, Bhanu Prakash Reddy Guda, Ashutosh Chaubey, Ishan Jindal, Avneet Jain |
| 2022 | Interspeech | Improved Relation Networks for End-to-End Speaker Verification and Identification. | Ashutosh Chaubey, Sparsh Sinha, Susmita Ghose |