| 2023 | ASRU | Multitask Learning Model with Text and Speech Representation for Fine-Grained Speech Scoring. | Seongjin Park, Rutuja Ubale |
| 2022 | ICMI | Predicting User Confidence in Video Recordings with Spatio-Temporal Multimodal Analytics. | Andrew Emerson, Patrick Houghton, Ke Chen, Vinay Basheerabad, Rutuja Ubale, Chee Wee Leong |
| 2019 | ASRU | Native Language Identification from Raw Waveforms Using Deep Convolutional Neural Networks with Attentive Pooling. | Rutuja Ubale, Vikram Ramanarayanan, Yao Qian, Keelan Evanini, Chee Wee Leong, Chong Min Lee |
| 2019 | ICASSP | Neural Approaches to Automated Speech Scoring of Monologue and Dialogue Responses. | Yao Qian, Patrick L. Lange, Keelan Evanini, Robert A. Pugh, Rutuja Ubale, Matthew Mulholland, Xinhao Wang |
| 2019 | ICMI | Are Humans Biased in Assessment of Video Interviews? | Chee Wee Leong, Katrina Roohr, Vikram Ramanarayanan, Michelle P. Martin-Raugh, Harrison Kell, Rutuja Ubale, Yao Qian, Zydrune Mladineo, Laura McCulla |
| 2018 | Interspeech | Improvements to an Automated Content Scoring System for Spoken CALL Responses: the ETS Submission to the Second Spoken CALL Shared Task. | Keelan Evanini, Matthew Mulholland, Rutuja Ubale, Yao Qian, Robert A. Pugh, Vikram Ramanarayanan, Aoife Cahill |
| 2018 | Interspeech | Toward Scalable Dialog Technology for Conversational Language Learning: Case Study of the TOEFL MOOC. | Vikram Ramanarayanan, David Pautler, Patrick L. Lange, Eugene Tsuprun, Rutuja Ubale, Keelan Evanini, David Suendermann-Oeft |
| 2017 | ASRU | Improving native language (L1) identifation with better VAD and TDNN trained separately on native and non-native English corpora. | Yao Qian, Keelan Evanini, Patrick L. Lange, Robert A. Pugh, Rutuja Ubale, Frank K. Soong |
| 2017 | ASRU | Exploring ASR-free end-to-end modeling to improve spoken language understanding in a cloud-based dialog system. | Yao Qian, Rutuja Ubale, Vikram Ramanarayanan, Patrick L. Lange, David Suendermann-Oeft, Keelan Evanini, Eugene Tsuprun |