Arushi Goel
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
14
Venues
10
Active years
2019–2026
Best venue rank
A*
Where they publish
Papers
14 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | ACL | Speech-Hands: A Self-Reflection Voice Agentic Approach to Speech Recognition and Audio Reasoning with Omni Perception. | Zhen Wan, Chao-Han Huck Yang, Jinchuan Tian, Hanrong Ye, Ankita Pasad, Szu-Wei Fu, Arushi Goel, Ryo Hachiuma, Shizhe Diao, Kunal Dhawan, Sreyan Ghosh, Yusuke Hirota, Zhehuai Chen, Rafael Valle, Chenhui Chu, Shinji Watanabe, Boris Ginsburg, Yu-Chiang Frank Wang |
| 2025 | CVPR | Visually Interpretable Subtask Reasoning for Visual Question Answering. | Yu Cheng, Arushi Goel, Hakan Bilen |
| 2025 | ICLR | Fugatto 1: Foundational Generative Audio Transformer Opus 1. | Rafael Valle, Rohan Badlani, Zhifeng Kong, Sang-gil Lee, Arushi Goel, Sungwon Kim, Joo Felipe Santos, Shuqi Dai, Siddharth Gururani, Aya Aljafari, Alexander H. Liu, Kevin J. Shih, Ryan Prenger, Wei Ping, Chao-Han Huck Yang, Bryan Catanzaro |
| 2025 | ICML | ETTA: Elucidating the Design Space of Text-to-Audio Models. | Sang-gil Lee, Zhifeng Kong, Arushi Goel, Sungwon Kim, Rafael Valle, Bryan Catanzaro |
| 2024 | ICML | Audio Flamingo: A Novel Audio Language Model with Few-Shot Learning and Dialogue Abilities. | Zhifeng Kong, Arushi Goel, Rohan Badlani, Wei Ping, Rafael Valle, Bryan Catanzaro |
| 2024 | ICRA | TiV-ODE: A Neural ODE-based Approach for Controllable Video Generation From Text-Image Pairs. | Yucheng Xu, Nanbo Li, Arushi Goel, Zonghai Yao, Zijian Guo, Hamidreza Kasaei, Mohammadreza Kasaei, Zhibin Li |
| 2023 | CoRL | Language-guided Robot Grasping: CLIP-based Referring Grasp Synthesis in Clutter. | Georgios Tziafas, Yucheng Xu, Arushi Goel, Mohammadreza Mohades Kasaei, Zhibin Li, Hamidreza Kasaei |
| 2023 | EMNLP | Semi-supervised multimodal coreference resolution in image narrations. | Arushi Goel, Basura Fernando, Frank Keller, Hakan Bilen |
| 2023 | ICCV | Who are you referring to? Coreference resolution in image narrations. | Arushi Goel, Basura Fernando, Frank Keller, Hakan Bilen |
| 2023 | ICCV | Encyclopedic VQA: Visual questions about detailed properties of fine-grained categories. | Thomas Mensink, Jasper R. R. Uijlings, Llus Castrejn, Arushi Goel, Felipe Cadar, Howard Zhou, Fei Sha, Andr Arajo, Vittorio Ferrari |
| 2022 | CVPR | Not All Relations are Equal: Mining Informative Labels for Scene Graph Generation. | Arushi Goel, Basura Fernando, Frank Keller, Hakan Bilen |
| 2020 | ECCV | Injecting Prior Knowledge into Image Caption Generation. | Arushi Goel, Basura Fernando, Thanh-Son Nguyen, Hakan Bilen |
| 2019 | CICLING | Semantic Roles in VerbNet and FrameNet: Statistical Analysis and Evaluation. | Aliaksandr Huminski, Fiona Liausvia, Arushi Goel |
| 2019 | CVPR | An End-To-End Network for Generating Social Relationship Graphs. | Arushi Goel, Keng Teck Ma, Cheston Tan |