Yale Song
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
42
Venues
15
Active years
2012–2026
Best venue rank
A*
Where they publish
Papers
42 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | WACV | Enhancing Visual Planning with Auxiliary Tasks and Multi-token Prediction. | Ce Zhang, Yale Song, Ruta Desai, Michael Louis Iuzzolino, Joseph Tighe, Gedas Bertasius, Satwik Kottur |
| 2025 | CVPR | VITED: Video Temporal Evidence Distillation. | Yujie Lu, Yale Song, William Wang, Lorenzo Torresani, Tushar Nagarajan |
| 2025 | ICCV | Streaming Videollms for Real-Time Procedural Video Understanding. | Dibyadip Chatterjee, Edoardo Remelli, Yale Song, Bugra Tekin, Abhay Mittal, Bharat Bhatnagar, Necati Cihan Camgz, Shreyas Hampali, Eric Sauser, Shugao Ma, Angela Yao, Fadime Sener |
| 2025 | ICCV | Enrich and Detect: Video Temporal Grounding With Multimodal Llms. | Shraman Pramanick, Effrosyni Mavroudi, Yale Song, Rama Chellappa, Lorenzo Torresani, Triantafyllos Afouras |
| 2024 | CVPR | Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives. | Kristen Grauman, Andrew Westbury, Lorenzo Torresani, Kris Kitani, Jitendra Malik, Triantafyllos Afouras, Kumar Ashutosh, Vijay Baiyya, Siddhant Bansal, Bikram Boote, Eugene Byrne, Zachary Chavis, Joya Chen, Feng Cheng, Fu-Jen Chu, Sean Crane, Avijit Dasgupta, Jing Dong, Mara Escobar, Cristhian Forigua, Abrham Gebreselasie, Sanjay Haresh, Jing Huang, Md Mohaiminul Islam, Suyog Dutt Jain, Rawal Khirodkar, Devansh Kukreja, Kevin J. Liang, Jia-Wei Liu, Sagnik Majumder, Yongsen Mao, Miguel Martin, Effrosyni Mavroudi, Tushar Nagarajan, Francesco Ragusa, Santhosh Kumar Ramakrishnan, Luigi Seminara, Arjun Somayazulu, Yale Song, Shan Su, Zihui Xue, Edward Zhang, Jinxu Zhang, Angela Castillo, Changan Chen, Xinzhu Fu, Ryosuke Furuta, Cristina Gonzlez, Prince Gupta, Jiabo Hu, Yifei Huang, Yiming Huang, Weslie Khoo, Anush Kumar, Robert Kuo, Sach Lakhavani, Miao Liu, Mi Luo, Zhengyi Luo, Brighid Meredith, Austin Miller, Oluwatumininu Oguntola, Xiaqing Pan, Penny Peng, Shraman Pramanick, Merey Ramazanova, Fiona Ryan, Wei Shan, Kiran K. Somasundaram, Chenan Song, Audrey Southerland, Masatoshi Tateno, Huiyu Wang, Yuchen Wang, Takuma Yagi, Mingfei Yan, Xitong Yang, Zecheng Yu, Shengxin Cindy Zha, Chen Zhao, Ziwei Zhao, Zhifan Zhu, Jeff Zhuo, Pablo Arbelez, Gedas Bertasius, Dima Damen, Jakob J. Engel, Giovanni Maria Farinella, Antonino Furnari, Bernard Ghanem, Judy Hoffman, C. V. Jawahar, Richard A. Newcombe, Hyun Soo Park, James M. Rehg, Yoichi Sato, Manolis Savva, Jianbo Shi, Mike Zheng Shout, Michael Wray |
| 2023 | CVPR | Egocentric Video Task Translation. | Zihui Xue, Yale Song, Kristen Grauman, Lorenzo Torresani |
| 2023 | ICCV | EgoVLPv2: Egocentric Video-Language Pre-training with Fusion in the Backbone. | Shraman Pramanick, Yale Song, Sayan Nag, Kevin Qinghong Lin, Hardik Shah, Mike Zheng Shou, Rama Chellappa, Pengchuan Zhang |
| 2023 | WACV | Scaling Novel Object Detection with Weakly Supervised Detection Transformers. | Tyler LaBonte, Yale Song, Xin Wang, Vibhav Vineet, Neel Joshi |
| 2022 | AAAI | DOC2PPT: Automatic Presentation Slides Generation from Scientific Documents. | Tsu-Jui Fu, William Yang Wang, Daniel McDuff, Yale Song |
| 2022 | CVPR | Robust Contrastive Learning against Noisy Views. | Ching-Yao Chuang, R. Devon Hjelm, Xin Wang, Vibhav Vineet, Neel Joshi, Antonio Torralba, Stefanie Jegelka, Yale Song |
| 2022 | ECCV | Neural-Sim: Learning to Generate Training Data with NeRF. | Yunhao Ge, Harkirat S. Behl, Jiashu Xu, Suriya Gunasekar, Neel Joshi, Yale Song, Xin Wang, Laurent Itti, Vibhav Vineet |
| 2022 | ICDE | Anomaly Detection in Time Series with Robust Variational Quasi-Recurrent Autoencoders. | Tung Kieu, Bin Yang, Chenjuan Guo, Razvan-Gabriel Cirstea, Yan Zhao, Yale Song, Christian S. Jensen |
| 2022 | ICML | Visual Attention Emerges from Recurrent Sparse Reconstruction. | Baifeng Shi, Yale Song, Neel Joshi, Trevor Darrell, Xin Wang |
| 2022 | IROS | COMPASS: Contrastive Multimodal Pretraining for Autonomous Systems. | Shuang Ma, Sai Vemprala, Wenshan Wang, Jayesh K. Gupta, Yale Song, Daniel McDuff, Ashish Kapoor |
| 2021 | ICCV | ACAV100M: Automatic Curation of Large-Scale Datasets for Audio-Visual Video Representation Learning. | Sangho Lee, Jiwan Chung, Youngjae Yu, Gunhee Kim, Thomas M. Breuel, Gal Chechik, Yale Song |
| 2021 | ICLR | Parameter Efficient Multimodal Transformers for Video Representation Learning. | Sangho Lee, Youngjae Yu, Gunhee Kim, Thomas M. Breuel, Jan Kautz, Yale Song |
| 2021 | ICLR | Active Contrastive Learning of Audio-Visual Video Representations. | Shuang Ma, Zhaoyang Zeng, Daniel McDuff, Yale Song |
| 2021 | ICLR | Self-Supervised Learning of Compressed Video Representations. | Youngjae Yu, Sangho Lee, Gunhee Kim, Yale Song |
| 2020 | ICPR | Attention-Based Deep Metric Learning for Near-Duplicate Video Retrieval. | Kuan-Hsun Wang, Chia-Chun Cheng, Yi-Ling Chen, Yale Song, Shang-Hong Lai |
| 2020 | Interspeech | Multi-Reference Neural TTS Stylization with Adversarial Cycle Consistency. | Matt Whitehill, Shuang Ma, Daniel McDuff, Yale Song |
| 2020 | WACV | Image to Video Domain Adaptation Using Web Supervision. | Andrew Kae, Yale Song |
| 2019 | CVPR | Polysemous Visual-Semantic Embedding for Cross-Modal Retrieval. | Yale Song, Mohammad Soleymani |
| 2019 | ICCV | Unpaired Image-to-Speech Synthesis With Multimodal Information Bottleneck. | Shuang Ma, Daniel McDuff, Yale Song |
| 2019 | ICLR | Neural TTS Stylization with Adversarial and Collaborative Games. | Shuang Ma, Daniel McDuff, Yale Song |
| 2018 | ICML | Video Prediction with Appearance and Motion Conditions. | Yunseok Jang, Gunhee Kim, Yale Song |
| 2018 | WACV | Image2GIF: Generating Cinemagraphs Using Recurrent Deep Q-Networks. | Yipin Zhou, Yale Song, Tamara L. Berg |
| 2017 | CVPR | TGIF-QA: Toward Spatio-Temporal Reasoning in Visual Question Answering. | Yunseok Jang, Yale Song, Youngjae Yu, Youngjin Kim, Gunhee Kim |
| 2017 | CVPR | Improving Pairwise Ranking for Multi-label Image Classification. | Yuncheng Li, Yale Song, Jiebo Luo |
| 2017 | ICCV | Learning from Noisy Labels with Distillation. | Yuncheng Li, Jianchao Yang, Yale Song, Liangliang Cao, Jiebo Luo, Li-Jia Li |
| 2016 | CHI | Fast, Cheap, and Good: Why Animated GIFs Engage Us. | Saeideh Bakhshi, David A. Shamma, Lyndon Kennedy, Yale Song, Paloma de Juan, Joseph Jofish Kaye |
| 2016 | CIKM | To Click or Not To Click: Automatic Selection of Beautiful Thumbnails from Videos. | Yale Song, Miriam Redi, Jordi Vallmitjana, Alejandro Jaimes |
| 2016 | CVPR | Video2GIF: Automatic Generation of Animated GIFs from Video. | Michael Gygli, Yale Song, Liangliang Cao |
| 2016 | CVPR | TGIF: A New Dataset and Benchmark on Animated GIF Description. | Yuncheng Li, Yale Song, Liangliang Cao, Joel R. Tetreault, Larry Goldberg, Alejandro Jaimes, Jiebo Luo |
| 2016 | IJCAI | Balancing Appearance and Context in Sketch Interpretation. | Yale Song, Randall Davis, Kaichen Ma, Dana L. Penney |
| 2015 | CVPR | Video co-summarization: Video summarization by visual co-occurrence. | Wen-Sheng Chu, Yale Song, Alejandro Jaimes |
| 2015 | CVPR | TVSum: Summarizing web videos using titles. | Yale Song, Jordi Vallmitjana, Amanda Stent, Alejandro Jaimes |
| 2015 | IJCAI | Continuous Body and Hand Gesture Recognition for Natural Human-Computer Interaction: Extended Abstract. | Yale Song, Randall Davis |
| 2013 | CVPR | Action Recognition by Hierarchical Sequence Summarization. | Yale Song, Louis-Philippe Morency, Randall Davis |
| 2013 | ICMI | Learning a sparse codebook of facial and body microexpressions for emotion recognition. | Yale Song, Louis-Philippe Morency, Randall Davis |
| 2013 | IJCAI | One-Class Conditional Random Fields for Sequential Anomaly Detection. | Yale Song, Zhen Wen, Ching-Yung Lin, Randall Davis |
| 2012 | CVPR | Multi-view latent variable discriminative models for action recognition. | Yale Song, Louis-Philippe Morency, Randall Davis |
| 2012 | ICMI | Multimodal human behavior analysis: learning correlation and interaction across modalities. | Yale Song, Louis-Philippe Morency, Randall Davis |