Skip to content

Yale Song

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

42

Venues

15

Active years

2012–2026

Best venue rank

A*

Where they publish

Papers

42 indexed papers, newest first.

YearVenueTitleAuthors
2026WACVEnhancing Visual Planning with Auxiliary Tasks and Multi-token Prediction.Ce Zhang, Yale Song, Ruta Desai, Michael Louis Iuzzolino, Joseph Tighe, Gedas Bertasius, Satwik Kottur
2025CVPRVITED: Video Temporal Evidence Distillation.Yujie Lu, Yale Song, William Wang, Lorenzo Torresani, Tushar Nagarajan
2025ICCVStreaming Videollms for Real-Time Procedural Video Understanding.Dibyadip Chatterjee, Edoardo Remelli, Yale Song, Bugra Tekin, Abhay Mittal, Bharat Bhatnagar, Necati Cihan Camgz, Shreyas Hampali, Eric Sauser, Shugao Ma, Angela Yao, Fadime Sener
2025ICCVEnrich and Detect: Video Temporal Grounding With Multimodal Llms.Shraman Pramanick, Effrosyni Mavroudi, Yale Song, Rama Chellappa, Lorenzo Torresani, Triantafyllos Afouras
2024CVPREgo-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives.Kristen Grauman, Andrew Westbury, Lorenzo Torresani, Kris Kitani, Jitendra Malik, Triantafyllos Afouras, Kumar Ashutosh, Vijay Baiyya, Siddhant Bansal, Bikram Boote, Eugene Byrne, Zachary Chavis, Joya Chen, Feng Cheng, Fu-Jen Chu, Sean Crane, Avijit Dasgupta, Jing Dong, Mara Escobar, Cristhian Forigua, Abrham Gebreselasie, Sanjay Haresh, Jing Huang, Md Mohaiminul Islam, Suyog Dutt Jain, Rawal Khirodkar, Devansh Kukreja, Kevin J. Liang, Jia-Wei Liu, Sagnik Majumder, Yongsen Mao, Miguel Martin, Effrosyni Mavroudi, Tushar Nagarajan, Francesco Ragusa, Santhosh Kumar Ramakrishnan, Luigi Seminara, Arjun Somayazulu, Yale Song, Shan Su, Zihui Xue, Edward Zhang, Jinxu Zhang, Angela Castillo, Changan Chen, Xinzhu Fu, Ryosuke Furuta, Cristina Gonzlez, Prince Gupta, Jiabo Hu, Yifei Huang, Yiming Huang, Weslie Khoo, Anush Kumar, Robert Kuo, Sach Lakhavani, Miao Liu, Mi Luo, Zhengyi Luo, Brighid Meredith, Austin Miller, Oluwatumininu Oguntola, Xiaqing Pan, Penny Peng, Shraman Pramanick, Merey Ramazanova, Fiona Ryan, Wei Shan, Kiran K. Somasundaram, Chenan Song, Audrey Southerland, Masatoshi Tateno, Huiyu Wang, Yuchen Wang, Takuma Yagi, Mingfei Yan, Xitong Yang, Zecheng Yu, Shengxin Cindy Zha, Chen Zhao, Ziwei Zhao, Zhifan Zhu, Jeff Zhuo, Pablo Arbelez, Gedas Bertasius, Dima Damen, Jakob J. Engel, Giovanni Maria Farinella, Antonino Furnari, Bernard Ghanem, Judy Hoffman, C. V. Jawahar, Richard A. Newcombe, Hyun Soo Park, James M. Rehg, Yoichi Sato, Manolis Savva, Jianbo Shi, Mike Zheng Shout, Michael Wray
2023CVPREgocentric Video Task Translation.Zihui Xue, Yale Song, Kristen Grauman, Lorenzo Torresani
2023ICCVEgoVLPv2: Egocentric Video-Language Pre-training with Fusion in the Backbone.Shraman Pramanick, Yale Song, Sayan Nag, Kevin Qinghong Lin, Hardik Shah, Mike Zheng Shou, Rama Chellappa, Pengchuan Zhang
2023WACVScaling Novel Object Detection with Weakly Supervised Detection Transformers.Tyler LaBonte, Yale Song, Xin Wang, Vibhav Vineet, Neel Joshi
2022AAAIDOC2PPT: Automatic Presentation Slides Generation from Scientific Documents.Tsu-Jui Fu, William Yang Wang, Daniel McDuff, Yale Song
2022CVPRRobust Contrastive Learning against Noisy Views.Ching-Yao Chuang, R. Devon Hjelm, Xin Wang, Vibhav Vineet, Neel Joshi, Antonio Torralba, Stefanie Jegelka, Yale Song
2022ECCVNeural-Sim: Learning to Generate Training Data with NeRF.Yunhao Ge, Harkirat S. Behl, Jiashu Xu, Suriya Gunasekar, Neel Joshi, Yale Song, Xin Wang, Laurent Itti, Vibhav Vineet
2022ICDEAnomaly Detection in Time Series with Robust Variational Quasi-Recurrent Autoencoders.Tung Kieu, Bin Yang, Chenjuan Guo, Razvan-Gabriel Cirstea, Yan Zhao, Yale Song, Christian S. Jensen
2022ICMLVisual Attention Emerges from Recurrent Sparse Reconstruction.Baifeng Shi, Yale Song, Neel Joshi, Trevor Darrell, Xin Wang
2022IROSCOMPASS: Contrastive Multimodal Pretraining for Autonomous Systems.Shuang Ma, Sai Vemprala, Wenshan Wang, Jayesh K. Gupta, Yale Song, Daniel McDuff, Ashish Kapoor
2021ICCVACAV100M: Automatic Curation of Large-Scale Datasets for Audio-Visual Video Representation Learning.Sangho Lee, Jiwan Chung, Youngjae Yu, Gunhee Kim, Thomas M. Breuel, Gal Chechik, Yale Song
2021ICLRParameter Efficient Multimodal Transformers for Video Representation Learning.Sangho Lee, Youngjae Yu, Gunhee Kim, Thomas M. Breuel, Jan Kautz, Yale Song
2021ICLRActive Contrastive Learning of Audio-Visual Video Representations.Shuang Ma, Zhaoyang Zeng, Daniel McDuff, Yale Song
2021ICLRSelf-Supervised Learning of Compressed Video Representations.Youngjae Yu, Sangho Lee, Gunhee Kim, Yale Song
2020ICPRAttention-Based Deep Metric Learning for Near-Duplicate Video Retrieval.Kuan-Hsun Wang, Chia-Chun Cheng, Yi-Ling Chen, Yale Song, Shang-Hong Lai
2020InterspeechMulti-Reference Neural TTS Stylization with Adversarial Cycle Consistency.Matt Whitehill, Shuang Ma, Daniel McDuff, Yale Song
2020WACVImage to Video Domain Adaptation Using Web Supervision.Andrew Kae, Yale Song
2019CVPRPolysemous Visual-Semantic Embedding for Cross-Modal Retrieval.Yale Song, Mohammad Soleymani
2019ICCVUnpaired Image-to-Speech Synthesis With Multimodal Information Bottleneck.Shuang Ma, Daniel McDuff, Yale Song
2019ICLRNeural TTS Stylization with Adversarial and Collaborative Games.Shuang Ma, Daniel McDuff, Yale Song
2018ICMLVideo Prediction with Appearance and Motion Conditions.Yunseok Jang, Gunhee Kim, Yale Song
2018WACVImage2GIF: Generating Cinemagraphs Using Recurrent Deep Q-Networks.Yipin Zhou, Yale Song, Tamara L. Berg
2017CVPRTGIF-QA: Toward Spatio-Temporal Reasoning in Visual Question Answering.Yunseok Jang, Yale Song, Youngjae Yu, Youngjin Kim, Gunhee Kim
2017CVPRImproving Pairwise Ranking for Multi-label Image Classification.Yuncheng Li, Yale Song, Jiebo Luo
2017ICCVLearning from Noisy Labels with Distillation.Yuncheng Li, Jianchao Yang, Yale Song, Liangliang Cao, Jiebo Luo, Li-Jia Li
2016CHIFast, Cheap, and Good: Why Animated GIFs Engage Us.Saeideh Bakhshi, David A. Shamma, Lyndon Kennedy, Yale Song, Paloma de Juan, Joseph Jofish Kaye
2016CIKMTo Click or Not To Click: Automatic Selection of Beautiful Thumbnails from Videos.Yale Song, Miriam Redi, Jordi Vallmitjana, Alejandro Jaimes
2016CVPRVideo2GIF: Automatic Generation of Animated GIFs from Video.Michael Gygli, Yale Song, Liangliang Cao
2016CVPRTGIF: A New Dataset and Benchmark on Animated GIF Description.Yuncheng Li, Yale Song, Liangliang Cao, Joel R. Tetreault, Larry Goldberg, Alejandro Jaimes, Jiebo Luo
2016IJCAIBalancing Appearance and Context in Sketch Interpretation.Yale Song, Randall Davis, Kaichen Ma, Dana L. Penney
2015CVPRVideo co-summarization: Video summarization by visual co-occurrence.Wen-Sheng Chu, Yale Song, Alejandro Jaimes
2015CVPRTVSum: Summarizing web videos using titles.Yale Song, Jordi Vallmitjana, Amanda Stent, Alejandro Jaimes
2015IJCAIContinuous Body and Hand Gesture Recognition for Natural Human-Computer Interaction: Extended Abstract.Yale Song, Randall Davis
2013CVPRAction Recognition by Hierarchical Sequence Summarization.Yale Song, Louis-Philippe Morency, Randall Davis
2013ICMILearning a sparse codebook of facial and body microexpressions for emotion recognition.Yale Song, Louis-Philippe Morency, Randall Davis
2013IJCAIOne-Class Conditional Random Fields for Sequential Anomaly Detection.Yale Song, Zhen Wen, Ching-Yung Lin, Randall Davis
2012CVPRMulti-view latent variable discriminative models for action recognition.Yale Song, Louis-Philippe Morency, Randall Davis
2012ICMIMultimodal human behavior analysis: learning correlation and interaction across modalities.Yale Song, Louis-Philippe Morency, Randall Davis