Skip to content

David A. Ross

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

31

Venues

14

Active years

1997–2025

Best venue rank

A*

Where they publish

Papers

31 indexed papers, newest first.

YearVenueTitleAuthors
2025CVPRLanguage-Guided Image Tokenization for Generation.Kaiwen Zha, Lijun Yu, Alireza Fathi, David A. Ross, Cordelia Schmid, Dina Katabi, Xiuye Gu
2024COLINGDistribution Aware Metrics for Conditional Natural Language Generation.David M. Chan, Yiming Ni, David A. Ross, Sudheendra Vijayanarasimhan, Austin Myers, John F. Canny
2024ICLRLanguage Model Beats Diffusion - Tokenizer is key to visual generation.Lijun Yu, Jos Lezama, Nitesh Bharadwaj Gundavarapu, Luca Versari, Kihyuk Sohn, David Minnen, Yong Cheng, Agrim Gupta, Xiuye Gu, Alexander G. Hauptmann, Boqing Gong, Ming-Hsuan Yang, Irfan Essa, David A. Ross, Lu Jiang
2024ICMLVideoPrism: A Foundational Visual Encoder for Video Understanding.Long Zhao, Nitesh Bharadwaj Gundavarapu, Liangzhe Yuan, Hao Zhou, Shen Yan, Jennifer J. Sun, Luke Friedman, Rui Qian, Tobias Weyand, Yue Zhao, Rachel Hornung, Florian Schroff, Ming-Hsuan Yang, David A. Ross, Huisheng Wang, Hartwig Adam, Mikhail Sirotenko, Ting Liu, Boqing Gong
2024ICMLSceneCraft: An LLM Agent for Synthesizing 3D Scenes as Blender Code.Ziniu Hu, Ahmet Iscen, Aashi Jain, Thomas Kipf, Yisong Yue, David A. Ross, Cordelia Schmid, Alireza Fathi
2024ICMLVideoPoet: A Large Language Model for Zero-Shot Video Generation.Dan Kondratyuk, Lijun Yu, Xiuye Gu, Jos Lezama, Jonathan Huang, Grant Schindler, Rachel Hornung, Vighnesh Birodkar, Jimmy Yan, Ming-Chang Chiu, Krishna Somandepalli, Hassan Akbari, Yair Alon, Yong Cheng, Joshua V. Dillon, Agrim Gupta, Meera Hahn, Anja Hauth, David Hendon, Alonso Martinez, David Minnen, Mikhail Sirotenko, Kihyuk Sohn, Xuan Yang, Hartwig Adam, Ming-Hsuan Yang, Irfan Essa, Huisheng Wang, David A. Ross, Bryan Seybold, Lu Jiang
2023CVPRReveal: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory.Ziniu Hu, Ahmet Iscen, Chen Sun, Zirui Wang, Kai-Wei Chang, Yizhou Sun, Cordelia Schmid, David A. Ross, Alireza Fathi
2023EMNLPIC3: Image Captioning by Committee Consensus.David Chan, Austin Myers, Sudheendra Vijayanarasimhan, David A. Ross, John F. Canny
2022CVPRWhat's in a Caption? Dataset-Specific Linguistic Diversity and Its Effect on Visual Description Models and Metrics.David M. Chan, Austin Myers, Sudheendra Vijayanarasimhan, David A. Ross, Bryan Seybold, John F. Canny
2021ICCVAI Choreographer: Music Conditioned 3D Dance Generation with AIST++.Ruilong Li, Shan Yang, David A. Ross, Angjoo Kanazawa
2020ACCVActive Learning for Video Description with Cluster-Regularized Ensemble Ranking.David M. Chan, Sudheendra Vijayanarasimhan, David A. Ross, John F. Canny
2020CVPRDOPS: Learning to Detect 3D Objects and Predict Their 3D Shapes.Mahyar Najibi, Guangda Lai, Abhijit Kundu, Zhichao Lu, Vivek Rathod, Thomas A. Funkhouser, Caroline Pantofaru, David A. Ross, Larry S. Davis, Alireza Fathi
2020ECCVAn LSTM Approach to Temporal 3D Object Detection in LiDAR Point Clouds.Rui Huang, Wanyue Zhang, Abhijit Kundu, Caroline Pantofaru, David A. Ross, Thomas A. Funkhouser, Alireza Fathi
2020ECCVVirtual Multi-view Fusion for 3D Semantic Segmentation.Abhijit Kundu, Xiaoqi Yin, Alireza Fathi, David A. Ross, Brian Brewington, Thomas A. Funkhouser, Caroline Pantofaru
2020ECCVPillar-Based Object Detection for Autonomous Driving.Yue Wang, Alireza Fathi, Abhijit Kundu, David A. Ross, Caroline Pantofaru, Thomas A. Funkhouser, Justin Solomon
2020WACVD3D: Distilled 3D Networks for Video Action Recognition.Jonathan C. Stroud, David A. Ross, Chen Sun, Jia Deng, Rahul Sukthankar
2018CVPRRethinking the Faster R-CNN Architecture for Temporal Action Localization.Yu-Wei Chao, Sudheendra Vijayanarasimhan, Bryan Seybold, David A. Ross, Jia Deng, Rahul Sukthankar
2018CVPRAVA: A Video Dataset of Spatio-Temporally Localized Atomic Visual Actions.Chunhui Gu, Chen Sun, David A. Ross, Carl Vondrick, Caroline Pantofaru, Yeqing Li, Sudheendra Vijayanarasimhan, George Toderici, Susanna Ricco, Rahul Sukthankar, Cordelia Schmid, Jitendra Malik
2014PSBBuilding the Next Generation of Quantitative Biologists.Kristine A. Pattin, Anna C. Greene, Russ B. Altman, Lawrence E. Hunter, David A. Ross, James A. Foster, Jason H. Moore
2011HCIHelping Hands versus ERSP Vision: Comparing Object Recognition Technologies for the Visually Impaired.Marc A. Lawson, Ellen Yi-Luen Do, James R. Marston, David A. Ross
2011ICASSPAutomatic Language Identification in music videos with low level audio and visual features.Vijay Chandrasekhar, Mehmet Emre Sargin, David A. Ross
2011ICCVThe power of comparative reasoning.Jay Yagnik, Dennis Strelow, David A. Ross, Ruei-Sung Lin
2010CVPRSPEC hashing: Similarity preserving algorithm for entropy-based coding.Ruei-Sung Lin, David A. Ross, Jay Yagnik
2008CVPRLearning stick-figure models using nonparametric Bayesian priors over trees.Edward Meeds, David A. Ross, Richard S. Zemel, Sam T. Roweis
2008ECCVUnsupervised Learning of Skeletons from Motion.David A. Ross, Daniel Tarlow, Richard S. Zemel
2006ICMLCombining discriminative features to infer complex trajectories.David A. Ross, Simon Osindero, Richard S. Zemel
2005ASSETSTalking braille: a wireless ubiquitous computing network for orientation and wayfinding.David A. Ross, Alexander Lightman
2004ECCVAdaptive Probabilistic Visual Tracking with Incremental Subspace Update.David A. Ross, Jongwoo Lim, Ming-Hsuan Yang
2000ASSETSWearable interfaces for orientation and wayfinding.David A. Ross, Bruce B. Blasch
2000ISWCEvaluation of Orientation Interfaces for Wearable Computers.David A. Ross, Bruce B. Blasch
1997ISWCThe Wearable Computer as a Remote Interface for People with Disabilities.David A. Ross, Jon A. Sanford