David A. Ross
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
31
Venues
14
Active years
1997–2025
Best venue rank
A*
Where they publish
Papers
31 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | CVPR | Language-Guided Image Tokenization for Generation. | Kaiwen Zha, Lijun Yu, Alireza Fathi, David A. Ross, Cordelia Schmid, Dina Katabi, Xiuye Gu |
| 2024 | COLING | Distribution Aware Metrics for Conditional Natural Language Generation. | David M. Chan, Yiming Ni, David A. Ross, Sudheendra Vijayanarasimhan, Austin Myers, John F. Canny |
| 2024 | ICLR | Language Model Beats Diffusion - Tokenizer is key to visual generation. | Lijun Yu, Jos Lezama, Nitesh Bharadwaj Gundavarapu, Luca Versari, Kihyuk Sohn, David Minnen, Yong Cheng, Agrim Gupta, Xiuye Gu, Alexander G. Hauptmann, Boqing Gong, Ming-Hsuan Yang, Irfan Essa, David A. Ross, Lu Jiang |
| 2024 | ICML | VideoPrism: A Foundational Visual Encoder for Video Understanding. | Long Zhao, Nitesh Bharadwaj Gundavarapu, Liangzhe Yuan, Hao Zhou, Shen Yan, Jennifer J. Sun, Luke Friedman, Rui Qian, Tobias Weyand, Yue Zhao, Rachel Hornung, Florian Schroff, Ming-Hsuan Yang, David A. Ross, Huisheng Wang, Hartwig Adam, Mikhail Sirotenko, Ting Liu, Boqing Gong |
| 2024 | ICML | SceneCraft: An LLM Agent for Synthesizing 3D Scenes as Blender Code. | Ziniu Hu, Ahmet Iscen, Aashi Jain, Thomas Kipf, Yisong Yue, David A. Ross, Cordelia Schmid, Alireza Fathi |
| 2024 | ICML | VideoPoet: A Large Language Model for Zero-Shot Video Generation. | Dan Kondratyuk, Lijun Yu, Xiuye Gu, Jos Lezama, Jonathan Huang, Grant Schindler, Rachel Hornung, Vighnesh Birodkar, Jimmy Yan, Ming-Chang Chiu, Krishna Somandepalli, Hassan Akbari, Yair Alon, Yong Cheng, Joshua V. Dillon, Agrim Gupta, Meera Hahn, Anja Hauth, David Hendon, Alonso Martinez, David Minnen, Mikhail Sirotenko, Kihyuk Sohn, Xuan Yang, Hartwig Adam, Ming-Hsuan Yang, Irfan Essa, Huisheng Wang, David A. Ross, Bryan Seybold, Lu Jiang |
| 2023 | CVPR | Reveal: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory. | Ziniu Hu, Ahmet Iscen, Chen Sun, Zirui Wang, Kai-Wei Chang, Yizhou Sun, Cordelia Schmid, David A. Ross, Alireza Fathi |
| 2023 | EMNLP | IC3: Image Captioning by Committee Consensus. | David Chan, Austin Myers, Sudheendra Vijayanarasimhan, David A. Ross, John F. Canny |
| 2022 | CVPR | What's in a Caption? Dataset-Specific Linguistic Diversity and Its Effect on Visual Description Models and Metrics. | David M. Chan, Austin Myers, Sudheendra Vijayanarasimhan, David A. Ross, Bryan Seybold, John F. Canny |
| 2021 | ICCV | AI Choreographer: Music Conditioned 3D Dance Generation with AIST++. | Ruilong Li, Shan Yang, David A. Ross, Angjoo Kanazawa |
| 2020 | ACCV | Active Learning for Video Description with Cluster-Regularized Ensemble Ranking. | David M. Chan, Sudheendra Vijayanarasimhan, David A. Ross, John F. Canny |
| 2020 | CVPR | DOPS: Learning to Detect 3D Objects and Predict Their 3D Shapes. | Mahyar Najibi, Guangda Lai, Abhijit Kundu, Zhichao Lu, Vivek Rathod, Thomas A. Funkhouser, Caroline Pantofaru, David A. Ross, Larry S. Davis, Alireza Fathi |
| 2020 | ECCV | An LSTM Approach to Temporal 3D Object Detection in LiDAR Point Clouds. | Rui Huang, Wanyue Zhang, Abhijit Kundu, Caroline Pantofaru, David A. Ross, Thomas A. Funkhouser, Alireza Fathi |
| 2020 | ECCV | Virtual Multi-view Fusion for 3D Semantic Segmentation. | Abhijit Kundu, Xiaoqi Yin, Alireza Fathi, David A. Ross, Brian Brewington, Thomas A. Funkhouser, Caroline Pantofaru |
| 2020 | ECCV | Pillar-Based Object Detection for Autonomous Driving. | Yue Wang, Alireza Fathi, Abhijit Kundu, David A. Ross, Caroline Pantofaru, Thomas A. Funkhouser, Justin Solomon |
| 2020 | WACV | D3D: Distilled 3D Networks for Video Action Recognition. | Jonathan C. Stroud, David A. Ross, Chen Sun, Jia Deng, Rahul Sukthankar |
| 2018 | CVPR | Rethinking the Faster R-CNN Architecture for Temporal Action Localization. | Yu-Wei Chao, Sudheendra Vijayanarasimhan, Bryan Seybold, David A. Ross, Jia Deng, Rahul Sukthankar |
| 2018 | CVPR | AVA: A Video Dataset of Spatio-Temporally Localized Atomic Visual Actions. | Chunhui Gu, Chen Sun, David A. Ross, Carl Vondrick, Caroline Pantofaru, Yeqing Li, Sudheendra Vijayanarasimhan, George Toderici, Susanna Ricco, Rahul Sukthankar, Cordelia Schmid, Jitendra Malik |
| 2014 | PSB | Building the Next Generation of Quantitative Biologists. | Kristine A. Pattin, Anna C. Greene, Russ B. Altman, Lawrence E. Hunter, David A. Ross, James A. Foster, Jason H. Moore |
| 2011 | HCI | Helping Hands versus ERSP Vision: Comparing Object Recognition Technologies for the Visually Impaired. | Marc A. Lawson, Ellen Yi-Luen Do, James R. Marston, David A. Ross |
| 2011 | ICASSP | Automatic Language Identification in music videos with low level audio and visual features. | Vijay Chandrasekhar, Mehmet Emre Sargin, David A. Ross |
| 2011 | ICCV | The power of comparative reasoning. | Jay Yagnik, Dennis Strelow, David A. Ross, Ruei-Sung Lin |
| 2010 | CVPR | SPEC hashing: Similarity preserving algorithm for entropy-based coding. | Ruei-Sung Lin, David A. Ross, Jay Yagnik |
| 2008 | CVPR | Learning stick-figure models using nonparametric Bayesian priors over trees. | Edward Meeds, David A. Ross, Richard S. Zemel, Sam T. Roweis |
| 2008 | ECCV | Unsupervised Learning of Skeletons from Motion. | David A. Ross, Daniel Tarlow, Richard S. Zemel |
| 2006 | ICML | Combining discriminative features to infer complex trajectories. | David A. Ross, Simon Osindero, Richard S. Zemel |
| 2005 | ASSETS | Talking braille: a wireless ubiquitous computing network for orientation and wayfinding. | David A. Ross, Alexander Lightman |
| 2004 | ECCV | Adaptive Probabilistic Visual Tracking with Incremental Subspace Update. | David A. Ross, Jongwoo Lim, Ming-Hsuan Yang |
| 2000 | ASSETS | Wearable interfaces for orientation and wayfinding. | David A. Ross, Bruce B. Blasch |
| 2000 | ISWC | Evaluation of Orientation Interfaces for Wearable Computers. | David A. Ross, Bruce B. Blasch |
| 1997 | ISWC | The Wearable Computer as a Remote Interface for People with Disabilities. | David A. Ross, Jon A. Sanford |