| 2025 | CVPR | Pos3R: 6D Pose Estimation for Unseen Objects Made Easy. | Weijian Deng, Dylan Campbell, Chunyi Sun, Jiahao Zhang, Shubham Kanitkar, Matthew E. Shaffer, Stephen Gould |
| 2025 | CVPR | VI^3NR: Variance Informed Initialization for Implicit Neural Representations. | Chamin Hewa Koneputugodage, Yizhak Ben-Shabat, Sameera Ramasinghe, Stephen Gould |
| 2025 | ICCV | Leaps and Bounds: An Improved Point Cloud Winding Number Formulation for Fast Normal Estimation and Surface Reconstruction. | Chamin Hewa Koneputugodage, Dylan Campbell, Stephen Gould |
| 2025 | ICCV | Manual-PA: Learning 3D Part Assembly from Instruction Diagrams. | Jiahao Zhang, Anoop Cherian, Cristian Rodriguez, Weijian Deng, Stephen Gould |
| 2025 | ICML | Can We Predict Performance of Large Models across Vision-Language Tasks? | Qinyu Zhao, Ming Xu, Kartik Gupta, Akshay Asthana, Liang Zheng, Stephen Gould |
| 2025 | IJCNN | Scaling Prompt Instructed Zero Shot Composed Image Retrieval with Image-Only Data. | Yiqun Duan, Sameera Ramasinghe, Stephen Gould, Thalaiyasingam Ajanthan |
| 2025 | SIGGRAPH | Unsupervised Decomposition of 3D Shapes into Expressive and Editable Extruded Profile Primitives. | Chunyi Sun, Junlin Han, Runjia Li, Weijian Deng, Dylan Campbell, Stephen Gould |
| 2025 | WACV | Temporally Grounding Instructional Diagrams in Unconstrained Videos. | Jiahao Zhang, Frederic Z. Zhang, Cristian Rodriguez, Yizhak Ben-Shabat, Anoop Cherian, Stephen Gould |
| 2024 | CVPR | 3DInAction: Understanding Human Actions in 3D Point Clouds. | Yizhak Ben-Shabat, Oren Shrout, Stephen Gould |
| 2024 | CVPR | Differentiable Neural Surface Refinement for Modeling Transparent Objects. | Weijian Deng, Dylan Campbell, Chunyi Sun, Shubham Kanitkar, Matthew E. Shaffer, Stephen Gould |
| 2024 | CVPR | Learning to Select Views for Efficient Multi-View Understanding. | Yunzhong Hou, Stephen Gould, Liang Zheng |
| 2024 | CVPR | Small Steps and Level Sets: Fitting Neural Surface Models with Point Guidance. | Chamin Hewa Koneputugodage, Yizhak Ben-Shabat, Dylan Campbell, Stephen Gould |
| 2024 | CVPR | Temporally Consistent Unbalanced Optimal Transport for Unsupervised Action Segmentation. | Ming Xu, Stephen Gould |
| 2024 | ECCV | Unsupervised Dense Prediction Using Differentiable Normalized Cuts. | Yanbin Liu, Stephen Gould |
| 2024 | ECCV | The First to Know: How Token Distributions Reveal Hidden Knowledge in Large Vision-Language Models? | Qinyu Zhao, Ming Xu, Kartik Gupta, Akshay Asthana, Liang Zheng, Stephen Gould |
| 2024 | ICAPS | Neuro-Symbolic Learning of Lifted Action Models from Visual Traces. | Kai Xi, Stephen Gould, Sylvie Thibaux |
| 2024 | ICLR | Towards Optimal Feature-Shaping Methods for Out-of-Distribution Detection. | Qinyu Zhao, Ming Xu, Kartik Gupta, Akshay Asthana, Liang Zheng, Stephen Gould |
| 2024 | ICML | An Empirical Study Into What Matters for Calibrating Vision-Language Models. | Weijie Tu, Weijian Deng, Dylan Campbell, Stephen Gould, Tom Gedeon |
| 2024 | WACV | Bi-directional Training for Composed Image Retrieval via Text Prompt Learning. | Zheyuan Liu, Weixuan Sun, Yicong Hong, Damien Teney, Stephen Gould |
| 2024 | WACV | IKEA Ego 3D Dataset: Understanding furniture assembly actions from ego-view 3D Point Clouds. | Yizhak Ben-Shabat, Jonathan Paul, Eviatar Segev, Oren Shrout, Stephen Gould |
| 2024 | WACV | Ray Deformation Networks for Novel View Synthesis of Refractive Objects. | Weijian Deng, Dylan Campbell, Chunyi Sun, Shubham Kanitkar, Matthew E. Shaffer, Stephen Gould |
| 2024 | WACV | LipAT: Beyond Style Transfer for Controllable Neural Simulation of Lipstick using Cosmetic Attributes. | Amila Silva, Olga Moskvyak, Alexander Long, Ravi Garg, Stephen Gould, Gil Avraham, Anton van den Hengel |
| 2024 | WACV | NeRFEditor: Differentiable Style Decomposition for 3D Scene Editing. | Chunyi Sun, Yanbin Liu, Junlin Han, Stephen Gould |
| 2023 | CVPR | Octree Guided Unoriented Surface Reconstruction. | Chamin Hewa Koneputugodage, Yizhak Ben-Shabat, Stephen Gould |
| 2023 | CVPR | High-Fidelity Guided Image Synthesis with Latent Diffusion Models. | Jaskirat Singh, Stephen Gould, Liang Zheng |
| 2023 | CVPR | Aligning Step-by-Step Instructional Diagrams to Video Demonstrations. | Jiahao Zhang, Anoop Cherian, Yanbin Liu, Yizhak Ben-Shabat, Cristian Rodriguez Opazo, Stephen Gould |
| 2023 | ICCV | Learning Navigational Visual Representations with Semantic Map Supervision. | Yicong Hong, Yang Zhou, Ruiyi Zhang, Franck Dernoncourt, Trung Bui, Stephen Gould, Hao Tan |
| 2023 | ICCV | Semi-Supervised Semantic Segmentation under Label Noise via Diverse Learning Groups. | Peixia Li, Pulak Purkait, Thalaiyasingam Ajanthan, Majid Abdolshah, Ravi Garg, Hisham Husain, Chenchen Xu, Stephen Gould, Wanli Ouyang, Anton van den Hengel |
| 2023 | ICCV | Scaling Data Generation in Vision-and-Language Navigation. | Zun Wang, Jialu Li, Yicong Hong, Yi Wang, Qi Wu, Mohit Bansal, Stephen Gould, Hao Tan, Yu Qiao |
| 2023 | ICCV | Exploring Predicate Visual Context in Detecting of Human-Object Interactions. | Frederic Z. Zhang, Yuhui Yuan, Dylan Campbell, Zhuoyao Zhong, Stephen Gould |
| 2023 | ICLR | Deep Declarative Dynamic Time Warping for End-to-End Learning of Alignment Paths. | Ming Xu, Sourav Garg, Michael Milford, Stephen Gould |
| 2023 | ICML | Confidence and Dispersity Speak: Characterizing Prediction Matrix for Unsupervised Accuracy Estimation. | Weijian Deng, Yumin Suh, Stephen Gould, Liang Zheng |
| 2022 | CVPR | DiGS : Divergence guided shape implicit neural representation for unoriented point clouds. | Yizhak Ben-Shabat, Chamin Hewa Koneputugodage, Stephen Gould |
| 2022 | CVPR | Bridging the Gap Between Learning in Discrete and Continuous Environments for Vision-and-Language Navigation. | Yicong Hong, Zun Wang, Qi Wu, Stephen Gould |
| 2022 | CVPR | Efficient Two-Stage Detection of Human-Object Interactions with a Novel Unary-Pairwise Transformer. | Frederic Z. Zhang, Dylan Campbell, Stephen Gould |
| 2022 | IROS | GoferBot: A Visual Guided Human-Robot Collaborative Assembly System. | Zheyu Zhuang, Yizhak Ben-Shabat, Jiahao Zhang, Stephen Gould, Robert E. Mahony |
| 2021 | CVPR | VLN BERT: A Recurrent Vision-and-Language BERT for Navigation. | Yicong Hong, Qi Wu, Yuankai Qi, Cristian Rodriguez Opazo, Stephen Gould |
| 2021 | CVPR | Probabilistic Tracklet Scoring and Inpainting for Multiple Object Tracking. | Fatemeh Sadat Saleh, Sadegh Aliakbarian, Hamid Rezatofighi, Mathieu Salzmann, Stephen Gould |
| 2021 | ICCV | Image Retrieval on Real-life Images with Pre-trained Vision-and-Language Models. | Zheyuan Liu, Cristian Rodriguez Opazo, Damien Teney, Stephen Gould |
| 2021 | ICCV | Contextually Plausible and Diverse 3D Human Motion Prediction. | Sadegh Aliakbarian, Fatemeh Sadat Saleh, Lars Petersson, Stephen Gould, Mathieu Salzmann |
| 2021 | ICCV | Spatially Conditioned Graphs for Detecting Human-Object Interactions. | Frederic Z. Zhang, Dylan Campbell, Stephen Gould |
| 2021 | ICDM | A Regularized Wasserstein Framework for Graph Kernels. | Asiri Wijesinghe, Qing Wang, Stephen Gould |
| 2021 | ICLR | Conditional Generative Modeling via Learning the Latent Space. | Sameera Ramasinghe, Kanchana Nisal Ranasinghe, Salman H. Khan, Nick Barnes, Stephen Gould |
| 2021 | ICML | What Does Rotation Prediction Tell Us about Classifier Accuracy under Varying Testing Environments? | Weijian Deng, Stephen Gould, Liang Zheng |
| 2021 | WACV | The IKEA ASM Dataset: Understanding People Assembling Furniture through Actions, Objects and Pose. | Yizhak Ben-Shabat, Xin Yu, Fatemeh Sadat Saleh, Dylan Campbell, Cristian Rodriguez Opazo, Hongdong Li, Stephen Gould |
| 2021 | WACV | DORi: Discovering Object Relationships for Moment Localization of a Natural Language Query in a Video. | Cristian Rodriguez Opazo, Edison Marrese-Taylor, Basura Fernando, Hongdong Li, Stephen Gould |
| 2020 | CVPR | A Stochastic Conditioning Scheme for Diverse Human Motion Prediction. | Mohammad Sadegh Aliakbarian, Fatemeh Sadat Saleh, Mathieu Salzmann, Lars Petersson, Stephen Gould |
| 2020 | CVPR | Inferring Temporal Compositions of Actions Using Probabilistic Automata. | Rodrigo Santa Cruz, Anoop Cherian, Basura Fernando, Dylan Campbell, Stephen Gould |
| 2020 | CVPR | Learning to Structure an Image With Few Colors. | Yunzhong Hou, Liang Zheng, Stephen Gould |
| 2020 | ECCV | DeepFit: 3D Surface Fitting via Neural Network Weighted Least Squares. | Yizhak Ben-Shabat, Stephen Gould |
| 2020 | ECCV | Solving the Blind Perspective-n-Point Problem End-to-End with Robust Differentiable Geometric Optimization. | Dylan Campbell, Liu Liu, Stephen Gould |
| 2020 | ECCV | Multiview Detection with Feature Perspective Transformation. | Yunzhong Hou, Liang Zheng, Stephen Gould |
| 2020 | EMNLP | Sub-Instruction Aware Vision-and-Language Navigation. | Yicong Hong, Cristian Rodriguez Opazo, Qi Wu, Stephen Gould |
| 2020 | ICLR | A Signal Propagation Perspective for Pruning Neural Networks at Initialization. | Namhoon Lee, Thalaiyasingam Ajanthan, Stephen Gould, Philip H. S. Torr |
| 2020 | IROS | Spectral-GANs for High-Resolution 3D Point-cloud Generation. | Sameera Ramasinghe, Salman H. Khan, Nick Barnes, Stephen Gould |
| 2020 | WACV | Proposal-free Temporal Moment Localization of a Natural-Language Query in Video using Guided Attention. | Cristian Rodriguez Opazo, Edison Marrese-Taylor, Fatemeh Sadat Saleh, Hongdong Li, Stephen Gould |
| 2020 | WACV | Blended Convolution and Synthesis for Efficient Discrimination of 3D Shapes. | Sameera Ramasinghe, Salman H. Khan, Nick Barnes, Stephen Gould |
| 2019 | CVPR | The Alignment of the Spheres: Globally-Optimal Spherical Mixture Alignment for Camera Pose Estimation. | Dylan Campbell, Lars Petersson, Laurent Kneip, Hongdong Li, Stephen Gould |
| 2019 | ICCV | Learning to Find Common Objects Across Few Image Collections. | Amirreza Shaban, Amir Rahimi, Shray Bansal, Stephen Gould, Byron Boots, Richard Hartley |
| 2018 | CVPR | Bottom-Up and Top-Down Attention for Image Captioning and Visual Question Answering. | Peter Anderson, Xiaodong He, Chris Buehler, Damien Teney, Mark Johnson, Stephen Gould, Lei Zhang |
| 2018 | CVPR | Vision-and-Language Navigation: Interpreting Visually-Grounded Navigation Instructions in Real Environments. | Peter Anderson, Qi Wu, Damien Teney, Jake Bruce, Mark Johnson, Niko Snderhauf, Ian D. Reid, Stephen Gould, Anton van den Hengel |
| 2018 | CVPR | Non-Linear Temporal Subspace Representations for Activity Recognition. | Anoop Cherian, Suvrit Sra, Stephen Gould, Richard Hartley |
| 2018 | CVPR | Video Representation Learning Using Discriminative Pooling. | Jue Wang, Anoop Cherian, Fatih Porikli, Stephen Gould |
| 2018 | WACV | Neural Algebra of Classifiers. | Rodrigo Santa Cruz, Basura Fernando, Anoop Cherian, Stephen Gould |
| 2017 | CVPR | Generalized Rank Pooling for Activity Recognition. | Anoop Cherian, Basura Fernando, Mehrtash Harandi, Stephen Gould |
| 2017 | CVPR | DeepPermNet: Visual Permutation Learning. | Rodrigo Santa Cruz, Basura Fernando, Anoop Cherian, Stephen Gould |
| 2017 | CVPR | Self-Supervised Video Representation Learning with Odd-One-Out Networks. | Basura Fernando, Hakan Bilen, Efstratios Gavves, Stephen Gould |
| 2017 | CVPR | Unsupervised Human Action Detection by Action Matching. | Basura Fernando, Sareh Shirazi, Stephen Gould |
| 2017 | DICTA | Human Pose Forecasting via Deep Markov Models. | Sam Toyer, Anoop Cherian, Tengda Han, Stephen Gould |
| 2017 | EMNLP | Guided Open Vocabulary Image Captioning with Constrained Beam Search. | Peter Anderson, Basura Fernando, Mark Johnson, Stephen Gould |
| 2017 | WACV | Higher-Order Pooling of CNN Features via Kernel Linearization for Action Recognition. | Anoop Cherian, Piotr Koniusz, Stephen Gould |
| 2016 | CVPR | Dynamic Image Networks for Action Recognition. | Hakan Bilen, Basura Fernando, Efstratios Gavves, Andrea Vedaldi, Stephen Gould |
| 2016 | CVPR | Discriminative Hierarchical Rank Pooling for Activity Recognition. | Basura Fernando, Peter Anderson, Marcus Hutter, Stephen Gould |
| 2016 | DICTA | Depth Dropout: Efficient Training of Residual Convolutional Neural Networks. | Jian Guo, Stephen Gould |
| 2016 | ECCV | SPICE: Semantic Propositional Image Caption Evaluation. | Peter Anderson, Basura Fernando, Mark Johnson, Stephen Gould |
| 2016 | ECCV | Deep Convolutional Neural Networks for Human Embryonic Cell Counting. | Aisha Khan, Stephen Gould, Mathieu Salzmann |
| 2016 | ECCV | Built-in Foreground/Background Prior for Weakly-Supervised Semantic Segmentation. | Fatemehsadat Saleh, Mohammad Sadegh Ali Akbarian, Mathieu Salzmann, Lars Petersson, Stephen Gould, Jos M. lvarez |
| 2016 | ICML | Learning End-to-end Video Classification with Rank-Pooling. | Basura Fernando, Stephen Gould |
| 2015 | ICCV | Hierarchical Higher-Order Regression Forest Fields: An Application to 3D Indoor Scene Labelling. | Trung-Thanh Pham, Ian D. Reid, Yasir Latif, Stephen Gould |
| 2015 | MICCAI | Detecting Abnormal Cell Division Patterns in Early Stage Human Embryo Development. | Aisha Khan, Stephen Gould, Mathieu Salzmann |
| 2015 | WACV | A Linear Chain Markov Model for Detection and Localization of Cells in Early Stage Embryo Development. | Aisha Khan, Stephen Gould, Mathieu Salzmann |
| 2015 | WACV | Multi-class Semantic Video Segmentation with Exemplar-Based Object Reasoning. | Buyu Liu, Xuming He, Stephen Gould |
| 2014 | ACCV | Reliable Point Correspondences in Scenes Dominated by Highly Reflective and Largely Homogeneous Surfaces. | Srimal Jayawardena, Stephen Gould, Hongdong Li, Marcus Hutter, Richard I. Hartley |
| 2014 | ACCV | Determining Interacting Objects in Human-Centric Activities via Qualitative Spatio-Temporal Reasoning. | Hajar Sadeghi Sokeh, Stephen Gould, Jochen Renz |
| 2014 | CVPR | An Exemplar-Based CRF for Multi-instance Object Segmentation. | Xuming He, Stephen Gould |
| 2014 | DICTA | Reflective Features Detection and Hierarchical Reflections Separation in Image Sequences. | Di Yang, Srimal Jayawardena, Stephen Gould, Marcus Hutter |
| 2014 | ECCV | Superpixel Graph Label Transfer with Learned Distance Metric. | Stephen Gould, Jiecheng Zhao, Xuming He, Yuhang Zhang |
| 2014 | WACV | Joint semantic and geometric segmentation of videos with a stage model. | Buyu Liu, Xuming He, Stephen Gould |
| 2013 | IJCAI | Efficient Extraction and Representation of Spatial Information from Video Data. | Hajar Sadeghi Sokeh, Stephen Gould, Jochen Renz |
| 2012 | ACCV | A Noise Tolerant Watershed Transformation with Viscous Force for Seeded Image Segmentation. | Di Yang, Stephen Gould, Marcus Hutter |
| 2012 | CVPR | Multiclass pixel labeling with non-local matching constraints. | Stephen Gould |
| 2012 | ECCV | PatchMatchGraph: Building a Graph of Dense Patch Correspondences for Label Transfer. | Stephen Gould, Yuhang Zhang |
| 2012 | ECCV | On Learning Higher-Order Consistency Potentials for Multi-class Pixel Labeling. | Kyoungup Park, Stephen Gould |
| 2012 | IVCNZ | Towards unsupervised semantic segmentation of street scenes from motion cues. | Hajar Sadeghi Sokeh, Stephen Gould |
| 2012 | MICCAI | Application of the IMM-JPDA Filter to Multiple Target Tracking in Total Internal Reflection Fluorescence Microscopy Images. | Seyed Hamid Rezatofighi, Stephen Gould, Richard I. Hartley, Katarina Mele, William E. Hughes |
| 2011 | DICTA | Simultaneous Multi-class Pixel Labeling over Coherent Image Sets. | Paul Rivera, Stephen Gould |
| 2011 | ICML | Max-margin Learning for Lower Linear Envelope Potentials in Binary Markov Random Fields. | Stephen Gould |
| 2010 | CVPR | Single image depth estimation from predicted semantic labels. | Beyang Liu, Stephen Gould, Daphne Koller |
| 2010 | ECCV | A Unified Contour-Pixel Model for Figure-Ground Segmentation. | Benjamin Packer, Stephen Gould, Daphne Koller |
| 2010 | ECCV | Discriminative Learning with Latent Variables for Cluttered Indoor Scene Understanding. | Huayan Wang, Stephen Gould, Daphne Koller |
| 2010 | ECCV | Discriminative Learning with Latent Variables for Cluttered Indoor Scene Understanding. | Huayan Wang, Stephen Gould, Daphne Koller |
| 2010 | ICML | Accelerated dual decomposition for MAP inference. | Vladimir Jojic, Stephen Gould, Daphne Koller |
| 2009 | CVPR | Alphabet SOUP: A framework for approximate energy minimization. | Stephen Gould, Fernando Amat, Daphne Koller |
| 2009 | ICCV | Decomposing a scene into geometric and semantically consistent regions. | Stephen Gould, Richard Fulton, Daphne Koller |
| 2009 | ICRA | High-accuracy 3D sensing for mobile manipulation: Improving object detection and door opening. | Morgan Quigley, Siddharth Batra, Stephen Gould, Ellen Klingbeil, Quoc V. Le, Ashley Wellman, Andrew Y. Ng |
| 2008 | UAI | Projected Subgradient Methods for Learning Sparse Gaussians. | John C. Duchi, Stephen Gould, Daphne Koller |
| 2007 | IJCAI | Peripheral-Foveal Vision for Real-time Object Recognition and Tracking in Video. | Stephen Gould, Joakim Arfvidsson, Adrian Kaehler, Benjamin Sapp, Marius Messner, Gary R. Bradski, Paul Baumstarck, Sukwon Chung, Andrew Y. Ng |