Skip to content

Stephen Gould

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

107

Venues

19

Active years

2007–2025

Best venue rank

A*

Where they publish

Papers

107 indexed papers, newest first.

YearVenueTitleAuthors
2025CVPRPos3R: 6D Pose Estimation for Unseen Objects Made Easy.Weijian Deng, Dylan Campbell, Chunyi Sun, Jiahao Zhang, Shubham Kanitkar, Matthew E. Shaffer, Stephen Gould
2025CVPRVI^3NR: Variance Informed Initialization for Implicit Neural Representations.Chamin Hewa Koneputugodage, Yizhak Ben-Shabat, Sameera Ramasinghe, Stephen Gould
2025ICCVLeaps and Bounds: An Improved Point Cloud Winding Number Formulation for Fast Normal Estimation and Surface Reconstruction.Chamin Hewa Koneputugodage, Dylan Campbell, Stephen Gould
2025ICCVManual-PA: Learning 3D Part Assembly from Instruction Diagrams.Jiahao Zhang, Anoop Cherian, Cristian Rodriguez, Weijian Deng, Stephen Gould
2025ICMLCan We Predict Performance of Large Models across Vision-Language Tasks?Qinyu Zhao, Ming Xu, Kartik Gupta, Akshay Asthana, Liang Zheng, Stephen Gould
2025IJCNNScaling Prompt Instructed Zero Shot Composed Image Retrieval with Image-Only Data.Yiqun Duan, Sameera Ramasinghe, Stephen Gould, Thalaiyasingam Ajanthan
2025SIGGRAPHUnsupervised Decomposition of 3D Shapes into Expressive and Editable Extruded Profile Primitives.Chunyi Sun, Junlin Han, Runjia Li, Weijian Deng, Dylan Campbell, Stephen Gould
2025WACVTemporally Grounding Instructional Diagrams in Unconstrained Videos.Jiahao Zhang, Frederic Z. Zhang, Cristian Rodriguez, Yizhak Ben-Shabat, Anoop Cherian, Stephen Gould
2024CVPR3DInAction: Understanding Human Actions in 3D Point Clouds.Yizhak Ben-Shabat, Oren Shrout, Stephen Gould
2024CVPRDifferentiable Neural Surface Refinement for Modeling Transparent Objects.Weijian Deng, Dylan Campbell, Chunyi Sun, Shubham Kanitkar, Matthew E. Shaffer, Stephen Gould
2024CVPRLearning to Select Views for Efficient Multi-View Understanding.Yunzhong Hou, Stephen Gould, Liang Zheng
2024CVPRSmall Steps and Level Sets: Fitting Neural Surface Models with Point Guidance.Chamin Hewa Koneputugodage, Yizhak Ben-Shabat, Dylan Campbell, Stephen Gould
2024CVPRTemporally Consistent Unbalanced Optimal Transport for Unsupervised Action Segmentation.Ming Xu, Stephen Gould
2024ECCVUnsupervised Dense Prediction Using Differentiable Normalized Cuts.Yanbin Liu, Stephen Gould
2024ECCVThe First to Know: How Token Distributions Reveal Hidden Knowledge in Large Vision-Language Models?Qinyu Zhao, Ming Xu, Kartik Gupta, Akshay Asthana, Liang Zheng, Stephen Gould
2024ICAPSNeuro-Symbolic Learning of Lifted Action Models from Visual Traces.Kai Xi, Stephen Gould, Sylvie Thibaux
2024ICLRTowards Optimal Feature-Shaping Methods for Out-of-Distribution Detection.Qinyu Zhao, Ming Xu, Kartik Gupta, Akshay Asthana, Liang Zheng, Stephen Gould
2024ICMLAn Empirical Study Into What Matters for Calibrating Vision-Language Models.Weijie Tu, Weijian Deng, Dylan Campbell, Stephen Gould, Tom Gedeon
2024WACVBi-directional Training for Composed Image Retrieval via Text Prompt Learning.Zheyuan Liu, Weixuan Sun, Yicong Hong, Damien Teney, Stephen Gould
2024WACVIKEA Ego 3D Dataset: Understanding furniture assembly actions from ego-view 3D Point Clouds.Yizhak Ben-Shabat, Jonathan Paul, Eviatar Segev, Oren Shrout, Stephen Gould
2024WACVRay Deformation Networks for Novel View Synthesis of Refractive Objects.Weijian Deng, Dylan Campbell, Chunyi Sun, Shubham Kanitkar, Matthew E. Shaffer, Stephen Gould
2024WACVLipAT: Beyond Style Transfer for Controllable Neural Simulation of Lipstick using Cosmetic Attributes.Amila Silva, Olga Moskvyak, Alexander Long, Ravi Garg, Stephen Gould, Gil Avraham, Anton van den Hengel
2024WACVNeRFEditor: Differentiable Style Decomposition for 3D Scene Editing.Chunyi Sun, Yanbin Liu, Junlin Han, Stephen Gould
2023CVPROctree Guided Unoriented Surface Reconstruction.Chamin Hewa Koneputugodage, Yizhak Ben-Shabat, Stephen Gould
2023CVPRHigh-Fidelity Guided Image Synthesis with Latent Diffusion Models.Jaskirat Singh, Stephen Gould, Liang Zheng
2023CVPRAligning Step-by-Step Instructional Diagrams to Video Demonstrations.Jiahao Zhang, Anoop Cherian, Yanbin Liu, Yizhak Ben-Shabat, Cristian Rodriguez Opazo, Stephen Gould
2023ICCVLearning Navigational Visual Representations with Semantic Map Supervision.Yicong Hong, Yang Zhou, Ruiyi Zhang, Franck Dernoncourt, Trung Bui, Stephen Gould, Hao Tan
2023ICCVSemi-Supervised Semantic Segmentation under Label Noise via Diverse Learning Groups.Peixia Li, Pulak Purkait, Thalaiyasingam Ajanthan, Majid Abdolshah, Ravi Garg, Hisham Husain, Chenchen Xu, Stephen Gould, Wanli Ouyang, Anton van den Hengel
2023ICCVScaling Data Generation in Vision-and-Language Navigation.Zun Wang, Jialu Li, Yicong Hong, Yi Wang, Qi Wu, Mohit Bansal, Stephen Gould, Hao Tan, Yu Qiao
2023ICCVExploring Predicate Visual Context in Detecting of Human-Object Interactions.Frederic Z. Zhang, Yuhui Yuan, Dylan Campbell, Zhuoyao Zhong, Stephen Gould
2023ICLRDeep Declarative Dynamic Time Warping for End-to-End Learning of Alignment Paths.Ming Xu, Sourav Garg, Michael Milford, Stephen Gould
2023ICMLConfidence and Dispersity Speak: Characterizing Prediction Matrix for Unsupervised Accuracy Estimation.Weijian Deng, Yumin Suh, Stephen Gould, Liang Zheng
2022CVPRDiGS : Divergence guided shape implicit neural representation for unoriented point clouds.Yizhak Ben-Shabat, Chamin Hewa Koneputugodage, Stephen Gould
2022CVPRBridging the Gap Between Learning in Discrete and Continuous Environments for Vision-and-Language Navigation.Yicong Hong, Zun Wang, Qi Wu, Stephen Gould
2022CVPREfficient Two-Stage Detection of Human-Object Interactions with a Novel Unary-Pairwise Transformer.Frederic Z. Zhang, Dylan Campbell, Stephen Gould
2022IROSGoferBot: A Visual Guided Human-Robot Collaborative Assembly System.Zheyu Zhuang, Yizhak Ben-Shabat, Jiahao Zhang, Stephen Gould, Robert E. Mahony
2021CVPRVLN BERT: A Recurrent Vision-and-Language BERT for Navigation.Yicong Hong, Qi Wu, Yuankai Qi, Cristian Rodriguez Opazo, Stephen Gould
2021CVPRProbabilistic Tracklet Scoring and Inpainting for Multiple Object Tracking.Fatemeh Sadat Saleh, Sadegh Aliakbarian, Hamid Rezatofighi, Mathieu Salzmann, Stephen Gould
2021ICCVImage Retrieval on Real-life Images with Pre-trained Vision-and-Language Models.Zheyuan Liu, Cristian Rodriguez Opazo, Damien Teney, Stephen Gould
2021ICCVContextually Plausible and Diverse 3D Human Motion Prediction.Sadegh Aliakbarian, Fatemeh Sadat Saleh, Lars Petersson, Stephen Gould, Mathieu Salzmann
2021ICCVSpatially Conditioned Graphs for Detecting Human-Object Interactions.Frederic Z. Zhang, Dylan Campbell, Stephen Gould
2021ICDMA Regularized Wasserstein Framework for Graph Kernels.Asiri Wijesinghe, Qing Wang, Stephen Gould
2021ICLRConditional Generative Modeling via Learning the Latent Space.Sameera Ramasinghe, Kanchana Nisal Ranasinghe, Salman H. Khan, Nick Barnes, Stephen Gould
2021ICMLWhat Does Rotation Prediction Tell Us about Classifier Accuracy under Varying Testing Environments?Weijian Deng, Stephen Gould, Liang Zheng
2021WACVThe IKEA ASM Dataset: Understanding People Assembling Furniture through Actions, Objects and Pose.Yizhak Ben-Shabat, Xin Yu, Fatemeh Sadat Saleh, Dylan Campbell, Cristian Rodriguez Opazo, Hongdong Li, Stephen Gould
2021WACVDORi: Discovering Object Relationships for Moment Localization of a Natural Language Query in a Video.Cristian Rodriguez Opazo, Edison Marrese-Taylor, Basura Fernando, Hongdong Li, Stephen Gould
2020CVPRA Stochastic Conditioning Scheme for Diverse Human Motion Prediction.Mohammad Sadegh Aliakbarian, Fatemeh Sadat Saleh, Mathieu Salzmann, Lars Petersson, Stephen Gould
2020CVPRInferring Temporal Compositions of Actions Using Probabilistic Automata.Rodrigo Santa Cruz, Anoop Cherian, Basura Fernando, Dylan Campbell, Stephen Gould
2020CVPRLearning to Structure an Image With Few Colors.Yunzhong Hou, Liang Zheng, Stephen Gould
2020ECCVDeepFit: 3D Surface Fitting via Neural Network Weighted Least Squares.Yizhak Ben-Shabat, Stephen Gould
2020ECCVSolving the Blind Perspective-n-Point Problem End-to-End with Robust Differentiable Geometric Optimization.Dylan Campbell, Liu Liu, Stephen Gould
2020ECCVMultiview Detection with Feature Perspective Transformation.Yunzhong Hou, Liang Zheng, Stephen Gould
2020EMNLPSub-Instruction Aware Vision-and-Language Navigation.Yicong Hong, Cristian Rodriguez Opazo, Qi Wu, Stephen Gould
2020ICLRA Signal Propagation Perspective for Pruning Neural Networks at Initialization.Namhoon Lee, Thalaiyasingam Ajanthan, Stephen Gould, Philip H. S. Torr
2020IROSSpectral-GANs for High-Resolution 3D Point-cloud Generation.Sameera Ramasinghe, Salman H. Khan, Nick Barnes, Stephen Gould
2020WACVProposal-free Temporal Moment Localization of a Natural-Language Query in Video using Guided Attention.Cristian Rodriguez Opazo, Edison Marrese-Taylor, Fatemeh Sadat Saleh, Hongdong Li, Stephen Gould
2020WACVBlended Convolution and Synthesis for Efficient Discrimination of 3D Shapes.Sameera Ramasinghe, Salman H. Khan, Nick Barnes, Stephen Gould
2019CVPRThe Alignment of the Spheres: Globally-Optimal Spherical Mixture Alignment for Camera Pose Estimation.Dylan Campbell, Lars Petersson, Laurent Kneip, Hongdong Li, Stephen Gould
2019ICCVLearning to Find Common Objects Across Few Image Collections.Amirreza Shaban, Amir Rahimi, Shray Bansal, Stephen Gould, Byron Boots, Richard Hartley
2018CVPRBottom-Up and Top-Down Attention for Image Captioning and Visual Question Answering.Peter Anderson, Xiaodong He, Chris Buehler, Damien Teney, Mark Johnson, Stephen Gould, Lei Zhang
2018CVPRVision-and-Language Navigation: Interpreting Visually-Grounded Navigation Instructions in Real Environments.Peter Anderson, Qi Wu, Damien Teney, Jake Bruce, Mark Johnson, Niko Snderhauf, Ian D. Reid, Stephen Gould, Anton van den Hengel
2018CVPRNon-Linear Temporal Subspace Representations for Activity Recognition.Anoop Cherian, Suvrit Sra, Stephen Gould, Richard Hartley
2018CVPRVideo Representation Learning Using Discriminative Pooling.Jue Wang, Anoop Cherian, Fatih Porikli, Stephen Gould
2018WACVNeural Algebra of Classifiers.Rodrigo Santa Cruz, Basura Fernando, Anoop Cherian, Stephen Gould
2017CVPRGeneralized Rank Pooling for Activity Recognition.Anoop Cherian, Basura Fernando, Mehrtash Harandi, Stephen Gould
2017CVPRDeepPermNet: Visual Permutation Learning.Rodrigo Santa Cruz, Basura Fernando, Anoop Cherian, Stephen Gould
2017CVPRSelf-Supervised Video Representation Learning with Odd-One-Out Networks.Basura Fernando, Hakan Bilen, Efstratios Gavves, Stephen Gould
2017CVPRUnsupervised Human Action Detection by Action Matching.Basura Fernando, Sareh Shirazi, Stephen Gould
2017DICTAHuman Pose Forecasting via Deep Markov Models.Sam Toyer, Anoop Cherian, Tengda Han, Stephen Gould
2017EMNLPGuided Open Vocabulary Image Captioning with Constrained Beam Search.Peter Anderson, Basura Fernando, Mark Johnson, Stephen Gould
2017WACVHigher-Order Pooling of CNN Features via Kernel Linearization for Action Recognition.Anoop Cherian, Piotr Koniusz, Stephen Gould
2016CVPRDynamic Image Networks for Action Recognition.Hakan Bilen, Basura Fernando, Efstratios Gavves, Andrea Vedaldi, Stephen Gould
2016CVPRDiscriminative Hierarchical Rank Pooling for Activity Recognition.Basura Fernando, Peter Anderson, Marcus Hutter, Stephen Gould
2016DICTADepth Dropout: Efficient Training of Residual Convolutional Neural Networks.Jian Guo, Stephen Gould
2016ECCVSPICE: Semantic Propositional Image Caption Evaluation.Peter Anderson, Basura Fernando, Mark Johnson, Stephen Gould
2016ECCVDeep Convolutional Neural Networks for Human Embryonic Cell Counting.Aisha Khan, Stephen Gould, Mathieu Salzmann
2016ECCVBuilt-in Foreground/Background Prior for Weakly-Supervised Semantic Segmentation.Fatemehsadat Saleh, Mohammad Sadegh Ali Akbarian, Mathieu Salzmann, Lars Petersson, Stephen Gould, Jos M. lvarez
2016ICMLLearning End-to-end Video Classification with Rank-Pooling.Basura Fernando, Stephen Gould
2015ICCVHierarchical Higher-Order Regression Forest Fields: An Application to 3D Indoor Scene Labelling.Trung-Thanh Pham, Ian D. Reid, Yasir Latif, Stephen Gould
2015MICCAIDetecting Abnormal Cell Division Patterns in Early Stage Human Embryo Development.Aisha Khan, Stephen Gould, Mathieu Salzmann
2015WACVA Linear Chain Markov Model for Detection and Localization of Cells in Early Stage Embryo Development.Aisha Khan, Stephen Gould, Mathieu Salzmann
2015WACVMulti-class Semantic Video Segmentation with Exemplar-Based Object Reasoning.Buyu Liu, Xuming He, Stephen Gould
2014ACCVReliable Point Correspondences in Scenes Dominated by Highly Reflective and Largely Homogeneous Surfaces.Srimal Jayawardena, Stephen Gould, Hongdong Li, Marcus Hutter, Richard I. Hartley
2014ACCVDetermining Interacting Objects in Human-Centric Activities via Qualitative Spatio-Temporal Reasoning.Hajar Sadeghi Sokeh, Stephen Gould, Jochen Renz
2014CVPRAn Exemplar-Based CRF for Multi-instance Object Segmentation.Xuming He, Stephen Gould
2014DICTAReflective Features Detection and Hierarchical Reflections Separation in Image Sequences.Di Yang, Srimal Jayawardena, Stephen Gould, Marcus Hutter
2014ECCVSuperpixel Graph Label Transfer with Learned Distance Metric.Stephen Gould, Jiecheng Zhao, Xuming He, Yuhang Zhang
2014WACVJoint semantic and geometric segmentation of videos with a stage model.Buyu Liu, Xuming He, Stephen Gould
2013IJCAIEfficient Extraction and Representation of Spatial Information from Video Data.Hajar Sadeghi Sokeh, Stephen Gould, Jochen Renz
2012ACCVA Noise Tolerant Watershed Transformation with Viscous Force for Seeded Image Segmentation.Di Yang, Stephen Gould, Marcus Hutter
2012CVPRMulticlass pixel labeling with non-local matching constraints.Stephen Gould
2012ECCVPatchMatchGraph: Building a Graph of Dense Patch Correspondences for Label Transfer.Stephen Gould, Yuhang Zhang
2012ECCVOn Learning Higher-Order Consistency Potentials for Multi-class Pixel Labeling.Kyoungup Park, Stephen Gould
2012IVCNZTowards unsupervised semantic segmentation of street scenes from motion cues.Hajar Sadeghi Sokeh, Stephen Gould
2012MICCAIApplication of the IMM-JPDA Filter to Multiple Target Tracking in Total Internal Reflection Fluorescence Microscopy Images.Seyed Hamid Rezatofighi, Stephen Gould, Richard I. Hartley, Katarina Mele, William E. Hughes
2011DICTASimultaneous Multi-class Pixel Labeling over Coherent Image Sets.Paul Rivera, Stephen Gould
2011ICMLMax-margin Learning for Lower Linear Envelope Potentials in Binary Markov Random Fields.Stephen Gould
2010CVPRSingle image depth estimation from predicted semantic labels.Beyang Liu, Stephen Gould, Daphne Koller
2010ECCVA Unified Contour-Pixel Model for Figure-Ground Segmentation.Benjamin Packer, Stephen Gould, Daphne Koller
2010ECCVDiscriminative Learning with Latent Variables for Cluttered Indoor Scene Understanding.Huayan Wang, Stephen Gould, Daphne Koller
2010ECCVDiscriminative Learning with Latent Variables for Cluttered Indoor Scene Understanding.Huayan Wang, Stephen Gould, Daphne Koller
2010ICMLAccelerated dual decomposition for MAP inference.Vladimir Jojic, Stephen Gould, Daphne Koller
2009CVPRAlphabet SOUP: A framework for approximate energy minimization.Stephen Gould, Fernando Amat, Daphne Koller
2009ICCVDecomposing a scene into geometric and semantically consistent regions.Stephen Gould, Richard Fulton, Daphne Koller
2009ICRAHigh-accuracy 3D sensing for mobile manipulation: Improving object detection and door opening.Morgan Quigley, Siddharth Batra, Stephen Gould, Ellen Klingbeil, Quoc V. Le, Ashley Wellman, Andrew Y. Ng
2008UAIProjected Subgradient Methods for Learning Sparse Gaussians.John C. Duchi, Stephen Gould, Daphne Koller
2007IJCAIPeripheral-Foveal Vision for Real-time Object Recognition and Tracking in Video.Stephen Gould, Joakim Arfvidsson, Adrian Kaehler, Benjamin Sapp, Marius Messner, Gary R. Bradski, Paul Baumstarck, Sukwon Chung, Andrew Y. Ng