| 2025 | ICLR | SSLAM: Enhancing Self-Supervised Models with Audio Mixtures for Polyphonic Soundscapes. | Tony Alex, Sara Atito, Armin Mustafa, Muhammad Awais, Philip J. B. Jackson |
| 2025 | ICLR | NarrativeBridge: Enhancing Video Captioning with Causal-Temporal Narrative. | Asmar Nadeem, Faegheh Sardari, Robert Dawes, Syed Sameed Husain, Adrian Hilton, Armin Mustafa |
| 2024 | AAAI | DTF-AT: Decoupled Time-Frequency Audio Transformer for Event Classification. | Tony Alex, Sara Ahmed, Armin Mustafa, Muhammad Awais, Philip J. B. Jackson |
| 2024 | CVPR | S3R-Net: A Single-Stage Approach to Self-Supervised Shadow Removal. | Nikolina Kubiak, Armin Mustafa, Graeme Phillipson, Stephen Jolly, Simon Hadfield |
| 2024 | ECCV | Attend-Fusion: Efficient Audio-Visual Fusion for Video Classification. | Mahrukh Awan, Asmar Nadeem, Muhammad Junaid Awan, Armin Mustafa, Syed Sameed Husain |
| 2024 | ECCV | ViscoNet: Bridging and Harmonizing Visual and Textual Conditioning for ControlNet. | Soon Yau Cheong, Armin Mustafa, Andrew Gilbert |
| 2024 | ECCV | RenDetNet: Weakly-Supervised Shadow Detection with Shadow Caster Verification. | Nikolina Kubiak, Elliot Wortman, Armin Mustafa, Graeme Phillipson, Stephen Jolly, Simon Hadfield |
| 2024 | ECCV | CoLeaF: A Contrastive-Collaborative Learning Framework for Weakly Supervised Audio-Visual Video Parsing. | Faegheh Sardari, Armin Mustafa, Philip J. B. Jackson, Adrian Hilton |
| 2024 | ICASSP | Max-AST: Combining Convolution, Local and Global Self-Attentions for Audio Event Classification. | Tony Alex, Sara Ahmed, Armin Mustafa, Muhammad Awais, Philip J. B. Jackson |
| 2024 | WACV | CAD - Contextual Multi-modal Alignment for Dynamic AVQA. | Asmar Nadeem, Adrian Hilton, Robert Dawes, Graham Thomas, Armin Mustafa |
| 2023 | CVPR | SEM-POS: Grammatically and Semantically Correct Video Captioning. | Asmar Nadeem, Adrian Hilton, Robert Dawes, Graham A. Thomas, Armin Mustafa |
| 2022 | BMVC | KPE: Keypoint Pose Encoding for Transformer-based Image Generation. | Soon Yau Cheong, Armin Mustafa, Andrew Gilbert |
| 2021 | BMVC | SILT: Self-supervised Lighting Transfer Using Implicit Image Decomposition. | Nikolina Kubiak, Armin Mustafa, Graeme Phillipson, Stephen Jolly, Simon Hadfield |
| 2021 | CVPR | Temporal Consistency Loss for High Resolution Textured and Clothed 3D Human Reconstruction From Monocular Video. | Akin Caliskan, Armin Mustafa, Adrian Hilton |
| 2021 | CVPR | Multi-Person Implicit Reconstruction From a Single Image. | Armin Mustafa, Akin Caliskan, Lourdes Agapito, Adrian Hilton |
| 2020 | ACCV | Multi-view Consistency Loss for Improved Single-Image 3D Reconstruction of Clothed People. | Akin Caliskan, Armin Mustafa, Evren Imre, Adrian Hilton |
| 2020 | ICRA | A*3D Dataset: Towards Autonomous Driving in Challenging Environments. | Quang-Hieu Pham, Pierre Sevestre, Ramanpreet Singh Pahwa, Huijing Zhan, Chun Ho Pang, Yuda Chen, Armin Mustafa, Vijay Chandrasekhar, Jie Lin |
| 2019 | ICCV | U4D: Unsupervised 4D Dynamic Scene Understanding. | Armin Mustafa, Chris Russell, Adrian Hilton |
| 2017 | CVPR | Semantically Coherent Co-Segmentation and Reconstruction of Dynamic Scenes. | Armin Mustafa, Adrian Hilton |
| 2016 | CVPR | Temporally Coherent 4D Reconstruction of Complex Dynamic Scenes. | Armin Mustafa, Hansung Kim, Jean-Yves Guillemaut, Adrian Hilton |
| 2016 | ECCV | 4D Match Trees for Non-rigid Surface Alignment. | Armin Mustafa, Hansung Kim, Adrian Hilton |
| 2015 | ICCV | General Dynamic Scene Reconstruction from Multiple View Video. | Armin Mustafa, Hansung Kim, Jean-Yves Guillemaut, Adrian Hilton |