| 2025 | ICCV | VGGSounder: Audio-Visual Evaluations for Foundation Models. | Daniil Zverev, Thaddus Wiedemer, Ameya Prabhu, Matthias Bethge, Wieland Brendel, A. Sophia Koepke |
| 2024 | CVPR | Audio-Visual Generalized Zero-Shot Learning using Pre-Trained Large Multi-Modal Models. | David Kurzendrfer, Otniel-Bogdan Mercea, A. Sophia Koepke, Zeynep Akata |
| 2024 | ICASSP | A Sound Approach: Using Large Language Models to Generate Audio Descriptions for Egocentric Text-Audio Retrieval. | Andreea-Maria Oncescu, Joo F. Henriques, Andrew Zisserman, Samuel Albanie, A. Sophia Koepke |
| 2024 | ICLR | Fantastic Gains and Where to Find Them: On the Existence and Prospect of General Knowledge Transfer between Any Pretrained Model. | Karsten Roth, Lukas Thede, A. Sophia Koepke, Oriol Vinyals, Olivier J. Hnaff, Zeynep Akata |
| 2023 | BMVC | Video-adverb retrieval with compositional adverb-action embeddings. | Thomas Hummel, Otniel-Bogdan Mercea, A. Sophia Koepke, Zeynep Akata |
| 2023 | CVPR | Exposing and Mitigating Spurious Correlations for Cross-Modal Retrieval. | Jae-Myung Kim, A. Sophia Koepke, Cordelia Schmid, Zeynep Akata |
| 2023 | ICCV | Image-free Classifier Injection for Zero-Shot Classification. | Anders Christensen, Massimiliano Mancini, A. Sophia Koepke, Ole Winther, Zeynep Akata |
| 2023 | ICCV | Waffling around for Performance: Visual Classification with Random Words and Broad Concepts. | Karsten Roth, Jae-Myung Kim, A. Sophia Koepke, Oriol Vinyals, Cordelia Schmid, Zeynep Akata |
| 2022 | CoRL | PlanT: Explainable Planning Transformers via Object-Level Representations. | Katrin Renz, Kashyap Chitta, Otniel-Bogdan Mercea, A. Sophia Koepke, Zeynep Akata, Andreas Geiger |
| 2022 | CVPR | Audiovisual Generalised Zero-shot Learning with Cross-modal Attention and Language. | Otniel-Bogdan Mercea, Lukas Riesch, A. Sophia Koepke, Zeynep Akata |
| 2022 | ECCV | Temporal and Cross-modal Attention for Audio-Visual Zero-Shot Learning. | Otniel-Bogdan Mercea, Thomas Hummel, A. Sophia Koepke, Zeynep Akata |
| 2021 | CVPR | Distilling Audio-Visual Knowledge by Compositional Contrastive Learning. | Yanbei Chen, Yongqin Xian, A. Sophia Koepke, Ying Shan, Zeynep Akata |
| 2021 | Interspeech | Audio Retrieval with Natural Language Queries. | Andreea-Maria Oncescu, A. Sophia Koepke, Joo F. Henriques, Zeynep Akata, Samuel Albanie |
| 2020 | ICASSP | Sight to Sound: An End-to-End Approach for Visual Piano Transcription. | A. Sophia Koepke, Olivia Wiles, Yael Moses, Andrew Zisserman |
| 2020 | ICML | CLEVR-X: A Visual Reasoning Dataset for Natural Language Explanations. | Leonard Salewski, A. Sophia Koepke, Hendrik P. A. Lensch, Zeynep Akata |
| 2018 | BMVC | Self-supervised learning of a facial attribute embedding from video. | A. Sophia Koepke, Olivia Wiles, Andrew Zisserman |
| 2018 | ECCV | X2Face: A Network for Controlling Face Generation Using Images, Audio, and Pose Codes. | Olivia Wiles, A. Sophia Koepke, Andrew Zisserman |