Skip to content

Honglie Chen

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

12

Venues

4

Active years

2019–2025

Best venue rank

A*

Where they publish

Papers

12 indexed papers, newest first.

YearVenueTitleAuthors
2025ICASSPLarge Language Models are Strong Audio-Visual Speech Recognition Learners.Umberto Cappellazzo, Minsu Kim, Honglie Chen, Pingchuan Ma, Stavros Petridis, Daniele Falavigna, Alessio Brutti, Maja Pantic
2025ICASSPContextual Speech Extraction: Leveraging Textual History as an Implicit Cue for Target Speech Extraction.Minsu Kim, Rodrigo Mira, Honglie Chen, Stavros Petridis, Maja Pantic
2025InterspeechRevival with Voice: Multi-modal Controllable Text-to-Speech Synthesis.Minsu Kim, Pingchuan Ma, Honglie Chen, Stavros Petridis, Maja Pantic
2024InterspeechRT-LA-VocE: Real-Time Low-SNR Audio-Visual Speech Enhancement.Honglie Chen, Rodrigo Mira, Stavros Petridis, Maja Pantic
2024InterspeechMSRS: Training Multimodal Speech Recognition Models from Scratch with Sparse Mask Optimization.Adriana Fernandez-Lopez, Honglie Chen, Pingchuan Ma, Lu Yin, Qiao Xiao, Stavros Petridis, Shiwei Liu, Maja Pantic
2023CVPRSynthVSR: Scaling Up Visual Speech RecognitionWith Synthetic Supervision.Xubo Liu, Egor Lakomkin, Konstantinos Vougioukas, Pingchuan Ma, Honglie Chen, Ruiming Xie, Morrie Doulaty, Niko Moritz, Jchym Kolr, Stavros Petridis, Maja Pantic, Christian Fuegen
2023ICASSPAuto-AVSR: Audio-Visual Speech Recognition with Automatic Labels.Pingchuan Ma, Alexandros Haliassos, Adriana Fernandez-Lopez, Honglie Chen, Stavros Petridis, Maja Pantic
2023InterspeechSparseVSR: Lightweight and Noise Robust Visual Speech Recognition.Adriana Fernandez-Lopez, Honglie Chen, Pingchuan Ma, Alexandros Haliassos, Stavros Petridis, Maja Pantic
2021BMVCAudio-Visual Synchronisation in the wild.Triantafyllos Afouras, Honglie Chen, Weidi Xie, Arsha Nagrani, Andrea Vedaldi, Andrew Zisserman
2021CVPRLocalizing Visual Sounds the Hard Way.Honglie Chen, Weidi Xie, Triantafyllos Afouras, Arsha Nagrani, Andrea Vedaldi, Andrew Zisserman
2020ICASSPVggsound: A Large-Scale Audio-Visual Dataset.Honglie Chen, Weidi Xie, Andrea Vedaldi, Andrew Zisserman
2019BMVCAutoCorrect: Deep Inductive Alignment of Noisy Geometric Annotations.Honglie Chen, Weidi Xie, Andrea Vedaldi, Andrew Zisserman