Honglie Chen
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
12
Venues
4
Active years
2019–2025
Best venue rank
A*
Where they publish
Papers
12 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICASSP | Large Language Models are Strong Audio-Visual Speech Recognition Learners. | Umberto Cappellazzo, Minsu Kim, Honglie Chen, Pingchuan Ma, Stavros Petridis, Daniele Falavigna, Alessio Brutti, Maja Pantic |
| 2025 | ICASSP | Contextual Speech Extraction: Leveraging Textual History as an Implicit Cue for Target Speech Extraction. | Minsu Kim, Rodrigo Mira, Honglie Chen, Stavros Petridis, Maja Pantic |
| 2025 | Interspeech | Revival with Voice: Multi-modal Controllable Text-to-Speech Synthesis. | Minsu Kim, Pingchuan Ma, Honglie Chen, Stavros Petridis, Maja Pantic |
| 2024 | Interspeech | RT-LA-VocE: Real-Time Low-SNR Audio-Visual Speech Enhancement. | Honglie Chen, Rodrigo Mira, Stavros Petridis, Maja Pantic |
| 2024 | Interspeech | MSRS: Training Multimodal Speech Recognition Models from Scratch with Sparse Mask Optimization. | Adriana Fernandez-Lopez, Honglie Chen, Pingchuan Ma, Lu Yin, Qiao Xiao, Stavros Petridis, Shiwei Liu, Maja Pantic |
| 2023 | CVPR | SynthVSR: Scaling Up Visual Speech RecognitionWith Synthetic Supervision. | Xubo Liu, Egor Lakomkin, Konstantinos Vougioukas, Pingchuan Ma, Honglie Chen, Ruiming Xie, Morrie Doulaty, Niko Moritz, Jchym Kolr, Stavros Petridis, Maja Pantic, Christian Fuegen |
| 2023 | ICASSP | Auto-AVSR: Audio-Visual Speech Recognition with Automatic Labels. | Pingchuan Ma, Alexandros Haliassos, Adriana Fernandez-Lopez, Honglie Chen, Stavros Petridis, Maja Pantic |
| 2023 | Interspeech | SparseVSR: Lightweight and Noise Robust Visual Speech Recognition. | Adriana Fernandez-Lopez, Honglie Chen, Pingchuan Ma, Alexandros Haliassos, Stavros Petridis, Maja Pantic |
| 2021 | BMVC | Audio-Visual Synchronisation in the wild. | Triantafyllos Afouras, Honglie Chen, Weidi Xie, Arsha Nagrani, Andrea Vedaldi, Andrew Zisserman |
| 2021 | CVPR | Localizing Visual Sounds the Hard Way. | Honglie Chen, Weidi Xie, Triantafyllos Afouras, Arsha Nagrani, Andrea Vedaldi, Andrew Zisserman |
| 2020 | ICASSP | Vggsound: A Large-Scale Audio-Visual Dataset. | Honglie Chen, Weidi Xie, Andrea Vedaldi, Andrew Zisserman |
| 2019 | BMVC | AutoCorrect: Deep Inductive Alignment of Noisy Geometric Annotations. | Honglie Chen, Weidi Xie, Andrea Vedaldi, Andrew Zisserman |