Skip to content

Alexander H. Liu

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

24

Venues

6

Active years

2019–2025

Best venue rank

A*

Where they publish

Papers

24 indexed papers, newest first.

YearVenueTitleAuthors
2025ACLSHuBERT: Self-Supervised Sign Language Representation Learning via Multi-Stream Cluster Prediction.Shester Gueuwou, Xiaodan Du, Greg Shakhnarovich, Karen Livescu, Alexander H. Liu
2025ASRUUSAD: Universal Speech and Audio Representation via Distillation.Heng-Jui Chang, Saurabhchand Bhati, James R. Glass, Alexander H. Liu
2025ASRUFull-Duplex-Bench: A Benchmark to Evaluate Full-Duplex Spoken Dialogue Models on Turn-taking Capabilities.Guan-Ting Lin, Jiachen Lian, Tingle Li, Qirui Wang, Gopala Anumanchipalli, Alexander H. Liu, Hung-Yi Lee
2025ICASSPGenerative Speech Foundation Model Pretraining for High-Quality Speech Extraction and Restoration.Pin-Jui Ku, Alexander H. Liu, Roman Korostik, Sung-Feng Huang, Szu-Wei Fu, Ante Jukic
2025ICLRUniWav: Towards Unified Pre-training for Speech Representation Learning and Generation.Alexander H. Liu, Sang-gil Lee, Chao-Han Huck Yang, Yuan Gong, Yu-Chiang Frank Wang, James R. Glass, Rafael Valle, Bryan Catanzaro
2025ICLRFugatto 1: Foundational Generative Audio Transformer Opus 1.Rafael Valle, Rohan Badlani, Zhifeng Kong, Sang-gil Lee, Arushi Goel, Sungwon Kim, Joo Felipe Santos, Shuqi Dai, Siddharth Gururani, Aya Aljafari, Alexander H. Liu, Kevin J. Shih, Ryan Prenger, Wei Ping, Chao-Han Huck Yang, Bryan Catanzaro
2024ACLCodec-SUPERB: An In-Depth Analysis of Sound Codec Models.Haibin Wu, Ho-Lam Chung, Yi-Cheng Lin, Yuan-Kuei Wu, Xuanjun Chen, Yu-Chi Pai, Hsiu-Hsuan Wang, Kai-Wei Chang, Alexander H. Liu, Hung-yi Lee
2024ICASSPRevisiting Self-supervised Learning of Speech Representation from a Mutual Information Perspective.Alexander H. Liu, Sung-Lin Yeh, James R. Glass
2024ICLRListen, Think, and Understand.Yuan Gong, Hongyin Luo, Alexander H. Liu, Leonid Karlinsky, James R. Glass
2024ICLRGenerative Pre-training for Speech with Flow Matching.Alexander H. Liu, Matthew Le, Apoorv Vyas, Bowen Shi, Andros Tjandra, Wei-Ning Hsu
2023ASRUJoint Audio and Speech Understanding.Yuan Gong, Alexander H. Liu, Hongyin Luo, Leonid Karlinsky, James R. Glass
2023ICLRContrastive Audio-Visual Masked Autoencoder.Yuan Gong, Andrew Rouditchenko, Alexander H. Liu, David Harwath, Leonid Karlinsky, Hilde Kuehne, James R. Glass
2023InterspeechSelf-supervised Fine-tuning for Improved Content Representations by Speaker-invariant Clustering.Heng-Jui Chang, Alexander H. Liu, James R. Glass
2022ACLCross-Modal Discrete Representation Learning.Alexander H. Liu, SouYoung Jin, Cheng-I Lai, Andrew Rouditchenko, Aude Oliva, James R. Glass
2022ICASSPOn the Interplay between Sparsity, Naturalness, Intelligibility, and Prosody in Speech Synthesis.Cheng-I Jeff Lai, Erica Cooper, Yang Zhang, Shiyu Chang, Kaizhi Qian, Yi-Lun Liao, Yung-Sung Chuang, Alexander H. Liu, Junichi Yamagishi, David D. Cox, James R. Glass
2022InterspeechSimple and Effective Unsupervised Speech Synthesis.Alexander H. Liu, Cheng-I Lai, Wei-Ning Hsu, Michael Auli, Alexei Baevski, James R. Glass
2021CVPRSpoken Moments: Learning Joint Audio-Visual Representations From Video Descriptions.Mathew Monfort, SouYoung Jin, Alexander H. Liu, David Harwath, Rogrio Feris, James R. Glass, Aude Oliva
2021InterspeechNon-Autoregressive Predictive Coding for Learning Speech Representations from Local Dependencies.Alexander H. Liu, Yu-An Chung, James R. Glass
2020ACLWorse WER, but Better BLEU? Leveraging Word Embedding as Intermediate in Multitask End-to-End Speech Translation.Shun-Po Chuang, Tzu-Wei Sung, Alexander H. Liu, Hung-yi Lee
2020ICASSPSequence-to-Sequence Automatic Speech Recognition with Word Embedding Regularization and Fused Decoding.Alexander H. Liu, Tzu-Wei Sung, Shun-Po Chuang, Hung-yi Lee, Lin-Shan Lee
2020ICASSPTowards Unsupervised Speech Recognition and Synthesis with Quantized Speech Representation Learning.Alexander H. Liu, Tao Tu, Hung-yi Lee, Lin-Shan Lee
2020InterspeechSemi-Supervised Learning for Multi-Speaker Text-to-Speech Synthesis Using Discrete Speech Representation.Tao Tu, Yuan-Jui Chen, Alexander H. Liu, Hung-yi Lee
2019CVPRTowards Scene Understanding: Unsupervised Monocular Depth Estimation With Semantic-Aware Representation.Po-Yi Chen, Alexander H. Liu, Yen-Cheng Liu, Yu-Chiang Frank Wang
2019ICASSPAdversarial Training of End-to-end Speech Recognition Using a Criticizing Language Model.Alexander H. Liu, Hung-yi Lee, Lin-Shan Lee