Skip to content

Hank Liao

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

24

Venues

4

Active years

2005–2024

Best venue rank

Multiconference

Where they publish

Papers

24 indexed papers, newest first.

YearVenueTitleAuthors
2024ICASSPConformer is All You Need for Visual Speech Recognition.Oscar Chang, Hank Liao, Dmitriy Serdyuk, Ankit Shahy, Olivier Siohan
2024ICASSPUSM-SCD: Multilingual Speaker Change Detection Based on Large Pretrained Foundation Models.Guanlong Zhao, Yongqiang Wang, Jason Pelecanos, Yu Zhang, Hank Liao, Yiling Huang, Han Lu, Quan Wang
2024InterspeechOn the Success and Limitations of Auxiliary Network Based Word-Level End-to-End Neural Speaker Diarization.Yiling Huang, Weiran Wang, Guanlong Zhao, Hank Liao, Wei Xia, Quan Wang
2024InterspeechDiarizationLM: Speaker Diarization Post-Processing with Large Language Models.Quan Wang, Yiling Huang, Guanlong Zhao, Evan Clark, Wei Xia, Hank Liao
2020ICASSPEnd-to-End Multi-Person Audio/Visual Automatic Speech Recognition.Otavio Braga, Takaki Makino, Olivier Siohan, Hank Liao
2019ASRUA Comparison of End-to-End Models for Long-Form Speech Recognition.Chung-Cheng Chiu, Anjuli Kannan, Rohit Prabhavalkar, Zhifeng Chen, Tara N. Sainath, Yonghui Wu, Wei Han, Yu Zhang, Ruoming Pang, Sergey Kishchenko, Patrick Nguyen, Arun Narayanan, Hank Liao, Shuyuan Zhang
2019ASRURecurrent Neural Network Transducer for Audio-Visual Speech Recognition.Takaki Makino, Hank Liao, Yannis M. Assael, Brendan Shillingford, Basilio Garcia, Otavio Braga, Olivier Siohan
2019InterspeechLarge-Scale Visual Speech Recognition.Brendan Shillingford, Yannis M. Assael, Matthew W. Hoffman, Thomas Paine, Can Hughes, Utsav Prabhu, Hank Liao, Hasim Sak, Kanishka Rao, Lorrayne Bennett, Marie Mulville, Misha Denil, Ben Coppin, Ben Laurie, Andrew W. Senior, Nando de Freitas
2018ICASSPRADMM: Recurrent Adaptive Mixture Model with Applications to Domain Robust Language Modeling.Kazuki Irie, Shankar Kumar, Michael Nirschl, Hank Liao
2017ASRULattice rescoring strategies for long short term memory language models in speech recognition.Shankar Kumar, Michael Nirschl, Daniel Niels Holtmann-Rice, Hank Liao, Ananda Theertha Suresh, Felix X. Yu
2017ASRUReducing the computational complexity for whole word models.Hagen Soltau, Hank Liao, Hasim Sak
2017InterspeechNeural Speech Recognizer: Acoustic-to-Word LSTM Model for Large Vocabulary Speech Recognition.Hagen Soltau, Hank Liao, Hasim Sak
2016InterspeechLearning N-Gram Language Models from Uncertain Data.Vitaly Kuznetsov, Hank Liao, Mehryar Mohri, Michael Riley, Brian Roark
2015ICASSPExemplar-based large vocabulary speech recognition using k-nearest neighbors.Yanbo Xu, Olivier Siohan, David Simcha, Sanjiv Kumar, Hank Liao
2015InterspeechLarge vocabulary automatic speech recognition for children.Hank Liao, Golan Pundak, Olivier Siohan, Melissa K. Carroll, Noah Coccaro, Qi-Ming Jiang, Tara N. Sainath, Andrew W. Senior, Franoise Beaufays, Michiel Bacchiani
2014ICASSPGMM-free DNN acoustic model training.Andrew W. Senior, Georg Heigold, Michiel Bacchiani, Hank Liao
2013ASRULarge scale deep neural network acoustic modeling with semi-supervised training data for YouTube video transcription.Hank Liao, Erik McDermott, Andrew W. Senior
2013ICASSPSpeaker adaptation of context dependent deep neural networks.Hank Liao
2012ICMIICMI'12 grand challenge: haptic voice recognition.Khe Chai Sim, Shengdong Zhao, Kai Yu, Hank Liao
2010InterspeechDecision tree state clustering with word and syllable features.Hank Liao, Christopher Alberti, Michiel Bacchiani, Olivier Siohan
2009ICASSPAn audio indexing system for election video material.Christopher Alberti, Michiel Bacchiani, Ari Bezman, Ciprian Chelba, Anastassia Drofa, Hank Liao, Pedro J. Moreno, Ted Power, Arnaud Sahuguet, Maria Shugrina, Olivier Siohan
2007ICASSPAdaptive Training with Joint Uncertainty Decoding for Robust Recognition of Noisy Data.Hank Liao, Mark J. F. Gales
2006InterspeechIssues with uncertainty decoding for noise robust speech recognition.Hank Liao, Mark J. F. Gales
2005InterspeechJoint uncertainty decoding for noise robust speech recognition.Hank Liao, Mark J. F. Gales