Skip to content

Yifan Gong

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

209

Venues

26

Active years

1987–2026

Best venue rank

A*

Where they publish

Papers

209 indexed papers, newest first.

YearVenueTitleAuthors
2026ACLInfluence-based Online Experience Selection for Effective RLHF.Yifan Gong, Jing Yao, Xiting Wang, Xunlong Wang, Xiaoyuan Yi, Xing Xie
2026ACLLLM Inductive Reasoning Through Multi-Agent Enhanced Monte Carlo Tree Search.Xiang Li, Yucheng Zhou, Xiangzhi Wei, Zesheng Shi, Haiyuan Wan, Yifan Gong, Fangming Liu, Jing Li
2025AAAILazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers.Xuan Shen, Zhao Song, Yufa Zhou, Bo Chen, Yanyu Li, Yifan Gong, Kai Zhang, Hao Tan, Jason Kuen, Henghui Ding, Zhihao Shu, Wei Niu, Pu Zhao, Yanzhi Wang, Jiuxiang Gu
2025ACLReasoning is All You Need for Video Generalization: A Counterfactual Benchmark with Sub-question Evaluation.Qiji Zhou, Yifan Gong, Guangsheng Bao, Hongjie Qiu, Jinqiang Li, Xiangrong Zhu, Huajian Zhang, Yue Zhang
2025EMNLPRECALL: REpresentation-aligned Catastrophic-forgetting ALLeviation via Hierarchical Model Merging.Bowen Wang, Haiyuan Wan, Liwen Shi, Chen Yang, Peng He, Yue Ma, Haochen Han, Wenhao Li, Tiao Tan, Yongjian Li, Fangming Liu, Yifan Gong, Sheng Zhang
2025ICCADSquat: Quant Small Language Models on the Edge.Xuan Shen, Peiyan Dong, Zhenglun Kong, Yifan Gong, Changdi Yang, Zhaoyang Han, Yanyue Xie, Lei Lu, Cheng Lyu, Chao Wu, Yanzhi Wang, Pu Zhao
2025ICCVCompressed Diffusion: Pruning with Knowledge Distillation for Efficient Text-to-Image Generation.Nasrin Kalanat, Ohiremen Dibua, Yan Kang, Yifan Gong, Xiaowei Jia
2025ICLRSparse Learning for State Space Models on Mobile.Xuan Shen, Hangyu Zheng, Yifan Gong, Zhenglun Kong, Changdi Yang, Zheng Zhan, Yushu Wu, Xue Lin, Yanzhi Wang, Pu Zhao, Wei Niu
2025IJCAIFairSMOE: Mitigating Multi-Attribute Fairness Problem with Sparse Mixture-of-Experts.Changdi Yang, Zheng Zhan, Ci Zhang, Yifan Gong, Yize Li, Zichong Meng, Jun Liu, Xuan Shen, Hao Tang, Geng Yuan, Pu Zhao, Xue Lin, Yanzhi Wang
2025InterspeechImproving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios.Aswin Shanmugam Subramanian, Amit Das, Naoyuki Kanda, Jinyu Li, Xiaofei Wang, Yifan Gong
2025WACVCan Adversarial Examples be Parsed to Reveal Victim Model Information?Yuguang Yao, Jiancheng Liu, Yifan Gong, Xiaoming Liu, Yanzhi Wang, Xue Lin, Sijia Liu
2024DACLOTUS: learning-based online thermal and latency variation management for two-stage detectors on edge devices.Yifan Gong, Yushu Wu, Zheng Zhan, Pu Zhao, Liangkai Liu, Chao Wu, Xulong Tang, Yanzhi Wang
2024ECCVEfficient Training with Denoised Neural Weights.Yifan Gong, Zheng Zhan, Yanyu Li, Yerlan Idelbayev, Andrey Zharkov, Kfir Aberman, Sergey Tulyakov, Yanzhi Wang, Jian Ren
2024EMNLPRethinking Token Reduction for State Space Models.Zheng Zhan, Yushu Wu, Zhenglun Kong, Changdi Yang, Yifan Gong, Xuan Shen, Xue Lin, Pu Zhao, Yanzhi Wang
2024ICASSPAdapting Large Language Model with Speech for Fully Formatted End-to-End Speech Recognition.Shaoshi Ling, Yuxuan Hu, Shuangbei Qian, Guoli Ye, Yao Qian, Yifan Gong, Ed Lin, Michael Zeng
2024ICCADAyE-Edge: Automated Deployment Space Search Empowering Accuracy yet Efficient Real-Time Object Detection on the Edge.Chao Wu, Yifan Gong, Liangkai Liu, Mengquan Li, Yushu Wu, Xuan Shen, Zhimin Li, Geng Yuan, Weisong Shi, Yanzhi Wang
2024ICMLE2GAN: Efficient Training of Efficient GANs for Image-to-Image Translation.Yifan Gong, Zheng Zhan, Qing Jin, Yanyu Li, Yerlan Idelbayev, Xian Liu, Andrey Zharkov, Kfir Aberman, Sergey Tulyakov, Yanzhi Wang, Jian Ren
2024InterspeechNOTSOFAR-1 Challenge: New Datasets, Baseline, and Tasks for Distant Meeting Transcription.Alon Vinnikov, Amir Ivry, Aviv Hurvitz, Igor Abramovski, Sharon Koubi, Ilya Gurvich, Shai Peer, Xiong Xiao, Benjamin Martinez Elizalde, Naoyuki Kanda, Xiaofei Wang, Shalev Shaer, Stav Yagev, Yossi Asher, Sunit Sivasankaran, Yifan Gong, Min Tang, Huaming Wang, Eyal Krupka
2024NAACLValue FULCRA: Mapping Large Language Models to the Multidimensional Spectrum of Basic Human Value.Jing Yao, Xiaoyuan Yi, Yifan Gong, Xiting Wang, Xing Xie
2023ASRUMulti Transcription-Style Speech Transcription Using Attention-Based Encoder-Decoder Model.Yan Huang, Piyush Behre, Guoli Ye, Shawn Chang, Yifan Gong
2023ASRUBuilding High-Accuracy Multilingual ASR With Gated Language Experts and Curriculum Training.Eric Sun, Jinyu Li, Yuxuan Hu, Yimeng Zhu, Long Zhou, Jian Xue, Peidong Wang, Linquan Liu, Shujie Liu, Edward Lin, Yifan Gong
2023DACCondense: A Framework for Device and Frequency Adaptive Neural Network Models on the Edge.Yifan Gong, Pu Zhao, Zheng Zhan, Yushu Wu, Chao Wu, Zhenglun Kong, Minghai Qin, Caiwen Ding, Yanzhi Wang
2023ICCADMOC: Multi-Objective Mobile CPU-GPU Co-Optimization for Power-Efficient DNN Inference.Yushu Wu, Yifan Gong, Zheng Zhan, Geng Yuan, Yanyu Li, Qi Wang, Chao Wu, Yanzhi Wang
2023ICLRSelf-Ensemble Protection: Training Checkpoints Are Good Data Protectors.Sizhe Chen, Geng Yuan, Xinwen Cheng, Yifan Gong, Minghai Qin, Yanzhi Wang, Xiaolin Huang
2023ICMLDualHSIC: HSIC-Bottleneck and Alignment for Continual Learning.Zifeng Wang, Zheng Zhan, Yifan Gong, Yucai Shao, Stratis Ioannidis, Yanzhi Wang, Jennifer G. Dy
2022ECCVCompiler-Aware Neural Architecture Search for On-Mobile Real-time Super-Resolution.Yushu Wu, Yifan Gong, Pu Zhao, Yanyu Li, Zheng Zhan, Wei Niu, Hao Tang, Minghai Qin, Bin Ren, Yanzhi Wang
2022ICASSPEndpoint Detection for Streaming End-to-End Multi-Talker ASR.Liang Lu, Jinyu Li, Yifan Gong
2022ICASSPHave Best of Both Worlds: Two-Pass Hybrid and E2E Cascading Framework for Speech Recognition.Guoli Ye, Vadim Mazalov, Jinyu Li, Yifan Gong
2022ICCADAll-in-One: A Highly Representative DNN Pruning Framework for Edge Devices with Dynamic Power Management.Yifan Gong, Zheng Zhan, Pu Zhao, Yushu Wu, Chao Wu, Caiwen Ding, Weiwen Jiang, Minghai Qin, Yanzhi Wang
2022ICLRReverse Engineering of Imperceptible Adversarial Image Perturbations.Yifan Gong, Yuguang Yao, Yize Li, Yimeng Zhang, Xiaoming Liu, Xue Lin, Sijia Liu
2022InterspeechInternal Language Model Adaptation with Text-Only Data for End-to-End Speech Recognition.Zhong Meng, Yashesh Gaur, Naoyuki Kanda, Jinyu Li, Xie Chen, Yu Wu, Yifan Gong
2022IECONA Hardware Architecture of Feature Extraction for Real-Time Visual SLAM.Jialin Li, Liangji Zhang, Xuewei Shen, Yifan Gong, Ying Lei, Chen Yang, Li Geng
2022SBAC-PADTCUDA: A QoS-based GPU Sharing Framework for Autonomous Navigation Systems.Pangbo Sun, Hao Wu, Jiangming Jin, Ziyue Jiang, Yifan Gong
2021ASRUOn Addressing Practical Challenges for RNN-Transducer.Rui Zhao, Jian Xue, Jinyu Li, Wenning Wei, Lei He, Yifan Gong
2021CLUSTERAccelerating GPU Message Communication for Autonomous Navigation Systems.Hao Wu, Jiangming Jin, Jidong Zhai, Yifan Gong, Wei Liu
2021ICASSPInternal Language Model Training for Domain-Adaptive End-To-End Speech Recognition.Zhong Meng, Naoyuki Kanda, Yashesh Gaur, Sarangarajan Parthasarathy, Eric Sun, Liang Lu, Xie Chen, Jinyu Li, Yifan Gong
2021ICASSPSequence-Level Self-Teaching Regularization.Eric Sun, Liang Lu, Zhong Meng, Yifan Gong
2021ICASSPEnsemble Combination between Different Time Segmentations.Jeremy Heng Meng Wong, Dimitrios Dimitriadis, Ken'ichi Kumatani, Yashesh Gaur, George Polovets, Partha Parthasarathy, Eric Sun, Jinyu Li, Yifan Gong
2021ICASSPHidden Markov Model Diarisation with Speaker Location Information.Jeremy Heng Meng Wong, Xiong Xiao, Yifan Gong
2021ICASSPMicrosoft Speaker Diarization System for the Voxceleb Speaker Recognition Challenge 2020.Xiong Xiao, Naoyuki Kanda, Zhuo Chen, Tianyan Zhou, Takuya Yoshioka, Sanyuan Chen, Yong Zhao, Gang Liu, Yu Wu, Jian Wu, Shujie Liu, Jinyu Li, Yifan Gong
2021ICCVAchieving on-Mobile Real-Time Super-Resolution with Neural Architecture and Pruning Search.Zheng Zhan, Yifan Gong, Pu Zhao, Geng Yuan, Wei Niu, Yushu Wu, Tianyun Zhang, Malith Jayaweera, David R. Kaeli, Bin Ren, Xue Lin, Yanzhi Wang
2021InterspeechStreaming Multi-Talker Speech Recognition with Joint Speaker Identification.Liang Lu, Naoyuki Kanda, Jinyu Li, Yifan Gong
2021InterspeechOn Minimum Word Error Rate Training of the Hybrid Autoregressive Transducer.Liang Lu, Zhong Meng, Naoyuki Kanda, Jinyu Li, Yifan Gong
2021InterspeechRapid Speaker Adaptation for Conformer Transducer: Attention and Bias Are All You Need.Yan Huang, Guoli Ye, Jinyu Li, Yifan Gong
2021InterspeechImproving RNN-T for Domain Scaling Using Semi-Supervised Training with Neural TTS.Yan Deng, Rui Zhao, Zhong Meng, Xie Chen, Bing Liu, Jinyu Li, Yifan Gong, Lei He
2021InterspeechMultiple Softmax Architecture for Streaming Multilingual End-to-End ASR Systems.Vikas Joshi, Amit Das, Eric Sun, Rupesh R. Mehta, Jinyu Li, Yifan Gong
2021InterspeechMinimum Word Error Rate Training with Language Model Fusion for End-to-End Speech Recognition.Zhong Meng, Yu Wu, Naoyuki Kanda, Liang Lu, Xie Chen, Guoli Ye, Eric Sun, Jinyu Li, Yifan Gong
2021InterspeechImproving Multilingual Transformer Transducer Models by Reducing Language Confusions.Eric Sun, Jinyu Li, Zhong Meng, Yu Wu, Jian Xue, Shujie Liu, Yifan Gong
2020DACRTMobile: Beyond Real-Time Mobile Acceleration of RNNs for Speech Recognition.Peiyan Dong, Siyue Wang, Wei Niu, Chengming Zhang, Sheng Lin, Zhengang Li, Yifan Gong, Bin Ren, Xue Lin, Dingwen Tao
2020ICASSPAcoustic Model Adaptation for Presentation Transcription and Intelligent Meeting Assistant Systems.Yan Huang, Yifan Gong
2020ICASSPUsing Personalized Speech Synthesis and Neural Language Generator for Rapid Speaker Adaptation.Yan Huang, Lei He, Wenning Wei, William Gale, Jinyu Li, Yifan Gong
2020ICASSPExploring Pre-Training with Alignments for RNN Transducer Based End-to-End Speech Recognition.Hu Hu, Rui Zhao, Jinyu Li, Liang Lu, Yifan Gong
2020ICASSPMinimum Latency Training Strategies for Streaming Sequence-to-Sequence ASR.Hirofumi Inaguma, Yashesh Gaur, Liang Lu, Jinyu Li, Yifan Gong
2020ICASSPHigh-Accuracy and Low-Latency Speech Recognition with Two-Head Contextual Layer Trajectory LSTM Model.Jinyu Li, Rui Zhao, Eric Sun, Jeremy Heng Meng Wong, Amit Das, Zhong Meng, Yifan Gong
2020ICASSPL-Vector: Neural Label Embedding for Domain Adaptation.Zhong Meng, Hu Hu, Jinyu Li, Changliang Liu, Yan Huang, Yifan Gong, Chin-Hui Lee
2020ICASSPAdaptation of RNN Transducer with Text-To-Speech Technology for Keyword Spotting.Eva Sharma, Guoli Ye, Wenning Wei, Rui Zhao, Yao Tian, Jian Wu, Lei He, Ed Lin, Yifan Gong
2020ICDCSSafe Process Quitting for GPU Multi-Process Service (MPS).Hao Wu, Wei Liu, Yifan Gong, Jiangming Jin
2020ICPPMemory-Centric Communication Mechanism for Real-time Autonomous Navigation Applications.Wei Liu, Yifan Gong, Hao Wu, Jidong Zhai, Jiangming Jin
2020InterspeechRapid RNN-T Adaptation Using Personalized Speech Synthesis and Neural Language Generator.Yan Huang, Jinyu Li, Lei He, Wenning Wei, William Gale, Yifan Gong
2020Interspeech1-D Row-Convolution LSTM: Fast Streaming ASR at Accuracy Parity with LC-BLSTM.Kshitiz Kumar, Chaojun Liu, Yifan Gong, Jian Wu
2020InterspeechBandpass Noise Generation and Augmentation for Unified ASR.Kshitiz Kumar, Bo Ren, Yifan Gong, Jian Wu
2020InterspeechDeveloping RNN-T Models Surpassing High-Performance Hybrid Models with Customization Capability.Jinyu Li, Rui Zhao, Zhong Meng, Yanqing Liu, Wenning Wei, Sarangarajan Parthasarathy, Vadim Mazalov, Zhenghao Wang, Lei He, Sheng Zhao, Yifan Gong
2020InterspeechExploring Transformers for Large-Scale Speech Recognition.Liang Lu, Changliang Liu, Jinyu Li, Yifan Gong
2020InterspeechCombination of End-to-End and Hybrid Models for Speech Recognition.Jeremy Heng Meng Wong, Yashesh Gaur, Rui Zhao, Liang Lu, Eric Sun, Jinyu Li, Yifan Gong
2020SBAC-PADA Robotic Communication Middleware Combining High Performance and High Reliability.Wei Liu, Hao Wu, Ziyue Jiang, Yifan Gong, Jiangming Jin
2019ASRUImproving RNN Transducer Modeling for End-to-End Speech Recognition.Jinyu Li, Rui Zhao, Hu Hu, Yifan Gong
2019ASRUCharacter-Aware Attention-Based End-to-End Speech Recognition.Zhong Meng, Yashesh Gaur, Jinyu Li, Yifan Gong
2019ASRUDomain Adaptation via Teacher-Student Learning for End-to-End Speech Recognition.Zhong Meng, Jinyu Li, Yashesh Gaur, Yifan Gong
2019ASRUAdvances in Online Audio-Visual Meeting Transcription.Takuya Yoshioka, Yan Huang, Aviv Hurvitz, Li Jiang, Sharon Koubi, Eyal Krupka, Ido Leichter, Changliang Liu, Partha Parthasarathy, Alon Vinnikov, Lingfeng Wu, Igor Abramovski, Xiong Xiao, Wayne Xiong, Huaming Wang, Zhenghao Wang, Jun Zhang, Yong Zhao, Tianyan Zhou, Cem Aksoylar, Zhuo Chen, Moshe David, Dimitrios Dimitriadis, Yifan Gong, Ilya Gurvich, Xuedong Huang
2019ASRUCNN with Phonetic Attention for Text-Independent Speaker Verification.Tianyan Zhou, Yong Zhao, Jinyu Li, Yifan Gong, Jian Wu
2019ICASSPUniversal Acoustic Modeling Using Neural Mixture Models.Amit Das, Jinyu Li, Changliang Liu, Yifan Gong
2019ICASSPWord Characters and Phone Pronunciation Embedding for ASR Confidence Classifier.Kshitiz Kumar, Tasos Anastasakos, Yifan Gong
2019ICASSPStatic and Dynamic State Predictions for Acoustic Model Combination.Kshitiz Kumar, Yifan Gong
2019ICASSPImproving Layer Trajectory LSTM with Future Context Frames.Jinyu Li, Liang Lu, Changliang Liu, Yifan Gong
2019ICASSPTowards Code-switching ASR for End-to-end CTC Models.Ke Li, Jinyu Li, Guoli Ye, Rui Zhao, Yifan Gong
2019ICASSPAdversarial Speaker Adaptation.Zhong Meng, Jinyu Li, Yifan Gong
2019ICASSPAttentive Adversarial Learning for Domain-invariant Training.Zhong Meng, Jinyu Li, Yifan Gong
2019ICASSPConditional Teacher-student Learning.Zhong Meng, Jinyu Li, Yong Zhao, Yifan Gong
2019ICASSPAdversarial Speaker Verification.Zhong Meng, Yong Zhao, Jinyu Li, Yifan Gong
2019ICASSPSingle-channel Speech Extraction Using Speaker Inventory and Attention Network.Xiong Xiao, Zhuo Chen, Takuya Yoshioka, Hakan Erdogan, Changliang Liu, Dimitrios Dimitriadis, Jasha Droppo, Yifan Gong
2019ICASSPEncrypted Speech Recognition Using Deep Polynomial Networks.Shi-Xiong Zhang, Yifan Gong, Dong Yu
2019InterspeechAcoustic-to-Phrase Models for Speech Recognition.Yashesh Gaur, Jinyu Li, Zhong Meng, Yifan Gong
2019InterspeechSelf-Teaching Networks.Liang Lu, Eric Sun, Yifan Gong
2019InterspeechSpeaker Adaptation for Attention-Based End-to-End Speech Recognition.Zhong Meng, Yashesh Gaur, Jinyu Li, Yifan Gong
2019InterspeechLayer Trajectory BLSTM.Eric Sun, Jinyu Li, Yifan Gong
2019IWQoSUntitled recordYifan Gong, Baochun Li, Ben Liang, Zheng Zhan
2018ICASSPEfficient Integration of Fixed Beamformers and Speech Separation Networks for Multi-Channel Far-Field Speech Separation.Zhuo Chen, Takuya Yoshioka, Xiong Xiao, Linyu Li, Michael L. Seltzer, Yifan Gong
2018ICASSPAdvancing Connectionist Temporal Classification with Attention Modeling.Amit Das, Jinyu Li, Rui Zhao, Yifan Gong
2018ICASSPAdvancing Acoustic-to-Word CTC Model.Jinyu Li, Guoli Ye, Amit Das, Rui Zhao, Yifan Gong
2018ICASSPDeveloping Far-Field Speaker System Via Teacher-Student Learning.Jinyu Li, Rui Zhao, Zhuo Chen, Changliang Liu, Xiong Xiao, Guoli Ye, Yifan Gong
2018ICASSPSpeaker-Invariant Training Via Adversarial Learning.Zhong Meng, Jinyu Li, Zhuo Chen, Yang Zhao, Vadim Mazalov, Yifan Gong, Biing-Hwang Juang
2018ICASSPAdversarial Teacher-Student Learning for Unsupervised Domain Adaptation.Zhong Meng, Jinyu Li, Yifan Gong, Biing-Hwang Juang
2018ICASSPDomain and Speaker Adaptation for Cortana Speech Recognition.Yong Zhao, Jinyu Li, Shi-Xiong Zhang, Liping Chen, Yifan Gong
2018InterspeechLayer Trajectory LSTM.Jinyu Li, Changliang Liu, Yifan Gong
2018InterspeechCycle-Consistent Speech Enhancement.Zhong Meng, Jinyu Li, Yifan Gong, Biing-Hwang Fred Juang
2018InterspeechAdversarial Feature-Mapping for Speech Enhancement.Zhong Meng, Jinyu Li, Yifan Gong, Biing-Hwang Fred Juang
2017ASRUCracking the cocktail party problem by multi-beam deep attractor network.Zhuo Chen, Jinyu Li, Xiong Xiao, Takuya Yoshioka, Huaming Wang, Zhenghao Wang, Yifan Gong
2017ASRUAcoustic-to-word model without OOV.Jinyu Li, Guoli Ye, Rui Zhao, Jasha Droppo, Yifan Gong
2017ASRUUnsupervised adaptation with domain separation networks for robust speech recognition.Zhong Meng, Zhuo Chen, Vadim Mazalov, Jinyu Li, Yifan Gong
2017ICASSPImproved cepstra minimum-mean-square-error noise reduction algorithm for robust speech recognition.Jinyu Li, Yan Huang, Yifan Gong
2017ICASSPExtended low-rank plus diagonal adaptation for deep and recurrent neural networks.Yong Zhao, Jinyu Li, Kshitiz Kumar, Yifan Gong
2017InterspeechImproving Mask Learning Based Speech Enhancement System with Restoration Layers and Residual Connection.Zhuo Chen, Yan Huang, Jinyu Li, Yifan Gong
2017InterspeechDon't Count on ASR to Transcribe for You: Breaking Bias with Two Crowds.Michael Levit, Yan Huang, Shuangyu Chang, Yifan Gong
2017InterspeechLarge-Scale Domain Adaptation via Teacher-Student Learning.Jinyu Li, Michael L. Seltzer, Xi Wang, Rui Zhao, Yifan Gong
2017SCEfficient process mapping in geo-distributed cloud data centers.Amelie Chi Zhou, Yifan Gong, Bingsheng He, Jidong Zhai
2016ICASSPNon-negative intermediate-layer DNN adaptation for a 10-KB speaker adaptation profile.Kshitiz Kumar, Chaojun Liu, Yifan Gong
2016ICASSPExploring multidimensional lstms for large vocabulary ASR.Jinyu Li, Abdelrahman Mohamed, Geoffrey Zweig, Yifan Gong
2016ICASSPInvestigations on speaker adaptation of LSTM RNN models for speech recognition.Chaojun Liu, Yongqiang Wang, Kshitiz Kumar, Yifan Gong
2016ICASSPSimplifying long short-term memory acoustic models for fast training and decoding.Yajie Miao, Jinyu Li, Yongqiang Wang, Shi-Xiong Zhang, Yifan Gong
2016ICASSPGeo-location dependent deep neural network acoustic model for speech recognition.Guoli Ye, Chaojun Liu, Yifan Gong
2016ICASSPRecurrent support vector machines for speech recognition.Shi-Xiong Zhang, Rui Zhao, Chaojun Liu, Jinyu Li, Yifan Gong
2016ICASSPLow-rank plus diagonal adaptation for deep neural networks.Yong Zhao, Jinyu Li, Yifan Gong
2016InterspeechSemi-Supervised Training in Deep Learning Acoustic Model.Yan Huang, Yongqiang Wang, Yifan Gong
2015ASRULSTM time and frequency recurrence for automatic speech recognition.Jinyu Li, Abdelrahman Mohamed, Geoffrey Zweig, Yifan Gong
2015ICASSPAn analysis of convolutional neural networks for speech recognition.Jui-Ting Huang, Jinyu Li, Yifan Gong
2015ICASSPEstimating confidence scores on ASR results using recurrent neural networks.Kaustubh Kalgaonkar, Chaojun Liu, Yifan Gong, Kaisheng Yao
2015ICASSPSmall-footprint high-performance deep neural network-based speech recognition using split-VQ.Yongqiang Wang, Jinyu Li, Yifan Gong
2015ICASSPDeep neural support vector machines for speech recognition.Shi-Xiong Zhang, Chaojun Liu, Kaisheng Yao, Yifan Gong
2015ICASSPInvestigating online low-footprint speaker adaptation using generalized linear regression and click-through data.Yong Zhao, Jinyu Li, Jian Xue, Yifan Gong
2015InterspeechRegularized sequence-level deep neural network model adaptation.Yan Huang, Yifan Gong
2015InterspeechConfidence-features and confidence-scores for ASR applications in arbitration and DNN speaker adaptation.Kshitiz Kumar, Ziad Al Bawab, Yong Zhao, Chaojun Liu, Benot Dumoulin, Yifan Gong
2015InterspeechDelta-melspectra features for noise robustness to DNN-based ASR systems.Kshitiz Kumar, Chaojun Liu, Yifan Gong
2015InterspeechIntermediate-layer DNN adaptation for offline and session-based iterative speaker adaptation.Kshitiz Kumar, Chaojun Liu, Kaisheng Yao, Yifan Gong
2015InterspeechSVD-based universal DNN modeling for multiple scenarios.Changliang Liu, Jinyu Li, Yifan Gong
2015SCMonetary cost optimizations for MPI-based HPC applications on Amazon clouds: checkpoints and replicated execution.Yifan Gong, Bingsheng He, Amelie Chi Zhou
2014ICASSPFactorized adaptation for deep neural network.Jinyu Li, Jui-Ting Huang, Yifan Gong
2014ICASSPSingular value decomposition based low-footprint speaker adaptation and personalization for deep neural network.Jian Xue, Jinyu Li, Dong Yu, Mike Seltzer, Yifan Gong
2014InterspeechTowards better performance with heterogeneous training data in acoustic modeling using deep neural networks.Yan Huang, Malcolm Slaney, Michael L. Seltzer, Yifan Gong
2014InterspeechA comparative analytic study on the Gaussian mixture and context dependent deep neural network hidden Markov models.Yan Huang, Dong Yu, Chaojun Liu, Yifan Gong
2014InterspeechMulti-accent deep neural network acoustic model with accent-specific top layer using the KLD-regularized model adaptation.Yan Huang, Dong Yu, Chaojun Liu, Yifan Gong
2014InterspeechNormalization of ASR confidence classifier scores via confidence mapping.Kshitiz Kumar, Chaojun Liu, Yifan Gong
2014InterspeechLearning small-size DNN with output-distribution-based criteria.Jinyu Li, Rui Zhao, Jui-Ting Huang, Yifan Gong
2014InterspeechVariable-component deep neural network for robust speech recognition.Rui Zhao, Jinyu Li, Yifan Gong
2014SCFinding Constant from Change: Revisiting Network Performance Aware Optimizations on IaaS Clouds.Yifan Gong, Bingsheng He, Dan Li
2013ICASSPRecent advances in deep learning for speech research at Microsoft.Li Deng, Jinyu Li, Jui-Ting Huang, Kaisheng Yao, Dong Yu, Frank Seide, Michael L. Seltzer, Geoffrey Zweig, Xiaodong He, Jason D. Williams, Yifan Gong, Alex Acero
2013ICASSPPredicting speech recognition confidence using deep learning with word identity and score features.Po-Sen Huang, Kshitiz Kumar, Chaojun Liu, Yifan Gong, Li Deng
2013ICASSPCross-language knowledge transfer using multilingual deep neural network with shared hidden layers.Jui-Ting Huang, Jinyu Li, Dong Yu, Li Deng, Yifan Gong
2013InterspeechSemi-supervised GMM and DNN acoustic model training with multi-system combination and confidence re-calibration.Yan Huang, Dong Yu, Yifan Gong, Chaojun Liu
2013InterspeechRestructuring of deep neural network acoustic models with singular value decomposition.Jian Xue, Jinyu Li, Yifan Gong
2012ICASSPImprovements to VTS feature enhancement.Jinyu Li, Michael L. Seltzer, Yifan Gong
2012InterspeechEfficient VTS Adaptation Using Jacobian Approximation.Jinyu Li, Michael L. Seltzer, Yifan Gong
2012InterspeechA Feature Space Transformation Method for Personalization using Generalized I-Vector Clustering.Kaisheng Yao, Yifan Gong, Chaojun Liu
2012PPoPPAn overview of CMPI: network performance aware MPI in the cloud.Yifan Gong, Bingsheng He, Jianlong Zhong
2010InterspeechUnscented transform with online distortion estimation for HMM adaptation.Jinyu Li, Dong Yu, Yifan Gong, Li Deng
2009ICASSPA study on multilingual acoustic modeling for large vocabulary ASR.Hui Lin, Li Deng, Dong Yu, Yifan Gong, Alex Acero, Chin-Hui Lee
2009ICASSPCross-lingual speech recognition under runtime resource constraints.Dong Yu, Li Deng, Peng Liu, Jian Wu, Yifan Gong, Alex Acero
2008ICASSPHMM adaptation using a phase-sensitive acoustic distortion model for environment-robust speech recognition.Jinyu Li, Li Deng, Dong Yu, Yifan Gong, Alex Acero
2008ICASSPAdaptation of compressed HMM parameters for resource-constrained speech recognition.Jinyu Li, Li Deng, Dong Yu, Jian Wu, Yifan Gong, Alex Acero
2008ICASSPA minimum-mean-square-error noise reduction algorithm on Mel-frequency cepstra for robust speech recognition.Dong Yu, Li Deng, Jasha Droppo, Jian Wu, Yifan Gong, Alex Acero
2008InterspeechDiscriminative training of variable-parameter HMMs for noise robust speech recognition.Dong Yu, Li Deng, Yifan Gong, Alex Acero
2008InterspeechParameter clustering and sharing in variable-parameter HMMs for noise robust speech recognition.Dong Yu, Li Deng, Yifan Gong, Alex Acero
2007ASRUHigh-performance hmm adaptation with joint compensation of additive and convolutive distortions via Vector Taylor Series.Jinyu Li, Li Deng, Dong Yu, Yifan Gong, Alex Acero
2006ICASSPModeling Variance Variation in a Variable Parameter HMM Framework for Noise Robust Speech Recognition.Xiaodong Cui, Yifan Gong
2004ICASSPCan back-ends be more robust than front-ends? Investigation over the Aurora-2 database.Alexis Bernard, Yifan Gong, Xiaodong Cui
2003ICASSPVariable parameter Gaussian mixture hidden Markov modeling for speech recognition.Xiaodong Cui, Yifan Gong
2003ICASSPModel-space compensation of microphone and noise for speaker-independent speech recognition.Yifan Gong
2002ICASSPNoise-robust open-set speaker recognition using noise-dependent Gaussian mixture classifier.Yifan Gong
2002InterspeechA comparative study of approximations for parallel model combination of static and dynamic parameters.Yifan Gong
2002InterspeechExperiments on speaker-independent voice command recognition using in-vehicle hands free speech.Yifan Gong, Lorin Netsch
2002InterspeechThe effects of speech compression on speech recognition and text-to-speech synthesis.Yeshwant K. Muthusamy, Yifan Gong, Roshan Gupta
2000ICASSPImplementing a high accuracy speaker-independent continuous speech recognizer on a fixed-point DSP.Yifan Gong, Yu-Hung Kao
2000ICASSPHMM adaptation and microphone array processing for distant speech recognition.Jim Kleban, Yifan Gong
1999ICASSPTransforming HMMs for speaker-independent hands-free speech recognition in the car.Yifan Gong, John J. Godfrey
1999ICASSPSpeech-enabled information retrieval in the automobile environment.Yeshwant K. Muthusamy, Rajeev Agarwal, Yifan Gong, Vishu Viswanathan
1999ICASSPSpeaker-dependent name dialing in a car environment with out-of-vocabulary rejection.Coimbatore S. Ramalingam, Yifan Gong, Lorin Netsch, Wallace W. Anderson, John J. Godfrey, Yu-Hung Kao
1997ICASSPA unified maximum likelihood approach to acoustic mismatch compensation: application to noisy Lombard speech recognition.Mohamed Afify, Yifan Gong, Jean-Paul Haton
1997ICASSPElimination of trajectory folding phenomenon: HMM, trajectory mixture HMM and mixture stochastic trajectory model.Irina Illina, Yifan Gong
1997ICASSPThe importance of segmentation probability in segment based speech recognizers.Jan P. Verhasselt, Irina Illina, Jean-Pierre Martens, Yifan Gong, Jean-Paul Haton
1997InterspeechCorrelation based predictive adaptation of hidden Markov models.Mohamed Afify, Yifan Gong, Jean Paul Haton
1997InterspeechAn acoustic subword unit approach to non-linguistic speech feature identification.Mohamed Afify, Yifan Gong, Jean Paul Haton
1997InterspeechSource normalization training for HMM applied to noisy telephone speech recognition.Yifan Gong
1997InterspeechSpeaker normalization training for mixture stochastic trajectory model.Irina Illina, Yifan Gong
1996ICASSPProbabilistic mapping networks for speaker recognition.Haizhou Li, Yifan Gong, Jean-Paul Haton
1996ICASSPA semi-continuous stochastic trajectory model for phoneme-based continuous speech recognition.Olivier Siohan, Yifan Gong
1996InterspeechModelling long term variability information in mixture stochastic trajectory framework.Yifan Gong, Irina Illina, Jean Paul Haton
1996InterspeechStochastic trajectory model with state-mixture for continuous speech recognition.Irina Illina, Yifan Gong
1996InterspeechImprovement in n-best search for continuous speech recognition.Irina Illina, Yifan Gong
1996InterspeechA study on continuous Chinese speech recognition based on stochastic trajectory models.Xiaohui Ma, Yifan Gong, Yuqing Fu, Jiren Lu, Jean Paul Haton
1995ICDARStochastic trajectory modeling for recognition of unconstrained handwritten words.George Saon, Abdel Belad, Yifan Gong
1995InterspeechStochastic trajectory models for speech recognition: an extension to modelling time correlation.Mohamed Afify, Yifan Gong, Jean Paul Haton
1995InterspeechEvaluation of Bayes decision approach to automatic determination of thresholds for speaker verification.Yifan Gong
1995InterspeechOn MMI learning of Gaussian mixture for speaker models.Haizhou Li, Jean Paul Haton, Yifan Gong
1995InterspeechSpeaker recognition with temporal transition models.Haizhou Li, Jean Paul Haton, Jian Su, Yifan Gong
1995InterspeechNoise adaptation using linear regression for continuous noisy speech recognition.Olivier Siohan, Yifan Gong, Jean Paul Haton
1994ICASSPStochastic trajectory modeling for speech recognition.Yifan Gong, Jean Paul Haton
1994ICASSPNoise independent speech recognition for a variety of noise types.William C. Treurniet, Yifan Gong
1994InterspeechNonlinear time alignment in stochastic trajectory models for speech recognition.Mohamed Afify, Yifan Gong, Jean Paul Haton
1994InterspeechA comparison of three noisy speech recognition approaches.Olivier Siohan, Yifan Gong, Jean Paul Haton
1994MVAOff-line Handwriting Recognition by Statistical Correlation.George Saon, Abdel Belad, Yifan Gong
1993InterspeechBase transformation for environment adaptation in continuous speech recognition.Yifan Gong
1993InterspeechIterative transformation and alignment for speech labeling.Yifan Gong, Jean Paul Haton
1993InterspeechDuration of phones as function of utterance length and its use in automatic speech recognition.Yifan Gong, William C. Treurniet
1993InterspeechUse of explicit context-dependent phonemic model in continuous speech recognition.Feriel Mouria, Yifan Gong, Jean Paul Haton
1993InterspeechA Bayesian approach to phone duration adaptation for lombard speech recognition.Olivier Siohan, Yifan Gong, Jean Paul Haton
1992ICASSPNonlinear vectorial interpolation for speaker recognition.Yifan Gong, Jean-Paul Haton
1992ICPRHand-written text recognition based on a new formulation.Yifan Gong, Anne Boyer
1992InterspeechDTW-based phonetic labeling using explicit phoneme duration constraints.Yifan Gong, Jean Paul Haton
1992InterspeechMinimization of speech alignment error by iterative transformation for speaker adaptation.Yifan Gong, Olivier Siohan, Jean Paul Haton
1991ICASSPNeural network coupled with IIR sequential adapter for phoneme recognition in continuous speech.Yifan Gong, Ying Cheng, Jean-Paul Haton
1991ICASSPNon-linear vector interpolation by neural network for phoneme identification in continuous speech.Yifan Gong, Jean-Paul Haton
1991ICASSPContinuous speech recognition based on high plausibility regions.Yifan Gong, Jean-Paul Haton, Feriel Mouria
1991InterspeechComparing two phoneme identification methods using a continuous speech recognizer.Yifan Gong, Jean Paul Haton
1991InterspeechVINICS: a continuous speech recognizer based on a new robust formulation.Yifan Gong, Jean Paul Haton
1990ICASSPText-independent speaker recognition by trajectory space comparison.Yifan Gong, Jean-Paul Haton
1990ICPRTowards a general signal interpretation system-signal-to-symbol conversion level.Yifan Gong, Jean-Paul Haton
1990ICPRA multiknowledge base system for continuous speech understanding.Yifan Gong, Jean-Paul Haton
1989InterspeechParallel construction of syntactic structure for continuous speech recognition.Yifan Gong, Anne Boyer, Jean Paul Haton
1988ICASSPA specialist society for continuous speech understanding.Yifan Gong, Jean-Paul Haton
1987InterspeechPhoneme-based continuous speech recognition without pre-segmentation.Yifan Gong, Jean Paul Haton