Skip to content

Nanxin Chen

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

30

Venues

5

Active years

2015–2024

Best venue rank

A*

Where they publish

Papers

30 indexed papers, newest first.

YearVenueTitleAuthors
2024InterspeechText Injection for Neural Contextual Biasing.Zhong Meng, Zelin Wu, Rohit Prabhavalkar, Cal Peyser, Weiran Wang, Nanxin Chen, Tara N. Sainath, Bhuvana Ramabhadran
2023ASRUE3 TTS: Easy End-to-End Diffusion-Based Text To Speech.Yuan Gao, Nobuyuki Morioka, Yu Zhang, Nanxin Chen
2023ASRUSLM: Bridge the Thin Gap Between Speech and Text Foundation Models.Mingqiu Wang, Wei Han, Izhak Shafran, Zelin Wu, Chung-Cheng Chiu, Yuan Cao, Nanxin Chen, Yu Zhang, Hagen Soltau, Paul K. Rubenstein, Lukas Zilka, Dian Yu, Golan Pundak, Nikhil Siddhartha, Johan Schalkwyk, Yonghui Wu
2023ICASSPFrom English to More Languages: Parameter-Efficient Model Reprogramming for Cross-Lingual Speech Recognition.Chao-Han Huck Yang, Bo Li, Yu Zhang, Nanxin Chen, Rohit Prabhavalkar, Tara N. Sainath, Trevor Strohman
2023ICASSPA Quantum Kernel Learning Approach to Acoustic Modeling for Spoken Command Recognition.Chao-Han Huck Yang, Bo Li, Yu Zhang, Nanxin Chen, Tara N. Sainath, Sabato Marco Siniscalchi, Chin-Hui Lee
2023InterspeechHow to Estimate Model Transferability of Pre-Trained Speech Models?Zih-Ching Chen, Chao-Han Huck Yang, Bo Li, Yu Zhang, Nanxin Chen, Shuo-Yiin Chang, Rohit Prabhavalkar, Hung-yi Lee, Tara N. Sainath
2022InterspeechSpecGrad: Diffusion Probabilistic Model based Neural Vocoder with Adaptive Noise Spectral Shaping.Yuma Koizumi, Heiga Zen, Kohei Yatabe, Nanxin Chen, Michiel Bacchiani
2021ASRUA Comparative Study on Non-Autoregressive Modelings for Speech-to-Text Generation.Yosuke Higuchi, Nanxin Chen, Yuya Fujita, Hirofumi Inaguma, Tatsuya Komatsu, Jaesong Lee, Jumon Nozaki, Tianzi Wang, Shinji Watanabe
2021ICASSPFocus on the Present: A Regularization Method for the ASR Source-Target Attention Layer.Nanxin Chen, Piotr Zelasko, Jess Villalba, Najim Dehak
2021ICLRWaveGrad: Estimating Gradients for Waveform Generation.Nanxin Chen, Yu Zhang, Heiga Zen, Ron J. Weiss, Mohammad Norouzi, William Chan
2021InterspeechAlign-Denoise: Single-Pass Non-Autoregressive Speech Recognition.Nanxin Chen, Piotr Zelasko, Laureano Moro-Velzquez, Jess Villalba, Najim Dehak
2021InterspeechWaveGrad 2: Iterative Refinement for Text-to-Speech Synthesis.Nanxin Chen, Yu Zhang, Heiga Zen, Ron J. Weiss, Mohammad Norouzi, Najim Dehak, William Chan
2020ICASSPZero-Shot Multi-Speaker Text-To-Speech with State-Of-The-Art Neural Speaker Embeddings.Erica Cooper, Cheng-I Lai, Yusuke Yasuda, Fuming Fang, Xin Wang, Nanxin Chen, Junichi Yamagishi
2020ICASSPFeature Enhancement with Deep Feature Losses for Speaker Verification.Saurabh Kataria, Phani Sankar Nidadavolu, Jess Villalba, Nanxin Chen, L. Paola Garca-Perera, Najim Dehak
2020ICASSPX-Vectors Meet Emotions: A Study On Dependencies Between Emotion and Speaker Recognition.Raghavendra Pappagari, Tianzi Wang, Jess Villalba, Nanxin Chen, Najim Dehak
2020ICASSPImproving Language Identification for Multilingual Speakers.Andrew Titus, Jan Silovsk, Nanxin Chen, Roger Hsiao, Mary Young, Arnab Ghoshal
2020IJCNNRobust Training of Vector Quantized Bottleneck Models.Adrian Lancucki, Jan Chorowski, Guillaume Sanchez, Ricard Marxer, Nanxin Chen, Hans J. G. A. Dolfing, Sameer Khurana, Tanel Alume, Antoine Laurent
2020InterspeechMask CTC: Non-Autoregressive End-to-End ASR with CTC and Mask Predict.Yosuke Higuchi, Shinji Watanabe, Nanxin Chen, Tetsuji Ogawa, Tetsunori Kobayashi
2019ASRUA Comparative Study on Transformer vs RNN in Speech Applications.Shigeki Karita, Xiaofei Wang, Shinji Watanabe, Takenori Yoshimura, Wangyou Zhang, Nanxin Chen, Tomoki Hayashi, Takaaki Hori, Hirofumi Inaguma, Ziyan Jiang, Masao Someki, Nelson Enrique Yalta Soplin, Ryuichi Yamamoto
2019InterspeechTied Mixture of Factor Analyzers Layer to Combine Frame Level Representations in Neural Speaker Embeddings.Nanxin Chen, Jess Villalba, Najim Dehak
2019InterspeechASSERT: Anti-Spoofing with Squeeze-Excitation and Residual Networks.Cheng-I Lai, Nanxin Chen, Jess Villalba, Najim Dehak
2019InterspeechThe JHU Speaker Recognition System for the VOiCES 2019 Challenge.David Snyder, Jess Villalba, Nanxin Chen, Daniel Povey, Gregory Sell, Najim Dehak, Sanjeev Khudanpur
2019InterspeechState-of-the-Art Speaker Recognition for Telephone and Video Speech: The JHU-MIT Submission for NIST SRE18.Jess Villalba, Nanxin Chen, David Snyder, Daniel Garcia-Romero, Alan McCree, Gregory Sell, Jonas Borgstrom, Fred Richardson, Suwon Shon, Franois Grondin, Rda Dehak, Leibny Paola Garca-Perera, Daniel Povey, Pedro A. Torres-Carrasquillo, Sanjeev Khudanpur, Najim Dehak
2018ICASSPMeasuring Uncertainty in Deep Regression Models: The Case of Age Estimation from Speech.Nanxin Chen, Jess Villalba, Yishay Carmiel, Najim Dehak
2018InterspeechAn Investigation of Non-linear i-vectors for Speaker Verification.Nanxin Chen, Jess Villalba, Najim Dehak
2018InterspeechEnd-to-end Deep Neural Network Age Estimation.Pegah Ghahremani, Phani Sankar Nidadavolu, Nanxin Chen, Jess Villalba, Daniel Povey, Sanjeev Khudanpur, Najim Dehak
2018InterspeechESPnet: End-to-End Speech Processing Toolkit.Shinji Watanabe, Takaaki Hori, Shigeki Karita, Tomoki Hayashi, Jiro Nishitoba, Yuya Unno, Nelson Enrique Yalta Soplin, Jahn Heymann, Matthew Wiesner, Nanxin Chen, Adithya Renduchintala, Tsubasa Ochiai
2017ICASSPEnd-to-end spoofing detection with raw waveform CLDNNS.Heinrich Dinkel, Nanxin Chen, Yanmin Qian, Kai Yu
2015InterspeechRobust deep feature for spoofing detection - the SJTU system for ASVspoof 2015 challenge.Nanxin Chen, Yanmin Qian, Heinrich Dinkel, Bo Chen, Kai Yu
2015InterspeechMulti-task learning for text-dependent speaker verification.Nanxin Chen, Yanmin Qian, Kai Yu