Nanxin Chen
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
30
Venues
5
Active years
2015–2024
Best venue rank
A*
Where they publish
Papers
30 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2024 | Interspeech | Text Injection for Neural Contextual Biasing. | Zhong Meng, Zelin Wu, Rohit Prabhavalkar, Cal Peyser, Weiran Wang, Nanxin Chen, Tara N. Sainath, Bhuvana Ramabhadran |
| 2023 | ASRU | E3 TTS: Easy End-to-End Diffusion-Based Text To Speech. | Yuan Gao, Nobuyuki Morioka, Yu Zhang, Nanxin Chen |
| 2023 | ASRU | SLM: Bridge the Thin Gap Between Speech and Text Foundation Models. | Mingqiu Wang, Wei Han, Izhak Shafran, Zelin Wu, Chung-Cheng Chiu, Yuan Cao, Nanxin Chen, Yu Zhang, Hagen Soltau, Paul K. Rubenstein, Lukas Zilka, Dian Yu, Golan Pundak, Nikhil Siddhartha, Johan Schalkwyk, Yonghui Wu |
| 2023 | ICASSP | From English to More Languages: Parameter-Efficient Model Reprogramming for Cross-Lingual Speech Recognition. | Chao-Han Huck Yang, Bo Li, Yu Zhang, Nanxin Chen, Rohit Prabhavalkar, Tara N. Sainath, Trevor Strohman |
| 2023 | ICASSP | A Quantum Kernel Learning Approach to Acoustic Modeling for Spoken Command Recognition. | Chao-Han Huck Yang, Bo Li, Yu Zhang, Nanxin Chen, Tara N. Sainath, Sabato Marco Siniscalchi, Chin-Hui Lee |
| 2023 | Interspeech | How to Estimate Model Transferability of Pre-Trained Speech Models? | Zih-Ching Chen, Chao-Han Huck Yang, Bo Li, Yu Zhang, Nanxin Chen, Shuo-Yiin Chang, Rohit Prabhavalkar, Hung-yi Lee, Tara N. Sainath |
| 2022 | Interspeech | SpecGrad: Diffusion Probabilistic Model based Neural Vocoder with Adaptive Noise Spectral Shaping. | Yuma Koizumi, Heiga Zen, Kohei Yatabe, Nanxin Chen, Michiel Bacchiani |
| 2021 | ASRU | A Comparative Study on Non-Autoregressive Modelings for Speech-to-Text Generation. | Yosuke Higuchi, Nanxin Chen, Yuya Fujita, Hirofumi Inaguma, Tatsuya Komatsu, Jaesong Lee, Jumon Nozaki, Tianzi Wang, Shinji Watanabe |
| 2021 | ICASSP | Focus on the Present: A Regularization Method for the ASR Source-Target Attention Layer. | Nanxin Chen, Piotr Zelasko, Jess Villalba, Najim Dehak |
| 2021 | ICLR | WaveGrad: Estimating Gradients for Waveform Generation. | Nanxin Chen, Yu Zhang, Heiga Zen, Ron J. Weiss, Mohammad Norouzi, William Chan |
| 2021 | Interspeech | Align-Denoise: Single-Pass Non-Autoregressive Speech Recognition. | Nanxin Chen, Piotr Zelasko, Laureano Moro-Velzquez, Jess Villalba, Najim Dehak |
| 2021 | Interspeech | WaveGrad 2: Iterative Refinement for Text-to-Speech Synthesis. | Nanxin Chen, Yu Zhang, Heiga Zen, Ron J. Weiss, Mohammad Norouzi, Najim Dehak, William Chan |
| 2020 | ICASSP | Zero-Shot Multi-Speaker Text-To-Speech with State-Of-The-Art Neural Speaker Embeddings. | Erica Cooper, Cheng-I Lai, Yusuke Yasuda, Fuming Fang, Xin Wang, Nanxin Chen, Junichi Yamagishi |
| 2020 | ICASSP | Feature Enhancement with Deep Feature Losses for Speaker Verification. | Saurabh Kataria, Phani Sankar Nidadavolu, Jess Villalba, Nanxin Chen, L. Paola Garca-Perera, Najim Dehak |
| 2020 | ICASSP | X-Vectors Meet Emotions: A Study On Dependencies Between Emotion and Speaker Recognition. | Raghavendra Pappagari, Tianzi Wang, Jess Villalba, Nanxin Chen, Najim Dehak |
| 2020 | ICASSP | Improving Language Identification for Multilingual Speakers. | Andrew Titus, Jan Silovsk, Nanxin Chen, Roger Hsiao, Mary Young, Arnab Ghoshal |
| 2020 | IJCNN | Robust Training of Vector Quantized Bottleneck Models. | Adrian Lancucki, Jan Chorowski, Guillaume Sanchez, Ricard Marxer, Nanxin Chen, Hans J. G. A. Dolfing, Sameer Khurana, Tanel Alume, Antoine Laurent |
| 2020 | Interspeech | Mask CTC: Non-Autoregressive End-to-End ASR with CTC and Mask Predict. | Yosuke Higuchi, Shinji Watanabe, Nanxin Chen, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2019 | ASRU | A Comparative Study on Transformer vs RNN in Speech Applications. | Shigeki Karita, Xiaofei Wang, Shinji Watanabe, Takenori Yoshimura, Wangyou Zhang, Nanxin Chen, Tomoki Hayashi, Takaaki Hori, Hirofumi Inaguma, Ziyan Jiang, Masao Someki, Nelson Enrique Yalta Soplin, Ryuichi Yamamoto |
| 2019 | Interspeech | Tied Mixture of Factor Analyzers Layer to Combine Frame Level Representations in Neural Speaker Embeddings. | Nanxin Chen, Jess Villalba, Najim Dehak |
| 2019 | Interspeech | ASSERT: Anti-Spoofing with Squeeze-Excitation and Residual Networks. | Cheng-I Lai, Nanxin Chen, Jess Villalba, Najim Dehak |
| 2019 | Interspeech | The JHU Speaker Recognition System for the VOiCES 2019 Challenge. | David Snyder, Jess Villalba, Nanxin Chen, Daniel Povey, Gregory Sell, Najim Dehak, Sanjeev Khudanpur |
| 2019 | Interspeech | State-of-the-Art Speaker Recognition for Telephone and Video Speech: The JHU-MIT Submission for NIST SRE18. | Jess Villalba, Nanxin Chen, David Snyder, Daniel Garcia-Romero, Alan McCree, Gregory Sell, Jonas Borgstrom, Fred Richardson, Suwon Shon, Franois Grondin, Rda Dehak, Leibny Paola Garca-Perera, Daniel Povey, Pedro A. Torres-Carrasquillo, Sanjeev Khudanpur, Najim Dehak |
| 2018 | ICASSP | Measuring Uncertainty in Deep Regression Models: The Case of Age Estimation from Speech. | Nanxin Chen, Jess Villalba, Yishay Carmiel, Najim Dehak |
| 2018 | Interspeech | An Investigation of Non-linear i-vectors for Speaker Verification. | Nanxin Chen, Jess Villalba, Najim Dehak |
| 2018 | Interspeech | End-to-end Deep Neural Network Age Estimation. | Pegah Ghahremani, Phani Sankar Nidadavolu, Nanxin Chen, Jess Villalba, Daniel Povey, Sanjeev Khudanpur, Najim Dehak |
| 2018 | Interspeech | ESPnet: End-to-End Speech Processing Toolkit. | Shinji Watanabe, Takaaki Hori, Shigeki Karita, Tomoki Hayashi, Jiro Nishitoba, Yuya Unno, Nelson Enrique Yalta Soplin, Jahn Heymann, Matthew Wiesner, Nanxin Chen, Adithya Renduchintala, Tsubasa Ochiai |
| 2017 | ICASSP | End-to-end spoofing detection with raw waveform CLDNNS. | Heinrich Dinkel, Nanxin Chen, Yanmin Qian, Kai Yu |
| 2015 | Interspeech | Robust deep feature for spoofing detection - the SJTU system for ASVspoof 2015 challenge. | Nanxin Chen, Yanmin Qian, Heinrich Dinkel, Bo Chen, Kai Yu |
| 2015 | Interspeech | Multi-task learning for text-dependent speaker verification. | Nanxin Chen, Yanmin Qian, Kai Yu |