Yossi Adi
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
64
Venues
13
Active years
2016–2026
Best venue rank
A*
Where they publish
Papers
64 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | ACL | LaMI: Augmenting Large Language Models via Late Multi-Image Fusion. | Guy Yariv, Idan Schwartz, Yossi Adi, Sagie Benaim |
| 2026 | ACL | StressTest: Can YOUR Speech LM Handle the Stress? | Iddo Yosha, Gallil Maimon, Yossi Adi |
| 2025 | ACL | Slamming: Training a Speech Language Model on One GPU in a Day. | Gallil Maimon, Avishai Elmakies, Yossi Adi |
| 2025 | ASRU | Speech Synthesis From Continuous Features Using Per-Token Latent Diffusion. | Arnon Turetzky, Avihu Dekel, Nimrod Shabtay, Slava Shechtman, David Haws, Hagai Aronowitz, Ron Hoory, Yossi Adi |
| 2025 | CVPR | Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation. | Guy Yariv, Yuval Kirstain, Amit Zohar, Shelly Sheynin, Yaniv Taigman, Yossi Adi, Sagie Benaim, Adam Polyak |
| 2025 | EMNLP | GmSLM : Generative Marmoset Spoken Language Modeling. | Talia Sternberg, Michael London, David Omer, Yossi Adi |
| 2025 | ICASSP | Salmon: A Suite for Acoustic Language Model Evaluation. | Gallil Maimon, Amit Roth, Yossi Adi |
| 2025 | ICASSP | Latent Watermarking of Audio Generative Models. | Robin San Roman, Pierre Fernandez, Antoine Deleforge, Yossi Adi, Romain Serizel |
| 2025 | ICASSP | MusicGen-Stem: Multi-stem music generation and edition through autoregressive modeling. | Simon Rouard, Robin San Roman, Yossi Adi, Axel Roebel |
| 2025 | ICASSP | Enhancing TTS Stability in Hebrew using Discrete Semantic Units. | Ella Zeldes, Or Tal, Yossi Adi |
| 2025 | ICCV | CAFA: A Controllable Automatic Foley Artist. | Roi Benita, Michael Finkelson, Tavi Halperin, Gleb Sterkin, Yossi Adi |
| 2025 | Interspeech | PAST: Phonetic-Acoustic Speech Tokenizer. | Nadav Har-Tuv, Or Tal, Yossi Adi |
| 2025 | Interspeech | WhiStress: Enriching Transcriptions with Sentence Stress Detection. | Iddo Yosha, Dorin Shteyman, Yossi Adi |
| 2024 | AAAI | Layer Collaboration in the Forward-Forward Algorithm. | Guy Lorberbom, Itai Gat, Yossi Adi, Alexander G. Schwing, Tamir Hazan |
| 2024 | AAAI | Diverse and Aligned Audio-to-Video Generation via Text-to-Video Model Adaptation. | Guy Yariv, Itai Gat, Sagie Benaim, Lior Wolf, Idan Schwartz, Yossi Adi |
| 2024 | EMNLP | Transformers are Multi-State RNNs. | Matanel Oren, Michael Hassid, Yarden Nir, Yossi Adi, Roy Schwartz |
| 2024 | ICLR | Masked Audio Generation using a Single Non-Autoregressive Transformer. | Alon Ziv, Itai Gat, Gal Le Lan, Tal Remez, Felix Kreuk, Jade Copet, Alexandre Dfossez, Gabriel Synnaeve, Yossi Adi |
| 2024 | ICML | An Independence-promoting Loss for Music Generation with Language Models. | Jean-Marie Lemercier, Simon Rouard, Jade Copet, Yossi Adi, Alexandre Dfossez |
| 2024 | Interspeech | Audio Enhancement from Multiple Crowdsourced Recordings: A Simple and Effective Baseline. | Shiran Aziz, Yossi Adi, Shmuel Peleg |
| 2024 | Interspeech | The Interspeech 2024 Challenge on Speech Processing Using Discrete Units. | Xuankai Chang, Jiatong Shi, Jinchuan Tian, Yuning Wu, Yuxun Tang, Yihan Wu, Shinji Watanabe, Yossi Adi, Xie Chen, Qin Jin |
| 2024 | Interspeech | NAST: Noise Aware Speech Tokenization for Speech Language Models. | Shoval Messica, Yossi Adi |
| 2024 | Interspeech | A Language Modeling Approach to Diacritic-Free Hebrew TTS. | Amit Roth, Arnon Turetzky, Yossi Adi |
| 2024 | Interspeech | HebDB: a Weakly Supervised Dataset for Hebrew Speech Processing. | Arnon Turetzky, Or Tal, Yael Segal, Yehoshua Dissen, Ella Zeldes, Amit Roth, Eyal Cohen, Yosi Shrem, Bronya Roni Chernyak, Olga Seleznova, Joseph Keshet, Yossi Adi |
| 2023 | CVPR | ReVISE: Self-Supervised Speech Resynthesis with Visual Input for Universal and Generalized Speech Regeneration. | Wei-Ning Hsu, Tal Remez, Bowen Shi, Jacob Donley, Yossi Adi |
| 2023 | EMNLP | Generative Spoken Language Model based on continuous word-sized audio tokens. | Robin Algayres, Yossi Adi, Tu Anh Nguyen, Jade Copet, Gabriel Synnaeve, Benot Sagot, Emmanuel Dupoux |
| 2023 | EMNLP | Speaking Style Conversion in the Waveform Domain Using Discrete Self-Supervised Units. | Gallil Maimon, Yossi Adi |
| 2023 | ICASSP | Do Coarser Units Benefit Cluster Prediction-Based Speech Pre-Training? | Ali Elkahky, Wei-Ning Hsu, Paden Tomasello, Tu Anh Nguyen, Robin Algayres, Yossi Adi, Jade Copet, Emmanuel Dupoux, Abdelrahman Mohamed |
| 2023 | ICASSP | A Holistic Cascade System, Benchmark, and Human Evaluation Protocol for Expressive Speech-to-Speech Translation. | Wen-Chin Huang, Benjamin N. Peloquin, Justine Kao, Changhan Wang, Hongyu Gong, Elizabeth Salesky, Yossi Adi, Ann Lee, Peng-Jen Chen |
| 2023 | ICASSP | AERO: Audio Super Resolution in the Spectral Domain. | Moshe Mandel, Or Tal, Yossi Adi |
| 2023 | ICASSP | I Hear Your True Colors: Image Guided Audio Generation. | Roy Sheffer, Yossi Adi |
| 2023 | ICASSP | Analysing Discrete Self Supervised Speech Representation For Spoken Language Modeling. | Amitay Sicherman, Yossi Adi |
| 2023 | ICLR | AudioGen: Textually Guided Audio Generation. | Felix Kreuk, Gabriel Synnaeve, Adam Polyak, Uriel Singer, Alexandre Dfossez, Jade Copet, Devi Parikh, Yaniv Taigman, Yossi Adi |
| 2023 | Interspeech | Expresso: A Benchmark and Analysis of Discrete Expressive Speech Resynthesis. | Tu Anh Nguyen, Wei-Ning Hsu, Antony D'Avirro, Bowen Shi, Itai Gat, Maryam Fazel-Zarandi, Tal Remez, Jade Copet, Gabriel Synnaeve, Michael Hassid, Felix Kreuk, Yossi Adi, Emmanuel Dupoux |
| 2023 | Interspeech | Adaptation of Text-Conditioned Diffusion Models for Audio-to-Image Generation. | Guy Yariv, Itai Gat, Lior Wolf, Yossi Adi, Idan Schwartz |
| 2022 | ACL | Text-Free Prosody-Aware Generative Spoken Language Modeling. | Eugene Kharitonov, Ann Lee, Adam Polyak, Yossi Adi, Jade Copet, Kushal Lakhotia, Tu Anh Nguyen, Morgane Rivire, Abdelrahman Mohamed, Emmanuel Dupoux, Wei-Ning Hsu |
| 2022 | ACL | Direct Speech-to-Speech Translation With Discrete Units. | Ann Lee, Peng-Jen Chen, Changhan Wang, Jiatao Gu, Sravya Popuri, Xutai Ma, Adam Polyak, Yossi Adi, Qing He, Yun Tang, Juan Pino, Wei-Ning Hsu |
| 2022 | EMNLP | Textless Speech Emotion Conversion using Discrete & Decomposed Representations. | Felix Kreuk, Adam Polyak, Jade Copet, Eugene Kharitonov, Tu Anh Nguyen, Morgane Rivire, Wei-Ning Hsu, Abdelrahman Mohamed, Emmanuel Dupoux, Yossi Adi |
| 2022 | ICASSP | Continual Self-Training With Bootstrapped Remixing For Speech Enhancement. | Efthymios Tzinis, Yossi Adi, Vamsi K. Ithapu, Buye Xu, Anurag Kumar |
| 2022 | ICLR | Learning Discrete Structured Variational Auto-Encoder using Natural Evolution Strategies. | Alon Berliner, Guy Rotman, Yossi Adi, Roi Reichart, Tamir Hazan |
| 2022 | Interspeech | Unsupervised Symbolic Music Segmentation using Ensemble Temporal Prediction Errors. | Shahaf Bassan, Yossi Adi, Jeffrey S. Rosenschein |
| 2022 | Interspeech | Enhanced Direct Speech-to-Speech Translation Using Self-supervised Pre-training and Data Augmentation. | Sravya Popuri, Peng-Jen Chen, Changhan Wang, Juan Pino, Yossi Adi, Jiatao Gu, Wei-Ning Hsu, Ann Lee |
| 2022 | Interspeech | Probing phoneme, language and speaker information in unsupervised speech representations. | Maureen de Seyssel, Marvin Lavechin, Yossi Adi, Emmanuel Dupoux, Guillaume Wisniewski |
| 2022 | Interspeech | A Systematic Comparison of Phonetic Aware Techniques for Speech Enhancement. | Or Tal, Moshe Mandel, Felix Kreuk, Yossi Adi |
| 2022 | Interspeech | Deep Audio Waveform Prior. | Arnon Turetzky, Tzvi Michelson, Yossi Adi, Shmuel Peleg |
| 2022 | NAACL | Textless Speech-to-Speech Translation on Real Data. | Ann Lee, Hongyu Gong, Paul-Ambroise Duquenne, Holger Schwenk, Peng-Jen Chen, Changhan Wang, Sravya Popuri, Yossi Adi, Juan Miguel Pino, Jiatao Gu, Wei-Ning Hsu |
| 2021 | AIES | Fairness in the Eyes of the Data: Certifying Machine-Learning Models. | Shahar Segal, Yossi Adi, Benny Pinkas, Carsten Baum, Chaya Ganesh, Joseph Keshet |
| 2021 | EMNLP | fairseq S\^2: A Scalable and Integrable Speech Synthesis Toolkit. | Changhan Wang, Wei-Ning Hsu, Yossi Adi, Adam Polyak, Ann Lee, Peng-Jen Chen, Jiatao Gu, Juan Pino |
| 2021 | ICASSP | Single Channel Voice Separation for Unknown Number of Speakers Under Reverberant and Noisy Settings. | Shlomo E. Chazan, Lior Wolf, Eliya Nachmani, Yossi Adi |
| 2021 | ICASSP | High Fidelity Speech Regeneration with Application to Speech Enhancement. | Adam Polyak, Lior Wolf, Yossi Adi, Ori Kabeli, Yaniv Taigman |
| 2021 | Interspeech | Speech Resynthesis from Discrete Disentangled Self-Supervised Representations. | Adam Polyak, Yossi Adi, Jade Copet, Eugene Kharitonov, Kushal Lakhotia, Wei-Ning Hsu, Abdelrahman Mohamed, Emmanuel Dupoux |
| 2020 | ICASSP | Phoneme Boundary Detection Using Learnable Segmental Features. | Felix Kreuk, Yaniv Sheena, Joseph Keshet, Yossi Adi |
| 2020 | ICML | Voice Separation with an Unknown Number of Multiple Speakers. | Eliya Nachmani, Yossi Adi, Lior Wolf |
| 2020 | Interspeech | Real Time Speech Enhancement in the Waveform Domain. | Alexandre Dfossez, Gabriel Synnaeve, Yossi Adi |
| 2020 | Interspeech | Hide and Speak: Towards Deep Neural Networks for Speech Steganography. | Felix Kreuk, Yossi Adi, Bhiksha Raj, Rita Singh, Joseph Keshet |
| 2020 | Interspeech | Self-Supervised Contrastive Learning for Unsupervised Phoneme Segmentation. | Felix Kreuk, Joseph Keshet, Yossi Adi |
| 2020 | Interspeech | Unsupervised Cross-Domain Singing Voice Conversion. | Adam Polyak, Lior Wolf, Yossi Adi, Yaniv Taigman |
| 2020 | LPAR | Minimal Modifications of Deep Neural Networks using Verification. | Ben Goldberger, Guy Katz, Yossi Adi, Joseph Keshet |
| 2019 | ICASSP | To Reverse the Gradient or Not: an Empirical Comparison of Adversarial and Multi-task Learning in Speech Recognition. | Yossi Adi, Neil Zeghidour, Ronan Collobert, Nicolas Usunier, Vitaliy Liptchinsky, Gabriel Synnaeve |
| 2018 | ICASSP | Fooling End-To-End Speaker Verification With Adversarial Examples. | Felix Kreuk, Yossi Adi, Moustapha Ciss, Joseph Keshet |
| 2017 | ICASSP | Sequence segmentation using joint RNN and structured prediction models. | Yossi Adi, Joseph Keshet, Emily Cibelli, Matthew Goldrick |
| 2017 | ICLR | Fine-grained Analysis of Sentence Embeddings Using Auxiliary Prediction Tasks. | Yossi Adi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, Yoav Goldberg |
| 2017 | Interspeech | Learning Similarity Functions for Pronunciation Variations. | Einat Naaman, Yossi Adi, Joseph Keshet |
| 2017 | Interspeech | Automatic Measurement of Pre-Aspiration. | Yaniv Sheena, Msa Hejn, Yossi Adi, Joseph Keshet |
| 2016 | Interspeech | Automatic Measurement of Voice Onset Time and Prevoicing Using Recurrent Neural Networks. | Yossi Adi, Joseph Keshet, Olga Dmitrieva, Matthew Goldrick |