Aren Jansen
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
59
Venues
11
Active years
2006–2025
Best venue rank
A*
Where they publish
Papers
59 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICLR | MELODI: Exploring Memory Compression for Long Contexts. | Yinpeng Chen, DeLesley Hutchins, Aren Jansen, Andrey Zhmoginov, David Racz, Jesper Sparre Andersen |
| 2025 | ICML | Long-Form Speech Generation with Spoken Language Models. | Se Jin Park, Julian Salazar, Aren Jansen, Keisuke Kinoshita, Yong Man Ro, R. J. Skerry-Ryan |
| 2024 | AAAI | V2Meow: Meowing to the Visual Beat via Video-to-Music Generation. | Kun Su, Judith Yue Li, Qingqing Huang, Dima Kuzmin, Joonseok Lee, Chris Donahue, Fei Sha, Aren Jansen, Yu Wang, Mauro Verzetti, Timo I. Denk |
| 2023 | ICASSP | Dataset Balancing Can Hurt Model Performance. | R. Channing Moore, Daniel P. W. Ellis, Eduardo Fonseca, Shawn Hershey, Aren Jansen, Manoj Plakal |
| 2022 | ICASSP | Universal Paralinguistic Speech Representations Using self-Supervised Conformers. | Joel Shor, Aren Jansen, Wei Han, Daniel S. Park, Yu Zhang |
| 2022 | Interspeech | Text-Driven Separation of Arbitrary Sounds. | Kevin Kilgour, Beat Gfeller, Qingqing Huang, Aren Jansen, Scott Wisdom, Marco Tagliasacchi |
| 2021 | ICASSP | The Benefit of Temporally-Strong Labels in Audio Event Classification. | Shawn Hershey, Daniel P. W. Ellis, Eduardo Fonseca, Aren Jansen, Caroline Liu, R. Channing Moore, Manoj Plakal |
| 2021 | ICLR | Into the Wild with AudioScope: Unsupervised Audio-Visual Separation of On-Screen Sounds. | Efthymios Tzinis, Scott Wisdom, Aren Jansen, Shawn Hershey, Tal Remez, Dan Ellis, John R. Hershey |
| 2020 | ICASSP | Large-Scale Weakly-Supervised Content Embeddings for Music Recommendation and Tagging. | Qingqing Huang, Aren Jansen, Li Zhang, Daniel P. W. Ellis, Rif A. Saurous, John R. Anderson |
| 2020 | ICASSP | Coincidence, Categorization, and Consolidation: Learning to Recognize Sounds with Minimal Supervision. | Aren Jansen, Daniel P. W. Ellis, Shawn Hershey, R. Channing Moore, Manoj Plakal, Ashok C. Popat, Rif A. Saurous |
| 2020 | ICASSP | Improving Universal Sound Separation Using Sound Classification. | Efthymios Tzinis, Scott Wisdom, John R. Hershey, Aren Jansen, Daniel P. W. Ellis |
| 2020 | Interspeech | Towards Learning a Universal Non-Semantic Representation of Speech. | Joel Shor, Aren Jansen, Ronnie Maor, Oran Lang, Omry Tuval, Flix de Chaumont Quitry, Marco Tagliasacchi, Ira Shavitt, Dotan Emanuel, Yinnon Haviv |
| 2018 | ICASSP | Unsupervised Learning of Semantic Audio Representations. | Aren Jansen, Manoj Plakal, Ratheet Pandya, Daniel P. W. Ellis, Shawn Hershey, Jiayang Liu, R. Channing Moore, Rif A. Saurous |
| 2017 | ICASSP | Audio Set: An ontology and human-labeled dataset for audio events. | Jort F. Gemmeke, Daniel P. W. Ellis, Dylan Freedman, Aren Jansen, Wade Lawrence, R. Channing Moore, Manoj Plakal, Marvin Ritter |
| 2017 | ICASSP | CNN architectures for large-scale audio classification. | Shawn Hershey, Sourish Chaudhuri, Daniel P. W. Ellis, Jort F. Gemmeke, Aren Jansen, R. Channing Moore, Manoj Plakal, Devin Platt, Rif A. Saurous, Bryan Seybold, Malcolm Slaney, Ron J. Weiss, Kevin W. Wilson |
| 2017 | ICASSP | Large-scale audio event discovery in one million YouTube videos. | Aren Jansen, Jort F. Gemmeke, Daniel P. W. Ellis, Xiaofeng Liu, Wade Lawrence, Dylan Freedman |
| 2016 | CogSci | A Framework for Evaluating Speech Representations. | Caitlin Richter, Naomi Feldman, Harini Salgado, Aren Jansen |
| 2016 | ICASSP | Context-dependent point process models for keyword search and detection-based ASR. | Chunxi Liu, Aren Jansen, Sanjeev Khudanpur |
| 2015 | ICASSP | Unsupervised neural network based feature extraction using weak top-down constraints. | Herman Kamper, Micha Elsner, Aren Jansen, Sharon Goldwater |
| 2015 | ICASSP | Segmental acoustic indexing for zero resource keyword search. | Keith D. Levin, Aren Jansen, Benjamin Van Durme |
| 2015 | ICASSP | Content-based recommender systems for spoken documents. | Jonathan Wintrode, Gregory Sell, Aren Jansen, Michelle Fox, Daniel Garcia-Romero, Alan McCree |
| 2015 | Interspeech | Fully unsupervised small-vocabulary speech recognition using a segmental Bayesian model. | Herman Kamper, Aren Jansen, Sharon Goldwater |
| 2015 | Interspeech | An evaluation of graph clustering methods for unsupervised term discovery. | Vince Lyzinski, Gregory Sell, Aren Jansen |
| 2015 | Interspeech | A comparison of neural network methods for unsupervised representation learning on the zero resource speech challenge. | Daniel Renshaw, Herman Kamper, Aren Jansen, Sharon Goldwater |
| 2015 | Interspeech | The zero resource speech challenge 2015. | Maarten Versteegh, Roland Thiollire, Thomas Schatz, Xuan-Nga Cao, Xavier Anguera, Aren Jansen, Emmanuel Dupoux |
| 2015 | NAACL | Using Zero-Resource Spoken Term Discovery for Ranked Retrieval. | Jerome White, Douglas W. Oard, Aren Jansen, Jiaul H. Paik, Rashmi Sankepally |
| 2015 | SIGIR | A Test Collection for Spoken Gujarati Queries. | Douglas W. Oard, Rashmi Sankepally, Jerome White, Aren Jansen, Craig Harman |
| 2014 | ICASSP | Unsupervised idiolect discovery for speaker recognition. | Aren Jansen, Daniel Garcia-Romero, Pascal Clark, Jaime Hernandez-Cordero |
| 2014 | ICASSP | Featherweight phonetic keyword search for conversational speech. | Keith Kintzley, Aren Jansen, Hynek Hermansky |
| 2014 | Interspeech | Low-resource open vocabulary keyword search using point process models. | Chunxi Liu, Aren Jansen, Guoguo Chen, Keith Kintzley, Jan Trmal, Sanjeev Khudanpur |
| 2014 | LREC | Bridging the gap between speech technology and natural language processing: an evaluation toolbox for term discovery systems. | Bogdan Ludusan, Maarten Versteegh, Aren Jansen, Guillaume Gravier, Xuan-Nga Cao, Mark Johnson, Emmanuel Dupoux |
| 2013 | ASRU | Fixed-dimensional acoustic embeddings of variable-length segments in low-resource settings. | Keith D. Levin, Katharine Henry, Aren Jansen, Karen Livescu |
| 2013 | ICASSP | Frequency offset correction in speech without detecting pitch. | Pascal Clark, Sri Harish Reddy Mallidi, Aren Jansen, Hynek Hermansky |
| 2013 | ICASSP | A summary of the 2012 JHU CLSP workshop on zero resource speech technologies and models of early language acquisition. | Aren Jansen, Emmanuel Dupoux, Sharon Goldwater, Mark Johnson, Sanjeev Khudanpur, Kenneth Church, Naomi Feldman, Hynek Hermansky, Florian Metze, Richard C. Rose, Mike Seltzer, Pascal Clark, Ian McGraw, Balakrishnan Varadarajan, Erin Bennett, Benjamin Brschinger, Justin T. Chiu, Ewan Dunbar, Abdellah Fourtassi, David Harwath, Chia-ying Lee, Keith D. Levin, Atta Norouzian, Vijayaditya Peddinti, Rachael Richardson, Thomas Schatz, Samuel Thomas |
| 2013 | ICASSP | Weak top-down constraints for unsupervised acoustic model training. | Aren Jansen, Samuel Thomas, Hynek Hermansky |
| 2013 | ICASSP | Zero resource graph-based confidence estimation for open vocabulary spoken term detection. | Atta Norouzian, Richard C. Rose, Sina Hamidi Ghalehjegh, Aren Jansen |
| 2013 | Interspeech | Text-to-speech inspired duration modeling for improved whole-word acoustic models. | Keith Kintzley, Aren Jansen, Hynek Hermansky |
| 2013 | Interspeech | Semi-supervised manifold learning approaches for spoken term verification. | Atta Norouzian, Richard C. Rose, Aren Jansen |
| 2013 | Interspeech | Evaluating speech features with the minimal-pair ABX task: analysis of the classical MFC/PLP pipeline. | Thomas Schatz, Vijayaditya Peddinti, Francis R. Bach, Aren Jansen, Hynek Hermansky, Emmanuel Dupoux |
| 2012 | Interspeech | Indexing Raw Acoustic Features for Scalable Zero Resource Search. | Aren Jansen, Benjamin Van Durme |
| 2012 | Interspeech | Intrinsic Spectral Analysis for Zero and High Resource Speech Recognition. | Aren Jansen, Samuel Thomas, Hynek Hermansky |
| 2012 | Interspeech | Inverting the Point Process Model for Fast Phonetic Keyword Search. | Keith Kintzley, Aren Jansen, Kenneth Church, Hynek Hermansky |
| 2012 | Interspeech | MAP Estimation of Whole-Word Acoustic Models with Dictionary Priors. | Keith Kintzley, Aren Jansen, Hynek Hermansky |
| 2012 | Interspeech | Exploiting Discriminative Point Process Models for Spoken Term Detection. | Atta Norouzian, Aren Jansen, Richard C. Rose, Samuel Thomas |
| 2012 | Interspeech | Data-driven Posterior Features for Low Resource Speech Recognition Applications. | Samuel Thomas, Sriram Ganapathy, Aren Jansen, Hynek Hermansky |
| 2011 | ASRU | Efficient spoken term discovery using randomized algorithms. | Aren Jansen, Benjamin Van Durme |
| 2011 | ASRU | Estimating document frequencies in a speech corpus. | Damianos Karakos, Mark Dredze, Ken Ward Church, Aren Jansen, Sanjeev Khudanpur |
| 2011 | ICASSP | Whole word discriminative point process models. | Aren Jansen |
| 2011 | ICASSP | Speech recognitionwith segmental conditional random fields: A summary of the JHU CLSP 2010 Summer Workshop. | Geoffrey Zweig, Patrick Nguyen, Dirk Van Compernolle, Kris Demuynck, Les E. Atlas, Pascal Clark, Gregory Sell, Meihong Wang, Fei Sha, Hynek Hermansky, Damianos Karakos, Aren Jansen, Samuel Thomas, Sivaram G. S. V. S., Samuel R. Bowman, Justine T. Kao |
| 2011 | Interspeech | Rapid Evaluation of Speech Representations for Spoken Term Discovery. | Michael A. Carlin, Samuel Thomas, Aren Jansen, Hynek Hermansky |
| 2011 | Interspeech | Towards Unsupervised Training of Speaker Independent Acoustic Models. | Aren Jansen, Kenneth Church |
| 2011 | Interspeech | Event Selection from Phone Posteriorgrams Using Matched Filters. | Keith Kintzley, Aren Jansen, Hynek Hermansky |
| 2010 | EMNLP | NLP on Spoken Documents Without ASR. | Mark Dredze, Aren Jansen, Glen Coppersmith, Ken Ward Church |
| 2010 | ICASSP | Detection-based speech recognition with sparse point process models. | Aren Jansen, Partha Niyogi |
| 2010 | Interspeech | Towards spoken term discovery at scale with zero resources. | Aren Jansen, Kenneth Church, Hynek Hermansky |
| 2009 | Interspeech | Robust keyword spotting with rapidly adapting point process models. | Aren Jansen, Partha Niyogi |
| 2008 | ICASSP | A hierarchical point process model for speech recognition. | Aren Jansen, Partha Niyogi |
| 2007 | Interspeech | Semi-supervised learning of speech sounds. | Aren Jansen, Partha Niyogi |
| 2006 | ICASSP | Intrinsic Fourier Analysis on the Manifold of Speech Sounds. | Aren Jansen, Partha Niyogi |