Skip to content

Aren Jansen

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

59

Venues

11

Active years

2006–2025

Best venue rank

A*

Where they publish

Papers

59 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRMELODI: Exploring Memory Compression for Long Contexts.Yinpeng Chen, DeLesley Hutchins, Aren Jansen, Andrey Zhmoginov, David Racz, Jesper Sparre Andersen
2025ICMLLong-Form Speech Generation with Spoken Language Models.Se Jin Park, Julian Salazar, Aren Jansen, Keisuke Kinoshita, Yong Man Ro, R. J. Skerry-Ryan
2024AAAIV2Meow: Meowing to the Visual Beat via Video-to-Music Generation.Kun Su, Judith Yue Li, Qingqing Huang, Dima Kuzmin, Joonseok Lee, Chris Donahue, Fei Sha, Aren Jansen, Yu Wang, Mauro Verzetti, Timo I. Denk
2023ICASSPDataset Balancing Can Hurt Model Performance.R. Channing Moore, Daniel P. W. Ellis, Eduardo Fonseca, Shawn Hershey, Aren Jansen, Manoj Plakal
2022ICASSPUniversal Paralinguistic Speech Representations Using self-Supervised Conformers.Joel Shor, Aren Jansen, Wei Han, Daniel S. Park, Yu Zhang
2022InterspeechText-Driven Separation of Arbitrary Sounds.Kevin Kilgour, Beat Gfeller, Qingqing Huang, Aren Jansen, Scott Wisdom, Marco Tagliasacchi
2021ICASSPThe Benefit of Temporally-Strong Labels in Audio Event Classification.Shawn Hershey, Daniel P. W. Ellis, Eduardo Fonseca, Aren Jansen, Caroline Liu, R. Channing Moore, Manoj Plakal
2021ICLRInto the Wild with AudioScope: Unsupervised Audio-Visual Separation of On-Screen Sounds.Efthymios Tzinis, Scott Wisdom, Aren Jansen, Shawn Hershey, Tal Remez, Dan Ellis, John R. Hershey
2020ICASSPLarge-Scale Weakly-Supervised Content Embeddings for Music Recommendation and Tagging.Qingqing Huang, Aren Jansen, Li Zhang, Daniel P. W. Ellis, Rif A. Saurous, John R. Anderson
2020ICASSPCoincidence, Categorization, and Consolidation: Learning to Recognize Sounds with Minimal Supervision.Aren Jansen, Daniel P. W. Ellis, Shawn Hershey, R. Channing Moore, Manoj Plakal, Ashok C. Popat, Rif A. Saurous
2020ICASSPImproving Universal Sound Separation Using Sound Classification.Efthymios Tzinis, Scott Wisdom, John R. Hershey, Aren Jansen, Daniel P. W. Ellis
2020InterspeechTowards Learning a Universal Non-Semantic Representation of Speech.Joel Shor, Aren Jansen, Ronnie Maor, Oran Lang, Omry Tuval, Flix de Chaumont Quitry, Marco Tagliasacchi, Ira Shavitt, Dotan Emanuel, Yinnon Haviv
2018ICASSPUnsupervised Learning of Semantic Audio Representations.Aren Jansen, Manoj Plakal, Ratheet Pandya, Daniel P. W. Ellis, Shawn Hershey, Jiayang Liu, R. Channing Moore, Rif A. Saurous
2017ICASSPAudio Set: An ontology and human-labeled dataset for audio events.Jort F. Gemmeke, Daniel P. W. Ellis, Dylan Freedman, Aren Jansen, Wade Lawrence, R. Channing Moore, Manoj Plakal, Marvin Ritter
2017ICASSPCNN architectures for large-scale audio classification.Shawn Hershey, Sourish Chaudhuri, Daniel P. W. Ellis, Jort F. Gemmeke, Aren Jansen, R. Channing Moore, Manoj Plakal, Devin Platt, Rif A. Saurous, Bryan Seybold, Malcolm Slaney, Ron J. Weiss, Kevin W. Wilson
2017ICASSPLarge-scale audio event discovery in one million YouTube videos.Aren Jansen, Jort F. Gemmeke, Daniel P. W. Ellis, Xiaofeng Liu, Wade Lawrence, Dylan Freedman
2016CogSciA Framework for Evaluating Speech Representations.Caitlin Richter, Naomi Feldman, Harini Salgado, Aren Jansen
2016ICASSPContext-dependent point process models for keyword search and detection-based ASR.Chunxi Liu, Aren Jansen, Sanjeev Khudanpur
2015ICASSPUnsupervised neural network based feature extraction using weak top-down constraints.Herman Kamper, Micha Elsner, Aren Jansen, Sharon Goldwater
2015ICASSPSegmental acoustic indexing for zero resource keyword search.Keith D. Levin, Aren Jansen, Benjamin Van Durme
2015ICASSPContent-based recommender systems for spoken documents.Jonathan Wintrode, Gregory Sell, Aren Jansen, Michelle Fox, Daniel Garcia-Romero, Alan McCree
2015InterspeechFully unsupervised small-vocabulary speech recognition using a segmental Bayesian model.Herman Kamper, Aren Jansen, Sharon Goldwater
2015InterspeechAn evaluation of graph clustering methods for unsupervised term discovery.Vince Lyzinski, Gregory Sell, Aren Jansen
2015InterspeechA comparison of neural network methods for unsupervised representation learning on the zero resource speech challenge.Daniel Renshaw, Herman Kamper, Aren Jansen, Sharon Goldwater
2015InterspeechThe zero resource speech challenge 2015.Maarten Versteegh, Roland Thiollire, Thomas Schatz, Xuan-Nga Cao, Xavier Anguera, Aren Jansen, Emmanuel Dupoux
2015NAACLUsing Zero-Resource Spoken Term Discovery for Ranked Retrieval.Jerome White, Douglas W. Oard, Aren Jansen, Jiaul H. Paik, Rashmi Sankepally
2015SIGIRA Test Collection for Spoken Gujarati Queries.Douglas W. Oard, Rashmi Sankepally, Jerome White, Aren Jansen, Craig Harman
2014ICASSPUnsupervised idiolect discovery for speaker recognition.Aren Jansen, Daniel Garcia-Romero, Pascal Clark, Jaime Hernandez-Cordero
2014ICASSPFeatherweight phonetic keyword search for conversational speech.Keith Kintzley, Aren Jansen, Hynek Hermansky
2014InterspeechLow-resource open vocabulary keyword search using point process models.Chunxi Liu, Aren Jansen, Guoguo Chen, Keith Kintzley, Jan Trmal, Sanjeev Khudanpur
2014LRECBridging the gap between speech technology and natural language processing: an evaluation toolbox for term discovery systems.Bogdan Ludusan, Maarten Versteegh, Aren Jansen, Guillaume Gravier, Xuan-Nga Cao, Mark Johnson, Emmanuel Dupoux
2013ASRUFixed-dimensional acoustic embeddings of variable-length segments in low-resource settings.Keith D. Levin, Katharine Henry, Aren Jansen, Karen Livescu
2013ICASSPFrequency offset correction in speech without detecting pitch.Pascal Clark, Sri Harish Reddy Mallidi, Aren Jansen, Hynek Hermansky
2013ICASSPA summary of the 2012 JHU CLSP workshop on zero resource speech technologies and models of early language acquisition.Aren Jansen, Emmanuel Dupoux, Sharon Goldwater, Mark Johnson, Sanjeev Khudanpur, Kenneth Church, Naomi Feldman, Hynek Hermansky, Florian Metze, Richard C. Rose, Mike Seltzer, Pascal Clark, Ian McGraw, Balakrishnan Varadarajan, Erin Bennett, Benjamin Brschinger, Justin T. Chiu, Ewan Dunbar, Abdellah Fourtassi, David Harwath, Chia-ying Lee, Keith D. Levin, Atta Norouzian, Vijayaditya Peddinti, Rachael Richardson, Thomas Schatz, Samuel Thomas
2013ICASSPWeak top-down constraints for unsupervised acoustic model training.Aren Jansen, Samuel Thomas, Hynek Hermansky
2013ICASSPZero resource graph-based confidence estimation for open vocabulary spoken term detection.Atta Norouzian, Richard C. Rose, Sina Hamidi Ghalehjegh, Aren Jansen
2013InterspeechText-to-speech inspired duration modeling for improved whole-word acoustic models.Keith Kintzley, Aren Jansen, Hynek Hermansky
2013InterspeechSemi-supervised manifold learning approaches for spoken term verification.Atta Norouzian, Richard C. Rose, Aren Jansen
2013InterspeechEvaluating speech features with the minimal-pair ABX task: analysis of the classical MFC/PLP pipeline.Thomas Schatz, Vijayaditya Peddinti, Francis R. Bach, Aren Jansen, Hynek Hermansky, Emmanuel Dupoux
2012InterspeechIndexing Raw Acoustic Features for Scalable Zero Resource Search.Aren Jansen, Benjamin Van Durme
2012InterspeechIntrinsic Spectral Analysis for Zero and High Resource Speech Recognition.Aren Jansen, Samuel Thomas, Hynek Hermansky
2012InterspeechInverting the Point Process Model for Fast Phonetic Keyword Search.Keith Kintzley, Aren Jansen, Kenneth Church, Hynek Hermansky
2012InterspeechMAP Estimation of Whole-Word Acoustic Models with Dictionary Priors.Keith Kintzley, Aren Jansen, Hynek Hermansky
2012InterspeechExploiting Discriminative Point Process Models for Spoken Term Detection.Atta Norouzian, Aren Jansen, Richard C. Rose, Samuel Thomas
2012InterspeechData-driven Posterior Features for Low Resource Speech Recognition Applications.Samuel Thomas, Sriram Ganapathy, Aren Jansen, Hynek Hermansky
2011ASRUEfficient spoken term discovery using randomized algorithms.Aren Jansen, Benjamin Van Durme
2011ASRUEstimating document frequencies in a speech corpus.Damianos Karakos, Mark Dredze, Ken Ward Church, Aren Jansen, Sanjeev Khudanpur
2011ICASSPWhole word discriminative point process models.Aren Jansen
2011ICASSPSpeech recognitionwith segmental conditional random fields: A summary of the JHU CLSP 2010 Summer Workshop.Geoffrey Zweig, Patrick Nguyen, Dirk Van Compernolle, Kris Demuynck, Les E. Atlas, Pascal Clark, Gregory Sell, Meihong Wang, Fei Sha, Hynek Hermansky, Damianos Karakos, Aren Jansen, Samuel Thomas, Sivaram G. S. V. S., Samuel R. Bowman, Justine T. Kao
2011InterspeechRapid Evaluation of Speech Representations for Spoken Term Discovery.Michael A. Carlin, Samuel Thomas, Aren Jansen, Hynek Hermansky
2011InterspeechTowards Unsupervised Training of Speaker Independent Acoustic Models.Aren Jansen, Kenneth Church
2011InterspeechEvent Selection from Phone Posteriorgrams Using Matched Filters.Keith Kintzley, Aren Jansen, Hynek Hermansky
2010EMNLPNLP on Spoken Documents Without ASR.Mark Dredze, Aren Jansen, Glen Coppersmith, Ken Ward Church
2010ICASSPDetection-based speech recognition with sparse point process models.Aren Jansen, Partha Niyogi
2010InterspeechTowards spoken term discovery at scale with zero resources.Aren Jansen, Kenneth Church, Hynek Hermansky
2009InterspeechRobust keyword spotting with rapidly adapting point process models.Aren Jansen, Partha Niyogi
2008ICASSPA hierarchical point process model for speech recognition.Aren Jansen, Partha Niyogi
2007InterspeechSemi-supervised learning of speech sounds.Aren Jansen, Partha Niyogi
2006ICASSPIntrinsic Fourier Analysis on the Manifold of Speech Sounds.Aren Jansen, Partha Niyogi