Skip to content

Berrak Sisman

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

35

Venues

8

Active years

2016–2026

Best venue rank

A*

Where they publish

Papers

35 indexed papers, newest first.

YearVenueTitleAuthors
2026ACLDiscovering and Causally Validating Emotion-Sensitive Neurons in Large Audio-Language Models.Xiutian Zhao, Bjrn W. Schuller, Berrak Sisman
2025EMNLPMultimodal Fine-grained Context Interaction Graph Modeling for Conversational Speech Synthesis.Zhenqi Jia, Rui Liu, Berrak Sisman, Haizhou Li
2025InterspeechTowards Emotionally Consistent Text-Based Speech Editing: Introducing EmoCorrector and The ECD-TSE Dataset.Rui Liu, Pu Gao, Jiatian Xi, Berrak Sisman, Carlos Busso, Haizhou Li
2025InterspeechEmotionRankCLAP: Bridging Natural Language Speaking Styles and Ordinal Speech Emotion via Rank-N-Contrast.Shreeram Suresh Chandra, Lucas Goncalves, Junchen Lu, Carlos Busso, Berrak Sisman
2025InterspeechCan Emotion Fool Anti-spoofing?Aurosweta Mahapatra, Ismail Rasim Ulgen, Abinay Reddy Naini, Carlos Busso, Berrak Sisman
2025InterspeechThe Interspeech 2025 Challenge on Speech Emotion Recognition in Naturalistic Conditions.Abinay Reddy Naini, Lucas Goncalves, Ali N. Salman, Pravin Mote, Ismail Rasim Ulgen, Thomas Thebaud, Laureano Moro-Velzquez, Leibny Paola Garca, Najim Dehak, Berrak Sisman, Carlos Busso
2025InterspeechAdvancing Pediatric ASR: The Role of Voice Generation in Disordered Speech.Karen Rosero, Ali N. Salman, Shreeram Suresh Chandra, Berrak Sisman, Cortney Van't Slot, Alex A. Kane, Rami R. Hallac, Carlos Busso
2024ICASSPRevealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition.Ismail Rasim Ulgen, Zongyang Du, Carlos Busso, Berrak Sisman
2024InterspeechUnsupervised Domain Adaptation for Speech Emotion Recognition using K-Nearest Neighbors Voice Conversion.Pravin Mote, Berrak Sisman, Carlos Busso
2024InterspeechTowards Naturalistic Voice Conversion: NaturalVoices Dataset with an Automatic Processing Pipeline.Ali N. Salman, Zongyang Du, Shreeram Suresh Chandra, Ismail Rasim lgen, Carlos Busso, Berrak Sisman
2024TenconSNIPER Training: Single-Shot Sparse Training for Text-to-Speech.Perry Lam, Huayun Zhang, Nancy F. Chen, Berrak Sisman, Dorien Herremans
2024TenconAccented Text-to-Speech Synthesis with a Conditional Variational Autoencoder.Jan Melechovsk, Ambuj Mehrish, Berrak Sisman, Dorien Herremans
2024TenconAccent Conversion in Text-to-Speech Using Multi-Level VAE and Adversarial Training.Jan Melechovsk, Ambuj Mehrish, Berrak Sisman, Dorien Herremans
2023InterspeechSlothSpeech: Denial-of-service Attack Against Speech Recognition Models.Mirazul Haque, Rutvij Shah, Simin Chen, Berrak Sisman, Cong Liu, Wei Yang
2023InterspeechHigh-Quality Automatic Voice Over with Accurate Alignment: Supervision through Self-Supervised Discrete Speech Units.Junchen Lu, Berrak Sisman, Mingyang Zhang, Haizhou Li
2022ICASSPVisualtts: TTS with Accurate Lip-Speech Synchronization for Automatic Voice Over.Junchen Lu, Berrak Sisman, Rui Liu, Mingyang Zhang, Haizhou Li
2022InterspeechAccurate Emotion Strength Assessment for Seen and Unseen Speech Based on Data-Driven Deep Learning.Rui Liu, Berrak Sisman, Bjrn W. Schuller, Guanglai Gao, Haizhou Li
2022InterspeechDisentanglement of Emotional Style and Speaker Identity for Expressive Voice Conversion.Zongyang Du, Berrak Sisman, Kun Zhou, Haizhou Li
2022InterspeechEPIC TTS Models: Empirical Pruning Investigations Characterizing Text-To-Speech Models.Perry Lam, Huayun Zhang, Nancy F. Chen, Berrak Sisman
2021ASRUExpressive Voice Conversion: A Joint Framework for Speaker Identity and Emotional Style Transfer.Zongyang Du, Berrak Sisman, Kun Zhou, Haizhou Li
2021ASRUDEEPA: A Deep Neural Analyzer for Speech and Singing Vocoding.Sergey Nikonorov, Berrak Sisman, Mingyang Zhang, Haizhou Li
2021ICASSPGraphspeech: Syntax-Aware Graph Attention Network for Neural Speech Synthesis.Rui Liu, Berrak Sisman, Haizhou Li
2021ICASSPSeen and Unseen Emotional Style Transfer for Voice Conversion with A New Emotional Speech Dataset.Kun Zhou, Berrak Sisman, Rui Liu, Haizhou Li
2021InterspeechReinforcement Learning for Emotional Text-to-Speech Synthesis with Improved Emotion Discriminability.Rui Liu, Berrak Sisman, Haizhou Li
2021InterspeechLimited Data Emotional Voice Conversion Leveraging Text-to-Speech: Two-Stage Sequence-to-Sequence Training.Kun Zhou, Berrak Sisman, Haizhou Li
2021SIGdialProceedings of the 22nd Annual Meeting of the Special Interest Group on Discourse and Dialogue.Haizhou Li, Gina-Anne Levow, Zhou Yu, Chitralekha Gupta, Berrak Sisman, Siqi Cai, David Vandyke, Nina Dethlefs, Yan Wu, Junyi Jessy Li
2020ICASSPTeacher-Student Training For Robust Tacotron-Based TTS.Rui Liu, Berrak Sisman, Jingdong Li, Feilong Bao, Guanglai Gao, Haizhou Li
2020InterspeechConverting Anyone's Emotion: Towards Speaker-Independent Emotional Voice Conversion.Kun Zhou, Berrak Sisman, Mingyang Zhang, Haizhou Li
2019ASRUOn the Study of Generative Adversarial Networks for Cross-Lingual Voice Conversion.Berrak Sisman, Mingyang Zhang, Minghui Dong, Haizhou Li
2019InterspeechVQVAE Unsupervised Unit Discovery and Multi-Scale Code2Spec Inverter for Zerospeech Challenge 2019.Andros Tjandra, Berrak Sisman, Mingyang Zhang, Sakriani Sakti, Haizhou Li, Satoshi Nakamura
2018InterspeechWavelet Analysis of Speaker Dependent and Independent Prosody for Voice Conversion.Berrak Sisman, Haizhou Li
2018InterspeechA Voice Conversion Framework with Tandem Feature Sparse Representation and Speaker-Adapted WaveNet Vocoder.Berrak Sisman, Mingyang Zhang, Haizhou Li
2017ASRUSparse representation of phonetic features for voice conversion with and without parallel data.Berrak Sisman, Haizhou Li, Kay Chen Tan
2016WCNCEnergy and data cooperation in energy harvesting multiple access channel.Berk Gurakan, Berrak Sisman, Onur Kaya, Sennur Ulukus
2016WCNCEnergy and data cooperation in energy harvesting multiple access channel.Berk Gurakan, Berrak Sisman, Onur Kaya, Sennur Ulukus