Sisman, Berrak
- Emotional Voice Conversion: Theory, Databases and ESD
2021/05/31 by Zhou, Kun, Sisman, Berrak, Liu, Rui +1 · 24 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- Seen and Unseen emotional style transfer for voice conversion with a new emotional speech dataset
2020/10/28 by Kun Zhou, Berrak Şişman, Zhou, Kun +5 · 17 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- An Overview of Voice Conversion and its Challenges: From Statistical Modeling to Deep Learning
2020/08/09 by Berrak Şişman, Sisman, Berrak, Junichi Yamagishi +5 · 13 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Disentanglement of Emotional Style and Speaker Identity for Expressive Voice Conversion
2021/10/20 by Du, Zongyang, Sisman, Berrak, Zhou, Kun +1 · 5 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Converting Anyone's Emotion: Towards Speaker-Independent Emotional Voice Conversion
2020/05/13 by Kun Zhou, Zhou, Kun, Berrak Şişman +5 · 5 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- VisualTTS: TTS with Accurate Lip-Speech Synchronization for Automatic Voice Over
2021/10/07 by Junchen Lu, Lu, Junchen, Berrak Şişman +7 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Face recognition and analysis #Speech and Audio Processing #Video Analysis and Summarization #electronic engineering #information engineering
- Speech Synthesis with Mixed Emotions
2022/08/11 by Zhou, Kun, Sisman, Berrak, Rana, Rajib +2 · 3 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
- Transforming Spectrum and Prosody for Emotional Voice Conversion with Non-Parallel Training Data
2020/02/01 by Kun Zhou, Berrak Şişman, Zhou, Kun +3 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Expressive Voice Conversion: A Joint Framework for Speaker Identity and Emotional Style Transfer
2021/07/08 by Zongyang Du, Du, Zongyang, Berrak Şişman +5 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Limited Data Emotional Voice Conversion Leveraging Text-to-Speech: Two-stage Sequence-to-Sequence Training
2021/03/31 by Kun Zhou, Berrak Şişman, Zhou, Kun +3 · 3 citations
Computer Science · Medicine · #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders
- Reinforcement Learning for Emotional Text-to-Speech Synthesis with Improved Emotion Discriminability
2021/04/03 by Liu, Rui, Sisman, Berrak, Li, Haizhou · 2 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- Style Mixture of Experts for Expressive Text-To-Speech Synthesis
2024/06/05 by Jawaid, Ahad, Chandra, Shreeram Suresh, Lu, Junchen +1 · 3 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- High-Quality Automatic Voice Over with Accurate Alignment: Supervision through Self-Supervised Discrete Speech Units
2023/06/29 by Lu, Junchen, Sisman, Berrak, Zhang, Mingyang +1 · 2 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- VQVAE Unsupervised Unit Discovery and Multi-scale Code2Spec Inverter for Zerospeech Challenge 2019
2019/05/27 by Tjandra, Andros, Sisman, Berrak, Zhang, Mingyang +3 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Versatile audio-visual learning for emotion recognition
2023/05/12 by Lucas Goncalves, Goncalves, Lucas, Seong-Gyun Leem +7 · 2 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multimedia (cs.MM) #Sentiment Analysis and Opinion Mining #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Teacher-Student Training for Robust Tacotron-based TTS
2019/11/07 by Liu, Rui, Sisman, Berrak, Li, Jingdong +3 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Expressive TTS Training with Frame and Style Reconstruction Loss
2020/08/04 by Liu, Rui, Sisman, Berrak, Gao, Guanglai +1 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- VAW-GAN for Disentanglement and Recomposition of Emotional Elements in\n Speech
2020/11/03 by Kun Zhou, Berrak Şişman, Zhou, Kun +3 · 1 citation
Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
- Towards Naturalistic Voice Conversion: NaturalVoices Dataset with an Automatic Processing Pipeline
2024/06/06 by Ali N. Salman, Salman, Ali N., Zongyang Du +9 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #electronic engineering #information engineering
- Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder
2022/11/07 by Jan Melechovský, Melechovsky, Jan, Ambuj Mehrish +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech
2024/10/17 by Melechovsky, Jan, Mehrish, Ambuj, Sisman, Berrak +1 · 2 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Exploring speech style spaces with language models: Emotional TTS without emotion labels
2024/05/18 by Shreeram Suresh Chandra, Chandra, Shreeram Suresh, Zongyang Du +3 · 2 citations
Computer Science · #Speech and dialogue systems
- Accent Conversion in Text-To-Speech Using Multi-Level VAE and Adversarial Training
2024/06/03 by Melechovsky, Jan, Mehrish, Ambuj, Sisman, Berrak +1 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- EmotionRankCLAP: Bridging Natural Language Speaking Styles and Ordinal Speech Emotion via Rank-N-Contrast
2025/05/29 by Shreeram Suresh Chandra, Chandra, Shreeram Suresh, Lucas Goncalves +7 · 3 citations
Computer Science · Psychology · #Emotion and Mood Recognition #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Sentiment Analysis and Opinion Mining
- Can Emotion Fool Anti-spoofing?
2025/05/29 by Amogh Mahapatra, Mahapatra, Aurosweta, İsmail Rasim Ülgen +7 · 4 citations
Social Sciences · #Audio and Speech Processing (eess.AS) #Cybersecurity and Cyber Warfare Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Freedom of Expression and Defamation #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Converting Anyone's Voice: End-to-End Expressive Voice Conversion with a Conditional Diffusion Model
2024/05/02 by Zongyang Du, Junchen Lu, Du, Zongyang +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering