vix.ing · top · new · best · stats · spec

Sisman, Berrak

  1. Emotional Voice Conversion: Theory, Databases and ESD
    2021/05/31 by Zhou, Kun, Sisman, Berrak, Liu, Rui +1 · 24 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  2. Seen and Unseen emotional style transfer for voice conversion with a new emotional speech dataset
    2020/10/28 by Kun Zhou, Berrak Şişman, Zhou, Kun +5 · 17 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. An Overview of Voice Conversion and its Challenges: From Statistical Modeling to Deep Learning
    2020/08/09 by Berrak Şişman, Sisman, Berrak, Junichi Yamagishi +5 · 13 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Disentanglement of Emotional Style and Speaker Identity for Expressive Voice Conversion
    2021/10/20 by Du, Zongyang, Sisman, Berrak, Zhou, Kun +1 · 5 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  5. Converting Anyone's Emotion: Towards Speaker-Independent Emotional Voice Conversion
    2020/05/13 by Kun Zhou, Zhou, Kun, Berrak Şişman +5 · 5 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  6. VisualTTS: TTS with Accurate Lip-Speech Synchronization for Automatic Voice Over
    2021/10/07 by Junchen Lu, Lu, Junchen, Berrak Şişman +7 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Face recognition and analysis #Speech and Audio Processing #Video Analysis and Summarization #electronic engineering #information engineering
  7. Speech Synthesis with Mixed Emotions
    2022/08/11 by Zhou, Kun, Sisman, Berrak, Rana, Rajib +2 · 3 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  8. Transforming Spectrum and Prosody for Emotional Voice Conversion with Non-Parallel Training Data
    2020/02/01 by Kun Zhou, Berrak Şişman, Zhou, Kun +3 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. Expressive Voice Conversion: A Joint Framework for Speaker Identity and Emotional Style Transfer
    2021/07/08 by Zongyang Du, Du, Zongyang, Berrak Şişman +5 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  10. Limited Data Emotional Voice Conversion Leveraging Text-to-Speech: Two-stage Sequence-to-Sequence Training
    2021/03/31 by Kun Zhou, Berrak Şişman, Zhou, Kun +3 · 3 citations
    Computer Science · Medicine · #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders
  11. Reinforcement Learning for Emotional Text-to-Speech Synthesis with Improved Emotion Discriminability
    2021/04/03 by Liu, Rui, Sisman, Berrak, Li, Haizhou · 2 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  12. Style Mixture of Experts for Expressive Text-To-Speech Synthesis
    2024/06/05 by Jawaid, Ahad, Chandra, Shreeram Suresh, Lu, Junchen +1 · 3 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  13. High-Quality Automatic Voice Over with Accurate Alignment: Supervision through Self-Supervised Discrete Speech Units
    2023/06/29 by Lu, Junchen, Sisman, Berrak, Zhang, Mingyang +1 · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  14. VQVAE Unsupervised Unit Discovery and Multi-scale Code2Spec Inverter for Zerospeech Challenge 2019
    2019/05/27 by Tjandra, Andros, Sisman, Berrak, Zhang, Mingyang +3 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  15. Versatile audio-visual learning for emotion recognition
    2023/05/12 by Lucas Goncalves, Goncalves, Lucas, Seong-Gyun Leem +7 · 2 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multimedia (cs.MM) #Sentiment Analysis and Opinion Mining #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  16. Teacher-Student Training for Robust Tacotron-based TTS
    2019/11/07 by Liu, Rui, Sisman, Berrak, Li, Jingdong +3 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  17. Expressive TTS Training with Frame and Style Reconstruction Loss
    2020/08/04 by Liu, Rui, Sisman, Berrak, Gao, Guanglai +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  18. VAW-GAN for Disentanglement and Recomposition of Emotional Elements in\n Speech
    2020/11/03 by Kun Zhou, Berrak Şişman, Zhou, Kun +3 · 1 citation
    Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
  19. Towards Naturalistic Voice Conversion: NaturalVoices Dataset with an Automatic Processing Pipeline
    2024/06/06 by Ali N. Salman, Salman, Ali N., Zongyang Du +9 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #electronic engineering #information engineering
  20. Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder
    2022/11/07 by Jan Melechovský, Melechovsky, Jan, Ambuj Mehrish +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  21. DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech
    2024/10/17 by Melechovsky, Jan, Mehrish, Ambuj, Sisman, Berrak +1 · 2 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  22. Exploring speech style spaces with language models: Emotional TTS without emotion labels
    2024/05/18 by Shreeram Suresh Chandra, Chandra, Shreeram Suresh, Zongyang Du +3 · 2 citations
    Computer Science · #Speech and dialogue systems
  23. Accent Conversion in Text-To-Speech Using Multi-Level VAE and Adversarial Training
    2024/06/03 by Melechovsky, Jan, Mehrish, Ambuj, Sisman, Berrak +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  24. EmotionRankCLAP: Bridging Natural Language Speaking Styles and Ordinal Speech Emotion via Rank-N-Contrast
    2025/05/29 by Shreeram Suresh Chandra, Chandra, Shreeram Suresh, Lucas Goncalves +7 · 3 citations
    Computer Science · Psychology · #Emotion and Mood Recognition #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Sentiment Analysis and Opinion Mining
  25. Can Emotion Fool Anti-spoofing?
    2025/05/29 by Amogh Mahapatra, Mahapatra, Aurosweta, İsmail Rasim Ülgen +7 · 4 citations
    Social Sciences · #Audio and Speech Processing (eess.AS) #Cybersecurity and Cyber Warfare Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Freedom of Expression and Defamation #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  26. Converting Anyone's Voice: End-to-End Expressive Voice Conversion with a Conditional Diffusion Model
    2024/05/02 by Zongyang Du, Junchen Lu, Du, Zongyang +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering