vix.ing · top · new · best · stats · spec

Junichi Yamagishi

  1. ASVspoof 2019: A large-scale public database of synthesized, converted and replayed speech
    2020/05/20 by Xin Wang, Junichi Yamagishi, Massimiliano Todisco +43 · 44 citations
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing
  2. Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmentation
    2022/02/24 by Hemlata Tak, Tak, Hemlata, Massimiliano Todisco +9 · 30 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  3. ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale
    2024/08/16 by Xin Wang, Héctor Delgado, Wang, Xin +23 · 42 citations
    Computer Science · #Hate Speech and Cyberbullying Detection
  4. OpenForensics: Large-Scale Challenging Dataset For Multi-Face Forgery Detection And Segmentation In-The-Wild
    2021/07/30 by Trung-Nghia Le, Le, Trung-Nghia, Huy H. Nguyen +5 · 1 voice · 5 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Digital Media Forensic Detection #FOS: Computer and information sciences #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #cs.CV
  5. Capsule-Forensics: Using Capsule Networks to Detect Forged Images and Videos
    2018/10/26 by Huy H. Nguyen, Junichi Yamagishi, Nguyen, Huy H. +3 · 13 citations
    Computer Science · #Digital Media Forensic Detection #Advanced Steganography and Watermarking Techniques #Generative Adversarial Networks and Image Synthesis
  6. Investigating self-supervised front ends for speech spoofing countermeasures
    2021/11/15 by Xin Wang, Wang, Xin, Junichi Yamagishi +1 · 17 citations
    Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
  7. An Overview of Voice Conversion and its Challenges: From Statistical Modeling to Deep Learning
    2020/08/09 by Berrak Şişman, Junichi Yamagishi, Sisman, Berrak +5 · 13 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  8. Generalization Ability of MOS Prediction Networks
    2021/10/06 by Erica Cooper, Wen-Chin Huang, Cooper, Erica +5 · 15 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #Topic Modeling #electronic engineering #information engineering
  9. How do Voices from Past Speech Synthesis Challenges Compare Today?
    2021/05/05 by Erica Cooper, Junichi Yamagishi, Cooper, Erica +1 · 11 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Topic Modeling #electronic engineering #information engineering
  10. Speaker Anonymization Using X-vector and Neural Waveform Models
    2019/05/30 by Fuming Fang, Fang, Fuming, Xin Wang +11 · 9 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  11. Voice Conversion Challenge 2020: Intra-lingual semi-parallel and cross-lingual voice conversion
    2020/08/28 by Yi Zhao, Zhao, Yi, Wen-Chin Huang +13 · 7 citations
    Computer Science · Medicine · #Speech Recognition and Synthesis #Topic Modeling #Voice and Speech Disorders
  12. An Initial Investigation for Detecting Partially Spoofed Audio
    2021/04/06 by Lin Zhang, Zhang, Lin, Xin Wang +9 · 6 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  13. Neural source-filter waveform models for statistical parametric speech synthesis
    2019/04/27 by Xin Wang, Wang, Xin, Shinji Takaki +3 · 5 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Natural Language Processing Techniques
  14. Spoofed training data for speech spoofing countermeasure can be efficiently created using neural vocoders
    2022/10/19 by Xin Wang, Junichi Yamagishi, Wang, Xin +1 · 7 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  15. Can large-scale vocoded spoofed data improve speech spoofing countermeasure with a self-supervised front end?
    2023/09/12 by Xin Wang, Junichi Yamagishi, Wang, Xin +1 · 8 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  16. Zero-Shot Multi-Speaker Text-To-Speech with State-of-the-art Neural Speaker Embeddings
    2019/10/23 by Erica Cooper, Cooper, Erica, Cheng-I Lai +11 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
  17. Investigating Range-Equalizing Bias in Mean Opinion Score Ratings of Synthesized Speech
    2023/05/17 by Erica Cooper, Cooper, Erica, Junichi Yamagishi +1 · 6 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing
  18. ASVspoof 2021: Automatic Speaker Verification Spoofing and Countermeasures Challenge Evaluation Plan
    2021/09/01 by Héctor Delgado, Nicholas Evans, Delgado, Héctor +19 · 4 citations
    Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
  19. The VoiceMOS Challenge 2024: Beyond Speech Quality Prediction
    2024/09/11 by Wen-Chin Huang, Szu‐Wei Fu, Huang, Wen-Chin +13 · 12 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  20. Use of a Capsule Network to Detect Fake Images and Videos
    2019/10/28 by Huy H. Nguyen, Nguyen, Huy H., Junichi Yamagishi +3 · 4 citations
    Computer Science · #Advanced Steganography and Watermarking Techniques #Computer Vision and Pattern Recognition (cs.CV) #Digital Media Forensic Detection #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis
  21. The VoicePrivacy 2022 Challenge: Progress and Perspectives in Voice Anonymisation
    2024/07/16 by Michele Panariello, Panariello, Michele, Natalia Tomashenko +17 · 9 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #electronic engineering #information engineering
  22. Language-Independent Speaker Anonymization Approach using Self-Supervised Pre-Trained Models
    2022/02/26 by Xiaoxiao Miao, Xin Wang, Miao, Xiaoxiao +7 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  23. The VoicePrivacy 2022 Challenge Evaluation Plan
    2022/03/23 by Natalia Tomashenko, Xin Wang, Tomashenko, Natalia +17 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
  24. Spoofing and countermeasures for speaker verification: A survey
    2015/02/01 by Zhizheng Wu, Nicholas Evans, Tomi Kinnunen +3 · 2 citations
  25. ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech
    2025/02/13 by Xin Wang, Wang, Xin, Héctor Delgado +54 · 9 citations
    Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Hate Speech and Cyberbullying Detection #Misinformation and Its Impacts #electronic engineering #information engineering
  26. Analyzing Language-Independent Speaker Anonymization Framework under Unseen Conditions
    2022/03/28 by Xiaoxiao Miao, Miao, Xiaoxiao, Xin Wang +7 · 3 citations
    Computer Science · #FOS: Computer and information sciences #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis
  27. Range-Based Equal Error Rate for Spoof Localization
    2023/05/28 by Lin Zhang, Zhang, Lin, X.J. Wang +7 · 2 citations
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Gait Recognition and Analysis #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  28. Post-training for Deepfake Speech Detection
    2025/06/26 by Wanying Ge, Ge, Wanying, Xin Wang +5 · 8 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  29. An Initial Investigation of Language Adaptation for TTS Systems under Low-resource Scenarios
    2024/06/13 by Gong Cheng, Erica Cooper, Gong, Cheng +21 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimodal Machine Learning Applications #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  30. GELP: GAN-Excited Linear Prediction for Speech Synthesis from\n Mel-spectrogram
    2019/04/08 by Lauri Juvela, Juvela, Lauri, Bajibabu Bollepalli +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  31. ASVspoof 5: Design, collection and validation of resources for spoofing, deepfake, and adversarial attack detection using crowdsourced speech
    2025/05/28 by Xin Wang, Héctor Delgado, Hemlata Tak +26 · 1 voice · 6 citations
    Computer Science · #Speech Recognition and Synthesis #Network Security and Intrusion Detection #Hate Speech and Cyberbullying Detection
  32. QualiSpeech: A Speech Quality Assessment Dataset with Natural Language Reasoning and Descriptions
    2025/03/26 by Siyin Wang, Wang, Siyin, Wenyi Yu +16 · 6 citations
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis
  33. Hiding speaker's sex in speech using zero-evidence speaker representation in an analysis/synthesis pipeline
    2022/11/29 by ‪Paul-Gauthier Noé‬, Xiaoxiao Miao, Noé, Paul-Gauthier +9 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  34. Predictions of Subjective Ratings and Spoofing Assessments of Voice Conversion Challenge 2020 Submissions
    2020/09/08 by Rohan Kumar Das, Tomi Kinnunen, Das, Rohan Kumar +13 · 1 citation
    Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
  35. Multi-Metric Optimization using Generative Adversarial Networks for Near-End Speech Intelligibility Enhancement
    2021/04/17 by Haoyu Li, Junichi Yamagishi, Li, Haoyu +1 · 1 citation
    Computer Science · Neuroscience · Engineering · #Speech and Audio Processing #Hearing Loss and Rehabilitation #Acoustic Wave Phenomena Research
  36. A Multi-Level Attention Model for Evidence-Based Fact Checking
    2021/06/02 by Canasai Kruengkrai, Kruengkrai, Canasai, Junichi Yamagishi +3 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Software Engineering Research #Topic Modeling
  37. Use of speaker recognition approaches for learning and evaluating embedding representations of musical instrument sounds
    2021/07/24 by Xuan Shi, Shi, Xuan, Erica Cooper +3 · 1 citation
    Computer Science · #Music and Audio Processing #Speech and Audio Processing #Music Technology and Sound Studies
  38. A Practical Guide to Logical Access Voice Presentation Attack Detection
    2022/01/10 by Xin Wang, Wang, Xin, Junichi Yamagishi +1 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and Audio Processing
  39. Towards An Integrated Approach for Expressive Piano Performance Synthesis from Music Scores
    2025/01/17 by Jingjing Tang, Tang, Jingjing, Erica Cooper +7 · 1 citation
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception #Sound (cs.SD) #electronic engineering #information engineering
  40. Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data
    2024/12/17 by Yun Liu, Xuechen Liu, Liu, Yun +5 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  41. Good practices for evaluation of synthesized speech
    2025/03/05 by Erica Cooper, Cooper, Erica, Sébastien Le Maguer +5 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and dialogue systems
  42. Spoofing-Aware Speaker Verification Robust Against Domain and Channel Mismatches
    2024/09/10 by Chang Zeng, Xiaoxiao Miao, Zeng, Chang +7 · 1 citation
    Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  43. Disentangling the Prosody and Semantic Information with Pre-trained Model for In-Context Learning based Zero-Shot Voice Conversion
    2024/09/08 by Zhengyang Chen, Shuai Wang, Chen, Zhengyang +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  44. Uncertainty as a Predictor: Leveraging Self-Supervised Learning for Zero-Shot MOS Prediction
    2023/12/25 by Aditya Ravuri, Erica Cooper, Ravuri, Aditya +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  45. It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model
    2024/12/03 by Mingyi Shi, Dafei Qin, Shi, Mingyi +11 · 1 citation
    Computer Science · #Speech and dialogue systems #Topic Modeling #Speech Recognition and Synthesis
  46. Explaining Speaker and Spoof Embeddings via Probing
    2024/12/24 by Xuechen Liu, Liu, Xuechen, Junichi Yamagishi +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  47. A Comparative Study on Proactive and Passive Detection of Deepfake Speech
    2025/06/17 by Chia-Hua Wu, Wanying Ge, Wu, Chia-Hua +9 · 3 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing
  48. Toward Interpretable Speech Deepfake Detection using Artifact-Specific Experts and Calibrated Detection Scores
    2026/07/23 by Viola Negroni, Xin Wang, Wanying Ge +3
    #cs.SD