Junichi Yamagishi
- ASVspoof 2019: A large-scale public database of synthesized, converted and replayed speech
2020/05/20 by Xin Wang, Junichi Yamagishi, Massimiliano Todisco +43 · 44 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing
- Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmentation
2022/02/24 by Hemlata Tak, Tak, Hemlata, Massimiliano Todisco +9 · 30 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale
2024/08/16 by Xin Wang, Héctor Delgado, Wang, Xin +23 · 42 citations
Computer Science · #Hate Speech and Cyberbullying Detection
- OpenForensics: Large-Scale Challenging Dataset For Multi-Face Forgery Detection And Segmentation In-The-Wild
2021/07/30 by Trung-Nghia Le, Le, Trung-Nghia, Huy H. Nguyen +5 · 1 voice · 5 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Digital Media Forensic Detection #FOS: Computer and information sciences #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #cs.CV
- Capsule-Forensics: Using Capsule Networks to Detect Forged Images and Videos
2018/10/26 by Huy H. Nguyen, Junichi Yamagishi, Nguyen, Huy H. +3 · 13 citations
Computer Science · #Digital Media Forensic Detection #Advanced Steganography and Watermarking Techniques #Generative Adversarial Networks and Image Synthesis
- Investigating self-supervised front ends for speech spoofing countermeasures
2021/11/15 by Xin Wang, Wang, Xin, Junichi Yamagishi +1 · 17 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
- An Overview of Voice Conversion and its Challenges: From Statistical Modeling to Deep Learning
2020/08/09 by Berrak Şişman, Junichi Yamagishi, Sisman, Berrak +5 · 13 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Generalization Ability of MOS Prediction Networks
2021/10/06 by Erica Cooper, Wen-Chin Huang, Cooper, Erica +5 · 15 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #Topic Modeling #electronic engineering #information engineering
- How do Voices from Past Speech Synthesis Challenges Compare Today?
2021/05/05 by Erica Cooper, Junichi Yamagishi, Cooper, Erica +1 · 11 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Topic Modeling #electronic engineering #information engineering
- Speaker Anonymization Using X-vector and Neural Waveform Models
2019/05/30 by Fuming Fang, Fang, Fuming, Xin Wang +11 · 9 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Voice Conversion Challenge 2020: Intra-lingual semi-parallel and cross-lingual voice conversion
2020/08/28 by Yi Zhao, Zhao, Yi, Wen-Chin Huang +13 · 7 citations
Computer Science · Medicine · #Speech Recognition and Synthesis #Topic Modeling #Voice and Speech Disorders
- An Initial Investigation for Detecting Partially Spoofed Audio
2021/04/06 by Lin Zhang, Zhang, Lin, Xin Wang +9 · 6 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Neural source-filter waveform models for statistical parametric speech synthesis
2019/04/27 by Xin Wang, Wang, Xin, Shinji Takaki +3 · 5 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Natural Language Processing Techniques
- Spoofed training data for speech spoofing countermeasure can be efficiently created using neural vocoders
2022/10/19 by Xin Wang, Junichi Yamagishi, Wang, Xin +1 · 7 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Can large-scale vocoded spoofed data improve speech spoofing countermeasure with a self-supervised front end?
2023/09/12 by Xin Wang, Junichi Yamagishi, Wang, Xin +1 · 8 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Zero-Shot Multi-Speaker Text-To-Speech with State-of-the-art Neural Speaker Embeddings
2019/10/23 by Erica Cooper, Cooper, Erica, Cheng-I Lai +11 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
- Investigating Range-Equalizing Bias in Mean Opinion Score Ratings of Synthesized Speech
2023/05/17 by Erica Cooper, Cooper, Erica, Junichi Yamagishi +1 · 6 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing
- ASVspoof 2021: Automatic Speaker Verification Spoofing and Countermeasures Challenge Evaluation Plan
2021/09/01 by Héctor Delgado, Nicholas Evans, Delgado, Héctor +19 · 4 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
- The VoiceMOS Challenge 2024: Beyond Speech Quality Prediction
2024/09/11 by Wen-Chin Huang, Szu‐Wei Fu, Huang, Wen-Chin +13 · 12 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Use of a Capsule Network to Detect Fake Images and Videos
2019/10/28 by Huy H. Nguyen, Nguyen, Huy H., Junichi Yamagishi +3 · 4 citations
Computer Science · #Advanced Steganography and Watermarking Techniques #Computer Vision and Pattern Recognition (cs.CV) #Digital Media Forensic Detection #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis
- The VoicePrivacy 2022 Challenge: Progress and Perspectives in Voice Anonymisation
2024/07/16 by Michele Panariello, Panariello, Michele, Natalia Tomashenko +17 · 9 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #electronic engineering #information engineering
- Language-Independent Speaker Anonymization Approach using Self-Supervised Pre-Trained Models
2022/02/26 by Xiaoxiao Miao, Xin Wang, Miao, Xiaoxiao +7 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- The VoicePrivacy 2022 Challenge Evaluation Plan
2022/03/23 by Natalia Tomashenko, Xin Wang, Tomashenko, Natalia +17 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
- Spoofing and countermeasures for speaker verification: A survey
2015/02/01 by Zhizheng Wu, Nicholas Evans, Tomi Kinnunen +3 · 2 citations
- ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech
2025/02/13 by Xin Wang, Wang, Xin, Héctor Delgado +54 · 9 citations
Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Hate Speech and Cyberbullying Detection #Misinformation and Its Impacts #electronic engineering #information engineering
- Analyzing Language-Independent Speaker Anonymization Framework under Unseen Conditions
2022/03/28 by Xiaoxiao Miao, Miao, Xiaoxiao, Xin Wang +7 · 3 citations
Computer Science · #FOS: Computer and information sciences #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis
- Range-Based Equal Error Rate for Spoof Localization
2023/05/28 by Lin Zhang, Zhang, Lin, X.J. Wang +7 · 2 citations
Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Gait Recognition and Analysis #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Post-training for Deepfake Speech Detection
2025/06/26 by Wanying Ge, Ge, Wanying, Xin Wang +5 · 8 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- An Initial Investigation of Language Adaptation for TTS Systems under Low-resource Scenarios
2024/06/13 by Gong Cheng, Erica Cooper, Gong, Cheng +21 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimodal Machine Learning Applications #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- GELP: GAN-Excited Linear Prediction for Speech Synthesis from\n Mel-spectrogram
2019/04/08 by Lauri Juvela, Juvela, Lauri, Bajibabu Bollepalli +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- ASVspoof 5: Design, collection and validation of resources for spoofing, deepfake, and adversarial attack detection using crowdsourced speech
2025/05/28 by Xin Wang, Héctor Delgado, Hemlata Tak +26 · 1 voice · 6 citations
Computer Science · #Speech Recognition and Synthesis #Network Security and Intrusion Detection #Hate Speech and Cyberbullying Detection
- QualiSpeech: A Speech Quality Assessment Dataset with Natural Language Reasoning and Descriptions
2025/03/26 by Siyin Wang, Wang, Siyin, Wenyi Yu +16 · 6 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis
- Hiding speaker's sex in speech using zero-evidence speaker representation in an analysis/synthesis pipeline
2022/11/29 by Paul-Gauthier Noé, Xiaoxiao Miao, Noé, Paul-Gauthier +9 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Predictions of Subjective Ratings and Spoofing Assessments of Voice Conversion Challenge 2020 Submissions
2020/09/08 by Rohan Kumar Das, Tomi Kinnunen, Das, Rohan Kumar +13 · 1 citation
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
- Multi-Metric Optimization using Generative Adversarial Networks for Near-End Speech Intelligibility Enhancement
2021/04/17 by Haoyu Li, Junichi Yamagishi, Li, Haoyu +1 · 1 citation
Computer Science · Neuroscience · Engineering · #Speech and Audio Processing #Hearing Loss and Rehabilitation #Acoustic Wave Phenomena Research
- A Multi-Level Attention Model for Evidence-Based Fact Checking
2021/06/02 by Canasai Kruengkrai, Kruengkrai, Canasai, Junichi Yamagishi +3 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Software Engineering Research #Topic Modeling
- Use of speaker recognition approaches for learning and evaluating embedding representations of musical instrument sounds
2021/07/24 by Xuan Shi, Shi, Xuan, Erica Cooper +3 · 1 citation
Computer Science · #Music and Audio Processing #Speech and Audio Processing #Music Technology and Sound Studies
- A Practical Guide to Logical Access Voice Presentation Attack Detection
2022/01/10 by Xin Wang, Wang, Xin, Junichi Yamagishi +1 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and Audio Processing
- Towards An Integrated Approach for Expressive Piano Performance Synthesis from Music Scores
2025/01/17 by Jingjing Tang, Tang, Jingjing, Erica Cooper +7 · 1 citation
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception #Sound (cs.SD) #electronic engineering #information engineering
- Libri2Vox Dataset: Target Speaker Extraction with Diverse Speaker Conditions and Synthetic Data
2024/12/17 by Yun Liu, Xuechen Liu, Liu, Yun +5 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Good practices for evaluation of synthesized speech
2025/03/05 by Erica Cooper, Cooper, Erica, Sébastien Le Maguer +5 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and dialogue systems
- Spoofing-Aware Speaker Verification Robust Against Domain and Channel Mismatches
2024/09/10 by Chang Zeng, Xiaoxiao Miao, Zeng, Chang +7 · 1 citation
Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Disentangling the Prosody and Semantic Information with Pre-trained Model for In-Context Learning based Zero-Shot Voice Conversion
2024/09/08 by Zhengyang Chen, Shuai Wang, Chen, Zhengyang +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Uncertainty as a Predictor: Leveraging Self-Supervised Learning for Zero-Shot MOS Prediction
2023/12/25 by Aditya Ravuri, Erica Cooper, Ravuri, Aditya +3 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model
2024/12/03 by Mingyi Shi, Dafei Qin, Shi, Mingyi +11 · 1 citation
Computer Science · #Speech and dialogue systems #Topic Modeling #Speech Recognition and Synthesis
- Explaining Speaker and Spoof Embeddings via Probing
2024/12/24 by Xuechen Liu, Liu, Xuechen, Junichi Yamagishi +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- A Comparative Study on Proactive and Passive Detection of Deepfake Speech
2025/06/17 by Chia-Hua Wu, Wanying Ge, Wu, Chia-Hua +9 · 3 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing
- Toward Interpretable Speech Deepfake Detection using Artifact-Specific Experts and Calibrated Detection Scores
2026/07/23 by Viola Negroni, Xin Wang, Wanying Ge +3
#cs.SD