ASVspoof 2019: A large-scale public database of synthesized, converted and replayed speech
2020/05/20 by Xin Wang, Junichi Yamagishi, Massimiliano Todisco +43 · 75 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing
paper · doi:10.1016/j.csl.2020.101114
openalex publication_date 2020/05/20 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/29
Citations
Cited by
- A Data-Centric Approach to Generalizable Speech Deepfake Detection
- BUT Systems for WildSpoof Challenge: SASV in the Wild
- The Affective Bridge: Preserving Speech Representations while Enhancing Deepfake Detection vian emotional Constraints
- DFALLM: Achieving Generalizable Multitask Deepfake Detection by Optimizing Audio LLM Components
- Physics-Guided Deepfake Detection for Voice Authentication Systems
- Curved Worlds, Clear Boundaries: Generalizing Speech Deepfake Detection using Hyperbolic and Spherical Geometry Spaces
- Multi-modal Deepfake Detection and Localization with FPN-Transformer
- NaturalVoices: A Large-Scale, Spontaneous and Emotional Podcast Dataset for Voice Conversion
- A Parameter-Efficient Multi-Scale Convolutional Adapter for Synthetic Speech Detection
- DeepfakeBench-MM: A Comprehensive Benchmark for Multimodal Deepfake Detection
- Can Current Detectors Catch Face-to-Voice Deepfake Attacks?
- EchoFake: A Replay-Aware Dataset for Practical Speech Deepfake Detection
- SpeechLLM-as-Judges: Towards General and Interpretable Speech Quality Evaluation
- Benchmarking Fake Voice Detection in the Fake Voice Generation Arms Race
- UR Channel-Robust Synthetic Speech Detection System for ASVspoof 2021
- On Deepfake Voice Detection -- It's All in the Presentation
- Advancing Zero-Shot Open-Set Speech Deepfake Source Tracing
- Generalizable Speech Deepfake Detection via Information Bottleneck Enhanced Adversarial Alignment
- HuLA: Prosody-Aware Anti-Spoofing with Multi-Task Learning for Expressive and Emotional Synthetic Speech
- The Impact of Audio Watermarking on Audio Anti-Spoofing Countermeasures
- SEA-Spoof: Bridging The Gap in Multilingual Audio Deepfake Detection for South-East Asian
- Teffic-Audio: Tell Fact from Fiction
- End-to-End Spectro-Temporal Graph Attention Networks for Speaker Verification Anti-Spoofing and Speech Deepfake Detection
- Attention-based Mixture of Experts for Robust Speech Deepfake Detection
- Are Multimodal Foundation Models All That Is Needed for Emofake Detection?
- How Does Instrumental Music Help SingFake Detection?
- Discrete optimal transport is a strong audio adversarial attack
- Bona fide Cross Testing Reveals Weak Spot in Audio Deepfake Detection Systems
- MoLEx: Mixture of LoRA Experts in Speech Self-Supervised Models for Audio Deepfake Detection
- AUDETER: A Large-scale Dataset for Deepfake Audio Detection in Open Worlds
- Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection
- Multi-level SSL Feature Gating for Audio Deepfake Detection
- Generalizable Audio Spoofing Detection using Non-Semantic Representations
- Multilingual Dataset Integration Strategies for Robust Audio Deepfake Detection: A SAFE Challenge System
- FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset
- Rapidly Adapting to New Voice Spoofing: Few-Shot Detection of Synthesized Speech Under Distribution Shifts
- Practical Attacks on Voice Spoofing Countermeasures
- Towards Reliable Audio Deepfake Attribution and Model Recognition: A Multi-Level Autoencoder-Based Framework
- Generalizable Audio Deepfake Detection via Hierarchical Structure Learning and Feature Whitening in Poincaré sphere
- Fusion of Modulation Spectrogram and SSL with Multi-head Attention for Fake Speech Detection
- Unraveling Hidden Representations: A Multi-Modal Layer Analysis for Better Synthetic Content Forensics
- Raw Differentiable Architecture Search for Speech Deepfake and Spoofing Detection
- Two Views, One Truth: Spectral and Self-Supervised Features Fusion for Robust Speech Deepfake Detection
- WaveVerify: A Novel Audio Watermarking Framework for Media Authentication and Combatting Deepfakes
- RW-Resnet: A Novel Speech Anti-Spoofing Model Using Raw Waveform
- Towards Scalable AASIST: Refining Graph Attention for Speech Deepfake Detection
- Collecting, Curating, and Annotating Good Quality Speech deepfake dataset for Famous Figures: Process and Challenges
- Post-training for Deepfake Speech Detection
- End-to-end anti-spoofing with RawNet2
- Manipulated Regions Localization For Partially Deepfake Audio: A Survey
- A Comparative Study on Proactive and Passive Detection of Deepfake Speech
- From Sharpness to Better Generalization for Speech Deepfake Detection
- A Comparative Study on Recent Neural Spoofing Countermeasures for Synthetic Speech Detection
- REIMU: Efficient Heterogeneous Hierarchical Reasoning for SSL-Based Speech Deepfake Detection
- Discrete Optimal Transport and Voice Conversion
- A Data-Driven Diffusion-based Approach for Audio Deepfake Explanations
- PartialEdit: Identifying Partial Deepfakes in the Era of Neural Speech Editing
- Unveiling Audio Deepfake Origins: A Deep Metric learning And Conformer Network Approach With Ensemble Fusion
- Source Tracing of Synthetic Speech Systems Through Paralinguistic Pre-Trained Representations
- Quality Assessment of Noisy and Enhanced Speech with Limited Data: UWB-NTIS System for VoiceMOS 2024
- XMAD-Bench: Cross-Domain Multilingual Audio Deepfake Benchmark
- Can Emotion Fool Anti-spoofing?
- Tell me Habibi, is it Real or Fake?
- ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech
- STOPA: A Database of Systematic VariaTion Of DeePfake Audio for Open-Set Source Tracing and Attribution
- Partially-Connected Differentiable Architecture Search for Deepfake and Spoofing Detection
- ATMM-SAGA: Alternating Training for Multi-Module with Score-Aware Gated Attention SASV system
- ASVspoof2019 vs. ASVspoof5: Assessment and Comparison
- BiCrossMamba-ST: Speech Deepfake Detection with Bidirectional Mamba Spectro-Temporal Cross-Attention
- Deepfake: definitions, performance metrics and standards, datasets, and a meta-review
- ALLM4ADD: Unlocking the Capabilities of Audio Large Language Models for Audio Deepfake Detection
- Beyond Identity: A Generalizable Approach for Deepfake Audio Detection
- Synth2Aug: Cross-domain speaker recognition with TTS synthesized speech
- ASVspoof 2019: A large-scale public database of synthesized, converted and replayed speech
- Tandem Assessment of Spoofing Countermeasures and Automatic Speaker Verification: Fundamentals
- Robustness in deepfake speech detection: A survey of failure mechanisms including an experimental case study
- Nes2Net: A Lightweight Nested Architecture for Foundation Model Driven Speech Anti-spoofing