2025/04/29 by Iwona Christop, Tomasz Kuczyński, Christop, Iwona +3
Computer Science · #Benchmark (surveying) #Cloning (programming) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Identification (biology) #Software #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
paper · pdf · doi:10.48550/arxiv.2504.20581
published in arXiv (Cornell University) (Cornell University)
openalex publication_date 2025/04/29 · openalex created_date 2025/10/10 · openalex updated_date 2026/08/05
We present a novel benchmark for voice cloning text-to-speech models. The benchmark consists of an evaluation protocol, an open-source library for assessing the performance of voice cloning models, and an accompanying leaderboard. The paper discusses design considerations and presents a detailed description of the evaluation procedure. The usage of the software library is explained, along with the organization of results on the leaderboard.