vix.ing · top · new · best · stats · spec

Leveraging Synthetic Audio Data for End-to-End Low-Resource Speech Translation

2024/06/25 by Yasmin Moslem, Moslem, Yasmin · 7 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering

paper · pdf · doi:10.48550/arxiv.2406.17363

openalex publication_date 2024/06/25 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

This paper describes our system submission to the International Conference on Spoken Language Translation (IWSLT 2024) for Irish-to-English speech translation. We built end-to-end systems based on Whisper, and employed a number of data augmentation techniques, such as speech back-translation and noise augmentation. We investigate the effect of using synthetic audio data and discuss several methods for enriching signal diversity.

Cited by

Related