2018/11/13 by Eleanor Chodroff, Chodroff, Eleanor · 6 citations
Computer Science · Psychology · #Computation and Language (cs.CL) #Computer science #FOS: Computer and information sciences #Linguistics #Natural language processing #Philosophy #Phonetics #Psychology #Second Language Acquisition and Learning #cs.CL
paper · pdf · doi:10.48550/arxiv.1811.05553
published in arXiv (Cornell University) (Cornell University)
arxiv created 2018/11/13 · openalex publication_date 2018/11/13 · arxiv updated 2018/11/15 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
Corpus phonetics has become an increasingly popular method of research in linguistic analysis. With advances in speech technology and computational power, large scale processing of speech data has become a viable technique. This tutorial introduces the speech scientist and engineer to various automatic speech processing tools. These include acoustic model creation and forced alignment using the Kaldi Automatic Speech Recognition Toolkit (Povey et al., 2011), forced alignment using FAVE-align (Rosenfelder et al., 2014), the Montreal Forced Aligner (McAuliffe et al., 2017), and the Penn Phonetics Lab Forced Aligner (Yuan & Liberman, 2008), as well as stop consonant burst alignment using AutoVOT (Keshet et al., 2014). The tutorial provides a general overview of each program, step-by-step instructions for running the program, as well as several tips and tricks.