vix.ing · top · new · best · stats · spec

Noise Robust IOA/CAS Speech Separation and Recognition System For The Third 'CHIME' Challenge

2015/09/21 by Xiaofei Wang, Chao Wu, Wang, Xiaofei +13
Computer Science · #Advanced Data Compression Techniques #Computation and Language (cs.CL) #FOS: Computer and information sciences #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing

paper · pdf · doi:10.48550/arxiv.1509.06103

openalex publication_date 2015/09/21 · openalex created_date 2016/06/24 · openalex updated_date 2026/07/28

Abstract

This paper presents the contribution to the third 'CHiME' speech separation and recognition challenge including both front-end signal processing and back-end speech recognition. In the front-end, Multi-channel Wiener filter (MWF) is designed to achieve background noise reduction. Different from traditional MWF, optimized parameter for the tradeoff between noise reduction and target signal distortion is built according to the desired noise reduction level. In the back-end, several techniques are taken advantage to improve the noisy Automatic Speech Recognition (ASR) performance including Deep Neural Network (DNN), Convolutional Neural Network (CNN) and Long short-term memory (LSTM) using medium vocabulary, Lattice rescoring with a big vocabulary language model finite state transducer, and ROVER scheme. Experimental results show the proposed system combining front-end and back-end is effective to improve the ASR performance.

Related