vix.ing · top · new · best · stats · spec

Machine Translation in Indian Languages: Challenges and Resolution

2017/08/26 by Raj Nath Patel, Patel, Raj Nath, Prakash B. Pimpale +3
Computer Science · #Algorithms and Data Compression #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Topic Modeling

paper · pdf · doi:10.48550/arxiv.1708.07950

openalex publication_date 2017/08/26 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

English to Indian language machine translation poses the challenge of structural and morphological divergence. This paper describes English to Indian language statistical machine translation using pre-ordering and suffix separation. The pre-ordering uses rules to transfer the structure of the source sentences prior to training and translation. This syntactic restructuring helps statistical machine translation to tackle the structural divergence and hence better translation quality. The suffix separation is used to tackle the morphological divergence between English and highly agglutinative Indian languages. We demonstrate that the use of pre-ordering and suffix separation helps in improving the quality of English to Indian Language machine translation.

Citations

Related