2004/05/10 by G. E. Miram, Miram, G. E., V. K. Petrov +1
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #I.2.7 #cs.CL
paper · pdf · doi:10.48550/arxiv.cs/0405037
arxiv created 2004/05/10 · arxiv updated 2009/12/01
A probabilistic model for computer-based generation of a machine translation system on the basis of English-Russian parallel text corpora is suggested. The model is trained using parallel text corpora with pre-aligned source and target sentences. The training of the model results in a bilingual dictionary of words and "word blocks" with relevant translation probability.