vix.ing · top · new · best · stats · spec

Hybrid language processing in the Spoken Language Translator

1997/01/02 by Manny Rayner, Rayner, Manny, David Carter +1 · 1 citation
Computer Science · #Algorithms and Data Compression #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #cmp-lg #cs.CL

paper · pdf · doi:10.48550/arxiv.cmp-lg/9701002

4 pages, uses icassp97.sty; to appear in ICASSP-97; see http://www.cam.sri.com for related material

arxiv created 1997/01/02 · openalex publication_date 1997/01/02 · arxiv updated 2009/11/30 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

The paper presents an overview of the Spoken Language Translator (SLT) system's hybrid language-processing architecture, focussing on the way in which rule-based and statistical methods are combined to achieve robust and efficient performance within a linguistically motivated framework. In general, we argue that rules are desirable in order to encode domain-independent linguistic constraints and achieve high-quality grammatical output, while corpus-derived statistics are needed if systems are to be efficient and robust; further, that hybrid architectures are superior from the point of view of portability to architectures which only make use of one type of information. We address the topics of ``multi-engine'' strategies for robust translation; robust bottom-up parsing using pruning and grammar specialization; rational development of linguistic rule-sets using balanced domain corpora; and efficient supervised training by interactive disambiguation. All work described is fully implemented in the current version of the SLT-2 system.

Cited by

Related