vix.ing · top · new · best · stats · spec

Automatic acquisition of hyponyms from large text corpora

1992/01/01 by Marti A. Hearst · 7 citations
Computer Science · #Advanced Text Analysis Techniques #Artificial intelligence #Computer science #Data mining #Information retrieval #Knowledge acquisition #Lexico #Lexicon #Natural Language Processing Techniques #Natural language processing #Programming language #Range (aeronautics) #Relation (database) #Set (abstract data type) #Topic Modeling

paper · pdf · doi:10.3115/992133.992154

openalex publication_date 1992/01/01 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/29

Abstract

We describe a method for the automatic acquisition of the hyponymy lexical relation from unrestricted text. Two goals motivate the approach: (i) avoidance of the need for pre-encoded knowledge and (ii) applicability across a wide range of text. We identify a set of lexico-syntactic patterns that are easily recognizable, that occur frequently and across text genre boundaries, and that indisputably indicate the lexical relation of interest. We describe a method for discovering these patterns and suggest that other lexical relations will also be acquirable in this way. A subset of the acquisition algorithm is implemented and the results are used to augment and critique the structure of a large hand-built thesaurus. Extensions and applications to areas such as information retrieval are suggested.

Cited by