1994/08/22 by Fernando Pereira, Pereira, Fernando, Naftali Tishby +3
Computer Science · #Bayesian Methods and Mixture Models #Computation and Language (cs.CL) #Data Mining Algorithms and Applications #FOS: Computer and information sciences #Natural Language Processing Techniques
paper · pdf · doi:10.48550/arxiv.cmp-lg/9408011
openalex publication_date 1994/08/22 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
We describe and experimentally evaluate a method for automatically clustering words according to their distribution in particular syntactic contexts. Deterministic annealing is used to find lowest distortion sets of clusters. As the annealing parameter increases, existing clusters become unstable and subdivide, yielding a hierarchical ``soft'' clustering of the data. Clusters are used as the basis for class models of word coocurrence, and the models evaluated with respect to held-out test data.