2019/05/14 by Loïc Vial, Vial, Loïc, Benjamin Lecouteux +3 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
paper · pdf · doi:10.48550/arxiv.1905.05677
openalex publication_date 2019/05/14 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
In this article, we tackle the issue of the limited quantity of manually\nsense annotated corpora for the task of word sense disambiguation, by\nexploiting the semantic relationships between senses such as synonymy,\nhypernymy and hyponymy, in order to compress the sense vocabulary of Princeton\nWordNet, and thus reduce the number of different sense tags that must be\nobserved to disambiguate all words of the lexical database. We propose two\ndifferent methods that greatly reduces the size of neural WSD models, with the\nbenefit of improving their coverage without additional training data, and\nwithout impacting their precision. In addition to our method, we present a WSD\nsystem which relies on pre-trained BERT word vectors in order to achieve\nresults that significantly outperform the state of the art on all WSD\nevaluation tasks.\n