2018/10/10 by Won Ik Cho, Young Ki Moon, Cho, Won Ik +5
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling #cs.CL
paper · pdf · doi:10.48550/arxiv.1810.04631
5 pages and 2 tables, Annotation guideline for Seoul Korean sentences
openalex publication_date 2018/10/10 · arxiv created 2019/07/09 · arxiv updated 2019/07/10 · openalex created_date 2019/07/23 · openalex updated_date 2026/07/28
Intention identification is a core issue in dialog management. However, due to the non-canonicality of the spoken language, it is difficult to extract the content automatically from the conversation-style utterances. This is much more challenging for languages like Korean and Japanese since the agglutination between morphemes make it difficult for the machines to parse the sentence and understand the intention. To suggest a guideline for this problem, and to merge the issue flexibly with the neural paraphrasing systems introduced recently, we propose a structured annotation scheme for Korean question/commands and the resulting corpus which are widely applicable to the field of argument mining. The scheme and dataset are expected to help machines understand the intention of natural language and grasp the core meaning of conversation-style instructions.