2021/12/02 by Chenxiao Liu, Liu, Chenxiao, Guanzhi Deng +7
Computer Science · #AI in Service Interactions #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
paper · pdf · doi:10.48550/arxiv.2112.01616
openalex publication_date 2021/12/02 · openalex created_date 2022/05/05 · openalex updated_date 2026/07/28
One challenge for evaluating current sequence- or dialogue-level chatbots, such as Empathetic Open-domain Conversation Models, is to determine whether the chatbot performs in an emotionally consistent way. The most recent work only evaluates on the aspects of context coherence, language fluency, response diversity, or logical self-consistency between dialogues. This work proposes training an evaluator to determine the emotional consistency of chatbots.