1997/04/15 by Marilyn A. Walker, Diane J. Litman, Candace A. Kamm +1 · 2 citations
Computer Science · #cmp-lg #cs.CL
published as Proceedings of the 35th Annual Meeting of the Association for Computational Linguistics · 10 pages, uses aclap, psfig, lingmacros, times
arxiv created 1997/04/15 · arxiv updated 2009/11/30
This paper presents PARADISE (PARAdigm for DIalogue System Evaluation), a general framework for evaluating spoken dialogue agents. The framework decouples task requirements from an agent's dialogue behaviors, supports comparisons among dialogue strategies, enables the calculation of performance over subdialogues and whole dialogues, specifies the relative contribution of various factors to performance, and makes it possible to compare agents performing different tasks by normalizing for task complexity.