Survey on Evaluation Methods for Dialogue Systems

Deriu, Jan, Rodrigo, Alvaro, Otegi, Arantxa, Echegoyen, Guillermo, Rosset, Sophie, Agirre, Eneko, Cieliebak, Mark

May-10-2019–arXiv.org Artificial Intelligence

In this paper we survey the methods and concepts developed for the evaluation of dialogue systems. Evaluation is a crucial part during the development process. Often, dialogue systems are evaluated by means of human evaluations and questionnaires. However, this tends to be very cost and time intensive. Thus, much work has been put into finding methods, which allow to reduce the involvement of human labour. In this survey, we present the main concepts and methods. For this, we differentiate between the various classes of dialogue systems (task-oriented dialogue systems, conversational dialogue systems, and question-answering dialogue systems). We cover each class by introducing the main technologies developed for the dialogue systems and then by presenting the evaluation methods regarding this class.

machine learning, question answering, reinforcement learning, (24 more...)

arXiv.org Artificial Intelligence

May-10-2019

arXiv.org PDF

Add feedback

Country:
- Europe (1.00)
- North America > United States
  - California (0.28)

Genre:
- Research Report > New Finding (1.00)
- Overview (1.00)

Industry:
- Health & Medicine (0.92)
- Consumer Products & Services > Restaurants (0.46)
- Media > News (0.46)
- Education > Educational Setting (0.46)

Technology:
- Information Technology
  - Data Science > Data Mining (1.00)
  - Communications > Social Media
    - Crowdsourcing (0.68)
  - Artificial Intelligence
    - Speech > Speech Recognition (1.00)
    - Representation & Reasoning
      - Personal Assistant Systems (1.00)
      - Rule-Based Reasoning (0.67)
      - Expert Systems (0.67)
    - Natural Language
      - Text Processing (1.00)
      - Discourse & Dialogue (1.00)
      - Chatbot (1.00)
      - Machine Translation (0.93)
      - Question Answering (0.87)
    - Machine Learning
      - Statistical Learning (1.00)
      - Neural Networks > Deep Learning (1.00)
      - Reinforcement Learning (0.68)
      - Learning Graphical Models > Undirected Networks
        Markov Models (0.68)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found