preprint Open access

A Diversity-Promoting Objective Function for Neural Conversation Models

  • arXiv (Cornell University)
  • Cornell University
Research footprint

At a glance

Citations
258
References
36
Comments
0
Paper overview

Öz

Sequence-to-sequence neural network models for generation of conversational responses tend to generate safe, commonplace responses (e.g., "I don't know") regardless of the input. We suggest that the traditional objective function, i.e., the likelihood of output (response) given input (message) is unsuited to response generation tasks. Instead we propose using Maximum Mutual Information (MMI) as the objective function in neural models. Experimental results demonstrate that the proposed MMI models produce more diverse, interesting, and appropriate responses, yielding substantive gains in BLEU scores on two conversational datasets and in human evaluations.

Record transparency

Publication details

DOI
10.48550/arxiv.1510.03055
OpenAlex
W1958706068
Document type
preprint
Language
EN
Source
arXiv (Cornell University)
Last metadata update
Community

Comments

Oturum Açın to join the discussion.

  1. No comments yet. Start the discussion.