Deep Reinforcement Learning For Modeling Chit-Chat Dialog With Discrete Attributes

Sankar, Chinnadhurai; Ravi, Sujith

Computer Science > Machine Learning

arXiv:1907.02848 (cs)

[Submitted on 5 Jul 2019 (v1), last revised 14 Sep 2019 (this version, v2)]

Title:Deep Reinforcement Learning For Modeling Chit-Chat Dialog With Discrete Attributes

Authors:Chinnadhurai Sankar, Sujith Ravi

View PDF

Abstract:Open domain dialog systems face the challenge of being repetitive and producing generic responses. In this paper, we demonstrate that by conditioning the response generation on interpretable discrete dialog attributes and composed attributes, it helps improve the model perplexity and results in diverse and interesting non-redundant responses. We propose to formulate the dialog attribute prediction as a reinforcement learning (RL) problem and use policy gradients methods to optimize utterance generation using long-term rewards. Unlike existing RL approaches which formulate the token prediction as a policy, our method reduces the complexity of the policy optimization by limiting the action space to dialog attributes, thereby making the policy optimization more practical and sample efficient. We demonstrate this with experimental and human evaluations.

Comments:	SIGDIAL 2019 - BEST PAPER AWARD
Subjects:	Machine Learning (cs.LG); Computation and Language (cs.CL)
Cite as:	arXiv:1907.02848 [cs.LG]
	(or arXiv:1907.02848v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1907.02848

Submission history

From: Chinnadhurai Sankar [view email]
[v1] Fri, 5 Jul 2019 14:19:15 UTC (177 KB)
[v2] Sat, 14 Sep 2019 17:03:04 UTC (177 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2019-07

Change to browse by:

cs
cs.CL

References & Citations

DBLP - CS Bibliography

listing | bibtex

Chinnadhurai Sankar
Sujith Ravi

export BibTeX citation

Computer Science > Machine Learning

Title:Deep Reinforcement Learning For Modeling Chit-Chat Dialog With Discrete Attributes

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Deep Reinforcement Learning For Modeling Chit-Chat Dialog With Discrete Attributes

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators