SUMBT+LaRL: Effective Multi-domain End-to-end Neural Task-oriented Dialog System

Lee, Hwaran; Jo, Seokhwan; Kim, HyungJun; Jung, Sangkeun; Kim, Tae-Yoon

doi:10.1109/ACCESS.2021.3105461

Computer Science > Computation and Language

arXiv:2009.10447 (cs)

[Submitted on 22 Sep 2020 (v1), last revised 26 Aug 2021 (this version, v3)]

Title:SUMBT+LaRL: Effective Multi-domain End-to-end Neural Task-oriented Dialog System

Authors:Hwaran Lee, Seokhwan Jo, HyungJun Kim, Sangkeun Jung, Tae-Yoon Kim

View PDF

Abstract:The recent advent of neural approaches for developing each dialog component in task-oriented dialog systems has remarkably improved, yet optimizing the overall system performance remains a challenge. Besides, previous research on modeling complicated multi-domain goal-oriented dialogs in end-to-end fashion has been limited. In this paper, we present an effective multi-domain end-to-end trainable neural dialog system SUMBT+LaRL that incorporates two previous strong models and facilitates them to be fully differentiable. Specifically, the SUMBT+ estimates user-acts as well as dialog belief states, and the LaRL models latent system action spaces and generates responses given the estimated contexts. We emphasize that the training framework of three steps significantly and stably increase dialog success rates: separately pretraining the SUMBT+ and LaRL, fine-tuning the entire system, and then reinforcement learning of dialog policy. We also introduce new reward criteria of reinforcement learning for dialog policy training. Then, we discuss experimental results depending on the reward criteria and different dialog evaluation methods. Consequently, our model achieved the new state-of-the-art success rate of 85.4% on corpus-based evaluation, and a comparable success rate of 81.40% on simulator-based evaluation provided by the DSTC8 challenge. To our best knowledge, our work is the first comprehensive study of a modularized E2E multi-domain dialog system that learning from each component to the entire dialog policy for task success.

Comments:	14 pages, 5 figures. This paper is accepted for publication in IEEE Access
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2009.10447 [cs.CL]
	(or arXiv:2009.10447v3 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2009.10447
Related DOI:	https://doi.org/10.1109/ACCESS.2021.3105461

Submission history

From: Hwaran Lee [view email]
[v1] Tue, 22 Sep 2020 11:02:21 UTC (1,374 KB)
[v2] Tue, 6 Oct 2020 02:17:27 UTC (1,371 KB)
[v3] Thu, 26 Aug 2021 08:55:20 UTC (2,971 KB)

Computer Science > Computation and Language

Title:SUMBT+LaRL: Effective Multi-domain End-to-end Neural Task-oriented Dialog System

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:SUMBT+LaRL: Effective Multi-domain End-to-end Neural Task-oriented Dialog System

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators