A Versatile Adaptive Curriculum Learning Framework for Task-oriented Dialogue Policy Learning

Zhao, Yangyang; Qin, Hua; Zhenyu, Wang; Zhu, Changxi; Wang, Shihan

doi:https://doi.org/10.18653/v1/2022.findings-naacl.54

A Versatile Adaptive Curriculum Learning Framework for Task-oriented Dialogue Policy Learning

Files

2022.findings_naacl.54.pdf (7.92 MB)

Publication date

2022-07-01

Authors

Zhao, Yangyang

Qin, Hua

Zhenyu, Wang

Zhu, Changxi

Wang, Shihan

DOI

https://doi.org/10.18653/v1/2022.findings-naacl.54

Document Type

Part of book

Metadata

Show full item record

Collections

Utrecht University Repository

License

cc_by

Abstract

Training a deep reinforcement learning-based dialogue policy with brute-force random sampling is costly. A new training paradigm was proposed to improve learning performance and efficiency by combining curriculum learning. However, attempts in the field of dialogue policy are very limited due to the lack of reliable evaluation of difficulty scores of dialogue tasks and the high sensitivity to the mode of progression through dialogue tasks. In this paper, we present a novel versatile adaptive curriculum learning (VACL) framework, which presents a substantial step toward applying automatic curriculum learning on dialogue policy tasks. It supports evaluating the difficulty of dialogue tasks only using the learning experiences of dialogue policy and skip-level selection according to their learning needs to maximize the learning efficiency. Moreover, an attractive feature of VACL is the construction of a generic, elastic global curriculum while training a good dialogue policy that could guide different dialogue policy learning without extra effort on re-training. The superiority and versatility of VACL are validated on three public dialogue datasets.

Citation

Zhao, Y, Qin, H, Zhenyu, W, Zhu, C & Wang, S 2022, A Versatile Adaptive Curriculum Learning Framework for Task-oriented Dialogue Policy Learning. in Findings of the Association for Computational Linguistics: NAACL 2022. Association for Computational Linguistics, pp. 711-723. https://doi.org/10.18653/v1/2022.findings-naacl.54

URI

https://dspace.library.uu.nl/handle/1874/423170

A Versatile Adaptive Curriculum Learning Framework for Task-oriented Dialogue Policy Learning

Files

Publication date

Authors

Editors

Advisors

Supervisors

DOI

Document Type

Metadata

Collections

License

Abstract

Keywords

Citation

URI