WeaSuLπ: Weakly Supervised Dialogue Policy Learning: Reward Estimation for Multi-turn Dialogue

被引:0
|
作者
Khandelwal, Anant [1 ]
机构
[1] Amazon, India Machine Learning, Bangalore, Karnataka, India
关键词
MODEL;
D O I
暂无
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
An intelligent dialogue system in a multi-turn setting should not only generate the responses which are of good quality, but it should also generate the responses which can lead to long-term success of the dialogue. Although, the current approaches improved the response quality, but they over-look the training signals present in the dialogue data. We can leverage these signals to generate the weakly supervised training data for learning dialog policy and reward estimator, and make the policy take actions (generates responses) which can foresee the future direction for a successful (rewarding) conversation. We simulate the dialogue between an agent and a user (modelled similar to an agent with supervised learning objective) to interact with each other. The agent uses dynamic blocking to generate ranked diverse responses and explorationexploitation to select among the Top-K responses. Each simulated state-action pair is evaluated (works as a weak annotation) with three quality modules: Semantic Relevant, Semantic Coherence and Consistent Flow. Empirical studies with two benchmarks indicate that our model can significantly out-perform the response quality and lead to a successful conversation on both automatic evaluation and human judgment.(1)
引用
收藏
页码:69 / 80
页数:12
相关论文
共 50 条
  • [41] EAGS: An extracting auxiliary knowledge graph model in multi-turn dialogue generation
    Ning, Bo
    Zhao, Deji
    Liu, Xinyi
    Li, Guanyu
    WORLD WIDE WEB-INTERNET AND WEB INFORMATION SYSTEMS, 2023, 26 (04): : 1545 - 1566
  • [42] HiBERT: Detecting the illogical patterns with hierarchical BERT for multi-turn dialogue reasoning
    Wang, Xu
    Zhang, Hainan
    Zhao, Shuai
    Chen, Hongshen
    Cheng, Bo
    Ding, Zhuoye
    Xu, Sulong
    Yan, Weipeng
    Lan, Yanyan
    NEUROCOMPUTING, 2023, 524 : 167 - 177
  • [43] Open-domain Multi-turn Dialogue Model Based on Knowledge Enhancement
    Xu F.
    Xu J.-M.
    Ma Y.
    Wang M.-W.
    Zhou G.-D.
    Ruan Jian Xue Bao/Journal of Software, 2024, 35 (02): : 758 - 772
  • [44] History-Adaption Knowledge Incorporation Mechanism for Multi-Turn Dialogue System
    Sun, Yajing
    Hu, Yue
    Xing, Luxi
    Yu, Jing
    Xie, Yuqiang
    THIRTY-FOURTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE, THE THIRTY-SECOND INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE CONFERENCE AND THE TENTH AAAI SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE, 2020, 34 : 8944 - 8951
  • [45] CHATEDIT: Towards Multi-turn Interactive Facial Image Editing via Dialogue
    Cui, Xing
    Li, Zekun
    Li, Peipei
    Hu, Yibo
    Shi, Hailin
    Cao, Chunshui
    He, Zhaofeng
    2023 CONFERENCE ON EMPIRICAL METHODS IN NATURAL LANGUAGE PROCESSING (EMNLP 2023), 2023, : 14567 - 14583
  • [46] HSAN: A HIERARCHICAL SELF-ATTENTION NETWORK FOR MULTI-TURN DIALOGUE GENERATION
    Kong, Yawei
    Zhang, Lu
    Ma, Can
    Cao, Cong
    2021 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP 2021), 2021, : 7433 - 7437
  • [47] Decoupled Dialogue Modeling and Semantic Parsing for Multi-Turn Text-to-SQL
    Chen, Zhi
    Chen, Lu
    Li, Hanqi
    Cao, Ruisheng
    Ma, Da
    Wu, Mengyue
    Yu, Kai
    FINDINGS OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS, ACL-IJCNLP 2021, 2021, : 3063 - 3074
  • [48] A Practical Dialogue-Act-Driven Conversation Model for Multi-Turn Response Selection
    Kumar, Harshit
    Agarwal, Arvind
    Joshi, Sachindra
    2019 CONFERENCE ON EMPIRICAL METHODS IN NATURAL LANGUAGE PROCESSING AND THE 9TH INTERNATIONAL JOINT CONFERENCE ON NATURAL LANGUAGE PROCESSING (EMNLP-IJCNLP 2019): PROCEEDINGS OF THE CONFERENCE, 2019, : 1980 - 1989
  • [49] NaturalConv: A Chinese Dialogue Dataset Towards Multi-turn Topic-driven Conversation
    Wang, Xiaoyang
    Li, Chen
    Zhao, Jianqiao
    Yu, Dong
    THIRTY-FIFTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE, THIRTY-THIRD CONFERENCE ON INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE AND THE ELEVENTH SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE, 2021, 35 : 14006 - 14014
  • [50] Topic-level knowledge sub-graphs for multi-turn dialogue generation
    Li, Jing
    Huang, Qingbao
    Cai, Yi
    Liu, Yongkang
    Fu, Mingyi
    Li, Qing
    KNOWLEDGE-BASED SYSTEMS, 2021, 234