Question-Answering Pair Matching Based on Question Classification and Ensemble Sentence Embedding

被引:0
|
作者
Jang J.-S. [1 ]
Kwon H.-Y. [2 ]
机构
[1] Department of Computer Science and Engineering, Seoul National University of Science and Technology, Seoul
[2] Department of Industrial Engineering/Graduate School of Data Science, Seoul National University of Science and Technology, Seoul
来源
基金
新加坡国家研究基金会;
关键词
data augmentation; Question-answering; text classification model; text embedding;
D O I
10.32604/csse.2023.035570
中图分类号
学科分类号
摘要
Question-answering (QA) models find answers to a given question. The necessity of automatically finding answers is increasing because it is very important and challenging from the large-scale QA data sets. In this paper, we deal with the QA pair matching approach in QA models, which finds the most relevant question and its recommended answer for a given question. Existing studies for the approach performed on the entire dataset or datasets within a category that the question writer manually specifies. In contrast, we aim to automatically find the category to which the question belongs by employing the text classification model and to find the answer corresponding to the question within the category. Due to the text classification model, we can effectively reduce the search space for finding the answers to a given question. Therefore, the proposed model improves the accuracy of the QA matching model and significantly reduces the model inference time. Furthermore, to improve the performance of finding similar sentences in each category, we present an ensemble embedding model for sentences, improving the performance compared to the individual embedding models. Using real-world QA data sets, we evaluate the performance of the proposed QA matching model. As a result, the accuracy of our final ensemble embedding model based on the text classification model is 81.18%, which outperforms the existing models by 9.81%~14.16% point. Moreover, in terms of the model inference speed, our model is faster than the existing models by 2.61~5.07 times due to the effective reduction of search spaces by the text classification model. © 2023 CRL Publishing. All rights reserved.
引用
收藏
页码:3471 / 3489
页数:18
相关论文
共 50 条
  • [1] CCQA: A Chinese Community Question-answering Tool Based on Sentence Similarity with Word Embedding
    Wang, Fei-xuan
    INTERNATIONAL CONFERENCE ON ARTIFICIAL INTELLIGENCE AND COMPUTER SCIENCE (AICS 2016), 2016, : 422 - 427
  • [2] Sentiment Classification towards Question-Answering with Hierarchical Matching Network
    Shen, Chenlin
    Sun, Changlong
    Wang, Jingjing
    Kang, Yangyang
    Li, Shoushan
    Liu, Xiaozhong
    Si, Luo
    Zhang, Min
    Zhou, Guodong
    2018 CONFERENCE ON EMPIRICAL METHODS IN NATURAL LANGUAGE PROCESSING (EMNLP 2018), 2018, : 3654 - 3663
  • [3] Named Entity-based Question-Answering Pair Generator
    Lahiri, Aritra Kumar
    Hu, Qinmin Vivian
    PROCEEDINGS OF THE 31ST ACM INTERNATIONAL CONFERENCE ON INFORMATION AND KNOWLEDGE MANAGEMENT, CIKM 2022, 2022, : 4902 - 4906
  • [4] Restricted-Domain Chinese Automatic Question-Answering System based on question sentence similarity
    Yu, ZT
    Fan, XZ
    Ji, PC
    PROCEEDINGS OF THE 2004 INTERNATIONAL CONFERENCE ON MACHINE LEARNING AND CYBERNETICS, VOLS 1-7, 2004, : 3023 - 3028
  • [5] Exploiting Sentence Embedding for Medical Question Answering
    Hao, Yu
    Liu, Xien
    Wu, Ji
    Lv, Ping
    THIRTY-THIRD AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE / THIRTY-FIRST INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE CONFERENCE / NINTH AAAI SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE, 2019, : 938 - 945
  • [6] Question-answering system
    Stupina, A. A.
    Zhukov, E. A.
    Ezhemanskaya, S. N.
    Karaseva, M. V.
    Korpacheva, L. N.
    XII INTERNATIONAL SCIENTIFIC AND RESEARCH CONFERENCE TOPICAL ISSUES IN AERONAUTICS AND ASTRONAUTICS, 2016, 155
  • [7] Deep Neural Network Models for Question Classification in Community Question-Answering Forums
    Upadhya, Akshay B.
    Udupa, Swastik
    Kamath, Sowmya S.
    2019 10TH INTERNATIONAL CONFERENCE ON COMPUTING, COMMUNICATION AND NETWORKING TECHNOLOGIES (ICCCNT), 2019,
  • [8] Machine learning based review on Development and Classification of Question-Answering Systems
    Uttarwar, Sayli
    Gambani, Simran
    Thakkar, Tej
    Mulla, Nikahat
    PROCEEDINGS OF THE 2019 3RD INTERNATIONAL CONFERENCE ON COMPUTING METHODOLOGIES AND COMMUNICATION (ICCMC 2019), 2019, : 359 - 366
  • [9] Financial FAQ Question-Answering System Based on Question Semantic Similarity
    Hong, Wenxing
    Li, Jun
    Li, Shuyan
    KNOWLEDGE SCIENCE, ENGINEERING AND MANAGEMENT, PT III, KSEM 2024, 2024, 14886 : 152 - 163
  • [10] Chinese question-answering system
    Huang, GT
    Yao, HH
    JOURNAL OF COMPUTER SCIENCE AND TECHNOLOGY, 2004, 19 (04) : 479 - 488