Analysis of LLMs for educational question classification and generation

被引:0
|
作者
Al Faraby, Said [1 ]
Romadhony, Ade [1 ]
Adiwijaya [1 ]
机构
[1] School of Computing, Telkom University, Jl. Telekomunikasi No.1, Terusan Buah Batu,Bandung,40257, Indonesia
关键词
Contrastive Learning - Question answering;
D O I
10.1016/j.caeai.2024.100298
中图分类号
学科分类号
摘要
Large language models (LLMs) like ChatGPT have shown promise in generating educational content, including questions. This study evaluates the effectiveness of LLMs in classifying and generating educational-type questions. We assessed ChatGPT's performance using a dataset of 4,959 user-generated questions labeled into ten categories, employing various prompting techniques and aggregating results with a voting method to enhance robustness. Additionally, we evaluated ChatGPT's accuracy in generating type-specific questions from 100 reading sections sourced from five online textbooks, which were manually reviewed by human evaluators. We also generated questions based on learning objectives and compared their quality to those crafted by human experts, with evaluations by experts and crowdsourced participants. Our findings reveal that ChatGPT achieved a macro-average F1-score of 0.57 in zero-shot classification, improving to 0.70 when combined with a Random Forest classifier using embeddings. The most effective prompting technique was zero-shot with added definitions, while few-shot and few-shot + Chain of Thought approaches underperformed. The voting method enhanced robustness in classification. In generating type-specific questions, ChatGPT's accuracy was lower than anticipated. However, quality differences between ChatGPT-generated and human-generated questions were not statistically significant, indicating ChatGPT's potential for educational content creation. This study underscores the transformative potential of LLMs in educational practices. By effectively classifying and generating high-quality educational questions, LLMs can reduce the workload on educators and enable personalized learning experiences. © 2024
引用
收藏
相关论文
共 50 条
  • [41] Generating Educational Materials with Different Levels of Readability using LLMs
    Huang, Chieh-Yang
    Wei, Jing
    Huang, Ting-Hao Kenneth
    PROCEEDINGS OF THE THIRD WORKSHOP ON INTELLIGENT AND INTERACTIVE WRITING ASSISTANTS, IN2WRITING 2024, 2024, : 16 - 22
  • [42] Educational policies in question
    Marchand, Philippe
    REVUE DU NORD, 2006, 88 (365) : 406 - 406
  • [43] Seeing Beyond Borders: Evaluating LLMs in Multilingual Ophthalmological Question Answering
    Restrepo, David
    Nakayama, Luis Filipe
    Dychiao, Robyn Gayle
    Wu, Chenwei
    McCoy, Liam G.
    Artiaga, Jose Carlo
    Cobanaj, Marisa
    Matos, Joao
    Gallifant, Jack
    Bitterman, Danielle S.
    Ferrer, Vincenz
    Aphinyanaphongs, Yindalon
    Celi, Leo Anthony
    2024 IEEE 12TH INTERNATIONAL CONFERENCE ON HEALTHCARE INFORMATICS, ICHI 2024, 2024, : 565 - 566
  • [44] Drilling Down into the Discourse Structure with LLMs for Long Document Question Answering
    Nair, Inderjeet
    Somasundaram, Shwetha
    Saxena, Apoory
    Goswami, Koustava
    FINDINGS OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS (EMNLP 2023), 2023, : 14593 - 14606
  • [45] Bias mitigation in text classification through cGAN and LLMs
    Kumar, Gunjan
    Singh, Jyoti Prakash
    PROCEEDINGS OF THE INDIAN NATIONAL SCIENCE ACADEMY, 2024,
  • [46] Improving educational web search for question-like queries through subject classification
    Yilmaz, Tolga
    Ozcan, Rifat
    Altingovde, Ismail Sengor
    Ulusoy, Ozgur
    INFORMATION PROCESSING & MANAGEMENT, 2019, 56 (01) : 228 - 246
  • [47] Knowledgeable Preference Alignment for LLMs in Domain-specific Question Answering
    Zhang, Yichi
    Chen, Zhuo
    Fang, Yin
    Lu, Yanxi
    Li, Fangming
    Zhang, Wen
    Chen, Huajun
    FINDINGS OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS: ACL 2024, 2024, : 891 - 904
  • [48] LOGIC OF CLASSIFICATION ACT - POSTULATE OR QUESTION FOR ANALYSIS OF MOBILITY
    GARONAUDY, M
    SOCIOLOGIE ET SOCIETES, 1976, 8 (02): : 37 - 59
  • [49] Towards Efficient DataWrangling with LLMs using Code Generation
    Li, Xue
    Dohmen, Till
    PROCEEDINGS OF THE 8TH WORKSHOP ON DATA MANAGEMENT FOR END-TO-END MACHINE LEARNING, DEEM 2024, 2024,
  • [50] Repair Is Nearly Generation: Multilingual Program Repair with LLMs
    Joshi, Harshit
    Sanchez, Jose Cambronero
    Gulwani, Sumit
    Le, Vu
    Radicek, Ivan
    Verbruggen, Gust
    THIRTY-SEVENTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE, VOL 37 NO 4, 2023, : 5131 - 5140