Analysis of LLMs for educational question classification and generation

被引:0
|
作者
Al Faraby, Said [1 ]
Romadhony, Ade [1 ]
Adiwijaya [1 ]
机构
[1] School of Computing, Telkom University, Jl. Telekomunikasi No.1, Terusan Buah Batu,Bandung,40257, Indonesia
关键词
Contrastive Learning - Question answering;
D O I
10.1016/j.caeai.2024.100298
中图分类号
学科分类号
摘要
Large language models (LLMs) like ChatGPT have shown promise in generating educational content, including questions. This study evaluates the effectiveness of LLMs in classifying and generating educational-type questions. We assessed ChatGPT's performance using a dataset of 4,959 user-generated questions labeled into ten categories, employing various prompting techniques and aggregating results with a voting method to enhance robustness. Additionally, we evaluated ChatGPT's accuracy in generating type-specific questions from 100 reading sections sourced from five online textbooks, which were manually reviewed by human evaluators. We also generated questions based on learning objectives and compared their quality to those crafted by human experts, with evaluations by experts and crowdsourced participants. Our findings reveal that ChatGPT achieved a macro-average F1-score of 0.57 in zero-shot classification, improving to 0.70 when combined with a Random Forest classifier using embeddings. The most effective prompting technique was zero-shot with added definitions, while few-shot and few-shot + Chain of Thought approaches underperformed. The voting method enhanced robustness in classification. In generating type-specific questions, ChatGPT's accuracy was lower than anticipated. However, quality differences between ChatGPT-generated and human-generated questions were not statistically significant, indicating ChatGPT's potential for educational content creation. This study underscores the transformative potential of LLMs in educational practices. By effectively classifying and generating high-quality educational questions, LLMs can reduce the workload on educators and enable personalized learning experiences. © 2024
引用
收藏
相关论文
共 50 条
  • [31] Question the educational practice Analysis of the context and the ways of teaching
    Cuomo, Nicola
    Imola, Alice
    REVISTA DE EDUCACION INCLUSIVA, 2008, 1 (01): : 49 - 58
  • [32] Educational Question Mining At Scale: Prediction, Analysis and Personalization
    Wang, Zichao
    Tschiatschek, Sebastian
    Woodhead, Simon
    Hernandez-Lobato, Jose Miguel
    Jones, Simon Peyton
    Baraniuk, Richard G.
    Zhang, Cheng
    THIRTY-FIFTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE, THIRTY-THIRD CONFERENCE ON INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE AND THE ELEVENTH SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE, 2021, 35 : 15669 - 15677
  • [33] Vector Graphics Generation with LLMs: Approaches and Models
    B. Timofeenko
    V. Efimova
    A. Filchenkov
    Journal of Mathematical Sciences, 2024, 285 (2) : 169 - 179
  • [34] Instruction Tuning with LLMs for Programming Exercise Generation
    Zeng, Guolong
    Xue, Qinchen
    Lu, Xuesong
    WEB INFORMATION SYSTEMS AND APPLICATIONS, WISA 2024, 2024, 14883 : 377 - 385
  • [35] Educational attainment: analysis by immigrant generation
    Chiswick, BR
    DebBurman, N
    ECONOMICS OF EDUCATION REVIEW, 2004, 23 (04) : 361 - 379
  • [36] Educational Question Generation of Children Storybooks via Question Type Distribution Learning and Event-Centric Summarization
    Zhao, Zhenjie
    Hou, Yufang
    Wang, Dakuo
    Yu, Mo
    Liu, Chengzhong
    Ma, Xiaojuan
    PROCEEDINGS OF THE 60TH ANNUAL MEETING OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS (ACL 2022), VOL 1: (LONG PAPERS), 2022, : 5073 - 5085
  • [37] EasyGen: Easing Multimodal Generation with BiDiffuser and LLMs
    Zhao, Xiangyu
    Liu, Bo
    Liu, Qijiong
    Shi, Guangyuan
    Wu, Xiao-Ming
    PROCEEDINGS OF THE 62ND ANNUAL MEETING OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS, VOL 1: LONG PAPERS, 2024, : 1351 - 1370
  • [38] LLMs for Test Input Generation for Semantic Applications
    Rasool, Zafaryab
    Barnett, Scott
    Willie, David
    Kurniawan, Stefanus
    Balugo, Sherwin
    Thudumu, Srikanth
    Abdelrazek, Mohamed
    PROCEEDINGS 2024 IEEE/ACM 3RD INTERNATIONAL CONFERENCE ON AI ENGINEERING-SOFTWARE ENGINEERING FOR AI, CAIN 2024, 2024, : 160 - 165
  • [39] Towards Human-Like Educational Question Generation with Large Language Models
    Wang, Zichao
    Valdez, Jakob
    Mallick, Debshila Basu
    Baraniuk, Richard G.
    ARTIFICIAL INTELLIGENCE IN EDUCATION, PT I, 2022, 13355 : 153 - 166
  • [40] Towards Human-Like Educational Question Generation with Small Language Models
    Fawzi, Fares
    Balan, Sarang
    Cukurova, Mutlu
    Yilmaz, Emine
    Bulathwela, Sahan
    ARTIFICIAL INTELLIGENCE IN EDUCATION: POSTERS AND LATE BREAKING RESULTS, WORKSHOPS AND TUTORIALS, INDUSTRY AND INNOVATION TRACKS, PRACTITIONERS, DOCTORAL CONSORTIUM AND BLUE SKY, AIED 2024, PT I, 2024, 2150 : 295 - 303