Item request has been placed! ×
Item request cannot be made. ×
loading  Processing Request

Keeping humans in the loop efficiently by generating question templates instead of questions using AI: Validity evidence on Hybrid AIG

Item request has been placed! ×
Item request cannot be made. ×
loading   Processing Request
  • Additional Information
    • Publication Information:
      Informa UK Limited, 2024.
    • Publication Date:
      2024
    • Abstract:
      Manually creating multiple-choice questions (MCQ) is inefficient. Automatic item generation (AIG) offers a scalable solution, with two main approaches: template-based and non-template-based (AI-driven). Template-based AIG ensures accuracy but requires significant expert input to develop templates. In contrast, AI-driven AIG can generate questions quickly but with inaccuracies. The Hybrid AIG combines the strengths of both methods. However, neither have MCQs been generated using the Hybrid AIG approach nor has any validity evidence been provided. We generated MCQs using the Hybrid AIG approach and investigated the validity evidence of these questions by determining whether experts could identify the correct answers. We used a custom ChatGPT to develop an item template, which were then fed into Gazitor, a template-based AIG (non-AI) software. A panel of medical doctors identified the answers. Of 105 decisions, 101 (96.2%) matched the software’s correct answer. In all MCQs (100%), the experts reached a consensus on the correct answer. The evidence corresponds to the ‘Relations to Other Variables’ in Messick’s validity framework. The Hybrid AIG approach can enhance the efficiency of MCQ generation while maintaining accuracy. It mitigates concerns about hallucinations while benefiting from AI.
    • ISSN:
      1466-187X
      0142-159X
    • Accession Number:
      10.1080/0142159x.2024.2430360
    • Accession Number:
      10.6084/m9.figshare.27923489.v1
    • Accession Number:
      10.6084/m9.figshare.27923489
    • Rights:
      CC BY
    • Accession Number:
      edsair.doi.dedup.....166625a2aee459b09f8d6c36b9034dad