GPT-4 versus human authors in clinically complex MCQ creation: A blinded analysis of item quality.

Purpose: To compare the structural quality of multiple choice questions (MCQs) generated by a large language model, a type of artificial intelligence (AI), GPT-4, against human-authored items at both novice and expert level. Methods: We conducted a blinded analysis of 124 MCQs: 40 generated by GPT-4...

Descripción completa

Detalles Bibliográficos
Publicado en:Medical Teacher Vol. 47; no. 12; pp. 1961 - 1975
Autores principales: Wu, Hannah, Zerner, Toby, Lee, Daniel, Court-Kowalski, Stefan, Devitt, Peter, Palmer, Edward
Formato: research tables/charts Journal Article
Publicado: Taylor & Francis Ltd Dec2025
Acceso en línea:Ver este registro en EBSCOhost