Research map: Validation of large language models for multiple-choice assessment in cranio-maxillofacial surgery – Part B: Student perception and didactic quality of AI- versus human-authored questions

Back to the article

Papers in this map

  1. World Medical Association Declaration of Helsinki · 2013 · 30960 citations · Cited by this paper
  2. Pulmonary embolism after radial forearm free flap reconstruction for oral squamous cell carcinoma: a multicenter retrospective cohort study · 2026 · Related
  3. FUTURE-AI: international consensus guideline for trustworthy and deployable artificial intelligence in healthcare · Karim Lekadir · 2025 · 609 citations · Cited by this paper
  4. Distinct synovial proinflammatory cytokine and hyaluronic acid profiles in temporomandibular joint of patients with Class Ⅱ and Class Ⅲ dentofacial deformities · 2026 · Related
  5. ChatGPT versus human in generating medical graduate exam multiple choice questions—A multinational prospective study (Hong Kong S.A.R., Singapore, Ireland, and the United Kingdom) · Billy Ho Hung Cheung · 2023 · 202 citations · Cited by this paper
  6. Superficial temporal artery perforator flap: an anatomical study and topographic mapping of cutaneous perforators · 2026 · Related
  7. Assessing ChatGPT’s Mastery of Bloom’s Taxonomy Using Psychosomatic Medicine Exam Questions: Mixed-Methods Study · Anne Herrmann‐Werner · 2024 · 103 citations · Cited by this paper
  8. Accuracy of fiducial-based augmented reality in auricular reconstruction · 2026 · Related
  9. AI versus human-generated multiple-choice questions for medical education: a cohort study in a high-stakes examination · Alex Kwok-Keung Law · 2025 · 87 citations · Cited by this paper
  10. Comparison of audiometric outcomes between modified two-flap and Furlow's palatoplasty techniques: A retrospective cohort study · 2026 · Related
  11. Prompt Engineering in Education: A Systematic Review of Approaches and Educational Applications · Yufeng Qian · 2025 · 71 citations · Cited by this paper
  12. The position of the mental foramen on panoramic radiographs of patients with neurofibromatosis type 1 · 2026 · Related
  13. ChatGPT 3.5 fails to write appropriate multiple choice practice exam questions · Alexander Ngo · 2023 · 58 citations · Cited by this paper
  14. Deep learning-automatic 3D analysis of regional condylar remodeling and skeletal relapse following bimaxillary surgery: a two-year follow-up study · 2026 · Related
  15. How well do large language model-based chatbots perform in oral and maxillofacial radiology? · Hui Jeong · 2024 · 46 citations · Cited by this paper
  16. Outcomes of the pectoralis major muscle flap for covering a mandibular reconstruction plate: Our experience at the Leiden University Medical Center · 2026 · Related
  17. Performance of large language models in oral and maxillofacial surgery examinations · Bernadette Quah · 2024 · 40 citations · Cited by this paper
  18. Announcements · 2026 · Related
  19. How does artificial intelligence master urological board examinations? A comparative analysis of different Large Language Models’ accuracy and reliability in the 2022 In-Service Assessment of the European Board of Urology · Lisa Kollitsch · 2024 · 34 citations · Cited by this paper
  20. LLM-Generated multiple choice practice quizzes for preclinical medical students · Troy Camarata · 2025 · 16 citations · Cited by this paper
  21. Comparison of AI-generated and clinician-designed multiple-choice questions in emergency medicine exam: a psychometric analysis · Murtaza Kaya · 2025 · 14 citations · Cited by this paper
  22. Large Language Model Clinical Vignettes and Multiple-Choice Questions for Postgraduate Medical Education · Frank Ian Jackson · 2025 · 11 citations · Cited by this paper
  23. AI-generated multiple-choice questions in health science education: Stakeholder perspectives and implementation considerations · Matthew Reid · 2025 · 10 citations · Cited by this paper
  24. Chatbots’ Role in Generating Single Best Answer Questions for Undergraduate Medical Student Assessment: Comparative Analysis · Enjy Abouzeid · 2025 · 9 citations · Cited by this paper
  25. Ten tips to harnessing generative AI for high-quality MCQS in medical education assessment · Mohi Eldin Magzoub · 2025 · 9 citations · Cited by this paper
  26. Performance of large language models (ChatGPT4-0, Grok2 and Gemini) in UK dentistry and dental hygiene and therapy assessments · Manàs Dave · 2025 · 8 citations · Cited by this paper
  27. How valuable are the questions and answers generated by large language models in oral and maxillofacial surgery? · Kyu Hyung Kim · 2025 · 7 citations · Cited by this paper
  28. The Generation and Use of Medical MCQs: A Narrative Review · Sinclair Steele · 2025 · 6 citations · Cited by this paper
  29. Comparative performance evaluation of ChatGPT-4 Omni and Gemini Advanced in the Turkish Dentistry Specialization Exam · Makbule Buse Dündar Sarı · 2026 · 5 citations · Cited by this paper
  30. A comparative analysis of the performance of large Language models in the dentistry specialty examination · Gediz Geduk · 2026 · 5 citations · Cited by this paper
  31. Comparison of applicability, difficulty, and discrimination indices of multiple-choice questions on medical imaging generated by different AI-based chatbots · Betül Nalan Karahan · 2025 · 4 citations · Cited by this paper
  32. Large language model use in oral and maxillofacial surgery training: a national resident survey · Nolan Kranc · 2026 · 2 citations · Cited by this paper
  33. Performance of five large language models in oral and maxillofacial surgery exam questions: a comparative study · Lisi Liu · 2026 · 2 citations · Cited by this paper

Source: OpenAlex (CC0)