Factual accuracy versus the health literacy barrier: a blinded neurotological assessment of ChatGPT in vertigo education
DOI:
https://doi.org/10.18203/issn.2454-5929.ijohns20262383Keywords:
ChatGPT, Artificial intelligence, Vertigo, Patient education, Otolaryngology, Vestibular disordersAbstract
Background: Patients increasingly utilize artificial intelligence (AI) models like ChatGPT to seek medical information regarding complex symptoms such as vertigo. The reliability of these AI platforms in providing accurate, comprehensible, and safe clinical information remains a subject of ongoing investigation. Objectives were to evaluate the clinical accuracy, comprehensiveness, and readability of ChatGPT’s responses to common patient inquiries regarding vertigo and vestibular disorders.
Methods: Fifty frequently asked questions (FAQs) concerning vertigo were curated from high-traffic patient forums, clinical encounters, and search engine trends. These questions were categorized into four domains: Diagnosis, treatment, home exercises/lifestyle, and prognosis. The FAQs were inputted into ChatGPT-4. Responses were independently evaluated using a 5-point Likert scale for accuracy and comprehensiveness. Readability was assessed using the Flesch-Kincaid grade level (FKGL) and Flesch reading ease (FRE) scores.
Results: ChatGPT demonstrated high overall accuracy (mean score 4.38±0.62) and comprehensiveness (4.21±0.74). The model performed best in the home exercises/lifestyle category (accuracy 4.65±0.45) and scored lowest in the diagnosis domain (accuracy 4.05±0.82), where it occasionally failed to distinguish atypical features of central vs. peripheral vertigo. The mean FKGL was 12.4±1.8, and the mean FRE was 41.2±10.5, indicating a reading level appropriate for college students but potentially difficult for the general public.
Conclusions: ChatGPT-4 generates highly accurate and comprehensive answers to common patient questions about vertigo, making it a valuable supplementary educational tool. However, its sophisticated vocabulary and occasional diagnostic ambiguity emphasize the necessity for physician oversight.
Metrics
References
Lechien JR, Rameau A. Applications of ChatGPT in Otolaryngology-Head Neck Surgery: A State of the art Review. Otolaryngol Head Neck Surg. 2024;171(3):667-77. DOI: https://doi.org/10.1002/ohn.807
Yanagita Y, Yokokawa D, Uchida S, Tawara J, Ikusaka M. Accuracy of ChatGPT on Medical Questions in the National Medical Licensing Examination in Japan: Evaluation Study. JMIR Form Res. 2023;7:e48023. DOI: https://doi.org/10.2196/48023
Zalzal HG, Abraham A, Cheng J, Shah RK. Can ChatGPT help patients answer their otolaryngology questions? Laryngoscope Investig Otolaryngol. 2023;9(1):e1193. DOI: https://doi.org/10.1002/lio2.1193
Langlie J, Kamrava B, Pasick LJ, Mei C, Hoffer ME. Artificial intelligence and ChatGPT: An otolaryngology patient's ally or foe? Am J Otolaryngol. 2024;45(3):104220. DOI: https://doi.org/10.1016/j.amjoto.2024.104220
Tessler I, Wolfovitz A, Alon EE, Gecel NA, Livneh N, Zimlichman E, et al. ChatGPT's adherence to otolaryngology clinical practice guidelines. Eur Arch Otorhinolaryngol. 2024;281(7):3829-34. DOI: https://doi.org/10.1007/s00405-024-08634-9
Liu X, Shi S, Zhang X, Qianwen G, Wuqing W. The role of ChatGPT-4o in differential diagnosis and management of vertigo-related disorders. Sci Rep. 2025;15(1):18688. DOI: https://doi.org/10.1038/s41598-025-96309-8
Bellinger JR, De La Chapa JS, Kwak MW, Ramos GA, Morrison D, Kesser BW. BPPV Information on Google Versus AI (ChatGPT). Otolaryngol Head Neck Surg. 2024;170(6):1504-11. DOI: https://doi.org/10.1002/ohn.506
Carnino JM, Chong NYK, Bayly H, Salvati LR, Tiwana HS, Levi JR. AI-generated text in otolaryngology publications: a comparative analysis before and after the release of ChatGPT. Eur Arch Otorhinolaryngol. 2024;281(11):6141-6. DOI: https://doi.org/10.1007/s00405-024-08834-3
Maksimoski M, Noble AR, Smith DF. Does ChatGPT Answer Otolaryngology Questions Accurately? Laryngoscope. 2024;134(9):4011-5. DOI: https://doi.org/10.1002/lary.31410
Buhr CR, Smith H, Huppertz T, Bahr-Hamm K, Matthias C, Blaikie A, et al. ChatGPT Versus Consultants: Blinded Evaluation on Answering Otorhinolaryngology Case-Based Questions. JMIR Med Educ. 2023;9:e49183. DOI: https://doi.org/10.2196/49183
Soon S, Perry B. Paging Dr. ChatGPT: safety, accuracy and readability of ChatGPT in ENT emergencies. Aust J Otolaryngol. 2025;8:8. DOI: https://doi.org/10.21037/ajo-24-56
Nielsen JPS, von Buchwald C, Grønhøj C. Validity of the large language model ChatGPT (GPT4) as a patient information source in otolaryngology by a variety of doctors in a tertiary otorhinolaryngology department. Acta Otolaryngol. 2023;143(9):779-82. DOI: https://doi.org/10.1080/00016489.2023.2254809