OpenAI's GPT-4 performs to a high degree on board-style dermatology questions.

Autor: Elias ML; Department of Dermatology, Donald and Barbara Zucker School of Medicine at Hofstra/Northwell, New Hyde Park, NY, USA., Burshtein J; Department of Dermatology, Donald and Barbara Zucker School of Medicine at Hofstra/Northwell, New Hyde Park, NY, USA., Sharon VR; Department of Dermatology, Donald and Barbara Zucker School of Medicine at Hofstra/Northwell, New Hyde Park, NY, USA.
Jazyk: angličtina
Zdroj: International journal of dermatology [Int J Dermatol] 2024 Jan; Vol. 63 (1), pp. 73-78. Date of Electronic Publication: 2023 Dec 22.
DOI: 10.1111/ijd.16913
Abstrakt: Background: Artificial intelligence tools such as OpenAI's GPT-4 have shown promise in medical education, but their potential in dermatology remains unexplored.
Objectives: To assess GPT-4's performance on dermatology board-style questions and determine its value as a supplementary educational tool for trainees and educators.
Methods: This cross-sectional study evaluated GPT-4's performance on 250 random dermatology board-style questions sampled from the American Academy of Dermatology's Board Prep Plus resource. Questions were divided into five subspecialties and various difficulty levels. GPT-4 responses were compared to the correct answers and evaluated by two physicians.
Results: GPT-4 achieved an overall accuracy of 75% on the 250 questions, with no significant variation based on subspecialty or question difficulty. The most common errors were factual and misunderstanding inaccuracies. Responses scored high in clarity, accuracy, and relevance but frequently lacked depth and completeness.
Conclusion: GPT-4 performed to a high degree and demonstrated promising performance as an educational adjunct in dermatology. Improvements in response depth and completeness are needed before its use as an unsupervised learning tool is established.
(© 2023 the International Society of Dermatology.)
Databáze: MEDLINE