EUVIMEDEuropean Health Evidence
Uhr 7/7Sources Journal Tree
Easy Demo

Lokaler Crossref-Datenbestand · journal-article

AnatomyGPT 5.2 Versus ChatGPT 5.2 in Solving Anatomical Questions: A Comparative Study

Weronika Chaba‐Karnaś, Eliza Tatarczyk, Natalia Kozioł, Natalia Kalita, Kacper Jaśkiewicz, Michał Karnaś, Grzegorz Fibiger, Kamil Możdżeń, Jerzy Walocha, Tomasz Kozioł

Clinical Anatomy · 2026

Vollständiger Abstract

Worum geht es in dieser Arbeit?

ABSTRACT Artificial intelligence is increasingly being used in education, including the field of anatomy, where it could support learning. The aim of this study was to evaluate and compare the performance of two large language models (LLMs), AnatomyGPT 5.2 and ChatGPT 5.2, in answering anatomy‐based questions. A standardized prompt was prepared, and anatomical questions, originally used in official anatomy tests, were posed separately to AnatomyGPT 5.2 and ChatGPT 5.2 in two independent attempts. A total of 550 questions were used, including 150 in Polish and 400 in English. ChatGPT 5.2 outperformed AnatomyGPT 5.2 in both trials, achieving accuracies of 84.36% and 86.36%, compared with 82.73% and 85.27%, respectively; however, the differences between the models were not statistically significant in either trial. In both trials, ChatGPT 5.2 and AnatomyGPT 5.2 performed better on the English question set than on the Polish question set, with the differences between the question sets being statistically significant ( p < 0.05). Questions on innervation and vascularization were most frequently answered correctly by both models, while multiple‐choice questions were least frequently correct; no statistically significant differences were found between models or trials across all categories of questions ( p > 0.05). Cohen's kappa analysis indicated substantial agreement between repeated responses for both models across the two trials ( p < 0.001). The intraclass correlation coefficient was 0.62 for AnatomyGPT 5.2 and 0.66 for ChatGPT 5.2, indicating moderate agreement between the original and repeated responses for both models. Overall, the studied models demonstrated generally high performance across most questions and relatively consistent responses over time. Higher accuracy was observed for the English question set than for the Polish question set; however, this difference cannot be attributed solely to language.

Bibliografischer Nachweis

Publikationsdaten

Autor:innen
Weronika Chaba‐Karnaś, Eliza Tatarczyk, Natalia Kozioł, Natalia Kalita, Kacper Jaśkiewicz, Michał Karnaś, Grzegorz Fibiger, Kamil Możdżeń, Jerzy Walocha, Tomasz Kozioł
Quelle
Clinical Anatomy
Publikation
2026-01-01
Band / Ausgabe
Nicht angegeben
Seiten
Nicht angegeben
ISSN / ISBN
0897-3806, 1098-2353
Zitationen
0 laut Crossref
Referenzen
0 hinterlegt

Zitieren

Zitierfähiger Nachweis

Weronika Chaba‐Karnaś, Eliza Tatarczyk, Natalia Kozioł, Natalia Kalita, Kacper Jaśkiewicz, Michał Karnaś, Grzegorz Fibiger, Kamil Możdżeń, Jerzy Walocha, Tomasz Kozioł (2026). AnatomyGPT 5.2 Versus ChatGPT 5.2 in Solving Anatomical Questions: A Comparative Study. Clinical Anatomy. https://doi.org/10.1002/ca.70206
RIS BibTeX CSL-JSON

Kontext

Themen, Förderung und Nutzung

Lizenzhinweise: Lizenz 1 · Lizenz 2