The Performance of Large Language Models in Bone Tumour Imaging: Comparative Analysis with Radiologists Using Text and Image-based Evaluation

Authors

DOI:

https://doi.org/10.7546/CRABS.2026.01.12

Keywords:

ChatGPT, artificial intelligence, bone tumours, large language models, radiology

Abstract

Large language models (LLMs) are emerging as transformative tools in radiology, with potential to enhance diagnostic workflows. However, their performance in bone tumour imaging – a domain requiring both knowledge-based reasoning and visual interpretation -- remains unclear. This study compares the diagnostic performance of LLMs with radiologists across text and image-based tasks.

In this cross-sectional study, two LLMs and two radiologists (a junior and a senior) were evaluated using fifty text-based multiple-choice questions (MCQs) and fifty radiographs with clinical vignettes from a public dataset. Participants classified lesions as benign or malignant, identified “don't-touch” lesions, and provided the most likely diagnosis. Responses were benchmarked against a reference standard using McNemar's tests.

In MCQs, ChatGPT-5 (92.0%) and Gemini 2.5 Pro (90.0%) achieved accuracies comparable to SR (88.0%) and JR (84.0%) (p > 0.05). For benign--malignant classification, LLMs (50.0%, 48.0%) were similar to JR (66.0%) but inferior to SR (94.0%) (p < 0.05). In identifying “don't-touch”' lesions, LLMs (46.0%) matched JR (64.0%) yet underperformed compared to SR (92.0%) (p < 0.05). For specific diagnosis, LLMs showed low accuracy (38.0%, 30.0%) versus JR (60.0%) and SR (86.0%) (p < 0.01).

LLMs may serve as useful adjuncts for clinicians and radiologists in text-based tasks and in distinguishing between benign and malignant bone tumours. However, their diagnostic accuracy remains limited.

Author Biographies

Eren Çamur, Ankara 29 Mayis State Hospital, Turkey

Mailing Address:
Department of Radiology
Ankara 29 Mayis State Hospital
Ankara, Türkiye

E-mail: eren.camur@outlook.com

Turay Cesur, Ankara Mamak State Hospital, Turkey

Mailing Address:
Department of Radiology
Ankara Mamak State Hospital
Ankara, Türkiye

E-mail: turaycesur93@gmail.com

Yasin Celal Güneş, Kirikkale Yuksek Ihtisas Hastanesi, Turkey

Mailing Address:
Department of Radiology
Kirikkale Yuksek Ihtisas Hastanesi
Kırıkkale, Türkiye

E-mail: gunesyasincelal@gmail.com

Semra Duran, Ankara Bilkent City Hospital, Turkey

Mailing Address:
Department of Radiology
Ankara Bilkent City Hospital
Ankara, Türkiye

E-mail: semraduran91@gmail.com

Downloads

Published

28-01-2026

How to Cite

[1]
E. Çamur, T. Cesur, Y. Güneş, and S. Duran, “The Performance of Large Language Models in Bone Tumour Imaging: Comparative Analysis with Radiologists Using Text and Image-based Evaluation”, C. R. Acad. Bulg. Sci., vol. 79, no. 1, pp. 95–102, Jan. 2026.

Issue

Section

Medicine