Original Article

Performance of Large Language Models on Diagnostic Radiology Board–Style Questions: A Comparative Evaluation of GPT-4o, Perplexity AI, and OpenEvidence

Objective: The objective of this study was to compare the diagnostic accuracy and internal consistency of GPT-4o (Generative Pre-Trained Transformer-4 omni), Perplexity AI (artificial intelligence), and OpenEvidence when applied to text-based, specialty-level radiology board questions. Methods: A total of 161 text-based multiple-choice questions from the American College of Radiology (ACR)…

Posted in: artificial intelligence 5 diagnostic accuracy 2 large language models 2 radiology 3

Original Article

The Value of Physical Examination: A New Conceptual Framework

The physical examination defines medical practice, yet its role is being questioned increasingly, with statistical comparisons of diagnostic accuracy often the sole metric used against newer technologies. We set out to highlight seven ways in which the physical examination has value beyond diagnostic accuracy to reaffirm its place in the…

Posted in: diagnostic accuracy 2 medical education 78 physical examination 5
SMA Menu