Original Article

Comparing Speed and Accuracy of Artificial Intelligence Large Language Models on the Orthopedic In-Training Examination

Objectives: Large language models (LLMs), such as Open AI’s Chat Generative Pre-Trained Transformer (GPT)-4 and Google Gemini, have gained significant attention for their ability to process complex language patterns and are being used increasingly in fields such as medicine, where they assist in learning, collaboration, and patient care. Although prior…

Posted in: ChatGPT 4 large language models 2

Original Article

Performance of Large Language Models on Diagnostic Radiology Board–Style Questions: A Comparative Evaluation of GPT-4o, Perplexity AI, and OpenEvidence

Objective: The objective of this study was to compare the diagnostic accuracy and internal consistency of GPT-4o (Generative Pre-Trained Transformer-4 omni), Perplexity AI (artificial intelligence), and OpenEvidence when applied to text-based, specialty-level radiology board questions. Methods: A total of 161 text-based multiple-choice questions from the American College of Radiology (ACR)…

Posted in: artificial intelligence 5 diagnostic accuracy 2 large language models 2 radiology 3
SMA Menu