Comparing the Performances of a 54-Year-Old Computer-Based Consultation to ChatGPT-4o.
This study aimed to evaluate and compare the diagnostic responses generated by two artificial intelligence (AI) models developed 54 years apart, and encourage physicians to explore the use of large language models (LLMs) like GPT-4o in clinical practice.A clinical case of metabolic acidosis was presented to GPT-4o, and the model's diagnostic reasoning, data interpretation, and management recommendations were recorded. These outputs were then compared with the responses from Schwartz's 1970 [...]
Author(s): Verdi, Elvan Burak, Akbilgic, Oguz
DOI: 10.1055/a-2628-8408