Assessing GPT-4’s Performance in Delivering Medical Advice: Comparative Analysis With Human Experts

Eunbeen Jo, Sanghoun Song, Jong-Ho Kim, Subin Lim, Ju Hyeon Kim, Jung‐Joon Cha, Young-Min Kim, Hyung Joon Joo

JMIR Medical Education · 2024 · 23 citations · 15 references

DOIFull text

Open access

Abstract

GPT-4 has shown promising potential in automated medical consultation, with comparable medical accuracy to human experts. However, challenges remain particularly in the realm of nuanced clinical judgment. Future improvements in LLMs may require the integration of specific clinical reasoning pathways and regulatory oversight for safe use. Further research is needed to understand the full potential of LLMs across various medical specialties and conditions.

References

15