Kan AI bestå den amerikanske medicinske licensprøve (USMLE)? — Status tjekket på 20. marts 2024 ?
Afgiv din stemme — læs så hvad vores redaktør og AI-modellerne fandt.
GPT-4 scorede over den beståede grænse i alle tre trin af United States Medical Licensing Exam. Medicinstuderende lærer nu "hvordan man bruger AI" som en klinisk færdighed.
Background
AI systems have made significant progress in processing and generating human-like language. Passing the USMLE medical licensing exam is a complex task that requires a deep understanding of medical concepts, clinical knowledge, and critical thinking skills. Currently, AI models can assist with certain aspects of medical education, such as providing study materials, practicing questions, and offering feedback, but they are not yet capable of replacing human judgment and expertise in a high-stakes exam like the USMLE. While AI can process vast amounts of medical information, its ability to apply this knowledge in a practical, real-world setting, such as a licensing exam, is still limited. The development of AI systems that can pass the USMLE exam would require significant advancements in areas like natural language understanding, common sense, and decision-making under uncertainty. GPT-4 scored above passing on all three steps of the United States Medical Licensing Exam. Med-schools now teach 'how to use AI' as a clinical skill.
Foreslå et tag
Mangler et begreb i dette emne? Foreslå det, admin gennemgår.
Status senest tjekket August 15, 2026.
Galleri
Kan AI bestå den amerikanske medicinske licensprøve (USMLE)? — Status tjekket på 20. marts 2024
Snævre demoer findes — men panelet var ikke enigt.
Juryen stod mellem to tætte lejre – den ene fejrede næsten-beståede point under optimerede testforhold, mens den anden bemærkede kløften mellem øvelsessessioner og reel klinisk dømmekraft. Til sidst landede de på ”Næsten”, idet de anerkendte de imponerende fremskridt mod medicinsk kompetence, samtidig med at de indså, at eksamenens rolle som portvagt for menneskelig ekspertise fortsat er uafsluttet. Udgang: AI har fortjent hæder i praksis, men endnu ikke den hvide kittel og stetoskopet.
The jury grappled between two close camps—one celebrating near-passing scores under optimized test conditions, the other noting the gap between practice sessions and real clinical judgment. Ultimately, they settled on “Almost,” acknowledging the impressive strides toward medical competence while recognizing the exam’s role in gatekeeping human expertise remains unfinished business. Verdict in: AI has earned honors in practice, but not yet the white coat and stethoscope.
But the data is real.
The Case File
Across 20 sessions, 51 jurors have heard this case. Combined tally: 24 YES · 23 ALMOST · 4 NO · 0 IN RESEARCH.
Note: cumulative includes older juror opinions. The current session tally above is the live verdict.
By a vote of 1 — 1 — 0, the panel returns a verdict of NæSTEN, with verdict confidence of 88%. The court so orders. Verdict downgraded from prior session.
"AI models can pass practice tests"
"Large language models (e.g., Med-PaLM 2) achieved approximate passing scores on USMLE under optimized conditions."
Individuelle nævningers udtalelser vises på originalengelsk for at bevare bevismæssig præcision.
Hvad publikum mener
Nej 18% · Ja 82% · Måske 0% 110 votesDiskussion
no comments⚖ 20 jury checks · seneste for 4 dage siden
Hver række er et separat jurytjek. Nævninger er AI-modeller (identiteter holdt neutrale med vilje). Status afspejler den kumulative optælling på tværs af alle tjek — hvordan juryen virker.