🚀NEW LABGetting Started with Claude AgentsStart lab
Evaluation · Reasoning

Towards Expert-Level Medical Question Answering (Med-PaLM 2)

First page
Towards Expert-Level Medical Question Answering (Med-PaLM 2)
Paper summary

Google's second-generation medical LLM.

Ask this paper

Key points
01

MedQA SOTA: Scored up to 86.5% on the MedQA dataset (USMLE-style questions) - a new state-of-the-art matching expert physicians.

02

Multi-benchmark leadership: Approaches or exceeds SOTA across MedMCQA, PubMedQA, and MMLU clinical topics datasets.

03

Human evaluation quality: Physician evaluators rated Med-PaLM 2 answers as comparable to those of other physicians on most axes.

04

Medical AI frontier: Set the bar for medical LLMs and informed FDA's thinking on AI-assisted clinical workflows.

Every Monday
Get next week’s papers.
Subscribe on Substack