🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
Evaluation · Reasoning

Towards Expert-Level Medical Question Answering (Med-PaLM 2)

Free while signed in. Answers cite the passages they came from.

First page
Towards Expert-Level Medical Question Answering (Med-PaLM 2)
The curator’s take

Google's second-generation medical LLM.

Key points
01

MedQA SOTA: Scored up to 86.5% on the MedQA dataset (USMLE-style questions) - a new state-of-the-art matching expert physicians.

02

Multi-benchmark leadership: Approaches or exceeds SOTA across MedMCQA, PubMedQA, and MMLU clinical topics datasets.

03

Human evaluation quality: Physician evaluators rated Med-PaLM 2 answers as comparable to those of other physicians on most axes.

04

Medical AI frontier: Set the bar for medical LLMs and informed FDA's thinking on AI-assisted clinical workflows.

Every Monday
Get next week’s papers.

The same picks and the same summaries, in your inbox. Free, and 176 issues deep.

Subscribe on Substack