Reading Group (+🧋): Measuring Physicians and AI Relevance Alignment with MedPAIR cover art

MEETUP · SAN FRANCISCO

Reading Group (+🧋): Measuring Physicians and AI Relevance Alignment with MedPAIR

Join the Snorkel AI Reading Group, a recurring forum to explore the latest frontier developments in AI while building meaningful connections within the community.

When
Thu, Oct 1
4:00 PM to 6:30 PM PT
Where
SoMa, San Francisco, CA
Host
AI Insiders
Entry
Free to attend

About this event

Join the Snorkel AI Reading Group, a recurring forum to explore the latest frontier developments in AI while building meaningful connections within the community.

LLMs now beat the human average on standardized medical exams, but a right answer doesn't mean a model reasoned its way there correctly: It might have latched onto an extreme lab value, stray detail, or piece of context a physician would immediately discount, and still landed on the correct choice by accident.

In this session, Yuexing Hao (Microsoft, MIT EECS) will present her work that introduces MedPAIR: Medical Dataset Comparing Physicians and AI Relevance Estimation and Question Answering to catch exactly that gap.

Among other things, you'll learn:

- Why a model can answer a medical question correctly while relying on completely different - and sometimes spurious - information than a physician would, and why accuracy alone can't catch it. - How MedPAIR's sentence-level annotation process surfaces exactly where physicians and LLMs part ways on what counts as clinically relevant. - Why models often overweight superficial signals, like an unusually extreme test result, while missing subtler cues that trainees flagged as decisive. - Across four medical QA benchmarks, how stripping out the context physicians deemed irrelevant lifted LLM accuracy, which in some cases was enough to beat the physicians' own average. A...

Read the full listing on Luma

Shooting Reading Group (+🧋): Measuring Physicians and AI Relevance Alignment with MedPAIR?

We staff meetups across San Francisco with vetted, background-checked professionals. Our studio edits every frame and delivers in 48 to 72 hours, so your recap goes out while the room is still talking about it.

  • Photo coverage for meetups from $200
  • A vetted professional confirms within 24 hours
  • Deposit is only charged once they confirm
  • Backup professional on standby for every date
Book coverage for Thu, Oct 1

Also coming up

See the full SF calendar

Details come from AI Insiders's public Luma listing and are refreshed every night. Times are San Francisco time. Confirm on the host's page before you go.