Arturo Loaiza-Bonilla: Benchmark Testing Frontier AI Against Physician Judgment in Hematology
Arturo LoAIza-Bonilla/LinkedIn

Arturo Loaiza-Bonilla: Benchmark Testing Frontier AI Against Physician Judgment in Hematology

Arturo LoAIza-Bonilla, Systemwide Chief of Hematology and Oncology at St. Luke’s University Health Network, shared on X:

“We’re testing the ones that keep hematologists up at night. In AML, CLL and mantle cell lymphoma there are decisions where NCCN, ELN and iwCLL genuinely stop short, and two excellent clinicians can defend opposite choices.

We built a key set of them into a preregistered benchmark, and we’re running frontier AI models and practicing physicians through the same items. Here’s the part that matters: the physicians aren’t the control arm. They’re the reference standard. If you treat these diseases, your judgment is the thing the models get measured against.

5-10 minutes. Anonymous. No patient data. Take one or take all three:

Questions: Yan Leyfman.”

Arturo Loaiza-Bonilla: Benchmark Testing Frontier AI Against Physician Judgment in Hematology

Other articles featuring Arturo LoAIza-Bonilla on OncoDaily.