Research

Explore our publications and preprints advancing healthcare through rigorous AI evaluation.

Preprint
Aug 27, 2026

Drivers of Oncologist Preference of AI-Generated Literature Review in a Randomized Mixed-Methods Study

Doctors increasingly rely on AI in the clinic, yet which report features make AI-generated responses useful and trustworthy remains unclear. […]

Nature Health
Aug 26, 2026

Patient factors in medical artificial intelligence: a systematic review

Medical artificial intelligence (AI) is increasingly developed, piloted and used in clinical practice, yet the translation of technical capability into […]

Jama Network Open
Aug 20, 2026

An Electronic Health Record–Integrated, Large Language Model–Powered Tool to Triage Surgical Patients

Surgical comanagement (SCM) is an evidence-based care model in which hospitalists jointly manage medically complex perioperative patients alongside surgical teams. […]

Preprint
Aug 20, 2026

HealMed: Multilingual Evaluation of Large Language Models in Medicine

We present HealMed, an expert-reviewed benchmark for multilingual evaluation of large language models in medicine. HealMed contains 1,000 examples in […]

Lancet Digital Health
Aug 19, 2026

Building safer clinical agents: the case for residency-level benchmarks in medical artificial intelligence

Advances in large language models (LLMs) have accelerated medical benchmarking, yet most evaluation of LLMs still relies on exam-style question […]

Preprint
Aug 11, 2026

RadFusion: Towards Threshold-Controllable Radiology Report Generation

Automated radiology report generation is advancing rapidly in response to the shortage of radiologists, yet unlike a perception model, existing […]

npj Digital Medicine
Aug 4, 2026

Toward expert-level medical text validation with language models

With the growing use of language models (LMs) in clinical environments, there is an immediate need to evaluate the accuracy […]

Preprint
Aug 2, 2026

High-Stakes Decisions with Language Models: Insights from Emergency Triage

High-stakes decisions under uncertainty, such as medical emergency triage, require more than accurate predictions. They depend on estimating the likelihood […]

Nature
Jul 28, 2026

Medical AI has a measurement problem

The development of two medical AI assistants highlights an unnerving challenge: as the technology races ahead, what is the best […]

Nature Medicine
Jul 27, 2026

Toward a test of medical AI superintelligence

Researchers urgently need a rigorous, task-based framework to define and measure medical AI ‘superintelligence’, because existing benchmarks are misleading and […]

Latest News

View all

Get the latest on our studies, grant awards, and media coverage.