Almanac — Retrieval-Augmented Language Models for Clinical Medicine

Zakka, Cyril; Shad, Rohan; Chaurasia, Akash; Dalal, Alex R.; Kim, Jennifer L.; Moor, Michael; Fong, Robyn; Phillips, C.; Alexander, Kevin; Ashley, Euan A.; Boyd, Jack; Boyd, Kathleen; Hirsch, Karen G.; Langlotz, Curt; Lee, Rita; Melia, Joanna; Nelson, Joanna; Sallam, Karim; Tullis, Stacey; Vogelsong, Melissa A.; Cunningham, John P.; Hiesinger, William

doi:10.1056/aioa2300068

articleNEJM AIJan 25, 2024GREEN OA

Almanac — Retrieval-Augmented Language Models for Clinical Medicine

CZCyril Zakka RSRohan Shad ACAkash Chaurasia ARAlex R. Dalal JLJennifer L. Kim

Stanford Medicine · Penn Center for AIDS Research · +6 more institutions

PubMed

Indexed incrossrefpubmed

Abstract

Background

Large language models (LLMs) have recently shown impressive zero-shot capabilities, whereby they can use auxiliary data, without the availability of task-specific training examples, to complete a variety of natural language tasks, such as summarization, dialogue generation, and question answering. However, despite many promising applications of LLMs in clinical medicine, adoption of these models has been limited by their tendency to generate incorrect and sometimes even harmful statements.

Methods

We tasked a panel of eight board-certified clinicians and two health care practitioners with evaluating Almanac, an LLM framework augmented with retrieval capabilities from curated medical resources for medical guideline and treatment recommendations. The panel compared responses from Almanac and standard LLMs (ChatGPT-4, Bing, and Bard) versus a novel data set of 314 clinical questions spanning nine medical specialties.

Citation impact

324

total citations

FWCI: 34.38
Percentile: 100%
References: 27

Citations per year

Authors

22

Topics & keywords

Topics

Keywords

Automatic summarization
Guideline
Health care
Precision medicine
Medicine
Medical education
Computer science
Artificial intelligence

UN Sustainable Development Goals

Peace, Justice and strong institutions

No related works found for this paper.