Large language models are few-shot clinical information extractors

Agrawal, Monica; Hegselmann, Stefan; Lang, Hunter; Kim, Yoon; Sontag, David

doi:10.18653/v1/2022.emnlp-main.130

articleJan 1, 2022GOLD OA

Large language models are few-shot clinical information extractors

MAMonica Agrawal SHStefan Hegselmann HLHunter Lang YKYoon Kim DSDavid Sontag

University of Münster

Indexed incrossref

Abstract

A long-running goal of the clinical NLP community is the extraction of important variables trapped in clinical notes. However, roadblocks have included dataset shift from the general domain and a lack of public clinical corpora and annotations. In this work, we show that large language models, such as InstructGPT (Ouyang et al., 2022), perform well at zero- and few-shot information extraction from clinical text despite not being trained specifically for the clinical domain. Whereas text classification and generation performance have already been studied extensively in such models, here we additionally demonstrate how to leverage them to tackle a diverse set of NLP tasks which require more structured outputs,…

Citation impact

288

total citations

FWCI: 36.35
Percentile: 100%
References: 77

Citations per year

Authors

5

Topics & keywords

Topics

Keywords

Computer science
Leverage (statistics)
Relationship extraction
Artificial intelligence
Natural language processing
Benchmarking
Information extraction
Annotation

UN Sustainable Development Goals

Quality Education

No related works found for this paper.