Knowledge-enhanced visual-language pre-training on chest radiology images

Zhang, Xiaoman; Wu, Chaoyi; Zhang, Ya; Xie, Weidi; Wang, Yanfeng

doi:10.1038/s41467-023-40260-7

articleNature CommunicationsJul 28, 2023GOLD OA

Knowledge-enhanced visual-language pre-training on chest radiology images

XZXiaoman Zhang CWChaoyi Wu YZYa Zhang WXWeidi Xie YWYanfeng Wang

Shanghai Jiao Tong University · Shandong Jiaotong University · +2 more institutions

PubMed

Indexed incrossrefdoajpubmed

Abstract

While multi-modal foundation models pre-trained on large-scale data have been successful in natural language understanding and vision recognition, their use in medical domains is still limited due to the fine-grained nature of medical tasks and the high demand for domain knowledge. To address this challenge, we propose an approach called Knowledge-enhanced Auto Diagnosis (KAD) which leverages existing medical domain knowledge to guide vision-language pre-training using paired chest X-rays and radiology reports. We evaluate KAD on four external X-ray datasets and demonstrate that its zero-shot performance is not only comparable to that of fully supervised models but also superior to the average of three expert…

Citation impact

175

total citations

FWCI: 38.97
Percentile: 100%
References: 32

Citations per year

Authors

5

Topics & keywords

Topics

Keywords

Computer science
Domain (mathematical analysis)
Artificial intelligence
Domain knowledge
Annotation
Machine learning
Natural language processing
Data science

UN Sustainable Development Goals

Quality Education

No related works found for this paper.

Funding

SA
Science and Technology Commission of Shanghai Municipality
Awards: 21DZ1100100, 22511106101, BP0719010, 18DZ2270700