Few-shot learning for medical text: A review of advances, trends, and opportunities

Yao Ge; Yuting Guo; Sudeshna Das; Mohammed Ali Al-Garadi; Abeed Sarker

doi:10.1016/j.jbi.2023.104458

Few-shot learning for medical text: A review of advances, trends, and opportunities

J Biomed Inform. 2023 Aug:144:104458. doi: 10.1016/j.jbi.2023.104458. Epub 2023 Jul 23.

Authors

Yao Ge¹, Yuting Guo¹, Sudeshna Das¹, Mohammed Ali Al-Garadi², Abeed Sarker³

Affiliations

¹ Department of Biomedical Informatics, School of Medicine, Emory University, Atlanta, GA, United States of America.
² Department of Biomedical Informatics, Vanderbilt University Medical Center, Vanderbilt University, Nashville, TN, United States of America.
³ Department of Biomedical Informatics, School of Medicine, Emory University, Atlanta, GA, United States of America; Department of Biomedical Engineering, Georgia Institute of Technology and Emory University, Atlanta, GA, United States of America. Electronic address: abeed@dbmi.emory.edu.

PMID: 37488023
PMCID: PMC10940971 (available on 2024-08-01)
DOI: 10.1016/j.jbi.2023.104458

Abstract

Background: Few-shot learning (FSL) is a class of machine learning methods that require small numbers of labeled instances for training. With many medical topics having limited annotated text-based data in practical settings, FSL-based natural language processing (NLP) holds substantial promise. We aimed to conduct a review to explore the current state of FSL methods for medical NLP.

Methods: We searched for articles published between January 2016 and October 2022 using PubMed/Medline, Embase, ACL Anthology, and IEEE Xplore Digital Library. We also searched the preprint servers (e.g., arXiv, medRxiv, and bioRxiv) via Google Scholar to identify the latest relevant methods. We included all articles that involved FSL and any form of medical text. We abstracted articles based on the data source, target task, training set size, primary method(s)/approach(es), and evaluation metric(s).

Results: Fifty-one articles met our inclusion criteria-all published after 2018, and most since 2020 (42/51; 82%). Concept extraction/named entity recognition was the most frequently addressed task (21/51; 41%), followed by text classification (16/51; 31%). Thirty-two (61%) articles reconstructed existing datasets to fit few-shot scenarios, and MIMIC-III was the most frequently used dataset (10/51; 20%). 77% of the articles attempted to incorporate prior knowledge to augment the small datasets available for training. Common methods included FSL with attention mechanisms (20/51; 39%), prototypical networks (11/51; 22%), meta-learning (7/51; 14%), and prompt-based learning methods, the latter being particularly popular since 2021. Benchmarking experiments demonstrated relative underperformance of FSL methods on biomedical NLP tasks.

Conclusion: Despite the potential for FSL in biomedical NLP, progress has been limited. This may be attributed to the rarity of specialized data, lack of standardized evaluation criteria, and the underperformance of FSL methods on biomedical topics. The creation of publicly-available specialized datasets for biomedical FSL may aid method development by facilitating comparative analyses.

Keywords: Biomedical informatics; Few-shot learning; Machine learning; Natural language processing.

Publication types

Review
Research Support, N.I.H., Extramural

MeSH terms

MEDLINE
Machine Learning*
Natural Language Processing*
PubMed
Publications

Grants and funding

R01 DA057599/DA/NIDA NIH HHS/United States