Assessing the performance of large language models in literature screening for pharmacovigilance: a comparative study

Dan Li; Leihong Wu; Mingfeng Zhang; Svitlana Shpyleva; Ying-Chi Lin; Ying-Chi Lin; Ho-Yin Huang; Ho-Yin Huang; Ting Li; Joshua Xu

doi:10.3389/fdsfr.2024.1379260

Frontiers in Drug Safety and Regulation (Jun 2024)

Assessing the performance of large language models in literature screening for pharmacovigilance: a comparative study

Dan Li,
Leihong Wu,
Mingfeng Zhang,
Svitlana Shpyleva,
Ying-Chi Lin,
Ying-Chi Lin,
Ho-Yin Huang,
Ho-Yin Huang,
Ting Li,
Joshua Xu

Affiliations

Dan Li: Division of Bioinformatics and Biostatistics, National Center for Toxicological Research, U.S. Food and Drug Administration, Jefferson, AR, United States
Leihong Wu: Division of Bioinformatics and Biostatistics, National Center for Toxicological Research, U.S. Food and Drug Administration, Jefferson, AR, United States
Mingfeng Zhang: Division of Epidemiology, Office of Surveillance and Epidemiology, Center for Drug Evaluation and Research, U.S. Food and Drug Administration, Silver Spring, MD, United States
Svitlana Shpyleva: Division of Biochemical Toxicology, National Center for Toxicological Research, U.S. Food and Drug Administration, Jefferson, AR, United States
Ying-Chi Lin: School of Pharmacy, College of Pharmacy, Kaohsiung Medical University, Kaohsiung, Taiwan
Ying-Chi Lin: Master/Doctoral Degree Program in Toxicology, College of Pharmacy, Kaohsiung Medical University, Kaohsiung, Taiwan
Ho-Yin Huang: School of Pharmacy, College of Pharmacy, Kaohsiung Medical University, Kaohsiung, Taiwan
Ho-Yin Huang: Department of Pharmacy, Kaohsiung Medical University Hospital, Kaohsiung Medical University, Kaohsiung, Taiwan
Ting Li: Division of Bioinformatics and Biostatistics, National Center for Toxicological Research, U.S. Food and Drug Administration, Jefferson, AR, United States
Joshua Xu: Division of Bioinformatics and Biostatistics, National Center for Toxicological Research, U.S. Food and Drug Administration, Jefferson, AR, United States

DOI: https://doi.org/10.3389/fdsfr.2024.1379260
Journal volume & issue: Vol. 4

Abstract

Read online

Pharmacovigilance plays a crucial role in ensuring the safety of pharmaceutical products. It involves the systematic monitoring of adverse events and the detection of potential safety concerns related to drugs. Manual literature screening for pharmacovigilance related articles is a labor-intensive and time-consuming task, requiring streamlined solutions to cope with the continuous growth of literature. The primary objective of this study is to assess the performance of Large Language Models (LLMs) in automating literature screening for pharmacovigilance, aiming to enhance the process by identifying relevant articles more effectively. This study represents a novel application of LLMs including OpenAI’s GPT-3.5, GPT-4, and Anthropic’s Claude2, in the field of pharmacovigilance, evaluating their ability to categorize medical publications as relevant or irrelevant for safety signal reviews. Our analysis encompassed N-shot learning, chain-of-thought reasoning, and evaluating metrics, with a focus on factors impacting accuracy. The findings highlight the promising potential of LLMs in literature screening, achieving a reproducibility of 93%, sensitivity of 97%, and specificity of 67% showcasing notable strengths in terms of reproducibility and sensitivity, although with moderate specificity. Notably, performance improved when models were provided examples consisting of abstracts, labels, and corresponding reasoning explanations. Moreover, our exploration identified several potential contributing factors influencing prediction outcomes. These factors encompassed the choice of key words and prompts, the balance of the examples, and variations in reasoning explanations. By configuring advanced LLMs for efficient screening of extensive literature databases, this study underscores the transformative potential of these models in drug safety monitoring. Furthermore, these insights gained from this study can inform the development of automated systems for pharmacovigilance, contributing to the ongoing efforts to ensure the safety and efficacy of pharmacovigilance products.

Published in Frontiers in Drug Safety and Regulation

ISSN: 2674-0869 (Online)
Publisher: Frontiers Media S.A.
Country of publisher: Switzerland
LCC subjects: Medicine: Therapeutics. Pharmacology
Website: https://www.frontiersin.org/journals/drug-safety-and-regulation/

About the journal

Abstract

Keywords