A Natural Language Processing System for Extracting Evidence of Drug Repurposing from Scientific Publications

Shivashankar Subramanian; Ioana Baldini; Sushma Ravichandran; Dmitriy A. Katz-Rogozhnikov; Karthikeyan Natesan Ramamurthy; Prasanna Sattigeri; Kush R. Varshney; Annmarie Wang; Pradeep Mangalath; Laura B. Kleiman

doi:10.1609/aaai.v34i08.7052

Authors

Shivashankar Subramanian The University of Melbourne
Ioana Baldini IBM Research
Sushma Ravichandran IBM Research
Dmitriy A. Katz-Rogozhnikov IBM Research
Karthikeyan Natesan Ramamurthy IBM Research
Prasanna Sattigeri IBM Research
Kush R. Varshney IBM Research
Annmarie Wang Massachusetts Institute of Technology
Pradeep Mangalath Harvard University
Laura B. Kleiman Cures Within Reach for Cancer

DOI:

https://doi.org/10.1609/aaai.v34i08.7052

Abstract

More than 200 generic drugs approved by the U.S. Food and Drug Administration for non-cancer indications have shown promise for treating cancer. Due to their long history of safe patient use, low cost, and widespread availability, repurposing of these drugs represents a major opportunity to rapidly improve outcomes for cancer patients and reduce healthcare costs. In many cases, there is already evidence of efficacy for cancer, but trying to manually extract such evidence from the scientific literature is intractable. In this emerging applications paper, we introduce a system to automate non-cancer generic drug evidence extraction from PubMed abstracts. Our primary contribution is to define the natural language processing pipeline required to obtain such evidence, comprising the following modules: querying, filtering, cancer type entity extraction, therapeutic association classification, and study type classification. Using the subject matter expertise on our team, we create our own datasets for these specialized domain-specific tasks. We obtain promising performance in each of the modules by utilizing modern language processing techniques and plan to treat them as baseline approaches for future improvement of individual components.

A Natural Language Processing System for Extracting Evidence of Drug Repurposing from Scientific Publications

Authors

DOI:

Abstract

Downloads

Published

How to Cite

Issue

Section

Information