DOE OSTI · 2000393
FrESCO: Framework for Exploring Scalable Computational Oncology
Abstract
The National Cancer Institute (NCI) monitors population level cancer trends as part of its Surveillance, Epidemiology, and End Results (SEER) program. This program consists of state or regional level cancer registries which collect, analyze, and annotate cancer pathology reports. From these annotated pathology reports, each individual registry aggregates cancer phenotype information from electronic health records. This data is then used to create summary statistics about cancer incidence and mortality to facilitate population health monitoring. Extracting phenotypic information from these reports is a labor intensive task, requiring specialized knowledge about the reports and cancer. Automating the information extraction process from cancer pathology reports has the potential to improve data quality by extracting information in a consistent manner across registries. It can also improve patient outcomes by reducing the time from diagnosis, enabling rapid case ascertainment for clinical trials. Here we present FrESCO, a modular deep-learning natural language processing (NLP) library initially designed for extracting pathology information from clinical text documents. This repository is not solely limited to clinical medical text, but may also be used by researchers just getting started with NLP methods and those looking for a robust solution for their classification problems.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Spannaus, Adam, Gounley, John, Shekar, Mayanka Chandra, Fox, Zachary R., Mohd-Yusof, Jamaludin, Schaefferkoetter, Noah, Hanson, Heidi A.. 2023-09-11. FrESCO: Framework for Exploring Scalable Computational Oncology. https://doi.org/10.21105/joss.05345
Cite the original work for its findings. Save a collection to share your selection of sources.