Hierarchical attention networks for information extraction from cancer pathology reports

Gao, Shang; Young, Michael T.; Qiu, John X.; Yoon, Hong-Jun; Christian, James B.; Fearn, Paul A.; Tourassi, Georgia D.; Ramanthan, Arvind

doi:10.1093/jamia/ocx131

Title: Hierarchical attention networks for information extraction from cancer pathology reports

Abstract

We explored how a deep learning (DL) approach based on hierarchical attention networks (HANs) can improve model performance for multiple information extraction tasks from unstructured cancer pathology reports compared to conventional methods that do not sufficiently capture syntactic and semantic contexts from free-text documents. Data for our analyses were obtained from 942 deidentified pathology reports collected by the National Cancer Institute Surveillance, Epidemiology, and End Results program. The HAN was implemented for 2 information extraction tasks: (1) primary site, matched to 12 International Classification of Diseases for Oncology topography codes (7 breast, 5 lung primary sites), and (2) histological grade classification, matched to G1–G4. Model performance metrics were compared to conventional machine learning (ML) approaches including naive Bayes, logistic regression, support vector machine, random forest, and extreme gradient boosting, and other DL models, including a recurrent neural network (RNN), a recurrent neural network with attention (RNN w/A), and a convolutional neural network. Our results demonstrate that for both information tasks, HAN performed significantly better compared to the conventional ML and DL techniques. In particular, across the 2 tasks, the mean micro and macro F-scores for the HAN with pretraining were (0.852,0.708), compared to naive Bayes (0.518, 0.213), logistic regression (0.682,more »« less

Authors:

Gao, Shang ^[1]; Young, Michael T. ^[1]; Qiu, John X. ^[1]; Yoon, Hong-Jun ^[1]; Christian, James B. ^[1]; Fearn, Paul A. ^[2]; Tourassi, Georgia D. ^[1]; Ramanthan, Arvind ^[1]

Computational Science and Engineering Division, Oak Ridge National Laboratory, Oak Ridge, TN, USA
Surveillance Informatics Branch, Division of Cancer Control and Population Sciences, National Cancer Institute, Bethesda, MD, USA

Publication Date:: Thu Nov 16 00:00:00 EST 2017

Research Org.:: Oak Ridge National Laboratory (ORNL), Oak Ridge, TN (United States). Oak Ridge Leadership Computing Facility (OLCF); National Cancer Inst., Bethesda, MD (United States)

Sponsoring Org.:: USDOE Office of Science (SC), Advanced Scientific Computing Research (ASCR); USDOE National Nuclear Security Administration (NNSA); National Inst. of Health (NIH) (United States)

OSTI Identifier:: 1430375

Alternate Identifier(s):: OSTI ID: 1474695

Grant/Contract Number:: AC05-00OR22725; AC02-06CH11357; AC52-07NA27344; AC52-06NA25396

Resource Type:: Published Article

Journal Name:: Journal of the American Medical Informatics Association

Additional Journal Information:: Journal Name: Journal of the American Medical Informatics Association Journal Volume: 25 Journal Issue: 3; Journal ID: ISSN 1067-5027

Publisher:: Oxford University Press

Country of Publication:: United Kingdom

Language:: English

Subject:: 97 MATHEMATICS AND COMPUTING; 60 APPLIED LIFE SCIENCES; clinical pathology reports; information retrieval; recurrent neural nets; attention networks; classification

Citation Formats


                    Gao, Shang, Young, Michael T., Qiu, John X., Yoon, Hong-Jun, Christian, James B., Fearn, Paul A., Tourassi, Georgia D., and Ramanthan, Arvind. Hierarchical attention networks for information extraction from cancer pathology reports.  United Kingdom: N. p., 2017. 
Web.  doi:10.1093/jamia/ocx131.

Copy to clipboard


                    Gao, Shang, Young, Michael T., Qiu, John X., Yoon, Hong-Jun, Christian, James B., Fearn, Paul A., Tourassi, Georgia D., & Ramanthan, Arvind. Hierarchical attention networks for information extraction from cancer pathology reports.  United Kingdom.  https://doi.org/10.1093/jamia/ocx131

Copy to clipboard


                    Gao, Shang, Young, Michael T., Qiu, John X., Yoon, Hong-Jun, Christian, James B., Fearn, Paul A., Tourassi, Georgia D., and Ramanthan, Arvind. Thu .  
"Hierarchical attention networks for information extraction from cancer pathology reports".  United Kingdom.  https://doi.org/10.1093/jamia/ocx131.

Copy to clipboard


                    
@article{osti_1430375,

  title        = {Hierarchical attention networks for information extraction from cancer pathology reports},

  author       = {Gao, Shang and Young, Michael T. and Qiu, John X. and Yoon, Hong-Jun and Christian, James B. and Fearn, Paul A. and Tourassi, Georgia D. and Ramanthan, Arvind},

  abstractNote = {We explored how a deep learning (DL) approach based on hierarchical attention networks (HANs) can improve model performance for multiple information extraction tasks from unstructured cancer pathology reports compared to conventional methods that do not sufficiently capture syntactic and semantic contexts from free-text documents. Data for our analyses were obtained from 942 deidentified pathology reports collected by the National Cancer Institute Surveillance, Epidemiology, and End Results program. The HAN was implemented for 2 information extraction tasks: (1) primary site, matched to 12 International Classification of Diseases for Oncology topography codes (7 breast, 5 lung primary sites), and (2) histological grade classification, matched to G1–G4. Model performance metrics were compared to conventional machine learning (ML) approaches including naive Bayes, logistic regression, support vector machine, random forest, and extreme gradient boosting, and other DL models, including a recurrent neural network (RNN), a recurrent neural network with attention (RNN w/A), and a convolutional neural network. Our results demonstrate that for both information tasks, HAN performed significantly better compared to the conventional ML and DL techniques. In particular, across the 2 tasks, the mean micro and macro F-scores for the HAN with pretraining were (0.852,0.708), compared to naive Bayes (0.518, 0.213), logistic regression (0.682, 0.453), support vector machine (0.634, 0.434), random forest (0.698, 0.508), extreme gradient boosting (0.696, 0.522), RNN (0.505, 0.301), RNN w/A (0.637, 0.471), and convolutional neural network (0.714, 0.460). HAN-based DL models show promise in information abstraction tasks within unstructured clinical pathology reports.},

  doi          = {10.1093/jamia/ocx131},

  journal      = {Journal of the American Medical Informatics Association},

  number       = 3,

  volume       = 25,

  place        = {United Kingdom},

  year         = {Thu Nov 16 00:00:00 EST 2017},

  month        = {Thu Nov 16 00:00:00 EST 2017}

}

Copy to clipboard

Journal Article:

Free Publicly Available Full Text

Publisher's Version of Record
https://doi.org/10.1093/jamia/ocx131

Other availability

Search WorldCat to find libraries that may hold this journal

Citation Metrics:

Cited by: 60 works

Citation information provided by
Web of Science

Save / Share:

Export Metadata

Save to My Library

Works referenced in this record:

Aiming High — Changing the Trajectory for Cancer
journal, May 2016

Lowy, Douglas R.; Collins, Francis S.
New England Journal of Medicine, Vol. 374, Issue 20
DOI: 10.1056/NEJMp1600894

Long Short-Term Memory
journal, November 1997

Hochreiter, Sepp; Schmidhuber, Jürgen
Neural Computation, Vol. 9, Issue 8
DOI: 10.1162/neco.1997.9.8.1735

Clinicians Are From Mars and Pathologists Are From Venus
journal, June 2000

Powsner, Seth M.; Costa, José; Homer, Robert J.
Archives of Pathology & Laboratory Medicine, Vol. 124, Issue 7
DOI: 10.5858/2000-124-1040-CAFMAP

Automated Classification of Free-text Pathology Reports for Registration of Incident Cases of Cancer
journal, January 2012

Jouhet, V.; Defossez, G.; Burgun, A.
Methods of Information in Medicine, Vol. 51, Issue 03
DOI: 10.3414/ME11-01-0005

Bootstrap confidence intervalsCommentCommentCommentCommentRejoinder
journal, September 1996

DiCiccio, Thomas J.; Efron, Bradley; Hall, Peter
Statistical Science, Vol. 11, Issue 3
DOI: 10.1214/ss/1032280214

Using Natural Language Processing to Improve Efficiency of Manual Chart Abstraction in Research: The Case of Breast Cancer Recurrence
journal, January 2014

Carrell, David S.; Halgrim, Scott; Tran, Diem-Thy
American Journal of Epidemiology, Vol. 179, Issue 6
DOI: 10.1093/aje/kwt441

Similar Records in DOE PAGES and OSTI.GOV collections:

Classifying Cancer Pathology Reports with Hierarchical Self-Attention Networks

Journal Article Gao, Shang ; Qiu, John X. ; Alawad, Mohammed ; ... - Artificial Intelligence in Medicine

We introduce a deep learning architecture, hierarchical self-attention networks (HiSANs), designed for classifying pathology reports and show how its unique architecture leads to a new state-of-the-art in accuracy, faster training, and clear interpretability. We evaluate performance on a corpus of 374,899 pathology reports obtained from the National Cancer Institute's (NCI) Surveillance, Epidemiology, and End Results (SEER) program. Each pathology report is associated with five clinical classification tasks – site, laterality, behavior, histology, and grade. We compare the performance of the HiSAN against other machine learning and deep learning approaches commonly used on medical text data – Naive Bayes, logistic regression,more »« less
https://doi.org/10.1016/j.artmed.2019.101726

Full Text Available
Deep Learning Based Superconducting Radio-Frequency Cavity Fault Classification at Jefferson Laboratory

Journal Article Vidyaratne, Lasitha ; Carpenter, Adam ; Powers, Tom ; ... - Frontiers in Artificial Intelligence

This work investigates the efficacy of deep learning (DL) for classifying C100 superconducting radio-frequency (SRF) cavity faults in the Continuous Electron Beam Accelerator Facility (CEBAF) at Jefferson Lab. CEBAF is a large, high-power continuous wave recirculating linac that utilizes 418 SRF cavities to accelerate electrons up to 12 GeV. Recent upgrades to CEBAF include installation of 11 new cryomodules (88 cavities) equipped with a low-level RF system that records RF time-series data from each cavity at the onset of an RF failure. Typically, subject matter experts (SME) analyze this data to determine the fault type and identify the cavity ofmore »« less
https://doi.org/10.3389/frai.2021.718950

Full Text Available
Jet charge and machine learning

Journal Article Fraser, Katherine ; Schwartz, Matthew D. - Journal of High Energy Physics (Online)

Modern machine learning techniques, such as convolutional, recurrent and recursive neural networks, have shown promise for jet substructure at the Large Hadron Collider. For example, they have demonstrated effectiveness at boosted top or W boson identification or for quark/gluon discrimination. We explore these methods for the purpose of classifying jets according to their electric charge. We find that both neural networks that incorporate distance within the jet as an input and boosted decision trees including radial distance information can provide significant improvement in jet charge extraction over current methods. Specifically, convolutional, recurrent, and recursive networks can provide the largest improvementmore »« less
Cited by 44
https://doi.org/10.1007/JHEP10(2018)093

Full Text Available
PathRepHAN: Hierarchical attention networks for pathology report classification

Software Ramanathan, Arvind ; Christian, James B. ; Tourassi, Georgia ; ...

Developers at ORNL explored how a deep learning (DL) approach based on hierarchical attention networks (HANs) can improve model performance for multiple information extraction tasks from unstructured cancer pathology reports compared to conventional methods that do not sufficiently capture syntactic and semantic contexts from free-text documents. The HAN was implemented for 2 information extraction tasks: (1) primary site, matched to 12 International Classification of Diseases for Oncology topography codes (7 breast, 5 lung primary sites), and (2) histological grade classification, matched to G1–G4. Our results demonstrate that for both information tasks, HAN performed significantly better compared to the conventional MLmore » and DL techniques.« less
https://doi.org/10.11578/dc.20201125.4

View Software
Multimodal Data Representation with Deep Learning for Extracting Cancer Characteristics from Clinical Text

Conference Alawad, Mohammed ; Gao, Shang ; Alamudun, Folami ; ...

This paper presents a multimodal data representation to improve the performance of deep learning models for extracting cancer key characteristics from unstructured text in pathology reports. Specifically, in addition to using the text as the input to deep learning models, we use concept unique identifiers (CUIs) as another source of information to the models. We analyze the performance of different text and CUI data representations, including word embeddings and bag of embeddings (BOE), with a convolutional neural network (CNN) and a fully connected multilayer perceptron neural network (MLP-NN). The high level document embeddings from text and CUI inputs are combinedmore »« less
Full Text Available

Similar Records

Title: Hierarchical attention networks for information extraction from cancer pathology reports

Abstract

Citation Formats

Aiming High — Changing the Trajectory for Cancer journal, May 2016

Long Short-Term Memory journal, November 1997

Clinicians Are From Mars and Pathologists Are From Venus journal, June 2000

Automated Classification of Free-text Pathology Reports for Registration of Incident Cases of Cancer journal, January 2012

Bootstrap confidence intervalsCommentCommentCommentCommentRejoinder journal, September 1996

Using Natural Language Processing to Improve Efficiency of Manual Chart Abstraction in Research: The Case of Breast Cancer Recurrence journal, January 2014

Aiming High — Changing the Trajectory for Cancer
journal, May 2016

Long Short-Term Memory
journal, November 1997

Clinicians Are From Mars and Pathologists Are From Venus
journal, June 2000

Automated Classification of Free-text Pathology Reports for Registration of Incident Cases of Cancer
journal, January 2012

Bootstrap confidence intervalsCommentCommentCommentCommentRejoinder
journal, September 1996

Using Natural Language Processing to Improve Efficiency of Manual Chart Abstraction in Research: The Case of Breast Cancer Recurrence
journal, January 2014