Thomas Potter scite author profile

Thomas Potter

2Publications

0Citation Statements Received

44Citation Statements Given

How they've been cited

How they cite others

Affiliations

Publications

Order By: Most citations

Extraction of Radiological Characteristics From Free-Text Imaging Reports Using Natural Language Processing Among Patients With Ischemic and Hemorrhagic Stroke: Algorithm Development and Validation (Preprint)

Hsu¹,

Bako²,

Potter³

et al. 2022

Preprint

View full text Add to dashboard Cite

BACKGROUND Neuroimaging is the gold standard diagnostic modality for all suspected stroke patients. However, the unstructured nature of imaging reports remains a major challenge to extracting useful information from electronic health records (EHR) systems. Despite the increasing adoption of natural language processing (NLP) for radiology reports, information extraction for many stroke imaging features has not been systematically evaluated. OBJECTIVE In this study, we propose an NLP pipeline, which adopts the state-of-the-art ClinicalBERT model with domain-specific pre-training to extract 13 stroke imaging features from head computed tomography (CT) imaging notes. METHODS We utilized the model to generate structured datasets with information on the presence or absence of common stroke features for 24,924 stroke patients. We compared the survival characteristics of patients with and without features of severe stroke (midline shift, perihematomal edema, or mass effect) using the Kaplan-Meier curve and log-rank test. RESULTS Pre-trained on 82,073 head CT notes with 61 million words and fine-tuned on 200 annotated notes, our HeadCT_BERT model achieved an average Area Under Receiver Operating Characteristic curve (AUROC) of 0.9831, F1 score of 0.8683, and accuracy of 97%. Among patients with acute ischemic stroke, admissions with any severe stroke feature in initial imaging notes were associated with lower probability of survival (P-value < .001). CONCLUSIONS Our proposed NLP pipeline achieved high performance and has the potential to improve medical research and patient safety.

show abstract

Extraction of Radiological Characteristics From Free-Text Imaging Reports Using Natural Language Processing Among Patients With Ischemic and Hemorrhagic Stroke: Algorithm Development and Validation

Hsu¹,

Bako²,

Potter³

et al. 2023

JMIR AI

View full text Add to dashboard Cite

Background Neuroimaging is the gold-standard diagnostic modality for all patients suspected of stroke. However, the unstructured nature of imaging reports remains a major challenge to extracting useful information from electronic health records systems. Despite the increasing adoption of natural language processing (NLP) for radiology reports, information extraction for many stroke imaging features has not been systematically evaluated. Objective In this study, we propose an NLP pipeline, which adopts the state-of-the-art ClinicalBERT model with domain-specific pretraining and task-oriented fine-tuning to extract 13 stroke features from head computed tomography imaging notes. Methods We used the model to generate structured data sets with information on the presence or absence of common stroke features for 24,924 patients with strokes. We compared the survival characteristics of patients with and without features of severe stroke (eg, midline shift, perihematomal edema, or mass effect) using the Kaplan-Meier curve and log-rank tests. Results Pretrained on 82,073 head computed tomography notes with 13.7 million words and fine-tuned on 200 annotated notes, our HeadCT_BERT model achieved an average area under receiver operating characteristic curve of 0.9831, F1-score of 0.8683, and accuracy of 97%. Among patients with acute ischemic stroke, admissions with any severe stroke feature in initial imaging notes were associated with a lower probability of survival (P<.001). Conclusions Our proposed NLP pipeline achieved high performance and has the potential to improve medical research and patient safety.

show abstract

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

customersupport@researchsolutions.com

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Thomas Potter

Extraction of Radiological Characteristics From Free-Text Imaging Reports Using Natural Language Processing Among Patients With Ischemic and Hemorrhagic Stroke: Algorithm Development and Validation (Preprint)

Extraction of Radiological Characteristics From Free-Text Imaging Reports Using Natural Language Processing Among Patients With Ischemic and Hemorrhagic Stroke: Algorithm Development and Validation

Contact Info

Product

Resources

About