Knowledge Extraction of Cohort Characteristics in Research Publications
No Thumbnail Available
Authors
Franklin, Jade S.
Chari, Shruthi
Foreman, Morgan A.
Seneviratne, Oshani
Gruen, Daniel M.
McCusker, Jamie
Das, Amar K.
McGuinness, Deborah L.
Issue Date
2020
Type
Article
Language
Keywords
Alternative Title
Abstract
When healthcare providers review the results of a clinical trial study to understand its applicability to their practice, they typically analyze how well the characteristics of the study cohort correspond to those of the patients they see. We have previously created a study cohort ontology to standardize this information and make it accessible for knowledge-based decision support. The extraction of this information from research publications is challenging, however, given the wide variance in reporting cohort characteristics in a tabular representation. To address this issue, we have developed an ontology-enabled knowledge extraction pipeline for automatically constructing knowledge graphs from the cohort characteristics found in PDF-formatted research papers. We evaluated our approach using a training and test set of 41 research publications and found an overall accuracy of 83.3% in correctly assembling the knowledge graphs. Our research provides a promising approach for extracting knowledge more broadly from tabular information in research publications.
Description
Full Citation
Franklin, J. S., Chari, S., Foreman, M. A., Seneviratne, O., Gruen, D. M., McCusker, J. P., Das, A. K., and McGuinness, D. L. (2020). Knowledge Extraction of Cohort Characteristics in Research Publications. In AMIA Annual Symposium Proceedings (Vol. 2020, p. 462). American Medical Informatics Association.
Publisher
AMIA