Format

Send to

Choose Destination
J Healthc Inform Res. 2017 Jun;1(1):1-18. doi: 10.1007/s41666-017-0005-6. Epub 2017 Jun 8.

An Interoperable Similarity-based Cohort Identification Method Using the OMOP Common Data Model version 5.0.

Author information

1
Department of Biomedical Informatics, Columbia University, New York NY 10032.
2
National Institute of Health, National Library of Medicine, Bethesda, MD 20892.
3
Department of Anesthesiology, Columbia University, New York NY 10032.

Abstract

Cohort identification for clinical studies tends to be laborious, time-consuming, and expensive. Developing automated or semi-automated methods for cohort identification is one of the "holy grails" in the field of biomedical informatics. We propose a high-throughput similarity-based cohort identification algorithm by applying numerical abstractions on Electronic Health Records (EHR) data. We implement this algorithm using the Observational Medical Outcomes Partnership (OMOP) Common Data Model (CDM), which enables sites using this standardized EHR data representation to avail this algorithm with minimum effort for local implementation. We validate its performance for a retrospective cohort identification task on six clinical trials conducted at the Columbia University Medical Center. Our algorithm achieves an average Area Under the Curve (AUC) of 0.966 and an average Precision at 5 of 0.983. This interoperable method promises to achieve efficient cohort identification in EHR databases. We discuss suitable applications of our method and its limitations and propose warranted future work.

KEYWORDS:

Case-based Reasoning (CBR); Cohort Identification; Electronic Health Records (EHR); Observational Medical Outcomes Partnership (OMOP); Phenotype; Similarity-based

Supplemental Content

Full text links

Icon for PubMed Central
Loading ...
Support Center