Electronic medical records for genetic research: results of the eMERGE consortium.

Kho AN, Pacheco JA, Peissig PL, Rasmussen L, Newton KM, Weston N, Crane PK, Pathak J, Chute CG, Bielinski SJ, Kullo IJ, Li R, Manolio TA, Chisholm RL, Denny JC
Sci Transl Med. 2011 3 (79): 79re1

PMID: 21508311 · PMCID: PMC3690272 · DOI:10.1126/scitranslmed.3001807

Clinical data in electronic medical records (EMRs) are a potential source of longitudinal clinical data for research. The Electronic Medical Records and Genomics Network (eMERGE) investigates whether data captured through routine clinical care using EMRs can identify disease phenotypes with sufficient positive and negative predictive values for use in genome-wide association studies (GWAS). Using data from five different sets of EMRs, we have identified five disease phenotypes with positive predictive values of 73 to 98% and negative predictive values of 98 to 100%. Most EMRs captured key information (diagnoses, medications, laboratory tests) used to define phenotypes in a structured format. We identified natural language processing as an important tool to improve case identification rates. Efforts and incentives to increase the implementation of interoperable EMRs will markedly improve the availability of clinical data for genomics research.

MeSH Terms (9)

Clinical Trials as Topic Data Collection Electronic Health Records Genetic Research Genome-Wide Association Study Genomics Genomics Humans Phenotype

Connections (1)

This publication is referenced by other Labnodes entities: