Unknown

Dataset Information

0

Exploring completeness in clinical data research networks with DQe-c.


ABSTRACT: Objective:To provide an open source, interoperable, and scalable data quality assessment tool for evaluation and visualization of completeness and conformance in electronic health record (EHR) data repositories. Materials and Methods:This article describes the tool's design and architecture and gives an overview of its outputs using a sample dataset of 200?000 randomly selected patient records with an encounter since January 1, 2010, extracted from the Research Patient Data Registry (RPDR) at Partners HealthCare. All the code and instructions to run the tool and interpret its results are provided in the Supplementary Appendix. Results:DQe-c produces a web-based report that summarizes data completeness and conformance in a given EHR data repository through descriptive graphics and tables. Results from running the tool on the sample RPDR data are organized into 4 sections: load and test details, completeness test, data model conformance test, and test of missingness in key clinical indicators. Discussion:Open science, interoperability across major clinical informatics platforms, and scalability to large databases are key design considerations for DQe-c. Iterative implementation of the tool across different institutions directed us to improve the scalability and interoperability of the tool and find ways to facilitate local setup. Conclusion:EHR data quality assessment has been hampered by implementation of ad hoc processes. The architecture and implementation of DQe-c offer valuable insights for developing reproducible and scalable data science tools to assess, manage, and process data in clinical data repositories.

SUBMITTER: Estiri H 

PROVIDER: S-EPMC6481389 | biostudies-literature | 2018 Jan

REPOSITORIES: biostudies-literature

altmetric image

Publications

Exploring completeness in clinical data research networks with DQe-c.

Estiri Hossein H   Stephens Kari A KA   Klann Jeffrey G JG   Murphy Shawn N SN  

Journal of the American Medical Informatics Association : JAMIA 20180101 1


<h4>Objective</h4>To provide an open source, interoperable, and scalable data quality assessment tool for evaluation and visualization of completeness and conformance in electronic health record (EHR) data repositories.<h4>Materials and methods</h4>This article describes the tool's design and architecture and gives an overview of its outputs using a sample dataset of 200 000 randomly selected patient records with an encounter since January 1, 2010, extracted from the Research Patient Data Regist  ...[more]

Similar Datasets

| S-EPMC8059005 | biostudies-literature
| S-EPMC5509661 | biostudies-literature
| S-EPMC4078279 | biostudies-literature
| S-EPMC6026022 | biostudies-literature
| S-EPMC4078282 | biostudies-literature
| S-EPMC8121837 | biostudies-literature
| S-EPMC8498340 | biostudies-literature
| S-EPMC5517261 | biostudies-other
| S-EPMC5862238 | biostudies-literature
| S-EPMC4306391 | biostudies-literature