Researchers Identify Healthcare Data Defects

Researchers at the University of Maryland, Baltimore County (UMBC) have developed a method to investigate the quality of healthcare data using a systematic approach, which is based on creating a taxonomy for data defects thorough literature review and examination of data. Using that taxonomy, the researchers developed software that automatically detects data defects effectively and efficiently.

The research is published in the Journal of the American Medical Informatics Association (JAMIA), and is led by Günes Koru, FAMIA, professor of information systems, and Yili Zhang, a former graduate student in Koru's lab who is now a postdoctoral fellow at Northwestern University. The paper stresses that the prevalence of defects in some of the existing healthcare data can be quite high. This must be addressed to better leverage the data to improve the quality of care, reduce costs, and achieve better healthcare outcomes. The team collaborated with an anonymous healthcare organization using real healthcare datasets.

Though many researchers today are involved in the analysis of healthcare data and are invested in its importance, there is very little research being done on the quality of the data being analyzed. Ultimately, this creates a far-reaching problem because important findings from the data may be less meaningful than assumed unless significant effort and money can be invested to deal with data quality problems with ad-hoc methods. For instance, much of the data that Koru's team analyzed contained errors of duplication, mismatched formatting and incorrect syntax.

Identifying these defects in healthcare data is deeply important when it comes to healthcare facilities providing essential services. Koru explains how healthcare facilities use the data collected. Healthcare organizations must "improve upon their services based on that data, and collect more data. If we can keep this cycle going, we can actually learn and improve more quickly, which is the main idea behind the concept of Learning Health Systems, and doing so is all the more important in the COVID-19 era," he says.

In the last decade, healthcare providers in the U.S. made a large leap from keeping patient records on paper to containing all patient information in computerized databases. This jump is significant because of the opportunity it provides for analysis, but researchers are still trying to learn how to effectively leverage the data as an asset.

Koru positions his team's research on data quality as being between the fields that are working to leverage data and the fields that are working to generate it. If the data itself--the bridge that connects the two fields - contains many inconsistencies and problems, then the relevant information cannot be used to provide better outcomes for patients and facilities.

In the future, Koru will continue to work with the partner facility's healthcare professionals to build a path forward. He will collaborate further to improve the quality of data and sustain an operation that bases much of its success on the data that it can gather from health services. His team will work with healthcare administration professionals when the software tools developed through this research are adopted in organizational settings to ensure the usability and usefulness of the tools.

"The taxonomy will help data stewards to identify, understand, and manage potential data quality problems in their future work," says Zhang.

Now more than ever, healthcare facilities are relying on strong data to support patients and the healthcare field as a whole. Koru and Zhang have found that collaborations between data researchers and healthcare organizations can generate effective solutions to the problem of data quality improvement.

Yili Zhang, Güneş Koru.
Understanding and detecting defects in healthcare administration data: Toward higher data quality to better support healthcare operations and decisions.
Journal of the American Medical Informatics Association, March 2020. doi: 10.1093/jamia/ocz201

Most Popular Now

Using Data and AI to Create Better Healt…

Academic medical centers could transform patient care by adopting principles from learning health systems principles, according to researchers from Weill Cornell Medicine and the University of California, San Diego. In...

AI Medical Receptionist Modernizing Doct…

A virtual medical receptionist named "Cassie," developed through research at Texas A&M University, is transforming the way patients interact with health care providers. Cassie is a digital-human assistant created by Humanate...

Northern Ireland Completes Nationwide Ro…

Go-lives at Western and Southern health and social care trusts mean every pathology service is using the same laboratory information management system; improving efficiency and quality. An ambitious technology project to...

AI Tool Set to Transform Characterisatio…

A multinational team of researchers, co-led by the Garvan Institute of Medical Research, has developed and tested a new AI tool to better characterise the diversity of individual cells within...

AI Detects Hidden Heart Disease Using Ex…

Mass General Brigham researchers have developed a new AI tool in collaboration with the United States Department of Veterans Affairs (VA) to probe through previously collected CT scans and identify...

Human-AI Collectives Make the Most Accur…

Diagnostic errors are among the most serious problems in everyday medical practice. AI systems - especially large language models (LLMs) like ChatGPT-4, Gemini, or Claude 3 - offer new ways...

MHP-Net: A Revolutionary AI Model for Ac…

Liver cancer is the sixth most common cancer globally and a leading cause of cancer-related deaths. Accurate segmentation of liver tumors is a crucial step for the management of the...

Highland Marketing Announced as Official…

Highland Marketing has been named, for the second year running, the official communications partner for HETT Show 2025, the UK's leading digital health conference and exhibition. Taking place 7-8 October...

Groundbreaking TACIT Algorithm Offers Ne…

Researchers at VCU Massey Comprehensive Cancer Center have developed a novel algorithm that could provide a revolutionary tool for determining the best options for patients - both in the treatment...

The Many Ways that AI Enters Rheumatolog…

High-resolution computed tomography (HRCT) is the standard to diagnose and assess progression in interstitial lung disease (ILD), a key feature in systemic sclerosis (SSc). But AI-assisted interpretation has the potential...