Researchers Identify Healthcare Data Defects

Researchers at the University of Maryland, Baltimore County (UMBC) have developed a method to investigate the quality of healthcare data using a systematic approach, which is based on creating a taxonomy for data defects thorough literature review and examination of data. Using that taxonomy, the researchers developed software that automatically detects data defects effectively and efficiently.

The research is published in the Journal of the American Medical Informatics Association (JAMIA), and is led by Günes Koru, FAMIA, professor of information systems, and Yili Zhang, a former graduate student in Koru's lab who is now a postdoctoral fellow at Northwestern University. The paper stresses that the prevalence of defects in some of the existing healthcare data can be quite high. This must be addressed to better leverage the data to improve the quality of care, reduce costs, and achieve better healthcare outcomes. The team collaborated with an anonymous healthcare organization using real healthcare datasets.

Though many researchers today are involved in the analysis of healthcare data and are invested in its importance, there is very little research being done on the quality of the data being analyzed. Ultimately, this creates a far-reaching problem because important findings from the data may be less meaningful than assumed unless significant effort and money can be invested to deal with data quality problems with ad-hoc methods. For instance, much of the data that Koru's team analyzed contained errors of duplication, mismatched formatting and incorrect syntax.

Identifying these defects in healthcare data is deeply important when it comes to healthcare facilities providing essential services. Koru explains how healthcare facilities use the data collected. Healthcare organizations must "improve upon their services based on that data, and collect more data. If we can keep this cycle going, we can actually learn and improve more quickly, which is the main idea behind the concept of Learning Health Systems, and doing so is all the more important in the COVID-19 era," he says.

In the last decade, healthcare providers in the U.S. made a large leap from keeping patient records on paper to containing all patient information in computerized databases. This jump is significant because of the opportunity it provides for analysis, but researchers are still trying to learn how to effectively leverage the data as an asset.

Koru positions his team's research on data quality as being between the fields that are working to leverage data and the fields that are working to generate it. If the data itself--the bridge that connects the two fields - contains many inconsistencies and problems, then the relevant information cannot be used to provide better outcomes for patients and facilities.

In the future, Koru will continue to work with the partner facility's healthcare professionals to build a path forward. He will collaborate further to improve the quality of data and sustain an operation that bases much of its success on the data that it can gather from health services. His team will work with healthcare administration professionals when the software tools developed through this research are adopted in organizational settings to ensure the usability and usefulness of the tools.

"The taxonomy will help data stewards to identify, understand, and manage potential data quality problems in their future work," says Zhang.

Now more than ever, healthcare facilities are relying on strong data to support patients and the healthcare field as a whole. Koru and Zhang have found that collaborations between data researchers and healthcare organizations can generate effective solutions to the problem of data quality improvement.

Yili Zhang, Güneş Koru.
Understanding and detecting defects in healthcare administration data: Toward higher data quality to better support healthcare operations and decisions.
Journal of the American Medical Informatics Association, March 2020. doi: 10.1093/jamia/ocz201

Most Popular Now

AI Tool Offers Deep Insight into the Imm…

Researchers explore the human immune system by looking at the active components, namely the various genes and cells involved. But there is a broad range of these, and observations necessarily...

Do Fitness Apps do More Harm than Good?

A study published in the British Journal of Health Psychology reveals the negative behavioral and psychological consequences of commercial fitness apps reported by users on social media. These impacts may...

AI Tool Beats Humans at Detecting Parasi…

Scientists at ARUP Laboratories have developed an artificial intelligence (AI) tool that detects intestinal parasites in stool samples more quickly and accurately than traditional methods, potentially transforming how labs diagnose...

Making Cancer Vaccines More Personal

In a new study, University of Arizona researchers created a model for cutaneous squamous cell carcinoma, a type of skin cancer, and identified two mutated tumor proteins, or neoantigens, that...

AI, Health, and Health Care Today and To…

Artificial intelligence (AI) carries promise and uncertainty for clinicians, patients, and health systems. This JAMA Summit Report presents expert perspectives on the opportunities, risks, and challenges of AI in health...

AI can Better Predict Future Risk for He…

A landmark study led by University' experts has shown that artificial intelligence can better predict how doctors should treat patients following a heart attack. The study, conducted by an international...

AI System Finds Crucial Clues for Diagno…

Doctors often must make critical decisions in minutes, relying on incomplete information. While electronic health records contain vast amounts of patient data, much of it remains difficult to interpret quickly...

Improved Cough-Detection Tech can Help w…

Researchers have improved the ability of wearable health devices to accurately detect when a patient is coughing, making it easier to monitor chronic health conditions and predict health risks such...

A New AI Model Improves the Prediction o…

Breast cancer is the most commonly diagnosed form of cancer in the world among women, with more than 2.3 million cases a year, and continues to be one of the...

Multimodal AI Poised to Revolutionize Ca…

Although artificial intelligence (AI) has already shown promise in cardiovascular medicine, most existing tools analyze only one type of data - such as electrocardiograms or cardiac images - limiting their...

New AI Tool Makes Medical Imaging Proces…

When doctors analyze a medical scan of an organ or area in the body, each part of the image has to be assigned an anatomical label. If the brain is...