AI Tools Predict DNA's Regulatory Role and 3D Structure

Newly developed artificial intelligence (AI) programs accurately predicted the role of DNA's regulatory elements and three-dimensional (3D) structure based solely on its raw sequence, according to two recent studies in Nature Genetics. These tools could eventually shed new light on how genetic mutations lead to disease and could lead to new understanding of how genetic sequence influences the spatial organization and function of chromosomal DNA in the nucleus, said study author Jian Zhou, Ph.D., Assistant Professor in the Lyda Hill Department of Bioinformatics at UTSW.

"Taken together, these two programs provide a more complete picture of how changes in DNA sequence, even in noncoding regions, can have dramatic effects on its spatial organization and function," said Dr. Zhou, a member of the Harold C. Simmons Comprehensive Cancer Center, a Lupe Murchison Foundation Scholar in Medical Research, and a Cancer Prevention and Research Institute of Texas (CPRIT) Scholar.

Only about 1% of human DNA encodes instructions for making proteins. Research in recent decades has shown that much of the remaining noncoding genetic material holds regulatory elements - such as promoters, enhancers, silencers, and insulators - that control how the coding DNA is expressed. How sequence controls the functions of most of these regulatory elements is not well understood, Dr. Zhou explained.

To better understand these regulatory components, he and colleagues at Princeton University and the Flatiron Institute developed a deep learning model they named Sei, which accurately sorts these snippets of noncoding DNA into 40 "sequence classes" or jobs - for example, as an enhancer for stem cell or brain cell gene activity. These 40 sequence classes, developed using nearly 22,000 data sets from previous studies studying genome regulation, cover more than 97% of the human genome. Moreover, Sei can score any sequence by its predicted activity in each of the 40 sequence classes and predict how mutations impact such activities.

By applying Sei to human genetics data, the researchers were able to characterize the regulatory architecture of 47 traits and diseases recorded in the UK Biobank database and explain how mutations in regulatory elements cause specific pathologies. Such capabilities can help gain a more systematic understanding of how genomic sequence changes are linked to diseases and other traits. The findings were published this month.

In May, Dr. Zhou reported the development of a different tool, called Orca, which predicts the 3D architecture of DNA in chromosomes based on its sequence. Using existing data sets of DNA sequences and structural data derived from previous studies that revealed the molecule’s folds, twists, and turns, Dr. Zhou trained the model to make connections and evaluated the model’s ability to predict structure at various length scales.

The findings showed that Orca predicted DNA structures both small and large based on their sequences with high accuracy, including for sequences carrying mutations associated with various health conditions including a form of leukemia and limb malformations. Orca also enabled the researchers to generate new hypotheses about how DNA sequence controls its local and large-scale 3D structure.

Dr. Zhou said that he and his colleagues plan to use Sei and Orca, which are both publicly available on web servers and as open-source code, to further explore the role of genetic mutations in causing the molecular and physical manifestations of diseases - research that could eventually lead to new ways to treat these conditions.

The Orca study was supported by grants from CPRIT (RR190071), the National Institutes of Health (DP2GM146336), and the UT Southwestern Endowed Scholars Program in Medical Science.

Zhou J.
Sequence-based modeling of three-dimensional genome architecture from kilobase to chromosome scale.
Nat Genet 54, 725-734, 2022. doi: 10.1038/s41588-022-01065-4

Most Popular Now

MEDICA 2024 + COMPAMED 2024: Adapted Hal…

11 - 14 November 2024, Düsseldorf, Germany. The final preparations for MEDICA 2024 and COMPAMED 2024 in Düsseldorf have begun. A total of more than 5,500 exhibitors from approximately 70 countries...

AI does Not Necessarily Lead to more Eff…

The use of artificial intelligence (AI) in hospitals and patient care is steadily increasing. Especially in specialist areas with a high proportion of imaging, such as radiology, AI has long...

Commission Joins Forces with Venture Cap…

The Commission has launched a Trusted Investors Network bringing together a group of investors ready to co-invest in innovative deep-tech companies in Europe together with the EU. The Union's investment...

Why the NHS is Seeking to Make Media Ser…

Opinion Article by Dean Moody, Healthcare Services Director, Airwave Healthcare. Tim Kelsey and Martha Lane Fox called for WiFi to be made available free of charge throughout the NHS back in...

An AI-Powered Pipeline for Personalized …

Ludwig Cancer Research scientists have developed a full, start-to-finish computational pipeline that integrates multiple molecular and genetic analyses of tumors and the specific molecular targets of T cells and harnesses...

Wearable Cameras Allow AI to Detect Medi…

A team of researchers says it has developed the first wearable camera system that, with the help of artificial intelligence (AI), detects potential errors in medication delivery. In a test whose...

Philips and Medtronic Advocacy Partnersh…

Royal Philips (NYSE: PHG, AEX: PHIA), a global leader in health technology, and Medtronic Neurovascular, a leading innovator in neurovascular therapies, today announced a strategic advocacy partnership. Delivering timely stroke...

AI could Transform How Hospitals Produce…

A pilot study led by researchers at University of California San Diego School of Medicine found that advanced artificial intelligence (AI) could potentially lead to easier, faster and more efficient...

New AI Tool Predicts Protein-Protein Int…

Scientists from Cleveland Clinic and Cornell University have designed a publicly-available software and web database to break down barriers to identifying key protein-protein interactions to treat with medication. The computational tool...

Great Start for Ideas and Innovations: D…

8 - 10 April 2025, Berlin, Germany. From 15 October to 15 November 2024, the DMEA invites experts from business, science, politics and practice to actively participate in shaping the congress...

Start-Ups will Once Again Have a Starrin…

11 - 14 November 2024, Düsseldorf, Germany. The finalists in the 16th Healthcare Innovation World Cup and the 13th MEDICA START-UP COMPETITION have advanced from around 550 candidates based in 62...

AI for Real-Rime, Patient-Focused Insigh…

A picture may be worth a thousand words, but still... they both have a lot of work to do to catch up to BiomedGPT. Covered recently in the prestigious journal Nature...