search

menu

  • Research Research
    • Where science meets inspired minds

    • Back
    • Research
    • Our Science
    • Research Groups
    • Facilities & Platforms
    • Clinical research
    • Find a researcher
    • Publications
    • Knowledge Transfer
  • Careers & study Careers & study
    • Become a leader in cancer research

    • Back
    • Careers & study
    • Vacancies
    • Faculty
    • Scientific staff
    • Scientific support staff
    • Postdoctoral fellows
    • PhD Students
    • Operational staff
    • Clinical fellows
    • Life in Amsterdam
    • Student internships
  • News & Events News & Events
    • Check out our stories and events

    • Back
    • News & Events
    • News
    • Media & Press
    • Calendar
  • About us About us
    • Maximum impact for cancer patients

    • Back
    • About us
    • Our vision
    • Organization
    • Collaborations
    • Responsible Research
    • Support us
    • Visit us
    • Contact us
  • Support us
Support us
  • Home
  • Publications
  • Research
  • Publications
  • Article

Predicting gene expression from DNA sequence using deep learning models.

Lucía Barbadilla-Martínez ,
Noud Klaassen ,
Bas van Steensel ,
Jeroen de Ridder

Abstract

Transcription of genes is regulated by DNA elements such as promoters and enhancers, the activity of which are in turn controlled by many transcription factors. Owing to the highly complex combinatorial logic involved, it has been difficult to construct computational models that predict gene activity from DNA sequence. Recent advances in deep learning techniques applied to data from epigenome mapping and high-throughput reporter assays have made substantial progress towards addressing this complexity. Such models can capture the regulatory grammar with remarkable accuracy and show great promise in predicting the effects of non-coding variants, uncovering detailed molecular mechanisms of gene regulation and designing synthetic regulatory elements for biotechnology. Here, we discuss the principles of these approaches, the types of training data sets that are available and the strengths and limitations of different approaches.

More about this publication

Nature reviews. Genetics

Volume 26
Issue nr. 10
Pages 666-680
Publication date 01-10-2025

Full text links

Publisher website (DOI) 10.1038/s41576-025-00841-2
Europe PubMed Central 40360798
Pubmed 40360798

Where science meets inspired minds

Contact

Plesmanlaan 121
1066CX Amsterdam

020 512 9111 communicatie@nki.nl

Quick links

  • Vacancies
  • News
  • Contact us
  • Media & Press

Follow us on

Disclaimer
Privacy statement
Cookies
Change cookie settings

This site uses cookies

This website uses cookies to ensure you get the best experience on our website.