SOURCE-LINKED INTELLIGENCE
Natural Language Processing to learn the language of the Human Genome
ailable multi-omics data. Throughout the project we will implement techniques for interpretable learning and strategies to observe, control, and prevent ethnic biases in our approach. We expect for large language models to change how we, as a scientific field, approach genomics data analysis and expect our models to establish how these techniques can be applied efficiently, transparently, and in a bias-reduced way. In addition to general understanding of genome biology, we plan to use our models in the future for technical improvements of data analysis, population genetics, and for translational uses with applications in cancer genomics and genome editing. Natural Language Processing, Genomics, Bioinformatics, Machine Learning, Genome Instability
Read original source ↗ Open in workspace
- recordType
- award
- status
- SIGNED
- region
- EU
- value
- 173847.36
- unit
- EUR
Evidence & attribution
European Commission, CORDIS Horizon Europe project dataset. Metadata adapted.
License: CORDIS reuse policy
First collected: 2026-09-20T01:21:06.728Z. This is not the publication date.