Simplicity within biological complexity
Przulj, Natasa, Malod-Dognin, Noel
–arXiv.org Artificial Intelligence
Heterogeneous, interconnected, systems-level, molecular data have become increasingly available and key in precision medicine. We need to utilize them to better stratify patients into risk groups, discover new biomarkers and targets, repurpose known and discover new drugs to personalize medical treatment. Existing methodologies are limited and a paradigm shift is needed to achieve quantitative and qualitative breakthroughs. In this perspective paper, we survey the literature and argue for the development of a comprehensive, general framework for embedding of multi-scale molecular network data that would enable their explainable exploitation in precision medicine in linear time. Network embedding methods map nodes to points in low-dimensional space, so that proximity in the learned space reflects the network's topology-function relationships. They have recently achieved unprecedented performance on hard problems of utilizing few omic data in various biomedical applications. However, research thus far has been limited to special variants of the problems and data, with the performance depending on the underlying topology-function network biology hypotheses, the biomedical applications and evaluation metrics. The availability of multi-omic data, modern graph embedding paradigms and compute power call for a creation and training of efficient, explainable and controllable models, having no potentially dangerous, unexpected behaviour, that make a qualitative breakthrough. We propose to develop a general, comprehensive embedding framework for multi-omic network data, from models to efficient and scalable software implementation, and to apply it to biomedical informatics. It will lead to a paradigm shift in computational and biomedical understanding of data and diseases that will open up ways to solving some of the major bottlenecks in precision medicine and other domains.
arXiv.org Artificial Intelligence
May-15-2024
- Country:
- North America > United States (0.28)
- Europe
- Switzerland (0.04)
- United Kingdom > England
- Greater London > London (0.04)
- Cambridgeshire > Cambridge (0.04)
- Spain > Catalonia
- Barcelona Province > Barcelona (0.04)
- Asia
- Genre:
- Research Report (1.00)
- Industry:
- Information Technology (1.00)
- Health & Medicine
- Pharmaceuticals & Biotechnology (1.00)
- Therapeutic Area
- Oncology (1.00)
- Hematology (1.00)
- Infections and Infectious Diseases (0.67)
- Neurology (0.67)
- Technology:
- Information Technology
- Information Management (1.00)
- Biomedical Informatics > Translational Bioinformatics (1.00)
- Communications > Networks (0.87)
- Data Science
- Data Mining (1.00)
- Data Integration (0.93)
- Artificial Intelligence
- Representation & Reasoning > Information Fusion (0.93)
- Vision (0.67)
- Natural Language
- Large Language Model (1.00)
- Machine Translation (0.92)
- Text Processing (0.67)
- Machine Learning
- Statistical Learning (1.00)
- Neural Networks > Deep Learning (1.00)
- Information Technology