已入深夜,您辛苦了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!祝你早点完成任务,早点休息,好梦!

Progress of bioinformatics studies for multi-omics and multi-modal data in complex diseases

哲学
作者
Xiaofan Liu,Zhi Lü
出处
期刊:Kexue tongbao [Science China Press]
卷期号:69 (30): 4432-4446 被引量:2
标识
DOI:10.1360/tb-2024-0416
摘要

With the advancement of high-throughput sequencing technologies, the integration of multi-omics and multi-modal data has become an important trend in the study of complex diseases. Multi-omics/multi-modal data provide new perspectives for a deeper understanding of the pathogenesis and development of diseases, offering crucial support for the precision diagnosis and treatment of complex diseases. This review first introduces various types of omics data and their contributions to complex disease research. Genomics reveals the genetic background and mutations associated with diseases by analyzing gene sequences; transcriptomics uncovers gene regulatory relationships related to diseases by studying expression patterns; proteomics focuses on the expression, modification, and interactions of proteins; metabolomics reflects adjustments in metabolic pathways before and after illness through changes in metabolites; radiomics shows disease-induced alterations via medical imaging. Integrating and analyzing these omics data can compensate for the information gaps of single omics data, enabling a more comprehensive understanding of the molecular mechanisms of complex diseases. Furthermore, this review introduces multi-omics databases related to complex diseases, covering diseases such as cancer, cardiovascular and cerebrovascular diseases, organ fibrosis, chronic kidney disease, Alzheimer’s disease, and inflammatory bowel disease. These databases facilitate researchers in obtaining and analyzing multi-omics data. Next, this review systematically categorizes the existing multi-omics integration methods into two types: correlation and network-based methods and data matrix and machine learning-based methods. These two approaches use different means to reveal the potential connections between data, thereby providing deeper insights into complex biological systems. Correlation and network-based methods involve using association analysis or complex network analysis to identify the intrinsic connections between different omics, thereby discovering biomarkers related to phenotypes. Data matrix and machine learning-based methods refer to utilizing statistical analysis, machine learning, and deep learning models to achieve data fusion for clustering or classification tasks, while revealing the inherent relationships between multi-omics data and identifying disease-related biomarkers. Data matrix and machine learning-based methods are further divided into early integration, intermediate integration, and late integration. Early integration method involves merging multi-omics data into a joint matrix and then applying machine learning or deep learning models for classification. Intermediate integration method involves modeling each omics data separately, followed by the integration of the transformed matrices or models. Late integration method independently models each omics data and then combines the model output results. Building on this, the review also discusses the applications of multi-omics integration models in complex diseases, such as disease screening, subtyping, prognosis, and drug response prediction. Finally, this review summarizes the current challenges in multi-omics/multi-modal data integration, which are divided into three levels: sample, data, and model. At the sample level, the absence of matching data limits the efficacy of integration methods, and researchers are addressing this issue through the development of data sharing and new algorithms. At the data level, the characteristics of high dimensionality, noise, and heterogeneity necessitate the use of more efficient deep learning methods for data integration. At the model level, the key challenges include lack of interpretability, low computational efficiency, and privacy concerns. Researchers are enhancing model interpretability through visualization tools and incorporating biological prior knowledge into deep learning models, while also exploring new technologies such as model acceleration and privacy-preserving computation to improve model efficiency and security. Despite the challenges, multi-omics integration has demonstrated significant potential in the diagnosis and treatment of complex diseases.

最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
隐形曼青应助小巧的电源采纳,获得10
1秒前
默默尔柳发布了新的文献求助10
1秒前
2秒前
方沅完成签到,获得积分10
3秒前
3秒前
4秒前
7秒前
因此不能发布了新的文献求助10
7秒前
9秒前
NexusExplorer应助xx采纳,获得10
10秒前
温柔的姿发布了新的文献求助10
10秒前
11秒前
kd1412完成签到 ,获得积分10
12秒前
12秒前
12秒前
kk发布了新的文献求助10
13秒前
Ava应助goo1069采纳,获得10
17秒前
zykk发布了新的文献求助10
19秒前
19秒前
伶俐的电灯胆完成签到,获得积分10
20秒前
小鱼爱吃肉应助温柔的姿采纳,获得10
21秒前
学术圈的泥石流完成签到 ,获得积分10
21秒前
Nole应助echo采纳,获得20
21秒前
CipherSage应助灝男采纳,获得10
22秒前
Akim应助ballia采纳,获得10
23秒前
23秒前
沉静弘文完成签到 ,获得积分10
23秒前
无花果应助隐形的西牛采纳,获得10
24秒前
酒药花醉鱼给酒药花醉鱼的求助进行了留言
25秒前
yn发布了新的文献求助10
25秒前
小马甲应助zykk采纳,获得10
26秒前
Xwei2051完成签到,获得积分10
26秒前
Fairy完成签到 ,获得积分10
27秒前
小费发布了新的文献求助10
27秒前
赘婿应助初景采纳,获得10
28秒前
科研通AI6.4应助1176517761采纳,获得10
28秒前
快乐小海带完成签到,获得积分10
31秒前
俭朴酒窝发布了新的文献求助10
33秒前
LihaoLin完成签到,获得积分10
34秒前
34秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Single Cell Analysis of the Tumor Microenvironment Landscape Across the Disease Spectrum of Multiple Myeloma 1000
2026年中国辛酸癸酸聚乙二醇甘油酯行业市场现状调查及投资机会研判报告 1000
2026年中国辛酸癸酸聚乙二醇甘油酯行业市场规模及竞争格局分析报告 1000
Fundamentals of Pharmaceutical and Biologics Regulations: A Global Perspective, Second Edition 700
The Cambridge History of China 英文版16册 600
作者名:Kristopher P. Plain,悉尼大学的,目前只能查到其四篇论文,想找到其博士论文 550
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7330293
求助须知:如何正确求助?哪些是违规求助? 8944609
关于积分的说明 18973640
捐赠科研通 6985347
什么是DOI,文献DOI怎么找? 3216703
关于科研通互助平台的介绍 2383293
邀请新用户注册赠送积分活动 2196276