计算机科学
任务(项目管理)
国家(计算机科学)
模拟生物系统
人工智能
计算模型
匹配(统计)
系统生物学
计算生物学
机器学习
生成模型
生成语法
计算基因组学
钥匙(锁)
流量(数学)
生物信息学
生物学数据
数据挖掘
计算复杂性理论
理论计算机科学
领域(数学)
作者
Alex Morehead,Lazar Atanackovic,Akshata Hegde,Yanli Wang,Frimpong Boadu,Joel Selvaraj,Alexander Tong,Aditi S. Krishnapriyan,Jianlin Cheng
标识
DOI:10.22541/au.175382408.89466370/v4
摘要
Numerous problems in bioinformatics and computational biology can be framed as a task of learning a mapping from one state of a biological system to another relevant state or to explore novel data points across biologically constrained spaces. However, manually deriving such mappings, e.g., to transform cells in a diseased state back into a healthy state, or extrapolating from existing datasets to create new data, is often nontrivial and can require extraordinary domain expertise and resources. Fortunately, the field of generative artificial intelligence (AI) has introduced a new training paradigm referred to as (conditional) flow matching, which has emerged as a promising solution to this problem, with broad applicability in computer vision, natural language processing, and the physical and life sciences. Flow matching is a powerful and principled, data-driven framework for efficiently learning a mapping between arbitrary pairs of high-dimensional data distributions, making it well-suited for addressing problems in molecular and cell biology. In this Review, we characterize the theoretical foundations of flow matching and its applications in biomolecular modeling for proteins, DNA/RNA, small molecules, and their interactions, as well as its uses in single/multi-cellular modeling for cell phenotyping and imaging, each contributing towards the development of an AI-based virtual cell. Lastly, this review highlights open-source flow matching methods and discusses future directions in flow-based generative modeling for bioinformatics and computational biology.
科研通智能强力驱动
Strongly Powered by AbleSci AI