课题基金 / 基金详情

Statistical Inference from Multiscale Biological Data: theory, algorithms, applications

Statistical Inference from Multiscale Biological Data: theory, algorithms, applications
多尺度生物数据的统计推断:理论、算法、应用
批准号:
EP/Y037375/1
负责人:
Luca Ferretti
金额:
$27.3万
依托单位:
依托单位国家:
英国
项目类别:
Research Grant
财政年份:
2023
资助国家:
英国
项目状态:
未结题
起止时间:
2023 至 --

项目摘要

项目成果

相似基金

相关文献

中文摘要
翻译
过去二十年见证了从基因组学到流行病学等生命科学不同领域的巨大实验突破。由于现代的高通量技术,跨多个尺度的生物系统——从单个分子到整个种群——现在可以在高空间和时间分辨率下进行定量探测。除了增强我们对系统组成的基本知识之外,这些数据潜在地编码了大量关于控制其演变的功能约束和限制其性能的物理约束的信息,以及关于组织级别、动态约束或设计原则的信息,这些信息很难从低吞吐量数据中识别出来。提取这些信息对于从具有期望功能的蛋白质设计到流行病期间接触重建等应用也至关重要。逆统计力学试图通过使用无序和随机系统的物理方法从数据中推断生成模型(玻尔兹曼分布)来做到这一点。然而,生物数据的特定特征,如严重的欠采样和异质性,限制了这些工具的有效性。SIMBAD旨在开发一类能够克服这些问题的统计推断技术。在SIMBAD中,理论工作将提供解决四个紧迫问题的概念和方法(学习蛋白质序列景观,逆建模代谢网络,从流行病学数据推断接触网络,改进生存分析模型),这反过来将指导理论与每个领域的现有标准相结合。这一努力有望为基础研究开辟新的途径,从而影响经济、技术和社会问题;SIMBAD中代表的高知名度的跨学科专业知识确保了可衡量和可实现的目标,将SIMBAD置于实现其目标的理想位置。
英文摘要
The last two decades have witnessed giant experimental breakthroughs in different areas of the life sciences, from genomics to epidemiology. Thanks to modern high-throughput techniques, biological systems across multiple scales -from single molecules up to entire populations- can now be probed quantitatively at high spatial and temporal resolutions. Besides enhancing our basic knowledge of a system's constituents, these data potentially encode a plethora of information about the functional constraints that govern its evolution and the physical constraints that limit its performance, as well as about levels of organization, dynamical constraints or design principles that would be hard to identify from low-throughput data. Extracting this information is also crucial for applications ranging from the design of proteins with a desired functionality to the reconstruction of contacts during an epidemics. Inverse statistical mechanics attempts to do it by inferring generative models (Boltzmann distributions) from data using methods from the physics of disordered and random systems. Specific characteristics of biological data however, like strong undersampling and heterogeneity, limit the effectiveness of these tools. SIMBAD aims at developing a class of statistical inference techniques capable of overcoming these issues. In SIMBAD, theoretical work will supply concepts and methods to address four pressing problems (learning protein sequence landscapes, inverse modeling metabolic networks, inferring contact networks from epidemiological data, and improving survival analysis models), which in turn will guide the theory towards integration with the existing standards of each field. This effort promises to open new pathways for basic research to impact economic, technological and societal issues; the high-profile cross-disciplinary expertise represented in SIMBAD ensures instead for measurable and achievable objectives, placing SIMBAD in an ideal position to achieve its goals.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
海外基金