Machine Learning Algorithms for Actionable Knowledge Discovery in Synthetic Biology
Machine Learning Algorithms for Actionable Knowledge Discovery in Synthetic Biology
批准号:
2132169
负责人:
金额:
$0.0万
依托单位:
依托单位国家:
英国
项目类别:
Studentship
财政年份:
2018
资助国家:
英国
项目状态:
已结题
起止时间:
2018 至 --
中文摘要
点击翻译按钮获取中文摘要
英文摘要
Synthetic biology applies engineering principles to design biological systems that do not exist in the natural world so as to achieve desired properties within a given organism. This approach is of great value to society since it can be used to produce high-value materials, such as fine chemicals, pharmaceuticals, bio-remediation, bio-fuels, etc. However, the inability to predict the behaviour of biological systems largely hinders progress in bioengineering applications. While domain knowledge fails to predict the effect of genotypes changes on phenotype, the development of machine learning techniques and tremendous amounts of data generated by omics technologies have made this possible. This project thus envisions innovative computational methods to discover actionable knowledge that can be fed into synthetic biology experiments and exploit in industry. The benefits are two-folds: (1) Meaningful biological findings deduced from omics information. (2) Novel machine learning model capable of extracting high-level information from high throughput dataset.More specifically, the main biological tasks are identifying biomarkers for a particular biological state, typically related to the risk, and constructing the biological network whose nodes representing gene, proteins, metabolites and edges indicating complex relations which can be functional or regulatory. For example, the first part of this project is to look at how bacteria, which are often used as the organism to design genetic circuits in synthetic biology, adjust their transciptomics to adapt to different environmental stimuli. This is very important as bacteria almost always experience a diverse range of stresses while growing in different conditions which may affect their own growth as well as the desired properties. To our knowledge no previous research has characterised genetic changes underpinning different biological states an organism may exhibit in various conditions, nor compensatory genetic circuit to relieve the stresses has been explored. This research will learn the phenotypical landscape of bacteria growing in various conditions, identify the genes responding to general stress conditions (i.e. the biomarkers), predict the cell state by looking at the gene expressions of these biomarkers (i.e. the gene fingerprint) and ultimately answer the question of how bacteria adjust their transcriptomics to adapt to different conditions in the form of a biological network. This work can be extended to any similar questions for different purpose while the same set of routines may be replicated.An automatic pipeline of data mining techniques will be designed to extract desired information from heavy noise, high dimension omics data. As a totally data-driven research to complement with detailed mechanistic understanding in domain knowledge, statistical tests and unsupervised learning algorithms such as differential expression analysis, dimension reduction and clustering methods will first be applied to effectively tackling the data dimensionality and extract interesting data patterns, based on which supervised learnings are followed. Biomarker identification will be achieved by devising feature selection methods embedded with classier that are robust to small sample size data. While biological networks can be much more flexible, the most widely studied ones are association networks where entities are only known to be functionally connected in some way. We aim to go beyond the mere association to causation by exploiting the structure and various representations of machine learning models being used to describe the biological processes, preferably in a probabilistic way.
期刊论文(1)
专著(0)
科研奖励(0)
会议论文
DOI:
10.3390/s21072436
发表时间:
2021-04-01
期刊:
Sensors (Basel, Switzerland)
影响因子:
--
作者:
[Huang Y, Smith W, Harwood C, Wipat A, Bacardit J]
通讯作者:
Bacardit J
国内基金
海外基金
登录
查看更多内容
Scalable Learning and Optimization: High-dimensional Models and Online Decision-Making Strategies for Big Data Analysis
-
批准号:--
-
项目类别:合作创新研究团队
-
资助金额:--
-
批准年份:2024
-
负责人:姚韬
-
依托单位:
Understanding structural evolution of galaxies with machine learning
-
批准号:
-
项目类别:省市级项目
-
资助金额:10.0万元
-
批准年份:2022
-
负责人:Nicola Rosario Napolitano
-
依托单位:
煤矿安全人机混合群智感知任务的约束动态多目标Q-learning进化分配
-
批准号:--
-
项目类别:青年科学基金项目
-
资助金额:30万元
-
批准年份:2022
-
负责人:吉建娇
-
依托单位:
基于领弹失效考量的智能弹药编队短时在线Q-learning协同控制机理
-
批准号:62003314
-
项目类别:青年科学基金项目
-
资助金额:24.0万元
-
批准年份:2020
-
负责人:沈剑
-
依托单位:
集成上下文张量分解的e-learning资源推荐方法研究
-
批准号:61902016
-
项目类别:青年科学基金项目
-
资助金额:24.0万元
-
批准年份:2019
-
负责人:万珊珊
-
依托单位:
具有时序迁移能力的Spiking-Transfer learning (脉冲-迁移学习)方法研究
-
批准号:61806040
-
项目类别:青年科学基金项目
-
资助金额:20.0万元
-
批准年份:2018
-
负责人:解修蕊
-
依托单位:
基于Deep-learning的三江源区冰川监测动态识别技术研究
-
批准号:51769027
-
项目类别:地区科学基金项目
-
资助金额:38.0万元
-
批准年份:2017
-
负责人:张大奇
-
依托单位:
具有时序处理能力的Spiking-Deep Learning(脉冲深度学习)方法研究
-
批准号:61573081
-
项目类别:面上项目
-
资助金额:64.0万元
-
批准年份:2015
-
负责人:屈鸿
-
依托单位:
基于有向超图的大型个性化e-learning学习过程模型的自动生成与优化
-
批准号:61572533
-
项目类别:面上项目
-
资助金额:66.0万元
-
批准年份:2015
-
负责人:孙雪冬
-
依托单位:
E-Learning中学习者情感补偿方法的研究
-
批准号:61402392
-
项目类别:青年科学基金项目
-
资助金额:26.0万元
-
批准年份:2014
-
负责人:秦继伟
-
依托单位: