Multi-modal data integration to identify kinase substrates
Multi-modal data integration to identify kinase substrates
批准号:
10451941
负责人:
Gaurav Pandey
金额:
$49.85万
依托单位国家:
美国
项目类别:
财政年份:
2022
资助国家:
美国
项目状态:
已结题
起止时间:
2022-07-05 至 2024-06-30
关键词:
AddressBindingBiochemicalBioinformaticsBiologyCell CycleCell Cycle RegulationCell physiologyCentral Nervous System DiseasesComputer softwareComputing MethodologiesConsensusDataData SourcesDevelopmentDiseaseDrug TargetingFDA approvedG-Protein-Coupled ReceptorsGene ExpressionGenetic TranscriptionGenomeGoalsHumanHuman BiologyIndividualInternetIon ChannelKnowledgeLightingMalignant NeoplasmsMetabolic DiseasesMethodologyMethodsModalityModelingNamesNeuraxisPharmacologyPhosphorylationPhosphotransferasesPhysiologicalProtein DynamicsProtein FamilyProtein KinaseProtein Kinase InteractionProteinsProteomeSignal TransductionSubstrate InteractionTestingTrainingWorkbasecomputer frameworkdata integrationdata repositorydata sharingdiverse datadrug developmentdrug discoveryexperienceflexibilityimprovedmachine learning methodmachine learning predictionmembermultidisciplinarymultimodal datanovelpredictive modelingprotein structuresmall moleculesoftware repositoryweb server
中文摘要
点击翻译按钮获取中文摘要
英文摘要
PROJECT SUMMARY
Kinases are involved in a variety of physiological functions, such as signal transduction, transcription,
development, and cell cycle regulation. Thus, dysregulation of protein kinases is associated with a range of
diseases, including cancer, metabolic diseases, and central nervous system disorders. More than 60 drugs
targeting kinases have been approved by the FDA, making them one of the most druggable protein families.
Despite their biomedical importance, a large group of human protein kinases remains highly understudied. These
proteins, often referred to as “dark kinases”, including by the Illuminating the Druggable Genome (IDG), have
limited knowledge of their substrate(s), which ultimately determine their cellular function. To address this
challenge, we will develop a novel computational framework to predict kinase-substrate interactions by
combining biologically relevant multi-modal data sources with cutting-edge machine learning methodologies.
Specifically, we will first derive features that quantify potential interactions between kinases and substrates from
diverse data sources, such as protein structure and dynamics, gene expression profiles, protein-protein and
protein-small molecule interaction networks, and evolutionary information (Aim 1). We will then develop
predictors of kinase-substrate interactions using an powerful machine learning methodology named Ensemble
Integration (EI; Aim 2). EI is based on the concept of heterogeneous ensembles that can aggregate an
unrestricted number and variety of base predictors derived from the above diverse data sources, and can benefit
from both the consensus and the diversity among these predictors. Due to its flexibility, EI is able to produce
more accurate predictions from multi-modal datasets than other established data integration methodologies, as
is expected for our project as well. Finally, we will evaluate the kinase-substrate interactions predicted by the EI-
based predictive model developed in Aim 2 using both computational and experimental methods (Aim 3). We
will also share the experimentally validated interactions, the most confident predictions from the EI model, and
all the data and software generated during this project through our KinaMetrix web server, as well as other public
data and software repositories. At its culmination, this project will produce novel and validated computational
methods and software to predict substrates of kinases, validated and high-confidence kinase-substrate
interactions for IDG dark kinases, and a public web server (KinaMetrix) to share these products. We expect that
these products will be highly useful for the study of dark kinases, especially in the IDG effort, as well as to better
understand kinase function and improve their utilization in drug development efforts. Our approach is also
expected to be generally applicable to other druggable protein families, such as ion channels and GPCRs.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Multi-modal data integration to identify kinase substrates
-
批准号:10659156
-
项目类别:
-
资助金额:$50.7万
-
财政年份:2022
-
负责人:Gaurav Pandey
-
依托单位:
Integrating genomic and clinical data to predict disease phenotypes using heterogeneous ensembles
-
批准号:10218766
-
项目类别:
-
资助金额:$54.0万
-
财政年份:2021
-
负责人:Gaurav Pandey
-
依托单位:
Integrating genomic and clinical data to predict disease phenotypes using heterogeneous ensembles
-
批准号:10589827
-
项目类别:
-
资助金额:$60.52万
-
财政年份:2021
-
负责人:Gaurav Pandey
-
依托单位:
Integrating genomic and clinical data to predict disease phenotypes using heterogeneous ensembles
-
批准号:10409755
-
项目类别:
-
资助金额:$53.55万
-
财政年份:2021
-
负责人:Gaurav Pandey
-
依托单位:
Boosting the Translational Impact of Scientific Competitions by Ensemble Learning
-
批准号:8864679
-
项目类别:
-
资助金额:$44.59万
-
财政年份:2015
-
负责人:Gaurav Pandey
-
依托单位:
国内基金
海外基金
登录
查看更多内容
帽结合蛋白(cap binding protein)调控乙烯信号转导的分子机制
-
批准号:32170319
-
项目类别:面上项目
-
资助金额:58.00万元
-
批准年份:2021
-
负责人:董春海
-
依托单位:
帽结合蛋白(cap binding protein)调控乙烯信号转导的分子机制
-
批准号:--
-
项目类别:--
-
资助金额:58万元
-
批准年份:2021
-
负责人:董春海
-
依托单位:
ID1 (Inhibitor of DNA binding 1) 在口蹄疫病毒感染中作用机制的研究
-
批准号:31672538
-
项目类别:面上项目
-
资助金额:62.0万元
-
批准年份:2016
-
负责人:孙跃峰
-
依托单位:
番茄EIN3-binding F-box蛋白2超表达诱导单性结实和果实成熟异常的机制研究
-
批准号:31372080
-
项目类别:面上项目
-
资助金额:80.0万元
-
批准年份:2013
-
负责人:杨迎伍
-
依托单位:
P53 binding protein 1 调控乳腺癌进展转移及化疗敏感性的机制研究
-
批准号:81172529
-
项目类别:面上项目
-
资助金额:58.0万元
-
批准年份:2011
-
负责人:杨其峰
-
依托单位:
DBP(Vitamin D Binding Protein)在多发性硬化中的作用和相关机制的蛋白质组学研究
-
批准号:81070952
-
项目类别:面上项目
-
资助金额:35.0万元
-
批准年份:2010
-
负责人:刘师莲
-
依托单位:
研究EB1(End-Binding protein 1)的癌基因特性及作用机制
-
批准号:30672361
-
项目类别:面上项目
-
资助金额:24.0万元
-
批准年份:2006
-
负责人:徐宁志
-
依托单位: