MARKOVIAN MODELS FOR PROTEIN IDENTIFICATION FROM TANDEM MASS SPECTROMETRY
MARKOVIAN MODELS FOR PROTEIN IDENTIFICATION FROM TANDEM MASS SPECTROMETRY
批准号:
8364375
负责人:
Vanathi Gopalakrishnan
金额:
$0.11万
依托单位国家:
美国
项目类别:
财政年份:
2011
资助国家:
美国
项目状态:
已结题
起止时间:
2011-09-15 至 2013-07-31
关键词:
AlgorithmsAmino Acid SequenceBiologicalBiological ProcessBiologyBiomedical ResearchComplexDataData SetDevelopmentDiagnosisDiseaseFundingFutureGrantHigh Performance ComputingLeadLearningMachine LearningMiningModelingMolecularNational Center for Research ResourcesOccupationsPatternPeptide Sequence DeterminationPrincipal InvestigatorProcessProgramming LanguagesProteinsPythonsResearchResearch InfrastructureResourcesRunningSamplingSourceTestingTrainingUnited States National Institutes of HealthWorkcostdata mininginsightmarkov modelmass spectrometernoveloutcome forecastresearch studytandem mass spectrometry
中文摘要
点击翻译按钮获取中文摘要
英文摘要
This subproject is one of many research subprojects utilizing the resources
provided by a Center grant funded by NIH/NCRR. Primary support for the subproject
and the subproject's principal investigator may have been provided by other sources,
including other NIH sources. The Total Cost listed for the subproject likely
represents the estimated amount of Center infrastructure utilized by the subproject,
not direct funding provided by the NCRR grant to the subproject or subproject staff.
Biomedical Research has been revolutionized with technological advances leading to massive accumulation of data. All this data now needs to be mined in order to draw actionable insights into the various biological processes. Complex machine learning algorithms are being developed to perform automated analyses of these large datasets and to come up with robust models that explain the observed data. Such models are then used to identify patterns in data that enable solving of challenging decision problems like diagnosis and prognosis of disease. Our research involves one such class of algorithms called Hidden Markov Models, which are used extensively in sequential data mining problems in Biology. Our particular focus is on development of novel algorithms for identification and quantification of protein sequences in complex biological samples using data that comes out of mass spectrometers. Such analysis will lead to molecular characterization of target conditions like diseased states. Our algorithms involve learning models from large training datasets and are computationally intensive. Additionally, in order to learn a robust model that will perform well across a variety of future test data, we are proposing to perform large-scale experiments with different model topologies and features, and require learning of hundreds of different models worth many days of number-crunching work. However, the entire experimentation can be parallelized trivially since all the models can be learned independently from each other and hence, the need for computing machines that can run multiple jobs in parallel. Our algorithms (homegrown) have been implemented using Python programming language and can take advantage of presence of multiple processing units or cores. After speaking with consultants at PSC, we were suggested that the Blacklight machines are most suitable for our needs.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Transfer Rule Learning for Knowledge Based Biomarker Discovery and Predictive Bio
-
批准号:8711497
-
项目类别:
-
资助金额:$30.22万
-
财政年份:2012
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Transfer Rule Learning with Functional Mapping for Integrative Modeling of Panomics Data
-
批准号:9246538
-
项目类别:
-
资助金额:$28.9万
-
财政年份:2012
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Transfer Rule Learning with Functional Mapping for Integrative Modeling of Panomics Data
-
批准号:9111473
-
项目类别:
-
资助金额:$29.6万
-
财政年份:2012
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Transfer Rule Learning for Knowledge Based Biomarker Discovery and Predictive Bio
-
批准号:8549840
-
项目类别:
-
资助金额:$28.93万
-
财政年份:2012
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Transfer Rule Learning for Knowledge Based Biomarker Discovery and Predictive Bio
-
批准号:8373065
-
项目类别:
-
资助金额:$29.97万
-
财政年份:2012
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Bayesian Rule Learning Methods for Disease Prediction and Biomarker Discovery
-
批准号:8318619
-
项目类别:
-
资助金额:$46.61万
-
财政年份:2011
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Bayesian Rule Learning Methods for Disease Prediction and Biomarker Discovery
-
批准号:8024941
-
项目类别:
-
资助金额:$28.25万
-
财政年份:2011
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Intelligent Aids for Proteomic Data Mining
-
批准号:7089794
-
项目类别:
-
资助金额:$12.74万
-
财政年份:2004
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Intelligent Aids for Proteomic Data Mining
-
批准号:6811846
-
项目类别:
-
资助金额:$12.27万
-
财政年份:2004
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Intelligent Aids for Proteomic Data Mining
-
批准号:7460715
-
项目类别:
-
资助金额:$13.27万
-
财政年份:2004
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Intelligent Aids for Proteomic Data Mining
-
批准号:6915489
-
项目类别:
-
资助金额:$12.39万
-
财政年份:2004
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
Intelligent Aids for Proteomic Data Mining
-
批准号:7254755
-
项目类别:
-
资助金额:$13.01万
-
财政年份:2004
-
负责人:Vanathi Gopalakrishnan
-
依托单位:
海外基金