Identifying genetic code reassignments in nucleotide sequence databases
Identifying genetic code reassignments in nucleotide sequence databases
批准号:
9907623
负责人:
Yekaterina Shulgina
金额:
$3.22万
依托单位:
依托单位国家:
美国
项目类别:
财政年份:
2019
资助国家:
美国
项目状态:
已结题
起止时间:
2019-12-16 至 2021-12-15
关键词:
AccidentsAddressAmino Acid SequenceAmino Acid Sequence DatabasesAmino AcidsBase SequenceBiologicalCandida albicansChargeCodon NucleotidesCommunicationComputer AnalysisComputing MethodologiesDNADNA SequenceDataData SetDevelopmentEnsureEnvironmentEvolutionFreezingGenbankGenesGenetic CodeGenetic ModelsGenetic VariationGenomeHomologous ProteinHumanLaboratoriesLifeMass Spectrum AnalysisMetagenomicsMethodsMitochondriaModernizationMycoplasma pneumoniaeNatureNorthern BlottingOralOrganismPeptide Sequence DeterminationProcessProteinsProteomeProteomicsResearch PersonnelResearch TrainingScienceTertiary Protein StructureTestingTrainingTransfer RNATransfer RNA AminoacylationTranslatingTranslationsTreesUpdateWritingYeastscareercomparativecomputerized toolscontigexperimental studyfollow-upgenetic predictorsgenetic variantgenome annotationin silicolaboratory equipmentmolecular sequence databasenovelopen sourcepathogenic bacteriapressureprogramsskills
中文摘要
点击翻译按钮获取中文摘要
英文摘要
Project Summary Abstract
Biological discoveries made in other organisms tell us about the functions of human
genes because of the ability to compare homologous protein sequences. Recent efforts to
sequence a greater diversity of species for comparative analysis have been primarily done on
the DNA level, and protein sequences are subsequently translated in silico assuming some
genetic code. However, there is currently no informed way of selecting the correct genetic code
for a newly sequenced organism, which is critical for the correct translation of predicted protein
sequences. As more diverse organisms are sequenced, species using variant genetic codes
continue to be found, suggesting that there may be a hidden diversity of alternative genetic
codes across the tree of life.
Aim 1 proposes building a computational tool to predict the genetic code used by an
organism from nucleotide sequence alone. This would fill in a critical missing step in genome
annotation pipelines and would ensure the accuracy of protein sequence databases, which are
predominantly composed of predicted protein sequences. In aim 2, the computational tool will
be used to infer the genetic code usage of all publicly available genomes and validate any new
genetic codes by computational analysis of tRNA genes, experimental confirmation of tRNA
expression via Northern blotting, and confirmation of altered codon translation via proteomic
mass spectrometry. In aim 3, the updated distribution of alternative genetic codes will be used
to address long-standing hypotheses in the field about how the genetic code is thought to
evolve.
This research training plan is intended to prepare the PI for a career as an independent
and interdisciplinary researcher. The training environment will be in a collaborative
computational laboratory, with access to a lab bench and shared lab equipment to do the
proposed experiments. The training plan will also include development of science
communication skills, including oral presentations and writing.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
海外基金