Unifying Templates, Ontologies and Tools to Achieve Effective Annotation of Bioassay Protocols
Unifying Templates, Ontologies and Tools to Achieve Effective Annotation of Bioassay Protocols
批准号:
9398728
负责人:
BARRY A BUNIN
金额:
$54.64万
依托单位国家:
美国
项目类别:
财政年份:
2017
资助国家:
美国
项目状态:
已结题
起止时间:
2017-08-01 至 2021-07-31
关键词:
AcademiaAddressAdoptedAdoptionAreaBig DataBiological AssayBiomedical ResearchChemicalsCommunicationCommunitiesCompetenceComplexComputer softwareComputersControlled VocabularyCustomDataData SetData Storage and RetrievalDevelopmentEcosystemEffectivenessElementsEnsureExerciseFAIR principlesFeedbackFoundationsHourJournalsLearningLibrariansMachine LearningManualsMapsMetadataMethodsOntologyOutputParticipantPharmaceutical PreparationsPolishesProblem SolvingProcessPropertyProtocols documentationPubChemPublishingReadabilityResearchResearch PersonnelRetrievalRiskScienceScientistSemanticsSiteSoftware EngineeringSoftware ToolsSpecialistSpecific qualifier valueStandardizationStructureSuggestionSystemTechnologyTestingTextTimeTranslatingTweensUpdateVocabularyWorkbasecost effectivedata modelingdesigndrug discoverydrug mechanismexperienceexperimental studyimprovedimproved functioningin vivoinformatics trainingnovelopen sourcepractical applicationpredictive modelingrepositorytooluser-friendly
中文摘要
点击翻译按钮获取中文摘要
英文摘要
Project Summary
Biological assays are the foundation for developing chemical probes and drugs, but new Big Data approaches
– which have revolutionized other areas of biomedical science – have not yet advanced this early step of
biomedical research: analysis of assay data. The obstacle is that scientists specify their assays through text
descriptions written in scientific English, which need to be translated into standardized annotations readable by
computers. This lack of standardized and machine-readable assay descriptions is a major impediment to
manage, find, aggregate, compare, re-use, and learn from the ever-growing corpus of assays (e.g., >1.2
million in PubChem). Thus, there is a critical need for better annotation and curation tools for drug discovery
assays. However, the process to go from a simple text protocol to highly detailed machine-readable semantic
annotations is not trivial. Multiple tools and technologies are required: ontologies or the structured controlled
vocabularies; templates that map specific vocabularies to properties that are to be captured; and software tools
to actually apply these ontologies to a given text. Currently, each of these exists in isolation; yet, a bottleneck
in any one tool or technology, or a gap between the different pieces, disrupts the overall process, resulting in
poor or no annotation of the datasets. Here we propose a project to combine and integrate these three
technologies (which are also the core competencies of the three groups collaborating on this proposal). We
will deliver a novel, comprehensive, user-friendly data annotation and curation system that is highly
interconnected, encompassing the full cycle, and real-world practice, of required tasks and decisions, by all
parties within the `bioassay annotation ecosystem' (researchers performing curation, dedicated curators, IT
specialists, ontology owners, and librarians/repositories). The alliance between academic and commercial
collaborators, who already work together, will greatly benefit the project and minimize execution risk. Our
specific aims are to: (1) Develop a bioassay-specific template editor and templates by adopting the Stanford
(Center for Expanded Data Annotation and Retrieval, CEDAR) data model to the machine learning-based
curation tool BioAssay Express, to exploit the broad functionality of its data structures, tools and interfaces; (2)
Define and create an ontology update process and tool (`OntoloBridge') to support rapid feedback between
curators/users and ontology experts and enable semi-automated incorporation of suggestions for updates to
existing published ontologies; (3) Develop new tools to export annotated data into public repositories such as
PubChem; and (4) Evaluate our solution across diverse audiences (pharma, academia, repositories). The
system will improve bioassay curation efficiency, quality, and effectiveness, enabling scientists to generate
standardized annotations for their experiments to make these data FAIR (Findable, Accessible, Interoperable,
Reusable). We envision this suite of tools will encourage annotation earlier in the data lifecycle while still
supporting annotation at later stages (e.g., submission to repositories or to journals).
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Virtual Approaches to New Chemistries
-
批准号:10447249
-
项目类别:
-
资助金额:$44.0万
-
财政年份:2022
-
负责人:BARRY A BUNIN
-
依托单位:
Virtual Approaches to New Chemistries
-
批准号:10636882
-
项目类别:
-
资助金额:$44.0万
-
财政年份:2022
-
负责人:BARRY A BUNIN
-
依托单位:
Automated Molecular Identity Disambiguator (AutoMID)
-
批准号:10357906
-
项目类别:
-
资助金额:$28.0万
-
财政年份:2020
-
负责人:BARRY A BUNIN
-
依托单位:
Automated Molecular Identity Disambiguator (AutoMID)
-
批准号:10569639
-
项目类别:
-
资助金额:$28.0万
-
财政年份:2020
-
负责人:BARRY A BUNIN
-
依托单位:
Intelligent Chemical Structure Browser for Drug Discovery and Optimization
-
批准号:10241834
-
项目类别:
-
资助金额:$72.73万
-
财政年份:2019
-
负责人:BARRY A BUNIN
-
依托单位:
A Robust, Secure Framework to Effortlessly Bind Distributed Databases and Analysis Tools into Tightly Integrated Translational Drug Discovery Computational Platforms
-
批准号:10484172
-
项目类别:
-
资助金额:$85.49万
-
财政年份:2019
-
负责人:BARRY A BUNIN
-
依托单位:
Digital representation of chemical mixtures to aid drug discovery and formulation
-
批准号:9902210
-
项目类别:
-
资助金额:$74.87万
-
财政年份:2019
-
负责人:BARRY A BUNIN
-
依托单位:
A Robust, Secure Framework to Effortlessly Bind Distributed Databases and Analysis Tools into Tightly Integrated Translational Drug Discovery Computational Platforms
-
批准号:10685358
-
项目类别:
-
资助金额:$85.49万
-
财政年份:2019
-
负责人:BARRY A BUNIN
-
依托单位:
Intelligent Chemical Structure Browser for Drug Discovery and Optimization
-
批准号:10386918
-
项目类别:
-
资助金额:$72.73万
-
财政年份:2019
-
负责人:BARRY A BUNIN
-
依托单位:
Novel deep learning strategy to better predict pharmacological properties of candidate drugs and focus discovery efforts
-
批准号:10133177
-
项目类别:
-
资助金额:$74.99万
-
财政年份:2018
-
负责人:BARRY A BUNIN
-
依托单位:
Novel deep learning strategy to better predict pharmacological properties of candidate drugs and focus discovery efforts
-
批准号:10004481
-
项目类别:
-
资助金额:$74.99万
-
财政年份:2018
-
负责人:BARRY A BUNIN
-
依托单位:
Comprehensive but simple encoding of bioassays to accelerate translational drug discovery
-
批准号:9464228
-
项目类别:
-
资助金额:$74.43万
-
财政年份:2017
-
负责人:BARRY A BUNIN
-
依托单位:
Unifying Templates, Ontologies and Tools to Achieve Effective Annotation of Bioassay Protocols
-
批准号:9979969
-
项目类别:
-
资助金额:$51.14万
-
财政年份:2017
-
负责人:BARRY A BUNIN
-
依托单位:
Simplifying encoding of bioassays to accelerate translational drug discovery
-
批准号:8901698
-
项目类别:
-
资助金额:$75.14万
-
财政年份:2013
-
负责人:BARRY A BUNIN
-
依托单位:
Simplifying encoding of bioassays to accelerate translational drug discovery
-
批准号:8591013
-
项目类别:
-
资助金额:$15.0万
-
财政年份:2013
-
负责人:BARRY A BUNIN
-
依托单位:
Biocomputation across distributed private datasets to enhance drug discovery
-
批准号:9345057
-
项目类别:
-
资助金额:$75.05万
-
财政年份:2013
-
负责人:BARRY A BUNIN
-
依托单位:
Biocomputation across distributed private datasets to enhance drug discovery
-
批准号:8198305
-
项目类别:
-
资助金额:$15.0万
-
财政年份:2011
-
负责人:BARRY A BUNIN
-
依托单位:
海外基金