Similarity of Precursors in Solid-State Synthesis as Text-Mined from Scientific Literature

Similarity of Precursors in Solid-State Synthesis as Text-Mined from Scientific Literature
复制标题

DOI:
10.1021/acs.chemmater.0c02553
复制
发表时间:
2020-09-22
影响因子:
8.6
通讯作者:
Ceder, Gerbrand
Ceder, Gerbrand
中科院分区:
材料科学2区
文献类型:
--
作者:
He, Tanjin;Sun, Wenhao;Ceder, Gerbrand

文献摘要

被引文献

相似文献

收集和分析固态化学文献中可用的大量信息可能会加速我们对材料合成的理解。但是,一个主要问题是难以识别合成段的哪些材料是前体或目标材料。在这项研究中,我们基于材料实体的上下文信息开发了一个两步化学命名的实体识别模型,以识别前体和目标。使用提取的数据,我们进行了荟萃分析,以研究固态合成背景下前体之间的相似性和差异。为了量化前体相似性,我们构建了一个替代模型,以计算一个在保留目标的同时,用另一个前体替换一个前体的生存能力。从前体的分层聚类中,我们证明了可以从文本数据中提取前体的“化学相似性”。量化前体的相似性有助于为预测合成模型中的候选反应物提供基础。
Collecting and analyzing the vast amount of information available in the solid-state chemistry literature may accelerate our understanding of materials synthesis. However, one major problem is the difficulty of identifying which materials from a synthesis paragraph are precursors or are target materials. In this study, we developed a two-step chemical named entity recognition model to identify precursors and targets, based on information from the context around material entities. Using the extracted data, we conducted a meta-analysis to study the similarities and differences between precursors in the context of solid-state synthesis. To quantify precursor similarity, we built a substitution model to calculate the viability of substituting one precursor with another while retaining the target. From a hierarchical clustering of the precursors, we demonstrate that the "chemical similarity" of precursors can be extracted from text data. Quantifying the similarity of precursors helps provide a foundation for suggesting candidate reactants in a predictive synthesis model.y