4.7 Article Data Paper

Dataset of solution-based inorganic materials synthesis procedures extracted from the scientific literature

Journal

SCIENTIFIC DATA
Volume 9, Issue 1, Pages -

Publisher

NATURE PORTFOLIO
DOI: 10.1038/s41597-022-01317-2

Keywords

-

Funding

  1. National Science Foundation [DMR-1922372]

Ask authors/readers for more resources

This study utilizes advanced machine learning and natural language processing techniques to construct a large-scale dataset of solution-based inorganic materials synthesis procedures. This dataset includes 35,675 synthesis procedures extracted from scientific literature. By utilizing this dataset, one can learn synthesis patterns and predict the synthesis of novel materials.
The development of a materials synthesis route is usually based on heuristics and experience. A possible new approach would be to apply data-driven approaches to learn the patterns of synthesis from past experience and use them to predict the syntheses of novel materials. However, this route is impeded by the lack of a large-scale database of synthesis formulations. In this work, we applied advanced machine learning and natural language processing techniques to construct a dataset of 35,675 solution-based synthesis procedures extracted from the scientific literature. Each procedure contains essential synthesis information including the precursors and target materials, their quantities, and the synthesis actions and corresponding attributes. Every procedure is also augmented with the reaction formula. Through this work, we are making freely available the first large dataset of solution-based inorganic materials synthesis procedures.

Authors

I am an author on this paper
Click your name to claim this paper and add it to your profile.

Reviews

Primary Rating

4.7
Not enough ratings

Secondary Ratings

Novelty
-
Significance
-
Scientific rigor
-
Rate this paper

Recommended

No Data Available
No Data Available