Data-driven Discovery of Photoactive Quaternary Oxides using First-principles Machine Learning

18 April 2019, Version 1
This content is a preprint and has not undergone peer review at the time of posting.


We present a low-cost, virtual high-throughput materials design workflow and use it to identify earth-abundant materials for solar energy applications from the quaternary oxide chemical space. A statistical model that predicts bandgap from chemical composition is built using supervised machine learning. The trained model forms the first in a hierarchy of screening steps. An ionic substitution algorithm is used to assign crystal structures, and an oxidation state probability model is used to discard unlikely chemistries. We demonstrate the utility of this process for screening over 1 million oxide compositions. We find that, despite the difficulties inherent to identifying stable multi-component inorganic materials, several compounds produced by our workflow are calculated to be thermodynamically stable or metastable and have desirable optoelectronic properties according to first-principles calculations. The predicted oxides are Li2MnSiO5, MnAg(SeO3)2 and two polymorphs of MnCdGe2O6, all four of which are found to have direct electronic bandgaps in the visible range of the solar spectrum.


Machine learning
Materials design
Materials screening
High-throughput screening

Supplementary weblinks


Comments are not moderated before they are posted, but they can be removed by the site moderators if they are found to be in contravention of our Commenting Policy [opens in a new tab] - please read this policy before you post. Comments should be used for scholarly discussion of the content in question. You can find more information about how to use the commenting feature here [opens in a new tab] .
This site is protected by reCAPTCHA and the Google Privacy Policy [opens in a new tab] and Terms of Service [opens in a new tab] apply.