-
How big is Big Data?
Authors:
Daniel T. Speckhard,
Tim Bechtel,
Luca M. Ghiringhelli,
Martin Kuban,
Santiago Rigamonti,
Claudia Draxl
Abstract:
Big data has ushered in a new wave of predictive power using machine learning models. In this work, we assess what {\it big} means in the context of typical materials-science machine-learning problems. This concerns not only data volume, but also data quality and veracity as much as infrastructure issues. With selected examples, we ask (i) how models generalize to similar datasets, (ii) how high-q…
▽ More
Big data has ushered in a new wave of predictive power using machine learning models. In this work, we assess what {\it big} means in the context of typical materials-science machine-learning problems. This concerns not only data volume, but also data quality and veracity as much as infrastructure issues. With selected examples, we ask (i) how models generalize to similar datasets, (ii) how high-quality datasets can be gathered from heterogenous sources, (iii) how the feature set and complexity of a model can affect expressivity, and (iv) what infrastructure requirements are needed to create larger datasets and train models on them. In sum, we find that big data present unique challenges along very different aspects that should serve to motivate further work.
△ Less
Submitted 18 May, 2024;
originally announced May 2024.
-
ITER-IA 3D MHD Simulations of Shattered Pellet Injection(SPI) -- D1.3 Code Validation (DIII-D)
Authors:
Charlson. C. Kim,
T. Bechtel,
J. L. Herfindal,
B. C. Lyons,
Y. Q. Liu,
P. B. Parks,
L. Lao
Abstract:
This report is in partial fulfillment of deliverable D1.3 Code Validation (DIII-D). These simulations focus on thermal quench phase of the SPI mitigation and are not typically carried beyond it to the current spike and subsequent current quench. NIMROD SPI simulations[1] are validated against DIII-D experiments. The target plasma for these simulations is DIII-D 160606@02990ms.
This report is in partial fulfillment of deliverable D1.3 Code Validation (DIII-D). These simulations focus on thermal quench phase of the SPI mitigation and are not typically carried beyond it to the current spike and subsequent current quench. NIMROD SPI simulations[1] are validated against DIII-D experiments. The target plasma for these simulations is DIII-D 160606@02990ms.
△ Less
Submitted 19 October, 2023;
originally announced October 2023.
-
Band-gap regression with architecture-optimized message-passing neural networks
Authors:
Tim Bechtel,
Daniel T. Speckhard,
Jonathan Godwin,
Claudia Draxl
Abstract:
Graph-based neural networks and, specifically, message-passing neural networks (MPNNs) have shown great potential in predicting physical properties of solids. In this work, we train an MPNN to first classify materials through density functional theory data from the AFLOW database as being metallic or semiconducting/insulating. We then perform a neural-architecture search to explore the model archi…
▽ More
Graph-based neural networks and, specifically, message-passing neural networks (MPNNs) have shown great potential in predicting physical properties of solids. In this work, we train an MPNN to first classify materials through density functional theory data from the AFLOW database as being metallic or semiconducting/insulating. We then perform a neural-architecture search to explore the model architecture and hyperparameter space of MPNNs to predict the band gaps of the materials identified as non-metals. The parameters in the search include the number of message-passing steps, latent size, and activation-function, among others. The top-performing models from the search are pooled into an ensemble that significantly outperforms existing models from the literature. Uncertainty quantification is evaluated with Monte-Carlo Dropout and ensembling, with the ensemble method proving superior. The domain of applicability of the ensemble model is analyzed with respect to the crystal systems, the inclusion of a Hubbard parameter in the density functional calculations, and the atomic species building up the materials.
△ Less
Submitted 12 September, 2023;
originally announced September 2023.