DeepFreak: Learning Crystallography Diffraction Patterns with Automated Machine Learning

Souza, Artur; Oliveira, Leonardo B.; Hollatz, Sabine; Feldman, Matt; Olukotun, Kunle; Holton, James M.; Cohen, Aina E.; Nardi, Luigi

Computer Science > Machine Learning

arXiv:1904.11834 (cs)

[Submitted on 26 Apr 2019 (v1), last revised 3 May 2019 (this version, v2)]

Title:DeepFreak: Learning Crystallography Diffraction Patterns with Automated Machine Learning

Authors:Artur Souza, Leonardo B. Oliveira, Sabine Hollatz, Matt Feldman, Kunle Olukotun, James M. Holton, Aina E. Cohen, Luigi Nardi

View PDF

Abstract:Serial crystallography is the field of science that studies the structure and properties of crystals via diffraction patterns. In this paper, we introduce a new serial crystallography dataset comprised of real and synthetic images; the synthetic images are generated through the use of a simulator that is both scalable and accurate. The resulting dataset is called DiffraNet, and it is composed of 25,457 512x512 grayscale labeled images. We explore several computer vision approaches for classification on DiffraNet such as standard feature extraction algorithms associated with Random Forests and Support Vector Machines but also an end-to-end CNN topology dubbed DeepFreak tailored to work on this new dataset. All implementations are publicly available and have been fine-tuned using off-the-shelf AutoML optimization tools for a fair comparison. Our best model achieves 98.5% accuracy on synthetic images and 94.51% accuracy on real images. We believe that the DiffraNet dataset and its classification methods will have in the long term a positive impact in accelerating discoveries in many disciplines, including chemistry, geology, biology, materials science, metallurgy, and physics.

Subjects:	Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (stat.ML)
Cite as:	arXiv:1904.11834 [cs.LG]
	(or arXiv:1904.11834v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1904.11834

Submission history

From: Artur Souza [view email]
[v1] Fri, 26 Apr 2019 13:12:40 UTC (760 KB)
[v2] Fri, 3 May 2019 15:11:32 UTC (760 KB)

Computer Science > Machine Learning

Title:DeepFreak: Learning Crystallography Diffraction Patterns with Automated Machine Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:DeepFreak: Learning Crystallography Diffraction Patterns with Automated Machine Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators