Automating 3DED data processing at eBIC
Petrovic, M. D.; Owen, D.; McDonagh, D.; Hatton, D.; Bragginton, E. C.; Nunes, P.; Crawshaw, A. D.; Waterman, D. G.
Show abstract
Three-dimensional electron diffraction (3DED) is an emerging and useful technique for solving molecular structures of small and biological macro-molecules from nanometre-sized crystals. We present our automated data processing workflow for 3DED datasets collected at Diamond Light Sources electron Bio-Imaging Centre (eBIC). For this purpose, we developed a package called AutoED. The processing pipeline includes data collection, analysis of the beam position, metadata gathering, file conversion, and finally data processing using xia2 (which supports both DIALS and XDS). The processing results are captured in a summary report produced by AutoED. Our main goal is to reduce the workload of electron diffraction scientists, but also to enforce good standards already used in macromolecular crystallography (MX). All the collected 3DED datasets are automatically converted into NeXus data format which is considered a Gold Standard for MX. This standardized data format allows for all the relevant metadata about the experiment to be kept together with diffraction images. We also discuss the methods used in AutoED to determine the electron beam position on diffraction images.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Decision-making in serial crystallography: a simple test to quickly determine whether sufficient data have been collected 94%
- A simple technique to classify diffraction data from dynamic proteins according to individual polymorphs 94%
- Shake-it-off: A simple ultrasonic cryo-EM specimen preparation device 93%
Similar papers in this journal
- Cryo2StructData: A Large Labeled Cryo-EM Density Map Dataset for AI-based Modeling of Protein Structures 91%
- The Brain Image Library: A Community-Contributed Microscopy Resource for Neuroscientists 91%
- Annotating Macromolecular Complexes in the Protein Data Bank: Improving the FAIRness of Structure Data 89%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.