Search NASA⌕ Search

SEARCH · Search NASA

Results for “Model Counting”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

Meta2DB: Curated Shotgun Metagenomic Feature Sets and Metadata for Health State Prediction

Meta2DB is a curated metagenomic and metadata database that provides structurally consistent microbiome taxonomy feature count tables for 13 897 samples across 84 studies, 23 disease states, and 34 geographical locations. All samples were uniformly processed using a streamlined metagenomic classification pipeline that employs a unique and comprehensive reference database indexed to contain all sequences across all kingdoms of life that were present in the NCBI Nucleotide (nt) database retrieved on 4 January 2023. This pipeline leverages high-performance computing (HPC) resources at Lawrence Livermore National Laboratory and was used to process 50TB of publicly available raw metagenomic sequence data. Extensive metadata curation was carried out through a combination of manual curation and automated parsing, producing a consistent inter-study metadata table specifically structured to facilitate training of ML models for prediction of human health.

Kok, C [Lawrence Livermore National Laboratory (LL↗

tite

The TITE library provides type erasure implementation utilizing the tag_invoke paradigm proposed for standardization here: https://open-std.org/JTC1/SC22/WG21/docs/papers/2019/p1895r0.pdf. The implementation contained herein is largely modeled after that provided in the standardization proposal and available at https://godbolt.org/z/3TvO4f. Significant modification have been made to the original implementation to improve its suitability to be utilized for GPU architectures. In particular the implementation: - only requires a C++14 standard. - has been extended to obtain vtables for GPU device architectures in addition to CPU host architectures. - provides a gpu_allocator class appropriate for allocation of the type-erased object to GPU memory - provides copy semantics omitted from the original implementation-- - generally the copy semantics are to completely copy the type-erased object - when the usage of the type-erased object guarantees immutability the copy semantics are altered to reference-counted shallow copies (copies of pointers) for improved performance

Solomon, CJ↗

Conservation laws and effective hadronization models

Hadronization models based on local string-breaking dynamics are typically Markovian by construction, yet the physical ensemble of final states is shaped by global constraints that couple the entire fragmentation trajectory. Recasting hadronization as a conditioned stochastic diffusion process provides a precise mathematical resolution to this tension. In particular, this language reveals explicitly that constraints stemming from conservation laws induce non-Markovian correlations between otherwise independent fragmentation steps, and that these correlations can be absorbed exactly into a renormalization of the local dynamics through a Doob $h$-transform. We develop this formalism for a $q\bar{q}$ string in the chiral limit, where the longitudinal-transverse factorization of the Lund kernel becomes exact, enabling systematic power counting and clean ultraviolet (UV)/infrared (IR) separation. The dynamics organize naturally into a tower of effective theories distinguished by the remaining string mass, spanning a UV fixed point with scale-invariant transport coefficients, an intermediate regime where transverse phase space induces controlled running, and an IR boundary layer where non-local effects enter at leading order. The tower exhibits genuine Wilsonian structure, including $β$-functions, anomalous dimensions, and systematic matching conditions. The resulting framework achieves a clean factorization of universal microscopic fragmentation dynamics from infrared constraint effects, and opens new directions for both the theoretical analysis and practical simulation of hadronization.

Menzo, Tony [Alabama U.; Fermilab] (ORCID:00000002↗

High-resolution lidar observations of sedimentation-induced size sorting of droplets near a laboratory cloud top

Cloud optical properties and precipitation, which are crucial to weather and climate, are strongly influenced by cloud microphysical properties that are still poorly understood. Here, we develop a high-resolution time-correlated single-photon-counting lidar and apply it to observe cloud microphysical properties at one-centimeter range resolution in a convection chamber under well-controlled conditions. Together with concurrent in-situ measurements and theoretical analysis, our lidar observations indicate that although turbulent mixing tends to homogenize the cloud in the bulk region, entrainment and sedimentation cause inhomogeneities in droplet concentrations near the cloud top. Specifically, the topmost region is directly affected by entrainment, and lidar profiles show clear evidence of entrained air and detrained cloud filament. The transition region below exhibits vertical size sorting of cloud droplets caused by sedimentation. Our results suggest that using a single sedimentation velocity for all cloud droplets, as is done in many atmospheric models, overlooks key physics relevant to the microphysical structure near the cloud top. In conclusion, our conceptual model used to describe these measurements can serve as a step toward improving the current modeling of processes in the cloud top region.

54 ENVIRONMENTAL SCIENCES↗

Author Correction: US oil and gas system emissions from nearly one million aerial site measurements

Correction to: Naturehttps://doi.org/10.1038/s41586-024-07117-5 Published online 13 March 2024 In the version of the article initially published, several errors were present and have been corrected in the HTML and PDF versions of the article and Supplementary Information. The main results, conclusions, and our interpretations of the data remain unchanged. See the new Supplementary Information Section S15 for a more detailed description of the errors corrected and the resulting effects on the analysis. Data processing and methods corrections Overflight count correction: We previously used pre-computed source coverage data for some Carbon Mapper campaigns that was computed differently than was required for our analysis. We have re-computed Carbon Mapper source coverage based on flightline polygons and source coordinates. Transition point computation, well sites: The updated version now correctly compares the cumulative emissions distribution of simulated well site emissions with that of aerially detected sources (rather than plumes) when computing the transition point. Transition point computation, midstream: Additionally, the transition point calculation has been corrected to exclude aerially detected midstream emissions below the transition point, which was previously leading to double counting of these emissions. This error was not present for upstream (well site) emissions. Calculation errors Unit error: We corrected a specific unit conversion error affecting well site emissions in the Kairos Fort Worth dataset. Across all datasets, we also correct the conversion factor for converting from standard volume to mass for midstream emissions. Sorting error: We correct code that was applying incorrect sorting when computing correction factors to account for partial detection at well sites. Small typographical corrections were made in Fig. 1b and SI Section S4.1. Data processing and methods corrections Overflight count correction: We previously used pre-computed source coverage data for some Carbon Mapper campaigns that was computed differently than was required for our analysis. We have re-computed Carbon Mapper source coverage based on flightline polygons and source coordinates. Transition point computation, well sites: The updated version now correctly compares the cumulative emissions distribution of simulated well site emissions with that of aerially detected sources (rather than plumes) when computing the transition point. Transition point computation, midstream: Additionally, the transition point calculation has been corrected to exclude aerially detected midstream emissions below the transition point, which was previously leading to double counting of these emissions. This error was not present for upstream (well site) emissions. Calculation errors Unit error: We corrected a specific unit conversion error affecting well site emissions in the Kairos Fort Worth dataset. Across all datasets, we also correct the conversion factor for converting from standard volume to mass for midstream emissions. Sorting error: We correct code that was applying incorrect sorting when computing correction factors to account for partial detection at well sites. Small typographical corrections were made in Fig. 1b and SI Section S4.1. The following practices may help researchers conducting similar analyses avoid making similar errors: 1, Clear, accessible documentation explaining the interpretation of all columns in data input tables and all internal variables within the model, 2, Simple cross-check calculations computed before and after unit conversions.

Sherwin, Evan D↗

Toward an AI-Powered Software Pipeline for Real-Time Tracking and Analysis of Wildfire and Smoke

Real-time tracking of wildfires and smoke is crucial for effective response, minimizing damage, protecting lives, and efficiently managing resources during fire emergencies. We develop a web-based AI-powered pipeline that detects wildfires in aerial video and estimates deployment-relevant behavior metrics, including cumulative burned area, burned-area growth rate, fire spread direction, and smoke dispersion. The system combines a YOLO-based detector with YCbCr-based fire segmentation, HSV-based smoke segmentation, Farneback optical flow, and centroid-based spatiotemporal tracking. Using ground sampling distance (GSD), pixel-level fire masks are converted to physical burned-area measurements by correlating fire pixel counts with camera altitude and tilt angle. We benchmark YOLO variants and non-YOLO baselines (GoogLeNet, CNN, DBN, Autoencoder, U-Net, and AlexNet) on the IEEE FLAME dataset and a newly created aerial frame dataset, Wildfire-DB. Cross-dataset evaluation uses a strict threshold-transfer protocol: decision thresholds are selected on FLAME validation and transferred unchanged to Wildfire-DB to quantify generalization under domain shift. YOLOv6 achieves the strongest cross-dataset frame-level fire detection on Wildfire-DB (ROC-AUC 0.8200, PR-AUC 0.8044, and transferred-threshold F1 0.7596). For tracking-oriented deployment requiring oriented localization, YOLO11-OBB provides the most reliable cross-dataset behavior among OBB-capable models while remaining computationally feasible. To analyze the feasibility of UAV deployment, we further measure inference efficiency using synchronized GPU and CPU power logs on a fixed workload of 1569 frames. YOLO-family models process the video in 5.73–12.47 seconds with net energy of 1247.28–1775.39 J, substantially lower latency and energy than heavier classification and reconstruction baselines. Overall, model optimality depends on operational objectives: YOLOv6 is best for cross-dataset detection robustness, whereas YOL...

Color segmentation↗

Denudation, solute export, landscape evolution modeling, and geographic information system data for the East River watershed, Colorado, USA (2020-2024)

This data package contains geographic information system (GIS) layers and tabular datasets associated with the study of lithologic controls on denudation, solute export, carbon-scaling relationships, and transient landscape evolution in the East River watershed near Crested Butte, Colorado, USA. The package includes GIS layers used to produce the Figure 2 map, including drainage, hillshade, lithology, sample locations, and basin polygons, together with comma-separated value (CSV) tables and matching CSV data dictionaries. One group of tables reports sample-level and catchment-level information for river-sediment samples analyzed for in situ-produced cosmogenic beryllium-10 (10Be), including sample names, outlet elevations, geographic coordinates, upstream drainage area, rock-type classes, production-rate scaling scheme, analyzed nuclide, catchment-averaged denudation rates, and associated lower and upper analytical uncertainties. Sample and catchment attributes provide the basis for comparing denudation rates across intrusive, shale, sedimentary, and mixed-lithology settings. A second group of tables reports supporting information for landscape-evolution modeling and the mapped geologic framework of the study area. Included files list parameter values and definitions for the two-phase landscape-evolution simulations, summarize full-domain model erosion fluxes and topographic metrics for different simulation configurations, provide a fixed-area carbon-model scaling table, and summarize mapped geologic units within the East River study domain, including geologic code, formation name, lithologic description, mapped area, and lithologic class grouping. Model outputs and geologic summaries support interpretation of transient landscape behavior and its relation to the mapped distribution of shale, intrusive, sedimentary, and surficial units. A third group of tables reports hydrologic and hydrochemical information used to quantify dissolved export from the watershed. Included files provide site-level values for drainage area, mean annual solute export, standard error of annual export, area-normalized solute yield, and equivalent weathering rate for five East River monitoring sites, along with metadata describing the number, sampling cadence, and date range of discharge records and partial and full total dissolved solids observations used in the solute-yield analyses. The package also contains a supplementary daily ion-load time series with daily mean discharge, discharge observation counts, dissolved concentrations, and daily loads for calcium, magnesium, sodium, potassium, chloride, sulfate, nitrate, fluoride, dissolved silica, charge-balance bicarbonate, and total dissolved solids. The package contains GIS files, comma-separated value files (.csv), CSV data dictionaries, a file-level metadata table, a package-tree text file, and a readme text file.

10Be↗

The performance of missing transverse momentum reconstruction and its significance with the ATLAS detector using 140 $\hbox {fb}^{-1}$ of $\sqrt{s}=13$ TeV pp collisions

This paper presents the reconstruction of missing transverse momentum ($p_{\text {T}}^{\text {miss}}$ ) in proton–proton collisions, at a center-of-mass energy of 13 TeV. This is a challenging task involving many detector inputs, combining fully calibrated electrons, muons, photons, hadronically decaying $\tau$ -leptons, hadronic jets, and soft activity from remaining tracks. Possible double counting of momentum is avoided by applying a signal ambiguity resolution procedure which rejects detector inputs that have already been used. Several $p_{\text {T}}^{\text {miss}}$ ‘working points’ are defined with varying stringency of selections, the tightest improving the resolution at high pile-up by up to 39% compared to the loosest. The $p_{\text {T}}^{\text {miss}}$ performance is evaluated using data and Monte Carlo simulation, with an emphasis on understanding the impact of pile-up, primarily using events consistent with leptonic Z decays. The studies use $140~\text {fb}^{-1}$ of data, collected by the ATLAS experiment at the Large Hadron Collider between 2015 and 2018. The results demonstrate that $p_{\text {T}}^{\text {miss}}$ reconstruction, and its associated significance, are well understood and reliably modelled by simulation. Finally, the systematic uncertainties on the soft $p_{\text {T}}^{\text {miss}}$ component are calculated. After various improvements the scale and resolution uncertainties are reduced by up to 76% and 51%, respectively, compared to the previous calculation at a lower luminosity.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Results from a multi-laboratory ocean metaproteomic intercomparison: effects of LC-MS acquisition and data analysis procedures

Metaproteomics is an increasingly popular methodology that provides information regarding the metabolic functions of specific microbial taxa and has potential for contributing to ocean ecology and biogeochemical studies. A blinded multi-laboratory intercomparison was conducted to assess comparability and reproducibility of taxonomic and functional results and their sensitivity to methodological variables. Euphotic zone samples from the Bermuda Atlantic Time-series Study (BATS) in the North Atlantic Ocean collected by in situ pumps and the autonomous underwater vehicle (AUV) Clio were distributed with a paired metagenome, and one-dimensional (1D) liquid chromatographic data-dependent acquisition mass spectrometry analysis was stipulated. Analysis of mass spectra from seven laboratories through a common bioinformatic pipeline identified a shared set of 1056 proteins from 1395 shared peptide constituents. Quantitative analyses showed good reproducibility: pairwise regressions of spectral counts between laboratories yielded R 2 values averaged 0.62±0.11, and a Sørensen similarity analysis of the top 1000 proteins revealed 70 %–80 % similarity between laboratory groups. Taxonomic and functional assignments showed good coherence between technical replicates and different laboratories. A bioinformatic intercomparison study, involving 10 laboratories using eight software packages, successfully identified thousands of peptides within the complex metaproteomic datasets, demonstrating the utility of these software tools for ocean metaproteomic research. Lessons learned and potential improvements in methods were described. Future efforts could examine reproducibility in deeper metaproteomes, examine accuracy in targeted absolute quantitation analyses, and develop standards for data output formats to improve data interoperability. Together, these results demonstrate the reproducibility of metaproteomic analyses and their suitability for microbial oceanography research, including integration into global-scale ocean surveys and ocean biogeochemical models.

59 BASIC BIOLOGICAL SCIENCES↗

Disturbance of hibernating bats due to researchers entering caves to conduct hibernacula surveys

Estimating population changes of bats is important for their conservation. Population estimates of hibernating bats are often calculated by researchers entering hibernacula to count bats; however, the disturbance caused by these surveys can cause bats to arouse unnaturally, fly, and lose body mass. We conducted 17 hibernacula surveys in 9 caves from 2013 to 2018 and used acoustic detectors to document cave-exiting bats the night following our surveys. We predicted that cave-exiting flights (i.e., bats flying out and then back into caves) of Townsend’s big-eared bats (Corynorhinus townsendii) and western small-footed myotis (Myotis ciliolabrum) would be higher the night following hibernacula surveys than on nights following no surveys. Those two species, however, did not fly out of caves more than predicted the night following 82% of surveys. Nonetheless, the activity of bats flying out of caves following surveys was related to a disturbance factor (i.e., number of researchers × total time in a cave). We produced a parsimonious model for predicting the probability of Townsend’s big-eared bats flying out of caves as a function of disturbance factor and ambient temperature. That model can be used to help biologists plan for the number of researchers, and the length of time those individuals are in a cave to minimize disturbing bats.

59 BASIC BIOLOGICAL SCIENCES↗

Synthesizing land use and demographic change in Southeast Asia’s smaller urbanized areas from 2000–2015

The majority of the human population now reside in urban areas today. The United Nations estimates that nearly half of all urban dwellers currently live in cities smaller than 500 000 persons and the majority of future urban growth will take place in Asia and Africa, likely in these smaller urban areas, not mega cities. Thus, understanding the factors that influence urban demographic trajectories in small urban areas is critical to address sustainable and equitable policy initiatives related to food security, changing climate hazard exposure, and economic opportunities. Here we focus on Southeast Asia—a region historically characterized by lower urban population proportions, yet with a rapidly shifting dynamic demographic—to examine correlates of demographic change among smaller cities. We combine two open-source satellite-informed datasets: GHS urban center database (2015) and age-sex gridded data from WorldPop to calculate socio-demographic characteristics to model drivers of change in annualized urban population growth from 2000–2015 for 505 urbanized places. We find a general pattern of decreasing dependency ratios as city-size increases for most urban areas in Southeast Asia. Higher rates of growth and more variation is observed for smaller cities—those with fewer than 300 000 persons, the lowest population limit for UN data on urbanization. When examining covariates of urban population growth, we find significant statistical associations of population change in smaller urbanized areas with climatic, economic, and land cover/land use variables, but with country-specific variations. Characterizing a continuum of urban population development in the context of changing environmental, economic and climate conditions has been an important sustainable development and equity issue for decades, but newer analysis of city-level drivers allows for systematic inquiry thus moving beyond total population counts for policy-relevant insight.

Southeast Asia synthesis↗

Low energy neutron light output characterization of EJ301D and deuterated stilbene with a comparison of light output characterization methods

The neutron-induced light yield of a 2.54 cm diameter by 2.54 cm long right circular cylinder of EJ301D and a (5.08 cm)3 custom made cube of deuterated trans-stilbene-d12 (d-stilbene) were measured over incident neutron energies from 300 keV to 2.2 MeV and 200 keV to 2.4 MeV, respectively. The measurements were performed using a time-of-flight experiment with a Cf-252 source and an approximately 1.5 m flight path. We compare three light output spectrum full energy deposition edge estimation methods: (1) simulating the neutron energy spectrum edge and fitting it to the light output spectrum, (2) using the inflection point of the light output spectrum edge (derivative method, a.k.a. Kornilov’s method), and (3) using an empirical model fit to the edge of the light output spectrum. Both the derivative and equation fit methods do not account for physical processes such as multiple neutron scattering in the detectors. They instead rely on assumptions about the linear shape continuum shape of the light output spectrum and the direct correlation between the location of the spectrum’s inflection point and maximum energy deposition. These assumptions were found to introduce bias into those methods when tested against simulated spectra with known edge locations. When tested against measured spectra the derivative method was found to differ from the simulation fit by greater than 30% at low energies with large discontinuities for adjacent data points above 800 keV incident neutron energy. The empirical equation fitting method was found to also exhibit bias of a similar magnitude, but with significantly more continuous behavior, especially with the lower count data of the smaller volumed EJ301D scintillator. Experimental light output yield for this neutron energy range is reported using the simulated spectrum fitting method because it includes physics neglected by the other methods, and did not exhibit the bias observed in the other methods

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Ion-Specific Effects on PuO 2 Nanoparticle Aggregation and Dissolution in Concentrated Electrolytes

Hydrolytic PuO 2 nanoparticles (NPs) are a dominant aqueous Pu-bearing phase in high ionic strength nuclear wastes, yet their reactivity in nonideal brines remains poorly constrained. We quantify how electrolyte identity and concentration control PuO 2 NP aggregation and ligand-assisted dissolution in acidic, high salinity solutions (NaCl, NaNO 3 , NaClO 4 , Na 2 SO 4 , Na 2 C 2 O 4 up to 5 M). A multitechnique workflow combining liquid scintillation counting (operationally defined aqueous [Pu]), scattering/electrokinetic measurements (aggregate size and zeta potential), and spectroscopy (UV–vis, XPS, Raman) resolves electrolyte-dependent partitioning between colloidal and molecular Pu species. Weakly coordinating anions (ClO 4 – , Cl – , NO 3 – ) largely preserve the (aggregated) nanoparticulate fraction but generate distinct dissolved species at high concentration, i.e., Pu(IV)–nitrato complexes in NaNO 3 and Pu(VI)–chloro complexes in NaCl. In contrast, stronger ligands substantially perturb PuO 2 NP stability: sulfate promotes partial dissolution to Pu(IV)–sulfate complexes at low concentration but reduces aqueous [Pu] at higher sulfate levels via secondary Pu(IV) sulfate formation, whereas oxalate drives strong dissolution to aqueous Pu–oxalate complexes. Aged NPs show similar trends with reduced aqueous fractions and more dominant aggregation mechanisms. In conclusion, these spectroscopically constrained speciation data provide a foundation for incorporating PuO 2 NP reactivity into thermodynamic and reactive transport models for high salinity waste and brine environments.

Aggregation↗

Peering into cloud physics using ultra-fine resolution radar and lidar systems

Cloud microphysical processes, such as droplet activation, condensational growth, and collisional growth, play a central role in the evolution of clouds and precipitation. Accurate representations of these processes in numerical models are challenging partially due to incomplete understanding of them at the process-level arising from limited systematic observations. Most surface-based active remote sensors, including today’s operational cloud radars and lidars, have a resolution on the order of tens of meters. This resolution is insufficient to resolve cloud microphysical processes that manifest at finer (meter and sub-meter) scales. A new set of ultra-high-resolution ground-based radar and lidar systems have been developed to address this observational gap. The newly developed 94-GHz cloud radar has a range resolution down to 2.8 m, or a factor of 10 finer than typical radars, using a large bandwidth and quadratic phase coding techniques. The lidar has a range resolution down to 10 cm, or a factor of 100 finer than typical lidars, using a time-gated time-correlated single photon counting technique. Such high-resolution observations were previously only achievable through in situ aircraft measurements. Even then, aircraft measurements do not permit continuous long-term cloud observation as is possible with ground-based remote sensing instruments. In this study, the first-light cloud observations from the new radar and lidar systems are shown to reveal detailed cloud structures that conventional sensors could only perceive in a bulk sense, thus providing new avenues to investigate cloud microphysical processes and their impact on weather and climate.

54 ENVIRONMENTAL SCIENCES↗

Sixteen multiple-amplifier sensing charge-coupled devices and characterization techniques targeting the next generation of astronomical instruments

We present a candidate sensor for future spectroscopic applications, such as a Stage-5 Spectroscopic Survey Experiment or the Habitable Worlds Observatory. This type of charge-coupled device (CCD) sensor features multiple in-line amplifiers at its output stage allowing multiple measurements of the same charge packet, either in each amplifier or in the different amplifiers. Recently, the operation of an eight-amplifier sensor has been experimentally demonstrated, and we present the operation of a 16-amplifier sensor. This new sensor enables a noise level of ∼1 erms− with a single sample per amplifier. In addition, it is shown that sub-electron noise can be achieved using multiple samples per amplifier. In addition to demonstrating the performance of the 16-amplifier sensor, we aim to create a framework for future analysis and performance optimization of this type of detectors. New models and techniques are presented to characterize specific parameters, which are absent in conventional CCDs and Skipper CCDs: charge transfer between amplifiers and independent and common noise in the amplifiers and their processing.

16 multiple-amplifer sensing CCD (MAS-CCD)↗

Exploring 𝛽 decay and 𝛽-delayed neutron emission in exotic 46,47 Cl isotopes

In this paper, 𝛽 − and 𝛽-delayed neutron decays of 46,47 Cl are reported from an experiment carried out at the National Superconducting Cyclotron Laboratory using the Beta Counting System. The half-lives of both 46 Cl and 47 Cl were extracted. Based on the delayed 𝛾-ray transitions observed, the level structure of 𝑁=28 46 Ar was determined. Completely different sets of excited states above the first 2 + state in 46 Ar were populated in the 46 Cl 𝛽⁢0⁢𝑛 and 47 Cl 𝛽⁢1⁢𝑛 decay channels. Two new 𝛾-ray transitions in 47 Ar were identified from the very weak 47 Cl 𝛽⁢0⁢𝑛 decay. Furthermore, 46 Cl 𝛽⁢1⁢𝑛 and 47 Cl 𝛽⁢2⁢𝑛 were also observed to yield different population patterns for levels in 45 Ar, including states of different parities. Here, the experimental results allow us to address some of the open questions related to the delayed neutron emission process. For isotopes with large neutron excess and high 𝑄 𝛽 values, delayed neutron emission remains an important decay mode and can be utilized as a powerful spectroscopic tool. Experimental results were compared with shell-model calculations using the FSU and 𝑉 MU effective interactions.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Towards time-resolved MicroED grid preparation using mix-and-inject gas dynamic virtual nozzles

Recent progress in gas dynamic virtual nozzle (GDVN) technologies in combination with high-brilliance synchrotron and X-ray free-electron lasers (XFELs) has allowed the visualization of protein dynamics in crystallo by mixing macromolecular protein crystals with a substrate using tunable mixing times on the order of milliseconds to seconds prior to serial X-ray diffraction data collection. This has become the method of choice for high-resolution structure determination of intermediate states. However, such experiments require large counts of crystals of proper sizes for high-resolution data collection, and premium beam times for screening efforts. Cryogenic microcrystal electron diffraction (MicroED) represents a complementary technique that may be a more accessible avenue for time-resolved nanocrystallography compared with serial X-ray diffraction experiments. MicroED can produce full diffraction datasets from just a few submicrometre-thick crystals, and the approach is more readily accessible, requiring standard cryogenic transmission electron microscopy (TEM) equipment available at many universities and institutes. Cryogenic MicroED, like other forms of cryo-EM, begins with rapidly freezing biological material on electron microscopy grids. In the case of MicroED, micro- to nano-crystals (<500 nm thick) are deposited onto electron microscopy grids and plunge-frozen for subsequent electron diffraction data collection. Here, we have incorporated GDVN technology developed originally for XFEL experiments into the freezing process as a first step towards time-resolved studies. We describe the limited deposition efficiency of the model MicroED protein proteinase K on TEM grids using GDVNs, preceding sample vitrification and successful MicroED data collection. We discuss both the initial results from such experiments and the methodological challenges in developing this approach into a reliable workflow for millisecond-to-second time-resolved structural studies of macromolecules. Our results promise a strategy to deposit crystals on grids using GDVNs and determine high-resolution structures by MicroED, constituting a first step towards development of time-resolved MicroED experiments.

MicroED↗

Hyper Spectral Anomaly Detection

The HSA is a statistics based anomaly detection model. The model performs unsupervised anomaly detection, based on a datapoint's density and similarity within a dataset. Density and similarity data are encoded into an affinity matrix. The affinity matrix is evolved to summarize the data's structure on greater topographical scales within the data's function space. The set of evolved affinity matrices and an anomaly score vector are passed to a user defined penalized objective function. The penalized objective function of anomaly scores is then minimized. Data points where the absolute value of the z-scores of anomaly scores greater than a specified threshold are predicted as anomalies. A novel multi-filter feature has also been implemented. To reduce false positive rates, the multi-filter records the indexes of the HSA predictions. A new dataset and data loader are instantiated consisting of all the initial HSA predictions and non-anomalous data points in a 10% and 90% split respectively. The HSA is then run through this data set and a count of number of times a data point is predicted is kept. In this way the initial predictions may be compared with data spanning the entire dataset. After the multi-filter is complete, all datapoints will have an associated anomaly score, as well as a multi-filter prediction count to further filter the anomalous predictions.

Rogers, DempseyD [Idaho National Laboratory (INL),↗