Search NASASearch

SEARCH · Search NASA

Results for “count data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Oak Ridge National Laboratory EAGLE-I TM : Modeling Electric Utility County Customers for Situational Awareness

During natural hazard events (hurricanes, wildfires, earthquakes, etc.) and recent man-made events (e.g., cyber attacks), the exchange of near real-time, spatially refined data within the response community is critical. The EAGLE-I$^{TM}$ platform is one tool that facilitates this data for decision makers within the energy sector. While much information can be collected and integrated into the system directly, other pertinent data must be augmented by other derived data products to enhance the information and allow for a consistent evaluation of on-the-ground conditions. One such data set that requires the addition of other derived data is the electric utility customer outage data that is aggregated to the county level within the EAGLE-I application. Without a county customer data set, outages can only be compared on total counts, which gives greater importance to higher population outages. Including an electric utility customer data set at the county level allows for these outage counts to be converted to percent outages and brings a consistent classification of outages and equal importance to all outages. To achieve this, several available data sets were combined and spatial disaggregation techniques were employed to model customer estimates at the county scale. This paper presents the approach to produce this data for the United States and lessons learned from working with these disparate data sets. Data validation is provided, where possible, and limitations of the model and possible improvements are discussed.

24 POWER TRANSMISSION AND DISTRIBUTION

Human Liver Epithelium Response to HCoV-229E Infection Epigenomics (ACS-DP4)

The purpose of this experiment was to evaluate how wild-type Human coronavirus strain 229E (HCoV-229E) infection alters chromatin accessibility in infected cells only. Sample data was obtained for mock and infected (standard and UV-inactivated) immortalized human liver cells (HuH-7) and collected 24 hrs. post infection. Samples were processed using assay for transposase-accessible chromatin using high-throughput sequencing (ATAC-Seq) and generated bar coded library samples were evaluated for RNA sequencing (RNA-Seq) expression analysis. Processed ATAC-Seq datasets are openly accessible from the download button and contain secondary processed RNA-Seq results files and supporting metadata materials. Data download includes a sample naming key, infection titer metadata, normalized counts, and relevant computational source code information supporting data transparency and reuse.

59 BASIC BIOLOGICAL SCIENCES

Detection of Isotopes in Urban Source Search Low-Count Gamma Spectra Using Hopfield Neural Networks

Source search campaigns involve measurements of background gamma-ray spectra with a mobile detector-spectrometer traveling along arbitrarily chosen trajectories over a wide screening area. Radiation counts are typically measured with a tellurium-doped sodium iodide [NaI(Tl)] scintillator detector-spectrometer in short acquisition intervals, usually 1 s. The objective is to detect orphan isotopes with half-lives shorter than those of the isotopes in the natural background. In principle, radioisotopes can be identified by their unique gamma emission spectrum. However, detecting orphan isotopes in search data is challenging because low counts measured in short acquisition intervals result in incomplete spectral lines. In this study, we investigate the performance of a Hopfield neural network (HNN) that implements an auto-associative memory for the detection of isotopes of interest in an urban search campaign. The HNN is trained on one example of gamma spectra with well-resolved spectral lines of each isotope of interest. During testing, the auto-associative memory implementation of the HNN processes low-count gamma spectra with partially complete isotopic lines by matching incoming measurements to the closest one of its memory-stored patterns. The testing database consisted of almost 10 000 1-s gamma spectra, including measurements of orphan isotopes 137 Cs, 241 Am, and 131 I, obtained during two urban search surveys with a NaI(Tl) detector. The performance of the HNN detection algorithm was evaluated using precision, recall, and F1 scores, and benchmarked with a multiple linear regression (MLR) identification algorithm. In conclusion, the test results demonstrate that HNN outperforms MLR in the detection of all the isotopes of interest.

Auto associative memory

Combining MicroED and native mass spectrometry for structural discovery of enzyme–small molecule complexes

With the goal of accelerating the discovery of small molecule–protein complexes, we leverage fast, low-dose, event-based electron counting microcrystal electron diffraction (MicroED) data collection and native mass spectrometry. This approach, which we term electron diffraction with native mass spectrometry (ED-MS), allows assignment of protein target structures bound to ligands with data obtained from crystal slurries soaked with mixtures of known inhibitors and crude biosynthetic reactions. This extends to libraries of printed ligands dispensed directly onto TEM grids for later soaking with microcrystal slurries, and complexes with noncovalent ligands. ED-MS resolves structures of the natural product, epoxide-based cysteine protease inhibitor E-64, and its biosynthetic analogs bound to the model cysteine protease, papain. It further identifies papain binding to its preferred natural products, by showing that two analogs of E-64 outcompete others in binding to papain crystals, and by detecting papain bound to E-64 and an analog from crude biosynthetic reactions, without purification. ED-MS also resolves binding of the CTX-M-14 β-lactamase, a target of active drug development, to the non-β-lactam inhibitor, avibactam, alone or in a cocktail of unrelated compounds. These results illustrate the utility of ED-MS for natural product ligand discovery and for structure-based screening of small molecule binders to macromolecular targets, promising utility for drug discovery.

MicroED

Laser Confocal Microscopy Uncertainty Quantification Study

At Los Alamos National Laboratory (LANL), the Storage Safety and Engineering (SSE) team completes annual surveillance on a subset of in-use interim nuclear material storage containers in fulfilment of requirements outlined in DOE Manual M 441.1-1. The containers are selected through several methods, such as subject matter expert judgement, random selection, and trending items. Following these selections, the SSE team has the capacity to complete surveillance on 15-20 containers each fiscal year, composed of a combination of SAVY-4000 and Hagan storage containers. Through previous work, the stainless-steel components of the containers have been identified as life limiting components, with an emphasis on the thin-walled bodies. The team is focused on understanding the extent of general and pitting corrosion, due to observations of extensive corrosion from stored contents and bag-out-bag degradation. Quantifying corrosion effects on the thin-walled stainless steel container bodies, and understanding potential impacts to the respective design release rates and design qualification release rates is paramount to the team. To date, destructive examination (DE) has proven to be the most insightful method for developing an understanding on the extent of corrosion on used containers. To standardize this process, the SSE team developed a destructive examination guide for analyzing stainless steel components of the containers. Corroded containers of interest are identified during surveillance activities and set aside for sectioning and characterization. Following sectioning, a major step in the DE workflow is the utilization of laser confocal microscopy for scanning corroded samples of interest and extracting data on pits, such as count, depth, and equivalent diameter. Adhering to the techniques outlined in the DE guide, analysis has been completed on two Hagans and one SAVY-4000 container, with the maximum pit depth recorded as 139.1 ± 22.82 μm on a 17.5 year old Hagan. The findings from the completed destructive examinations will be utilized to support lifetime extension efforts of the SAVY-4000 as the team can better estimate corrosion rates and effects over time based on stored contents and age. Due to the implications of observing extreme pit depths that approach the nominal container body thickness of .0299 inches (0.759 mm) or minimum container thickness of 0.236” (0.6 mm), high confidence in the LCM measurements is desired. Through testing outlined in, it was concluded that the total error ascribed to the 20x objective when conducting large image mapping on the Keyence VK-X3050 laser confocal microscope (LCM) relative to a 50x objective (reference) is 16.4% (± 8.73%). For shallow features on the order of pristine SAVY surface defects (i.e. 5 μm), this uncertainty is appropriate. However, this conservative estimate of total error poses a fundamental concern for pit depths that approach the thickness of the measured samples. That is, with the measurement uncertainty currently employed on all measurements, the LCM would be unable to resolve if a pit with a depth of 515 μm is through wall. Standard step height samples were procured and used in the present study to assess the resolution and repeatability of height measurements. Understanding the resolution and repeatability of height measurements was the first focus of the team as it relates directly to pit depth, which is of primary concern. Calibration gratings were procured to evaluate the resolution and repeatability of measurements in the X and Y axes of the LCM stage. The results of the depth uncertainty study were conducted first and presented in the subsequent sections. The planar uncertainty study is appended to the depth study with conclusions from both summarized at the end of the report.

36 MATERIALS SCIENCE

Massive compression for high data rate macromolecular crystallography (HDRMX): impact on diffraction data and subsequent structural analysis

New higher-count-rate, integrating, large-area X-ray detectors with framing rates as high as 17400 images per second are beginning to be available. These will soon be used for specialized macromolecular crystallography experiments but will require optimal lossy compression algorithms to enable systems to keep up with data throughput. Some information may be lost. Can we minimize this loss with acceptable impact on structural information? To explore this question, we have considered several approaches: summing short sequences of images, binning to create the effect of larger pixels, use of JPEG-2000 lossy wavelet-based compression, and use of Hcompress, which is a Haar-wavelet-based lossy compression borrowed from astronomy. We also explore the effect of the combination of summing, binning, and Hcompress or JPEG-2000. In each of these last two methods one can specify approximately how much one wants the result to be compressed from the starting file size. These provide particularly effective lossy compressions that retain essential information for structure solution from Bragg reflections.

47 OTHER INSTRUMENTATION

Engineered Membrane Vesicle Production via oprF or oprI Deletion Has Distinct Phenotypic Effects in Pseudomonas putida -putative knockouts table

Table S1, putative gene knockout targets in P. putida KT2440 to enhance vesiculation; Table S2, protein sequence identity of OmpA from E. coli K12 to P. putida KT2440 genes; Table S3, strains utilized in this study and corresponding construction details; Table S4, oligonucleotides utilized in this study; Table S5, plasmids utilized in this study; Table S6, sequences for mNeonGreen, tags, and codon-optimized genes; Figure S1, particle count per gCDW for KT2440 and knockout strains corresponding to data presented in Figure 1B; Figure S2, OD600 measurements of extracted MVs from KT2440 and knockout strains; Figure S3, particle count per gCDW for WT, ΔPP_4669, and ΔPP_1502; Figure S4, particle count per gCDW for KT2440 and knockout strains corresponding to data presented in Figure 3C; Figure S5, sizes of MVs corresponding to particle counts in Figure S4; Figure S6, particle count per gCDW for KT2440 grown on 20 mM glucose alone or 20 mM glucose plus 12.5 mM p-coumarate and 12.5 mM ferulate; Figure S7 and Figure S8, principal component analysis of the cellular fractions; Figure S9, heatmap of outer membrane proteins with differential abundance; and Figure S10, mNeonGreen (mNG) fluorescence signal for the cellular fraction and the extracellular fraction

hypervesiculation

Gamma Radiation Monitoring Data

This dataset contains gamma-ray total counts in counts/minute (cpm), every 15-minutes, from the Gamma Radiation Monitoring campaign at the ENA site (Graciosa, Azores).

54 ENVIRONMENTAL SCIENCES

Investigation of fast and efficient lossless compression algorithms for macromolecular crystallography experiments

Structural biology experiments benefit significantly from state-of-the-art synchrotron data collection. One can acquire macromolecular crystallography (MX) diffraction data on large-area photon-counting pixel-array detectors at framing rates exceeding 1000 frames per second, using 200 Gbps network connectivity, or higher when available. In extreme cases this represents a raw data throughput of about 25 GB s −1 , which is nearly impossible to deliver at reasonable cost without compression. Our field has used lossless compression for decades to make such data collection manageable. Many MX beamlines are now fitted with DECTRIS Eiger detectors, all of which are delivered with optimized compression algorithms by default, and they perform well with current framing rates and typical diffraction data. However, better lossless compression algorithms have been developed and are now available to the research community. Here one of the latest and most promising lossless compression algorithms is investigated on a variety of diffraction data like those routinely acquired at state-of-the-art MX beamlines.

36 MATERIALS SCIENCE

Reappraisal of SU(3)-flavor breaking in B → D P

In light of recently found deviations of the experimental data from predictions from QCD factorization for B ( s ) → D ( s ) P decays, where P = { π , K } , we systematically probe the current status of the SU ( 3 ) F expansion from a fit to experimental branching ratio data without any further theory input. We find that the current data are in agreement with the power counting of the SU ( 3 ) F expansion. While the SU ( 3 ) F limit is excluded at > 5 σ , amplitude-level SU ( 3 ) F -breaking contributions of ∼ 20 % suffice for an excellent description of the data. SU ( 3 ) F breaking is needed in tree ( > 5 σ ) and color-suppressed tree ( 2.4 σ ) diagrams. We are not yet sensitive to SU ( 3 ) F breaking in exchange diagrams. From the underlying SU ( 3 ) F parametrization we predict the unmeasured branching ratios B ( B ¯ s 0 → π − D + ) = 2 B ( B ¯ s 0 → π 0 D 0 ) = [ 0.3 , 7.2 ] × 10 − 6 of suppressed decays that can be searched for at the LHCb experiment. Published by the American Physical Society 2024

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

LandScan Global 30 Arcsecond Annual Global Gridded Population Datasets from 2000 to 2022

Abstract Oak Ridge National Laboratory (ORNL) annually develops the LandScan Global (LSG) dataset, a 30 arcsecond global gridded population dataset representing global ambient human population distribution. This multivariable dasymetric model disaggregates census counts within administrative boundaries using ancillary data. Each country’s distribution reflects cultural and socioeconomic patterns; manual validations yield a unique global dataset for assessing populations at risk. For over two decades, LSG has been a standard for estimating populations at risk, aiding U.S. federal government, academia and humanitarian organizations. During disasters such as the 2004 Indian Ocean tsunami and the 2010 Haiti earthquake and geopolitical crises such as the Syrian civil war and the 2022 Russian invasion of Ukraine, LSG supported scientific and operational communities in emergency response and recovery. In 2022, LSG datasets from 2000 onward were made publicly available through ORNL’s LandScan Portal. This data descriptor details our methodology and the application of geospatial science and machine learning to geographic and demographic data, highlighting uses in urban resiliency, emergency management, disaster response, and human health and security.

Science & Technology - Other Topics

Calibration of a broadband x-ray crystal spectrometer using a continuum x-ray source and photon-counting detectors

We report measurements of sensitivity of a broadband (≃20–30 keV) x-ray crystal spectrometer using a bremsstrahlung continuum x-ray source and several detectors, including an energy-discriminating photon-counting point detector (Si drift diode, Amptek Inc.), imaging hybrid photon-counting detectors (Eiger2-Si and Eiger2-CdTe, DECTRIS Ltd.), and image plates (SR type). Sensitivity is defined as a ratio of the crystal-reflected energy-position-dispersed spectrum measured by a given detector to the incident non-dispersed spectrum measured using the energy-discriminating point detector. The sensitivity derived from data measured exclusively by the point detector is considered detector-independent and serves as a reference. The sensitivities derived from data collected with the hybrid photon-counting detectors were matched to the reference using a scale factor of 1.05. The image plate-derived sensitivities required a scale factor of 1.16 to match the reference. Furthermore, the resulting mismatch in the shapes of all measured scaled sensitivities in the range of the spectrometer was ≲±10%, while the mismatch between the shapes of the sensitivities corresponding to the Si-based detectors was ≲±2.5%.

Crystal optics

Effect of Biomass Water Dynamics in Cosmic-Ray Neutron Sensor Observations: A Long-Term Analysis of Maize–Soybean Rotation in Nebraska

Precise soil water content (SWC) measurement is crucial for effective water resource management. This study utilizes the Cosmic-Ray Neutron Sensor (CRNS) for area-averaged SWC measurements, emphasizing the need to consider all hydrogen sources, including time-variable plant biomass and water content. Near Mead, Nebraska, three field sites (CSP1, CSP2, and CSP3) growing a maize–soybean rotation were monitored for 5 (CSP1 and CSP2) and 13 (CSP3) years. Data collection included destructive biomass water equivalent (BWE) biweekly sampling, epithermal neutron counts, atmospheric meteorological variables, and point-scale SWC from a sparse time domain reflectometry (TDR) network (four locations and five depths). In 2023, dense gravimetric SWC surveys were collected eight (CSP1 and CSP2) and nine (CSP3) times over the growing season (April to October). The N0 parameter exhibited a linear relationship with BWE, suggesting that a straightforward vegetation correction factor may be suitable (fb). Results from the 2023 gravimetric surveys and long-term TDR data indicated a neutron count rate reduction of about 1% for every 1 kg m−2 (or mm of water) increase in BWE. This reduction factor aligns with existing shorter-term row crop studies but nearly doubles the value previously reported for forests. This long-term study contributes insights into the vegetation correction factor for CRNS, helping resolve a long-standing issue within the CRNS community.

Chemistry

Leveraging 13C-Labeling to Assign Molecular Formulas to Unknown Yeast Metabolites

Mass spectrometry analyses have identified tens of thousands of unknown small molecule-associated peaks in different biological specimens. Notably, even the simplest and best studied organisms like Escherichia coli and Saccharomyces cerevisiae yield thousands of unknown peaks. A key question is how many of these reflect actual novel endogenous metabolites. To explore this, Mahieu and Patti used complete 13 C -labeling in E. coli to credential peaks as biological. This reduced the number of unknowns by more than 90%. Here, we carry out similar uniform 13 C-labeling in the Baker’s yeast S. cerevisiae and two less-studied bioenergy-relevant yeasts Rhodotorula toruloides (lipid producer) and Issatchenkia orientalis (organic acid producer). Identification of unknown metabolite peaks and their molecular formulas is facilitated through software tailored for 13 C labeling data and resulting knowledge of carbon atom count. A classification model evaluates the plausibility of each candidate formula, with peaks lacking plausible candidate formulas unlikely to reflect metabolite molecular ions. This approach prioritizes about one hundred candidate abundant unknown metabolites with logical molecular formulas. Most of these are species-specific rather than conserved across yeasts, and more are found in the nonmodel yeasts than S. cerevisiae. Thus, 13 C-labeling data on unknown metabolites highlights the potential for discovering new metabolites and pathways in nonmodel yeasts.

Carbon

Warming amplifies the variability of methane emissions from a coastal wetland, 2025, Maryland.

These data accompany the published paper Lewis et al., 202X and are from a brackish coastal wetland in situ soil warming experiment (GENX) equipped with automated flux chambers. Methane (CH4) and carbon dioxide (CO2) fluxes were measured in 12 automated chambers using custom-built automated chambers connected to an LI-7810 CH4/CO2 analyzer. The chambers are 1.5 m tall and contain the dominant vegetation species of the site (Schoenoplectus americanus, Spartina patens, and Distichlis spicata). The chambers are also distributed across a soil warming gradient, ranging from ambient to 6°C above ambient, that was started in February 2022. This dataset contains the following files: (1) CH4 and CO2 fluxes from each chamber for March to November 2025, statistics for each flux, and environmental data (water depth, salinity, air temperature) at the time of the flux measurement; (2) 15-minute soil temperature data for each chamber; (3) Aboveground vegetation biomass (total and by species) and stem counts and dimensions for S. americanus; (4) Elevation for each chamber. All data processing code is available on Github.

Coastal wetland

Real-time data processing for serial crystallography experiments

We report the use of streaming data interfaces to perform fully online data processing for serial crystallography experiments, without storing intermediate data on disk. The system produces Bragg reflection intensity measurements suitable for scaling and merging, with a latency of less than 1 s per frame. Our system uses the CrystFEL software in combination with the ASAP::O data framework. In a series of user experiments at PETRA III, frames from a 16 megapixel Dectris EIGER2 X detector were searched for peaks, indexed and integrated at the maximum full-frame readout speed of 133 frames per second. The computational resources required depend on various factors, most significantly the fraction of non-blank frames ('hits'). The average single-thread processing time per frame was 242 ms for blank frames and 455 ms for hits, meaning that a single 96-core computing node was sufficient to keep up with the data, with ample headroom for unexpected throughput reductions. Further significant improvements are expected, for example by binning pixel intensities together to reduce the pixel count. We discuss the implications of real-time data processing on the `data deluge' problem from recent and future photon-science experiments, in particular on calibration requirements, computing access patterns and the need for the preservation of raw data.

47 OTHER INSTRUMENTATION

Photon temporal-mode readout for inference of neutron star merger remnant gravitational waves

Gravitational waves emitted after neutron star binary coalescences and the information they carry about dense matter are a high-priority target for next-generation detectors. Even though such detectors are expected to observe millions of signals, detectable postmerger emission will remain rare. Here, in this work, we explore postmerger detectability and inference through an alternative detector readout scheme for data dominated by quantum-noise, which is the case above 1 kHz; photon-counting. In such a readout, signals and noise become quantized into discrete distributions corresponding to the detection of single photons measured in a chosen basis of modes. Through simulated data, we demonstrate that photon counting can be efficient even for weak signals. We find ∼1 in 100 signals with a postmerger signal-to-noise ratio of 0.2 can result in a single photon and thus be detected. Furthermore, after 2 ×10 4 signals—equivalent to 10 −2 to 1.5 years of observation—photon counting results in a twofold improvement in the measurement of the radius of a 1.6⁢𝑀 ⊙ neutron star. Constraints can be further tightened if the detector classical noise is reduced. Photon counting offers a promising alternative to traditional homodyne readout techniques for extracting information from low signal-to-noise ratio postmerger signals.

gravitational wave detection

Bayesian Framework for Bioburden Density Estimation in Planetary Protection

To comply with the international planetary protection policy set forth by the Committee on Space Research and NASA Agency level requirements, spacecraft destined to biologically sensitive planetary bodies have to minimize terrestrial biological contamination. Analysis, testing and inspection are the standard forward verification activities that are used to demonstrate compliance with the biological contamination requirements. For testing of spacecraft surface areas, a swab or wipe sample is collected from surfaces prior to last access and subsequently processed in the lab using NASA Approved Planetary Protection Methods for Culture Based Assays. Raw data resulting from this assay is then statistically treated employing a mathematical paradigm stemming from the 1970’s Viking Lander Project to generate the bioburden density and total microbial bioburden present. This standard approach arbitrarily accounts for error and provides an upper conservative bound as it reports the maximum number of spores estimated to be present on flight hardware surfaces. A bioburden density estimate factors in the following variables: the observed bioburden count, representative volume processed, sampling efficiencies. Notably, to account for error in the approach, a 0 observed count is arbitrarily changed to a count of 1 for each hardware grouping. The data generated by spacecraft bioburden verification campaigns in the past have resulted in <80% of wipes and <90% of swabs containing a bioburden count of 0. As such, having a robust and well documented statistical approach for dealing with the probability of low incident rates is necessary to be able to estimate spacecraft bioburden. Being able to statistically describe the bioburden distribution and associated confidence level is a gamechanger for the development of bioburden allocations during mission design and will allow for tighter management of risk throughout spacecraft build. Thus, Empirical Bayes statistical approach was evaluated to estimate the microbial bioburden on spacecraft to mitigate the aforementioned mathematical concerns and provide a probabilistic bioburden distribution of the flight hardware surface. For application of this approach to performing bioburden calculations, a range of non-informative prior assumptions on hardware surfaces are explored for Bayesian analyses while informative priors using posterior distributions from prior assays are utilized for Empirical Bayes analyses. Several non-informative priors are currently under investigation to assess fitness including use of these priors to serve as a foundation to build off of NASA specification values or a basis of risk to account for unknowns during the integration and testing process. Informative priors under consideration are generated using sampled bioburden values from hardware originating within like processing environments (e.g. vendor cleaning process or similar assembly process), temporal spacecraft status events as a prediction for hardware cleanliness of future samples, and heritage system bioburden actuals to predict allocation for subsequent missions. Informative priors and probabilistic bioburden distributions are then validated using data sets from the Mars Exploration Rover, Mars Science Laboratory, and InSight missions. Using Empirical Bayes approach to generate a probabilistic bioburden distribution as demonstrated through mission use cases provides a valid approach for use in the end-to-end requirements verification process.

97 - MATHEMATICS AND COMPUTING