Search NASA⌕ Search

SEARCH · Search NASA

Results for “data processing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 631 records · Page 35

Optimizing spin dressing sensitivity for the nEDMSF experiment

nEDMSF aims to measure the neutron electric dipole moment (d n ) with unprecedented precision. In this paper we explore the experiment's sensitivity when operating with an implementation of the critical dressing method in which the angle between the neutron and Helium-3 spins (ϕ 3n ) is subjected to a square modulation by an amount ϕ d (the “dressing angle”). Several parameters can be tuned to optimize sensitivity. We find roughly 10% improvement over a previous estimate, resulting primarily from the addition of a waiting period between the π/2 pulse that initiates d n -driven ϕ 3n growth and the start of ϕ3n modulation. We find negligible further improvement by allowing ϕ d to vary continuously over the course of a run, and no degradation resulting from the addition of an in situ background measurement into each ϕ3n modulation sequence. A complete simulation confirms a 300 live-day sensitivity ofσ = 1.45×10 -28 e ·cm. At this level of sensitivity, σ ϕ3n0 = 1 mrad precision on the initial n/ 3 He angle difference is not negligible.

47 OTHER INSTRUMENTATION↗

Tools for unbinned unfolding

Machine learning has enabled differential cross section measurements that are not discretized. Going beyond the traditional histogram-based paradigm, these unbinned unfolding methods are rapidly being integrated into experimental workflows. Here, in order to enable widespread adaptation and standardization, we develop methods, benchmarks, and software for unbinned unfolding. For methodology, we demonstrate the utility of boosted decision trees for unfolding with a relatively small number of high-level features. This complements state-of-the-art deep learning models capable of unfolding the full phase space. To benchmark unbinned unfolding methods, we develop an extension of existing dataset to include acceptance effects, a necessary challenge for real measurements. Additionally, we directly compare binned and unbinned methods using discretized inputs for the latter in order to control for the binning itself. Lastly, we have assembled two software packages for the OmniFold unbinned unfolding method that should serve as the starting point for any future analyses using this technique. One package is based on the widely-used RooUnfold framework and the other is a standalone package available through the Python Package Index (PyPI).

47 OTHER INSTRUMENTATION↗

SALSA: a new versatile readout chip for MPGD detectors

The SALSA chip is a future readout ASIC foreseen for the MPGD detectors, developed in the framework of the EIC collider project, to equip the MPGD trackers of the EPIC experiment. It is designed to be versatile, to be adapted to other usages of MPGD detectors like TPC or photon detectors. It integrates a frontend block and an ADC for each of the 64 channels, associated to a configurable DSP processor meant to correct data and reduce the raw data flux to limit the output bandwidth. It will be compatible with the continuous readout foreseen for the EPIC DAQ, but will also work in a triggered environment. Several prototypes are already produced in order to qualify the different blocks of the chip, in particular the frontend, the ADC and the clock generation. The next 32-channel prototype is currently under development and is planed to be produced in 2025. In conclusion, the final prototype will be produced and tested from 2026 for a production of the SALSA chip at the horizon of 2027.

Data processing↗

Neural posterior unfolding

Differential cross section measurements are the currency of scientific exchange in particle and nuclear physics. A key challenge for these analyses is the correction for detector distortions, known as deconvolution or unfolding. Binned unfolding of cross section measurements traditionally rely on the regularized inversion of the response matrix that represents the detector response, mapping pre-detector (`particle level') observables to post-detector (`detector level') observables. In this paper we introduce Neural Posterior Unfolding, a modern, Bayesian approach that leverages normalizing flows for unfolding. By using normalizing flows for neural posterior estimation, NPU offers several key advantages including implicit regularization through the neural network architecture, fast amortized inference that eliminates the need for repeated retraining, and direct access to the full uncertainty in the unfolded result. In addition to introducing NPU, we implement a classical Bayesian unfolding method called Fully Bayesian Unfolding (FBU) in modern Python so it can also be studied. These tools are validated on simple Gaussian examples and then tested on simulated jet substructure examples from the Large Hadron Collider (LHC). We find that the Bayesian methods are effective and worth additional development to be analysis ready for cross section measurements at the LHC and beyond.

Analysis and statistical methods↗

The Natural Products Magnetic Resonance Database (NP-MRD) for 2025

The Natural Products Magnetic Resonance Database or NP-MRD (https://np-mrd.org) is a comprehensive, freely accessible, web-based resource for the deposition, distribution, extraction and retrieval of nuclear magnetic resonance (NMR) data on natural products. The NP-MRD was initially established to support compound de-replication and data dissemination for the natural products community. However, that community has now grown to include many users from the metabolomics, microbiomics, foodomics and nutrition science fields. Indeed, since its launch in 2021, the NP-MRD has expanded enormously in size, scope and popularity. The current version of NP-MRD now contains nearly 7X more compounds (281,859 vs. 40,908) and 7X more NMR spectra (5.1 million vs. 817,000) than the first release. More specifically, an additional 4.6 million predicted spectra and another 11,000 spectra simulated from experimental chemical shifts were deposited into the database. Likewise, the number of NMR raw spectral data depositions has grown from a 165 spectra per year to more than 10,000 per year. As a result of this expansion, the number of monthly webpage views has grown from 55 to 20,000 and the number of monthly visitors has increased from 7 to 2500. To address this growth and to better support the expanding needs of its diverse community of users, many additional improvements to the NP-MRD have been made. These include significant enhancements to the data submission process, important improvements to the visualization and display of NMR spectra, notable updates to the database’s spectral search utilities and useful additions to support better NMR spectral analysis/prediction. Significant efforts have also been undertaken to remediate and update many of NP-MRD’s database entries. This manuscript describes these database improvements and expansion efforts, along with how they have been implemented and what future upgrades to the NP-MRD are planned.

Artifical Intelligence↗

Measurement of the A dependence of the ν μ charged-current quasielasticlike cross section as a function of muon and proton kinematics at ⟨E ν ⟩ ∼ 6 GeV

The first simultaneous measurements of the 𝜈 𝜇 quasielasticlike cross section on C, CH, H 2 ⁡O, Fe, and Pb targets as a function of kinematic imbalance variables in the plane transverse to the incoming neutrino direction are presented. These variables combine the muon and proton information to provide a new way to disentangle the effects of the nucleus in quasielasticlike processes. The data were obtained using a wideband 𝜈 𝜇 beam with ⟨E 𝜈 ⟩ ∼ 6 GeV. Cross-section ratios of the different target materials to CH are also shown. These measurements are used to explore the nature of the cross-section 𝐴 scaling, as well as initial and final state interaction effects. Comparisons are made to predictions from a number of commonly used neutrino Monte Carlo event generators. The range of predictions of the different models tends to cover the data but the degree and consistency of the agreement suffers in regions, and on higher 𝐴 targets, where the final state interactions are expected to be more pronounced.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Randomized low-rank decompositions of nuclear three-body interactions

First-principles simulations of many-fermion systems are commonly limited by the computational requirements of processing large data objects. As a remedy, we propose the use of low-rank approximations of three-body interactions, which are the dominant such limitation in nuclear physics. We introduce a randomized decomposition technique to handle the excessively large matrix dimensions and study the sensitivity of low-rank properties to interaction details. The developed low-rank three-nucleon interactions are benchmarked in ab initio simulations of few- and many-body systems. Exploiting low-rank properties provides a promising route to extend the microscopic description of atomic nuclei to large systems where storage requirements exceed the computational capacities of the most advanced high-performance computing facilities.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Traffic Signal Control for Large-Scale Urban Traffic Networks: Real-World Experiments using Vision-Based Sensors

Effective control of traffic signals plays a critical role in ensuring smooth vehicle flow in urban areas. Expertly engineered traffic signal controllers can considerably minimize travel delays and enhance sustainability. In this paper, the team proposes the Model Predictive Control (MPC) traffic signal control strategy using real-time traffic flow data from a vision-based camera as feedback information. Also, a realistic signal timing plan that considers National Electrical Manufacturers Association (NEMA) constraints has been developed to be applied to real-world scenarios. The primary aim is to reduce the number of vehicles across all links in the controlled area, thereby optimizing traffic flow and reducing energy consumption. To validate the proposed method, several real-life experiments were conducted at 24 intersections in Chattanooga, Tennessee, by collaborating with traffic field engineers. These experiments demonstrated significant performance improvements in comparison to the existing method.

data processing↗

PMU Data Quality and Sensor Health Monitoring

Phasor Measurement Units (PMUs) play a critical role in the evolution of the electric power industry by providing high-precision, real-time monitoring of essential power system metrics. However, effectively detecting abnormalities and critical events from PMU data is a complex task, complicated by intricate temporal patterns, a scarcity of labeled data for training algo- rithms, and constraints on online computational power. In this study, we apply TranAD, an innovative algorithm that combines transformer architectures with the refinement of adversarial learning, to both synthetic and real-world PMU datasets for developing a data quality and sensor online health monitoring platform for utilities. Our findings reveal that TranAD not only provides efficient detection and localization but also enhances the detail with which abnormalities are detected, marking a a significant step forward in the field of clean data acquisition processes for power system monitoring

deep neural network, machine learning (ML)↗

Enhanced Machine-Learning Flow for Microwave-Sensing Systems for Contaminant Detection in Food

The presence of foreign bodies in packaged food is a serious concern for both fnal consumers (allergies, injuries, choking) and food manufacturers (reputation and economic losses). In particular, low-density plastics, glass and wood splinters are hard to detect even by the most advanced X-ray imagers. One solution is Machine-Learning-based Microwave Sensing (MLMWS): a non-invasive, contactless, and real-time method which uses a machine-learning (ML) classifer to analyze the scattered microwaves from the irradiated target object. In this paper, we want to extend our previous work about contaminant detection in cocoa-hazelnut spread jars by proposing an enhanced ML flow to increase the accuracy of the ML classifier. For the first time in this case study, we use a multi-class classifier, we train it with scattering parameters measured at multiple microwave frequencies, with a new pre-processing scaler, data augmentation, quantization-aware training and a pruning schedule. The results show a contaminant detection multi-class accuracy of 94.167% with a latency of 26 µs when targeting an AMD/Xilinx Kria K26 FPGA. Finally, we released our datasets publicly to OpenML.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

On the anatomy of acoustic emission

Abrupt, local frictional fault failure comprises a displacement that is normally accompanied by acoustic emission (AE)—an impulsive elastic wave broadcast with an amplitude proportional to particle velocity. The aggregate of these displacements is the basic fault motion. In laboratory shear experiments, the examination of a sequence of laboratory earthquakes includes continuous measurements of fault motion and the associated AE that is broadcast. From these measurements, connections between the fault motion and cumulative sum of the AE amplitude can be identified. The composition of the AE broadcasts reveals inhomogeneity in the fault mechanical structure from which they arise. This inhomogeneity can be decomposed into a time invariant AE component and an articulated AE component. The articulated AE component serves as a “state of the fault diagnostic” that follows a distinctive pattern to fault failure. Thus, the articulated AE component can be used directly to monitor the state of the fault.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

MODAQ (Modular Offshore Data Acquisition System) [SWR-24-58]

The National Renewable Energy Laboratory (NREL) has developed a data acquisition system to support laboratory (component) and in-water validation and testing of Marine Renewable Energy (MRE) technologies. The Modular Ocean Data AcQuisition System (MODAQ) is a robust system built on industrial grade hardware that can acquire high-quality and summary data from many sensor and instrument types. MODAQ includes modules to perform real-time QC screening, system health monitoring, and send alerts. MODAQ is also part of a larger data management, processing, visualization, and archiving system.

Raye, Robert↗

CoreMS AutoQC Uploader

The invention is a self contained software utility that is deployed on the computer controlling a mass spectrometer. The purpose of the software is to monitor a given directory for files matching a user specified criteria and automatically upload matching files to a remote server as well as trigger a request that the data be processed by the cloud based CoreMS software

Rabus, Jordan [Pacific Northwest National Laborato↗

Seismic H2: Version 1.0

Seismic-H2 is an integrated software package for geological hydrogen reservoir simulation, optimization, and leakage monitoring. The package includes multiple components: (1) code used for modeling seismic wave propagation in 3D heterogeneous elastic media based on finite-difference method to support detection of geological hydrogen storage reservoir leakage; (2) 3D reservoir simulations of leaks from an underground reservoir and 3D simulations of saline aquifers and depleted gas reservoirs; (3) seismic monitoring costs of passive and active seismic monitoring required for UHS; (4) rock physics calculations and interpolations for converting the reservoir simulations from part (2) into the elastic media models in part (1); (5) pre-processing seismic data; and lastly (6), a GUI interface that combines these different components.

Creasy, Neala↗

sourcePy

Pollutant source identification techniques (of which there are many variations) are either locked behind researchers writing their own code for each use case or GUI platforms that are easy to use but inflexible and opaque. The Python package sourcePy brings together many of the pollutant source identification algorithms, giving the user full control out of the box. It aims to create a platform for source identification experiments where the full analysis from beginning to end can be done in Python, with a level of specificity in design that isn't available in the GUI options. sourcePy provides users with a few key features: -A Python interface with HYSPLIT, which can be used to generate trajectories and concentration plumes -Several Python classes which standardize the preparation and processing of data related to source identification experiments -Example scripts and notebooks that allow even new python users to get started with their own experiments quickly -Visualization methods

Arseneau, Isaac [Oak Ridge National Laboratory (OR↗

Iterative ML and Experiments for Emerging VOCs

SAND2026-17074O Iterative ML and Experiments for Emerging VOCs is a tool that analyzes and predicts the behaviors of SARS-CoV-2 variants. It processes experimental data on ACE2 (the receptor for the SARS-CoV-2 virus that allows it to infect the cell) and antibody binding using machine learning models, including neural networks, to forecast ACE2 interactions and variant expression. The tool employs transfer learning and global epistasis modeling, integrating public datasets with proprietary data to enhance prediction accuracy. Additionally, it fits concentration-response curves to determine dissociation constants and generates visualizations to support research findings, thereby aiding in the identification of new antibodies for emerging variants of concern. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Sheffield, Thomas [Sandia National Lab. (SNL-NM), ↗

Data for Pilot-Scale Processing of Miscanthus x giganteus for Recovery of Anthocyanins Integrated with Production of Microbial Lipids and Lignin-Rich Residue

Chemical-free hydrothermal pretreatment of Miscanthus x giganteus (Mxg) at the lab scale using high liquid-to-solid ratios resulted in the recovery of anthocyanins and enhanced enzymatic digestibility of residual biomass. In this study, the process is scaled up by using a continuous hydrothermal pretreatment reactor operated at a low liquid-to-solid ratio (50 % w/w solids) as an important step towards commercialization. Anthocyanin yield was 70 % w/w at the pilot scale (50 kg of Mxg), compared to the 94 % w/w yield achieved at the lab scale (0.5 g of Mxg). The pretreated biomass was subsequently refined mechanically using a disc mill to increase the accessibility of cellulose by cellulases. Enzymatic saccharification of the pretreated and disc-milled residue yielded 238 g/L sugar concentration by operating in fed-batch mode at 50 % w/v solids content. Two strains of Rhodosporidium toruloides were evaluated for converting the hydrolysate sugars into microbial lipids, and strain Y-6987 had the highest lipid titer (11.0 g/L). Further, the residue left after enzymatic saccharification was determined to be enriched 1.7-fold in the lignin content. This lignin-rich residue has value as a feedstock for the production of sustainable aviation fuel precursors and other high-value lignin-based chemicals. Hence the proposed biorefinery based on Mxg creates an opportunity for generating revenue from multiple high-value products. As the demand for biofuels and biobased products is rising, the biorefinery products from Mxg would create a niche in the industrial sector.

Conversion↗

High fidelity actuator line data from 9 turbine wind farm simulations using ExaWind

This data was generated with the ExaWind code suite (https://github.com/Exawind) to investigate the performance of different Active Wake Mixing turbine control in a wind farm situated in a stable atmospheric boundary layer. All cases correspond to a 3x3 wind farm in a 10km x 10km domain using a total mesh size that varied between 1.6 X 10^9 to 1.85 X 10^9 grid cells. The simulations were run across 1800-2000 GPUs on Frontier. The case description and data generation process is fully documented in Yalla, G. R., Brown, K., Cheung, L., Houck, D., deVelder, N., and Balaji, J. (2025). "Estimating annual energy production of wake mixing control strategies including comparisons to wake steering." Wind Energy Sciences (https://doi.org/10.5194/wes-2025-250).

17 WIND ENERGY↗