Search NASA⌕ Search

SEARCH · Search NASA

Results for “Data analysis methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 469 records · Page 26

GIS Supported Optimal Site Selection for Coastal Structure Integrated Wave Energy Converters: Preprint

There is an urgent need for adaptative engineering towards more resilient coastal communities, and Coastal Structure Integrated Wave Energy Converters (CSI-WECs) are a promising solution. CSI-WECs are wave energy converters (WECs) that are built into coastal protection structures, such as breakwaters. These devices provide the dual benefits of coastal protection and local energy production, and unlike other WECs, maximizing energy production is not always the main objective. CSI-WECs are located near the shore, where the wave resource is lower, thus site selection for these devices differs from the typical offshore WECs. Other attributes of a site that may be more important than wave power include existing coastal structures, port proximity, electric transmission line proximity, and location of disadvantaged communities. Geospatial information systems (GIS) interfaces can be used to easily visualize geospatial data that represents these difference kinds of criteria important for the determination of optimal marine energy sites. Multi-Criteria Decision Analysis (MCDA) is a geospatial analysis method that allows for the evaluation of multiple, usually overlapping, criteria. This project applies GIS-based MCDA methods to two distinct case studies in Puerto Rico and California for CSI-WEC site selection. The two study sites contrast in terms of wave resource, coastal hazards, and local energy needs. This research demonstrates the utility of applying an MCDA framework within GIS to facilitate efficient site selection for devices with unique characteristics in different use cases.

coastal protection↗

The Profiled Feldman-Cousins Method for Confidence Interval Construction for the Nova 3-Flavor Oscillation Analysis

The small interaction cross-section of neutrinos makes experimental neutrino physics particularly responsive to technological advancements. A significant development leveraged by the NOvA experiment is large-scale parallel processing, enabling novel computational approaches to longstanding experimental challenges. Central to managing the resulting high-throughput data is NOvA’s implementation of the Freight Train model, designed for efficient data production and handling.This dissertation details the methodology and execution of the NOvA 2024 3-Flavor Oscillation Analysis, supported by a comprehensive dataset spanning ten years. It emphasizes frequentist results refined through the Feldman-Cousins (FC) technique, specifically addressing confidence interval corrections in parameter estimation. The computational intensity associated with Feldman-Cousins arises from extensive Monte Carlo simulations, which were substantially mitigated through parallel computing on the Perlmutter supercomputer at the National Energy Research Scientific Computing Center (NERSC), employing the MPI framework.To further enhance computational efficiency, an Importance Sampling method is introduced and evaluated, demonstrating significant potential to reduce complexity, particularly in exploring extreme parameter space regions. This thesis presents both the successful application of advanced computational resources and the development of sophisticated statistical techniques, aiming to enhance the precision and scope of neutrino oscillation analyses.

Dye ajdye11190@gmail.com, Andrew Joseph [Mississip↗

PV Fleet Performance Data Initiative 2026 Update

We provide an update on the PV Fleet Performance Data Initiative at the 2026 PV Reliability Workshop. Our latest runs incorporate additional data sources and an integrated analysis pipeline run on our Kestrel HPC cluster. Initial degradation findings suggest that single-axis tracked PV systems exhibit higher performance loss rates than fixed-tilt systems, an increase of 0.5 %/yr, almost double. We discuss multiple methods for identifying stuck tracker rows, which are suspected to be a contributor to the enhanced degradation. Through satellite image detection and data-driven approaches we address the topic of identifying when stuck trackers are occuring and to what extent the problem exists. Preliminary results suggest that the increased performance loss detected for the tracked systems would be consistent with stuck tracker rows affecting on the order of 5% - 10% of the system.

14 SOLAR ENERGY↗

ZMPY3D: accelerating protein structure volume analysis through vectorized 3D Zernike moments and Python-based GPU integration

Abstract Motivation Volumetric 3D object analyses are being applied in research fields such as structural bioinformatics, biophysics, and structural biology, with potential integration of artificial intelligence/machine learning (AI/ML) techniques. One such method, 3D Zernike moments, has proven valuable in analyzing protein structures (e.g., protein fold classification, protein–protein interaction analysis, and molecular dynamics simulations). Their compactness and efficiency make them amenable to large-scale analyses. Established methods for deriving 3D Zernike moments, however, can be inefficient, particularly when higher order terms are required, hindering broader applications. As the volume of experimental and computationally-predicted protein structure information continues to increase, structural biology has become a “big data” science requiring more efficient analysis tools. Results This application note presents a Python-based software package, ZMPY3D, to accelerate computation of 3D Zernike moments by vectorizing the mathematical formulae and using graphical processing units (GPUs). The package offers popular GPU-supported libraries such as CuPy and TensorFlow together with NumPy implementations, aiming to improve computational efficiency, adaptability, and flexibility in future algorithm development. The ZMPY3D package can be installed via PyPI, and the source code is available from GitHub. Volumetric-based protein 3D structural similarity scores and transform matrix of superposition functionalities have both been implemented, creating a powerful computational tool that will allow the research community to amalgamate 3D Zernike moments with existing AI/ML tools, to advance research and education in protein structure bioinformatics. Availability and implementation ZMPY3D, implemented in Python, is available on GitHub (https://github.com/tawssie/ZMPY3D) and PyPI, released under the GPL License.

Lai, Jhih-Siang (ORCID:0000000156775890)↗

Efficient Signal Processing in BOTDA: Utilizing PCA and PCA-Based Neural Networks for Temperature Monitoring

This work presents a comparative analysis of the various signal processing techniques used in the Brillouin gain spectrum (BGS) peak estimation. Traditional fitting methods such as Lorentzian curve fitting (LCF) are slow and less effective in noisy data. PCA-based methods were tested on the experimental data: A Euclidian distance-based approach, and a probabilistic deep neural network (PDNN) based approach, both using 5 principal components to represent a single BGS. Both methods significantly reduce computational time with respect to LCF, whereas PDNN offers uncertainty insights along with the parameter value. Measuring a range of temperatures, analyzing accuracy, and speed, it can be concluded that PCA trained PDNN outperforms other methods, and appears to be helpful in scenario where large datasets are generated.

Brillouin optical time domain analysis↗

U.S. Average Market Carbon Dioxide Production Baseline Documentation For 45Q Life Cycle Analysis: Version 1.0 (2018-2022)

U.S. Average Benchmark life cycle inventory documentation for market carbon dioxide production for calendar years 2018 - 2022. The documentation includes a technical report describing data sources and statistical methods used and an Excel file showing calculations. The calculations performed are used in the 45Q benchmark json-ld dataset, compatible with the NETL CO2U LCA Guidance Toolkit. Summary impact results will be included in the NETL CO2U LCA Documentation Spreadsheet within the Toolkit. To access the model referenced in the report, please visit https://www.netl.doe.gov/energy-analysis/details?id=21915ba8-d2cf-43ea-b9b2-1d76363a46a7

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

A Computational Review of Privacy-Preserving Mechanisms for the Smart Grid

Smart grid technologies have rapidly become one of the largest and most comprehensive sources of data for the modern utility. For the most part, data streams are seen as an essential tool that enable utilities to carry their day-to-day business operations, but they also create the need for efficient and secure data management strategies. In the context of the smart grid, ensuring data privacy is becoming an increasing concern due to a combination of factors that range from shifts in operational paradigms and rapid technology evolution to changes in legislation. Furthermore, researchers have highlighted the risks associated with improperly protected energy records. For example, energy consumption data from homes could be used to infer the behaviors and habits of home occupants through activity recognition or user profiling (Fan, 2017), which may lead to unfair service pricing, targeted advertising, or other personal security violations. Similarly, Electric Vehicles’ (EVs) charging metadata could be used to reveal private information about the owner such as their payment methods, preferred charging stations, and other locational and timing information that could be used to reconstruct the vehicle owner’s behaviors. The privacy of user data, even when used for statistical analysis or machine learning training processes, also needs to be carefully considered, as an individual’s private traits may still be vulnerable if their inclusion/exclusion greatly impacts the result or could be linked to a public dataset through cross-reference. The breach of user privacy also has severe impacts for organizations that store, transmit, or work on the data in the form of diminishing the public’s trust in them while potentially incurring legal consequences (e.g., fines and suspensions under the European Union General Data Protection Regulation, Health Insurance Portability and Accountability Act, etc.). Because of these risks, several privacy-preserving mechanisms are available to help organizations comply with privacy legislations and prevent the unauthorized and malicious use of user data. In light of these concerns, this report focuses on performing a computational review of privacy-preserving mechanisms that have received a significant amount of interest in literature. It specifically focuses on 1) homomorphic encryption, 2) zero-knowledge proofs, 3) differential privacy, and 4) federated learning. It is worth noting that although many of the methods presented in this document rely on cryptographic primitives, their intent is not to provide perfect secrecy, but rather to enable users to maintain privacy, and thus they shall not be compared or equated to other constructs that are aimed to address cybersecurity constructs.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Gold L-shell emission spectroscopy as a temperature diagnostic in laser-driven experiments

Measurements of plasma conditions within gold hohlraums are challenging but provide important information about hohlraum energy transport. To date, there are no reliable passive diagnostic approaches that produce information about the hohlraum plasma conditions. Gold L-shell emission spectroscopy is a promising method to measure the electron temperature in multi-keV gold hohlraum plasmas. In this study, we compare time-resolved measurements of gold L-shell spectra against a more conventional electron temperature-sensitive diagnostic, K-shell emission from zinc, a mid-Z dopant. The Au L-shell and Zn K-shell spectra are taken simultaneously. The analysis of both sets of data exhibits good agreement with a singular rad-hydro model, suggesting a similar evolution in the mean electron temperature of the hohlraum plasma. These findings support the use of L-shell spectroscopy of gold as a method to study gold plasma conditions.

Emission spectroscopy↗

Effect of solvothermal synthesis parameters on the crystallite size and atomic structure of cobalt iron oxide nanoparticles

We here investigate how the synthesis method affects the crystallite size and atomic structure of cobalt iron oxide nanoparticles. By using a simple solvothermal method, we first synthesized cobalt ferrite nanoparticles of ca. 2 and 7 nm, characterized by Transmission Electron Microscopy (TEM), Small Angle X-ray scattering (SAXS), X-ray and neutron total scattering. The smallest particle size corresponds to only a few spinel unit cells. Nevertheless, Pair Distribution Function (PDF) analysis of X-ray and neutron total scattering data shows that the atomic structure, even in the smallest nanoparticles, is well described by the spinel structure, although with significant disorder and a contraction of the unit cell parameter. These effects can be explained by the surface oxidation of the small nanoparticles, which is confirmed by X-ray near edge absorption spectroscopy (XANES). Neutron total scattering data and PDF analysis reveal a higher degree of inversion in the spinel structure of the smallest nanoparticles. Neutron total scattering data also allow magnetic PDF (mPDF) analysis, which shows that the ferrimagnetic domains correspond to ca. 80% of the crystallite size in the larger particles. A similar but less well-defined magnetic ordering was observed for the smallest nanoparticles. Finally, we used a co-precipitation synthesis method at room temperature to synthesize ferrite nanoparticles similar in size to the smallest crystallites synthesized by the solvothermal method. Structural analysis with PDF demonstrates that the ferrite nanoparticles synthesized via this method exhibit a significantly more defective structure compared to those synthesized via a solvothermal method.

77 NANOSCIENCE AND NANOTECHNOLOGY↗

Circularity Futures Workshop Series: Summary Report

The aim of this report is to synthesize key feedback received from the three-part Circularity Futures workshop series held in Spring 2024. The workshop series was conducted by the National Renewable Energy Laboratory (NREL) on behalf of U.S. Department of Energy, Office Energy Efficiency and Renewable Energy (EERE), and was broken into three workshops: Workshop 1 - Circularity Analysis Needs and Priorities; Workshop 2 - Circularity Metrics and Indicators; and Workshop 3 - Circularity Data. Together, the workshops focused on identifying the existing priorities and gaps in the circularity modeling space, understanding different stakeholders' use and interpretation of circularity metrics and indicators, identifying common data gaps and data quality challenges, and assessing the robustness of available solutions. The workshop series brought a diverse group of stakeholders - including representatives from U.S. government offices, national labs, nonprofit organizations, industry, and academia - to collect first-hand feedback on needs, priorities, challenges and opportunities in the circularity modeling and analysis space. The workshop discussions highlighted numerous common needs, priorities and challenges among the interviewed groups. Several topics were frequently discussed, including: 1) Circularity as a pathway for sustainable economic growth: While circularity is generally defined in terms of resource conservation and reducing wasteful disposal of materials, participants agreed that circular strategies should serve broader economic, environmental, and social goals. It is therefore crucial for circularity analysis to look beyond waste reduction and instead evaluate a variety of impact metrics such as cost savings, job creation, air quality, and pollutant emissions. Mutli-criteria decision-making frameworks may be useful for making sense of disparate metrics and evaluating tradeoffs between impact categories.; 2) Economic and social factors are not well understood: Underdevelopment of existing end-of-life (EOL) management infrastructure, inconsistent standardization codes and policy space in reusing recycled content, and suboptimal collection and sorting strategies collectively contribute to uncertainty about the economic potential of circular pathways. The latter observation is consistent among all technologies but more emphasized for renewable energy systems. Social impacts of circularity practices are less understood and less researched than other sustainability aspects.; 3) Inconsistent methods for assessing emerging technologies: LCA and TEA results vary widely depending on the assumptions made with regards to market adoption of new technologies. Emerging technologies suffer limited availability of data needed to conduct a robust circularity analysis. Yet, understanding projected impacts of proposed nascent technology is a key need for different stakeholder groups.; and 4) Lack of temporally and geospatially explicit data: There is a need for open data that represents variations in circularity technologies over time and location. The lack thereof leads to aggregated and potentially misrepresented results in circularity analysis. Sensitivity analyses should be included to verify whether options perceived as more sustainable align with real-world practices.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Dense autoencoders, clustering techniques, and semi-supervised learning for HPGe $γ$-spectra

Classifying high-resolution gamma spectra by their isotopic content is an essential task in nuclear forensics and other applications. Traditional analysis methods are often time-intensive, but machine learning (ML) may help analysts quickly process many spectra. Such methods tend to rely on abundant, well-labeled data for training. Historical gamma data exists in various fields but is not uniformly useful for supervised ML due to inconsistent labeling. Here, to address some of these challenges, we present a method to classify and organize unlabeled data from high-purity germanium detectors using an autoencoding neural network (autoencoder). We trained dense autoencoders to compress gamma data into latent representations that enable efficient data characterization. By clustering the encoded spectra or lower-dimensional mappings of them, we identified and removed portions of over-abundant data categories, resulting in a more balanced dataset and improved autoencoder performance. This encoding and clustering pipeline also enabled the organization of spectra into self-consistent categories. Finally, we found that encoded representations showed potential as inputs for semi-supervised learning of nuclide identification (NID) labels, achieving an average F1 score of 0.85 ± 0.03 when mapping encodings to a set of 65 isotope labels.

Autoencoders↗

Evidence for B + → K + ν ν ¯ decays

We search for the rare decay B + → K + ν ν ¯ in a 362 fb − 1 sample of electron-positron collisions at the ϒ ( 4 S ) resonance collected with the Belle II detector at the SuperKEKB collider. We use the inclusive properties of the accompanying B meson in ϒ ( 4 S ) → B B ¯ events to suppress background from other decays of the signal B candidate and light-quark pair production. We validate the measurement with an auxiliary analysis based on a conventional hadronic reconstruction of the accompanying B meson. For background suppression, we exploit distinct signal features using machine learning methods tuned with simulated data. The signal-reconstruction efficiency and background suppression are validated through various control channels. The branching fraction is extracted in a maximum likelihood fit. Our inclusive and hadronic analyses yield consistent results for the B + → K + ν ν ¯ branching fraction of [ 2.7 ± 0.5 ( stat ) ± 0.5 ( syst ) ] × 10 − 5 and [ 1.1 − 0.8 + 0.9 ( stat ) − 0.5 + 0.8 ( syst ) ] × 10 − 5 , respectively. Combining the results, we determine the branching fraction of the decay B + → K + ν ν ¯ to be [ 2.3 ± 0.5 ( stat ) − 0.4 + 0.5 ( syst ) ] × 10 − 5 , providing the first evidence for this decay at 3.5 standard deviations. The combined result is 2.7 standard deviations above the standard model expectation. Published by the American Physical Society 2024

Astronomy & Astrophysics↗

Visual Analytics of Performance of Quantum Computing Systems and Circuit Optimization

Driven by potential exponential speedups in business, security, and scientific scenarios, interest in quantum computing is surging. This interest feeds the development of quantum computing hardware, but several challenges arise in optimizing application performance for hardware metrics (e.g., qubit coherence and gate fidelity). In this work, we describe a visual analytics approach for analyzing the performance properties of quantum devices and quantum circuit optimization. Our approach allows users to explore spatial and temporal patterns in quantum device performance data and it computes similarities and variances in key performance metrics. Detailed analysis of the error properties characterizing individual qubits is also supported. We also describe a method for visualizing the optimization of quantum circuits. The resulting visualization tool allows researchers to design more efficient quantum algorithms and applications by increasing the interpretability of quantum computations.

Chae, Junghoon↗

A new method for measuring refractory corrosion of ceramics in glass

Abstract Nuclear waste glass vitrification furnaces are lined with refractory ceramic blocks to contain the molten glass. The refractory liner is susceptible to corrosion and has a finite service lifetime. For this reason, predicting the refractory corrosion in contact with molten glass is integral to estimating melter service lifetime. Standardized laboratory tests varying time and temperature are commonly performed to estimate refractory material loss as a function of glass composition. These data are time and resource‐intensive to collect and are susceptible to considerable measurement error. In order to accelerate glass formulation and design for nuclear waste vitrification, methods are needed to increase laboratory‐scale throughput while maintaining data quality. In this work, a method to remove the residual glass from a corroded coupon using hydrofluoric acid is presented that accelerates the throughput of sample analysis while simultaneously facilitating more accurate measurements.

Amoroso, Jake W. [Savannah River National Laborato↗

Heating effects on jack pine pyrogenic organic matter properties from a pyrocosm study in 2022

This dataset contains data associated with the preprint “Fire removes preexisting pyrogenic organic matter from the ecosystem through the mechanisms of both direct combustion and increasing mineralizability” (Luo et al., 2025b), which is the complementary study to the published paper “Reburning pyrogenic organic matter: a laboratory method for dosing dynamic heat fluxes from above” (Luo et al., 2025a). We designed a full-factorial experiment with different burial depths of jack pine (Pinus banksiana Lamb) pyrogenic organic matter (PyOM) (Surface, 1 cm, and 5 cm) and different heat-flux profiles (High, Low, and Control) to examine how subsequent fires affect the properties of preexisting PyOM. We measured total carbon (C), pH, dissolved organic carbon (DOC), dissolved inorganic carbon (DIC), and mineralized C (as CO₂-C, from a 12-week incubation).We found that high heat flux and/or surface placement resulted in substantial direct C losses through combustion. Intermediate heat exposure produced both combustion losses and increases in DOC and mineralizability, which may have complex long-term implications: an increased dissolved fraction of PyOM may promote downward transport into mineral soils and potentially contribute to deeper, longer-term C storage, but it may also make PyOM more susceptible to microbial decomposition. Under the lowest heat flux and deepest burial, most PyOM was retained, and changes in DOC and C mineralization were minimal. Finally, PyOM pH, an important chemical property, decreased under low-temperature heating but increased under higher temperatures.We uploaded pH data for all samples (“pH_of_all_samples.csv”); pH and temperature-related data (peak temperature and degree hours) for samples in High and Low heat-flux treatments (“pH_vs_peakT_and_degree_hours_only_for_heated_samples.csv”); total C data (“CN_pct_C_stock_C_loss_in_samples.csv”); DOC and DIC data (“doc_dic.csv”); and mineralized C (CO₂-C) data (“CO2-C_all_original.csv”). Additional details can be found in the Methods & Sampling section.All datasets uploaded to ESS-DIVE are clearly labeled, cleaned, and include both raw and derived data, ready for reuse in other analyses. All analysis code and raw datasets are also available on GitHub: https://github.com/MengmengLuo/Fire-removes-preexisting-pyrogenic-organic-matter-from-the-ecosystem.

54 ENVIRONMENTAL SCIENCES↗

Overview of the SCEC/USGS Community Stress Drop Validation Study Using the 2019 Ridgecrest Earthquake Sequence

We present initial findings from the ongoing Community Stress Drop Validation Study to compare spectral stress-drop estimates for earthquakes in the 2019 Ridgecrest, California, sequence. This study uses a unified dataset to independently estimate earthquake source parameters through various methods. Stress drop, which denotes the change in average shear stress along a fault during earthquake rupture, is a critical parameter in earthquake science, impacting ground motion, rupture simulation, and source physics. Spectral stress drop is commonly derived by fitting the amplitude-spectrum shape, but estimates can vary substantially across studies for individual earthquakes. Sponsored jointly by the U.S. Geological Survey and the Statewide (previously, Southern) California Earthquake Center our community study aims to elucidate sources of variability and uncertainty in earthquake spectral stress-drop estimates through quantitative comparison of submitted results from independent analyses. The dataset includes nearly 13,000 earthquakes ranging from M 1 to 7 during a two-week period of the 2019 Ridgecrest sequence, recorded within a 1° radius. Here, in this article, we report on 56 unique submissions received from 20 different groups, detailing spectral corner frequencies (or source durations), moment magnitudes, and estimated spectral stress drops. Methods employed encompass spectral ratio analysis, spectral decomposition and inversion, finite-fault modeling, ground-motion-based approaches, and combined methods. Initial analysis reveals significant scatter across submitted spectral stress drops spanning over six orders of magnitude. However, we can identify between-method trends and offsets within the data to mitigate this variability. Averaging submissions for a prioritized subset of 56 events shows reduced variability of spectral stress drop, indicating overall consistency in recovered spectral stress-drop values.

58 GEOSCIENCES↗

Large-x partons from Jefferson Lab to the LHC

This project addressed the longitudinal quark and gluon structure of hadrons, a key topic in the 2015 and 2023 NSAC long range plans for nuclear science. The research encompassed and combined theoretical studies with advanced Quantum Chromo Dynamics (QCD) analysis of experimental data, from Jefferson Lab to the Large Hadron Collider. The results have pushed the boundaries in both aspects, and capitalized on the results and methods developed in previous grant renewals.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Improving the Productivity and Performance of Large-Scale Integrated Algal Systems for Wastewater Treatment and Biofuel Production

The goal of this project was to develop and demonstrate an integrated system for algal biofuel production system and wastewater treatment that can produce low-cost drop-in biofuels. Experimental data and techno-economic analysis showed the ability to produce drop-in biofuels from wastewater derived algal biomass at a cost of $3.32 and identified methods to further reduce costs. In particular, when accounting for wastewater treatment cost savings relative to conventional processes, the proposed integrated system can support a negative minimum fuel selling price. This means the normal costs of wastewater treatment are sufficient to cover all the costs of biofuel production with the integrated system.

09 BIOMASS FUELS↗