Search NASA⌕ Search

SEARCH · Search NASA

Results for “data analytics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

Developing Fluorescence-Based Sensors to Support Rare Earth Element Separation

Rare earth elements (REEs) are essential to most renewable energy technologies. Unfortunately, as we transition to sustainable energy production, the demand for REEs is rapidly growing well beyond current rates of production. As a result, novel means of efficient, scalable, and easily adaptable methods for processing primary and recycle feedstocks are needed. Development and integration of sensors for highly selective in-line monitoring can support more efficient design and testing of such novel separation processes, as well as more cost-effective deployment of those separation flowsheets. Work here will explore the application of fluorescence spectroscopy, a highly sensitive and selective technique, to quantify multiple lanthanides in complex mixtures including known interferents or quenching agents. Results include identification of the optimal excitation wavelength and the limit of detection of various rare earth elements as well as the performance of data-science-based quantification approaches in streams where “unknowns” are present. Overall, the data science tools in conjunction with optical sensor data were able to quantify analytes in the presence of other lanthanides which can be anticipated in the actual industrial stream. Here we include characterization of lanthanides in a microfluidic device similar to those used in new process development. This study demonstrates the capability of utilizing fluorescence spectroscopy to quantify analytes in a complicated solution matrix, suggesting this is a successful approach for in-line monitoring to optimize the separation efficiency in an industrial stream.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Analytical Identification Method of Generalized Short‐Circuit Ratio Using Phasor Measurement Units

This paper introduces a novel analytical approach for the identification of the admittance matrix and the generalized short-circuit ratio (gSCR) in power systems integrated with renewable energy sources. The proposed method leverages voltage and current measurements from phasor measurement units (PMUs) to construct a least squares objective function, which is then solved using matrix calculus and partial derivatives. Unlike conventional optimization algorithms, this approach provides an analytical solution that substantially reduces data requirements, enabling the efficient and accurate identification of the gSCR with smaller datasets. Additionally, its fixed computational complexity allows for real-time updates as new data are collected, ensuring continuous refinement of the system of equations and enabling rapid, precise gSCR calculations. The method also exhibits strong robustness against measurement noise, making it well-suited for practical applications in dynamic power systems. The combination of reduced data requirements, real-time adaptability, noise robustness and fixed computational load establishes this method as a highly efficient and reliable tool for real-time power system stability analysis. Case studies on an EPRI 36-bus system demonstrate the method's effectiveness, highlighting its accuracy in closely matching true gSCR values, even under diverse disturbances and noisy conditions.

Han, Zelei [Hohai University, Nanjing (China)] (OR↗

Calorimetric wire detector for measurement of atomic hydrogen beams

A calorimetric detector for minimally disruptive measurements of atomic hydrogen beams is described. The calorimeter measures heat released by the recombination of hydrogen atoms into molecules on a thin wire. As a demonstration, the angular distribution of a beam with a peak intensity of ≈ 10 16 atoms/(cm 2 s) is measured by translating the wire across the beam. The data agree well with an analytic model of the beam from the thermal hydrogen atom source. Using the beam shape model, the relative intensity of the beam can be determined to 5% precision or better at any angle.

Astaschov, M. [Johannes Gutenberg Univ., Mainz (Ge↗

A miniaturized feedstocks-to-fuels pipeline for screening the efficiency of deconstruction and microbial conversion of lignocellulosic biomass

Sustainably grown biomass is a promising alternative to produce fuels and chemicals and reduce the dependency on fossil energy sources. However, the efficient conversion of lignocellulosic biomass into biofuels and bioproducts often requires extensive testing of components and reaction conditions used in the pretreatment, saccharification, and bioconversion steps. This restriction can result in a significant and unwieldy number of combinations of biomass types, solvents, microbial strains, and operational parameters that need to be characterized, turning these efforts into a daunting and time-consuming task. Here we developed a high-throughput feedstocks-to-fuels screening platform to address these challenges. The result is a miniaturized semi-automated platform that leverages the capabilities of a solid handling robot, a liquid handling robot, analytical instruments, and a centralized data repository, adapted to operate as an ionic-liquid-based biomass conversion pipeline. The pipeline was tested by using sorghum as feedstock, the biocompatible ionic liquid cholinium phosphate as pretreatment solvent, a “one-pot” process configuration that does not require ionic liquid removal after pretreatment, and an engineered strain of the yeast Rhodosporidium toruloides that produces the jet-fuel precursor bisabolene as a conversion microbe. By the simultaneous processing of 48 samples, we show that this configuration and reaction conditions result in sugar yields (~70%) and bisabolene titers (~1500 mg/L) that are comparable to the efficiencies observed at larger scales but require only a fraction of the time. We expect that this Feedstocks-to-Fuels pipeline will become an effective tool to screen thousands of bioenergy crop and feedstock samples and assist process optimization efforts and the development of predictive deconstruction approaches.

09 BIOMASS FUELS↗

TEAMER - Field Demonstration of MarineSitu’s Marine Energy Monitoring Tools - CRADA 664 (Abstract)

In order to effectively monitor for marine life around marine energy devices and thus minimize the risk of collision, multiple sensors working in coordination and augmented with around-the-clock automated monitoring algorithms need to be installed in challenging high-energy tidal and wave environments. Such systems are often too expensive for widespread adoption, or lack sufficient sensors or smarts to enable around-the-clock, real-time monitoring without human involvement. MarineSitu has been working to tackle this problem by developing a low-cost, combined sonar and stereo camera sensor array with connected real-time AI-based algorithms for automatically detecting marine life in these marine energy suitable environments. In this TEAMER project with Pacific Northwest National Lab (PNNL), MarineSitu will be testing this novel sensor system for the first time in the high-energy tidal channel environment at PNNL’s Marine and Coastal Research Lab. Throughout this deployment, MarineSitu will be monitoring their system and running analytics on the sensor’s data in real-time. Meanwhile, PNNL Data Scientists and Ocean Engineers, will be evaluating the system’s effectiveness and ease of use both as a tool for plug-and-play environmental monitoring and novel environmental monitoring research. In doing so, the team will improve MarineSitu’s system and software, produce insightful data products, and develop novel visualizations and AI algorithms for combining and analyzing the data produced by systems like MarineSitu’s.

16 TIDAL AND WAVE POWER↗

TEAMER – Field Demonstration of MarineSitu’s Marine Energy Monitoring (Abstract)

In order to effectively monitor for marine life around marine energy devices and thus minimize the risk of collision, multiple sensors working in coordination and augmented with around-the-clock automated monitoring algorithms need to be installed in challenging high-energy tidal and wave environments. Such systems are often too expensive for widespread adoption, or lack sufficient sensors or smarts to enable around-the-clock, real-time monitoring without human involvement. MarineSitu has been working to tackle this problem by developing a low-cost, combined sonar and stereo camera sensor array with connected real-time AI-based algorithms for automatically detecting marine life in these marine energy suitable environments. In this TEAMER project with Pacific Northwest National Lab (PNNL), MarineSitu will be testing this novel sensor system for the first time in the high-energy tidal channel environment at PNNL’s Marine and Coastal Research Lab. Throughout this deployment, MarineSitu will be monitoring their system and running analytics on the sensor’s data in real-time. Meanwhile, PNNL Data Scientists and Ocean Engineers, will be evaluating the system’s effectiveness and ease of use both as a tool for plug-and-play environmental monitoring and novel environmental monitoring research. In doing so, the team will improve MarineSitu’s system and software, produce insightful data products, and develop novel visualizations and AI algorithms for combining and analyzing the data produced by systems like MarineSitu’s.

16 TIDAL AND WAVE POWER↗

Battery Life Prediction Using Reduced-Order Physics Models and Machine Learning (CRADA Final Report)

Phase 1 (Original CRADA, plus no-cost extension modifications #1-3, 6/1/2017 to 3/13/2021): The Australian Department of Defence (AUDoD) is performing accelerated aging tests of Li-ion batteries to benchmark their reliability and degradation characteristics. Using its previously developed battery lifetime predictive model framework, the National Laboratory of the Rockies (NLR) will develop analytical models based the AUDoD data to predict lifetime of the multiple Li-ion battery chemistries under real-world use scenarios of interest to AUDoD. The NLR model is based on physical degradation mechanisms encountered by Li-ion batteries and has been previously validated. Phase 2 (CRADA modification #4, plus no-cost extension modification #5, 2/22/2021 to 3/30/2025): Train and support Australian Department of Defence personnel to use NLR software for model-based estimation of Li-ion battery lifetime using accelerated battery aging data collected by the Australian Department of Defence. Under separate DOE funding from 2019 to 2021, NLR enhanced its battery life-prediction software using machine learning algorithms to automate portions of the model-fitting process, requiring significantly less labor and expert judgment and also adding uncertainty quantification, increasing statistical rigor. Under Phase 2, NLR will customize NLR Software and provide it to AuDoD. NLR will enhance its NLR Model to capture aging modes of AuDoD's multi-cell modules, including cell-balancing effects. NLR will develop example single-cell and multi-cell models based on one AuDoD battery aging dataset. NLR will train AuDoD personnel on NLR Software. By the conclusion of the project, NLR will have provided AuDoD the training materials, a user manual and software needed to perform their own analysis of additional and/or future battery aging datasets.

33 ADVANCED PROPULSION SYSTEMS↗

Pathway-based analyses of gene expression profiles at low doses of ionizing radiation

Radiation exposure poses a significant threat to human health. Emerging research indicates that even low-dose radiation once believed to be safe, may have harmful effects. This perception has spurred a growing interest in investigating the potential risks associated with low-dose radiation exposure across various scenarios. To comprehensively explore the health consequences of low-dose radiation, our study employs a robust statistical framework that examines whether specific groups of genes, belonging to known pathways, exhibit coordinated expression patterns that align with the radiation levels. Notably, our findings reveal the existence of intricate yet consistent signatures that reflect the molecular response to radiation exposure, distinguishing between low-dose and high-dose radiation. Moreover, we leverage a pathway-constrained variational autoencoder to capture the nonlinear interactions within gene expression data. By comparing these two analytical approaches, our study aims to gain valuable insights into the impact of low-dose radiation on gene expression patterns, identify pathways that are differentially affected, and harness the potential of machine learning to uncover hidden activity within biological networks. This comparative analysis contributes to a deeper understanding of the molecular consequences of low-dose radiation exposure.

63 RADIATION, THERMAL, AND OTHER ENVIRON. POLLUTAN↗

Value-based Insights from the Implementation of Hierarchical Control for Energy Savings and Demand Response in Residential Premises

As the adoption of distributed energy resources and electric vehicles at residential customer premises increases exponentially, behind-the-meter assets can be utilized to achieve energy cost reduction and demand response through coordination and control strategies. A hierarchical control architecture from the utility headend to residential premises is implemented to attain these objectives. This paper extracts the values from the development, implementation, and deployment of that control hierarchy. The development of the control philosophy is built upon the existing advanced metering infrastructure, communication protocols, and industry-compatible application programming interfaces. Results are presented visually with analytical insights by utilizing the data from hardware-in-the-loop testing and simulation analysis out of the collected data from the field.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Livewire: A Model Platform for Data Quality Assessment and AI Readiness Across DOE Missions

High-quality, well-governed data is essential for accelerating discovery and achieving operational excellence across DOE and national laboratory missions. The Livewire Data Platform is a DOE-supported platform that offers automated assessments of data quality, standardization, provenance, and Artificial Intelligence (AI) readiness. It allows researchers and data practitioners to systematically and easily evaluate datasets against established governance criteria and prepare them for advanced analytics. Livewire addresses critical challenges in DOE's data ecosystem with integrated capabilities for metadata validation, provenance tracking, and schema alignment. This platform's automated workflows assist users in identifying data quality gaps, enhancing interoperability between datasets collected from various stakeholders, and ensuring compliance with DOE data standards, all while reducing manual curation efforts. Additionally, we will discuss its AI readiness framework, which is being developed to prepare datasets for training models, developing advanced analytic tools, and machine learning applications. Using some of the more than one hundred tabular datasets on Livewire, processed with this open-source methodology, we will demonstrate how Livewire can serve as a model for scalable, standards-driven data management. This approach provides a pathway to leverage existing and future datasets within the DOE, boosting innovation and efficiency across national laboratories.

33 - ADVANCED PROPULSION SYSTEMS↗

Initial Calibration of Large Timing Arrays for the LHC

In preparation for HL-LHC operation a number of new detector systems are being constructed with timing precision on physics objects of <50 picoseconds. These time stamps will reduce the level of pileup induced backgrounds in this LHC phase where the number of interactions per crossing will reach of order 100-200. In the case of CMS, three new systems have initially to be corrected for the usual amplitude walk resulting from the effect of variations in signal size on leading edge timing. In these systems the resulting timing spread (ie walk) ranges from one to four nanoseconds. In the following note we advocate approaching this initial calibration for walk as a calculable correction given early calibration during commissioning -- rather than depending on special collider data to perform the calibration. We derive a simple analytic expression for the walk correction and confirm its effectiveness with lab data.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Using Best Basis Inventory Data to Direct Strategies for Real Time Monitoring of Hanford High Level Waste – 26226

The potential to accelerate the processing of low- and high-level tank waste by applying real-time monitoring (RTM) of chemical and physical properties has prompted research into the suitability of multiple analytical methods for that purpose. The broad variety of waste stream properties and the large number of analytes of interest (as evidenced by Waste Acceptance Criteria (WAC) and Process Control Limit (PCL) lists) lead to an overwhelming set of possible analytical scenarios. This report describes the use of Best Basis Inventory (BBI) data to find the most relevant analytical targets for the specific case of monitoring the blending of High Level Waste from multiple tanks prior to introduction into a vitrification facility. Campaigns for blending this waste to minimize the risk of exceeding WACs and PCLs have been proposed. However, the predicted compositions of the blended materials do not incorporate any uncertainties that may be associated with the representativeness of the waste layer samples or the laboratory analyses that generated the BBI data. Also, they do not include any uncertainty associated with the precision of collecting highly specific fractions of the layers during a blending campaign or any inhomogeneities that may exist in those layers. Monte Carlo methods are used to apply uncertainties to the compositions of the individual layers specified in the campaign recipes. The resulting variations in the compositions of the blended materials allow estimation of the risks of exceeding WACs and PCLs for each campaign. A critical subset of WACs/PCLs – NOx, NaK, AlFeZr, and S – are especially at risk of being exceeded in multiple campaigns. These analytes should be the focus of instrument development. We also have extracted the expected solid/supernate distribution for these analytes, which establishes important performance criteria for individual analytical methods. The BBI data also permits an understanding of the different chemical forms in which the analytes appear. Thus, the need to establish instrumental sensitivity to these forms can be gainfully addressed. Although concentrating on one specific application – the blending of tank waste - this approach should be generalizable for the analysis of other possible RTM applications for waste processing.

Lascola, Robert [Savannah River National Laborator↗

Data for: A hybrid biophysical-machine learning framework for diurnal surface energy flux estimation using proximal sensing

Thermal-based remote sensing of surface energy fluxes has traditionally relied on high spatial resolution satellite data with revisit frequencies on the order of weeks. In this study, we evaluate a biophysics-based analytical surface energy balance model for predicting latent energy (LE) and sensible heat (H) fluxes using proximal sensing observations. The Surface Temperature Initiated Closure (STIC1.2) model has been extensively validated across a wide range of spatial and temporal scales using various satellite-derived thermal datasets. Here we extend this validation by applying STIC at sub-hourly temporal resolution over multiple growing seasons for four distinct agricultural systems. We further develop and evaluate novel STIC variants that incorporate machine learning (ML) techniques to eliminate the need for specific surface energy balance observations, specifically net radiation and soil heat flux, thereby enhancing model applicability in data-sparse settings. The integration of an ML component to estimate surface available energy is shown to have strong predictive performance for both LE (R2 = 0.81-0.94) and H (R2 = 0.46-0.72) across all agricultural systems examined here, demonstrating the potential of hybrid biophysical – machine learning approaches for surface energy balance modeling with minimal data requirements. This study concludes with a novel application of explainable machine learning (exML) to diagnose sources of model error. This exML framework attributes residual prediction errors to both model input variables and environmental drivers not explicitly included in the simulation experiments. This approach provides a new pathway for improving model design and integrating previously overlooked yet influential variables into future model iterations.

Agricultural Sciences↗

MAGIC: M arching Cubes Isosurface Uncertainty Visualization for G auss i an Uncertain Data With Spatial C orrelation

Here, in this paper, we study the propagation of data uncertainty through the marching cubes algorithm for isosurface visualization for correlated uncertain data. Consideration of correlation has been shown paramount for avoiding errors in uncertainty quantification and visualization in multiple prior studies. Although the problem of isosurface uncertainty with spatial data correlation has been previously addressed, there are two major limitations to prior treatments. First, there are no analytical formulations for uncertainty quantification of isosurfaces when the data uncertainty is characterized by a Gaussian distribution with spatial correlation. Second, as a consequence of the lack of analytical formulations,existing techniques resort to a Monte Carlo sampling approach, which is expensive and difficult to integrate into visualization tools. To address these limitations, we present a closed-form framework to efficiently derive uncertainty in marching cubes level-sets for Gaussian uncertain data with spatial correlation (MAGIC). To derive closed-form solutions, we leverage the Hinkley's derivation on the ratio of Gaussian distributions. With our analytical framework, we achieve a significant speed-up and enhanced accuracy of uncertainty quantification over classical Monte Carlo methods. We further accelerate our analytical solutions using many-core processors to achieve speed-ups up to 585× and integrability with production visualization tools for broader impact. We demonstrate the effectiveness of our correlation-aware uncertainty framework through experiments on meteorology, urban flow, and astrophysics simulation datasets.

Gaussian↗

Parity-odd four-point correlation function from the DESI data release 1 luminous red galaxy sample

The parity-odd four-point function provides a unique probe of fundamental symmetries and potential new physics in the large-scale structure of the Universe. We present measurements of the parity-odd four-point function using the Dark Energy Spectroscopic Instrument (DESI) DR1 luminous red galaxy (LRG) sample and assess its detection significance. Our analysis considers both auto- and cross-correlations, using two complementary approaches to the covariance: (i) the full analytic covariance matrix applied to the uncompressed data vector, and (ii) a compressed data vector combined with a hybrid covariance matrix constructed from simulations and analytic estimates. When using the full analytic covariance matrix without corrections, we observe apparent auto-correlation signals with significance up to 4⁢𝜎. However, this excess is also consistent with a mismatch between the statistical fluctuations estimated from the simulations and those present in the real data. Our findings therefore suggest that the parity-odd signal in the current DESI DR1 LRG sample is consistent with zero. We note, however, that the low completeness of this sample may have a non-negligible impact on the detection sensitivity. Future data releases with improved completeness will be crucial for further investigation.

Hou, Jiamin [Ludwig-Maximilians-Universität; Unive↗

Filling the Gaps: A Bayesian Mixture Model for Imputing Missing Soil Water Content Data

ABSTRACT Soil water content (SWC) data are central to evaluating how soil moisture varies over time and space and influences critical plant and ecosystem functions, especially in water‐limited drylands. However, sensors that record SWC at high frequencies often malfunction, leading to incomplete timeseries and limiting our understanding of dryland ecosystem dynamics. We developed an analytical approach to impute missing SWC data, which we tested at six eddy flux tower sites along an elevation gradient in the southwestern United States. We impute missing data as a mixture of linearly interpolated SWC between the observed endpoints of a missing data gap and SWC simulated by an ecosystem water balance model (SOILWAT2). Within a Bayesian framework, we allowed the relative utility (mixture weight) of each component (linearly interpolated vs. SOILWAT2) to vary by depth, site and gap characteristics. We explored “fixed” weights versus “dynamic” weights that vary as a function of cumulative precipitation, average temperature, and time since the start of the gap. Both models estimated missing SWC data well ( R 2 = 0.70–0.88 vs. 0.75–0.91 for fixed vs. dynamic weights, respectively), but the utility of linearly interpolated versus SOILWAT2 values depended on site and depth. SOILWAT2 was more useful for more arid sites, shallower depths, longer and warmer gaps and gaps that received greater precipitation. Overall, the mixture model reliably gap‐fills SWC, while lending insight into processes governing SWC dynamics. This approach to impute missing data could be adapted to accommodate more than two mixture components and other types of environmental timeseries.

Ogle, Kiona [School of Informatics, Computing, and↗

PMDT: AI-Enabled Predictive Maintenance Digital Twins for Advanced Nuclear Reactors

Our team made substantial technical progress on various fronts during the course of the program. Multiple milestones were geared towards demonstrating the feasibility of machine learning based predictive maintenance digital twins towards reducing O&M costs, whereas some other milestones actually focused on identifying technical gaps and developing technologies such as humble AI to provide necessary robustness to the ML-based models. We were able to demonstrate in many cases that Machine learning-based methods can be successfully adapted for Nuclear plant environments especially for remote monitoring applications. Detailed analyses were carried out with plant and full scope simulation data along with capabilities of enhanced analytics to assess and set realistic expectations on cost reductions in O&M. These assessments are paving the way for investments towards reactor design improvements as well project planning for SMR projects as they develop and mature in the next few years. Technology developed under this program got direct visibility to GE Hitachi and their utility customers and resulted in positive intents to deploy some of the elements from design phase. The project additionally resulted in several reports, publications, software and data generation that will be useful in deployment and O&M services for BWRX300 fleets.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Integrated Hourly Meteorological Database of 20 Meteorological Stations (1981-2022) for Watershed Function SFA Hydrological Modeling

This dataset contains (a) a script “R_met_integrated_for_modeling.R”, and (b) associated input CSV files: 3 CSV files per location to create a 5-variable integrated meteorological dataset file (air temperature, precipitation, wind speed, relative humidity, and solar radiation) for 19 meteorological stations and 1 location within Trail Creek from the modeling team within the East River Community Observatory as part of the Watershed Function Scientific Focus Area (SFA). As meteorological forcings varied across the watershed, a high-frequency database is needed to ensure consistency in the data analysis and modeling. We evaluated several data sources, including gridded meteorological products and field data from meteorological stations. We determined that our modeling efforts required multiple data sources to meet all their needs. As output, this dataset contains (c) a single CSV data file (*_1981-2022.csv) for each location (20 CSV output files total) containing hourly time series data for 1981 to 2022 and (d) five PNG files of time series and density plots for each variable per location (100 PNG files). Detailed location metadata is contained within the Integrated_Met_Database_Locations.csv file for each point location included within this dataset, obtained from Varadharajan et al., 2023 doi:10.15485/1660962. This dataset also includes (e) a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and (f) a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type. Review the (g) ReadMe_Integrated_Met_Database.pdf file for additional details on the script, methods, and structure of the dataset.The script integrates Northwest Alliance for Computational Science and Engineering’s PRISM gridded data product, National Oceanic and Atmospheric Administration’s NCEP-NCAR Reanalysis 1 gridded data product (through the `RCNEP` R package, Kemp et al., doi:10.32614/CRAN.package.RNCEP), and analytical-based calculations. Further, this script downscales the input data into hourly frequency, which is necessary for the modeling efforts.

54 ENVIRONMENTAL SCIENCES↗