Search NASA⌕ Search

SEARCH · Search NASA

Results for “data standard”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 271 records · Page 15

Chemistry imaging and distribution analysis of rare earth elements in coal using LIBS and LA-ICP-MS instruments

Currently, demand for rare earth elements (REEs) increased significantly. Coal is actively evaluated as potential economic sources for extraction of REEs. Here, in this work, laser-induced breakdown spectroscopy (LIBS) was evaluated for rapid estimation of REEs content and their distribution in the natural coal samples. The results were compared with similar laser ablation–inductively coupled plasma–mass spectrometry (LA-ICP-MS) measurements. Thirteen coal samples (nine standard samples and five natural samples) were used in this study. Powder samples were pressed into pellets while coal chunks were directly ablated for data recording. Pellets of the powder standard samples were used to optimize the data acquisition system and then data recorded with this optimized system was used to identify the proper data acquisition and analysis models. After establishing the proper data acquisition system and analysis model using the standard samples, natural coal samples in powder form and their chunks were utilized to record LIBS and LA-ICP-MS spectra. Multivariate calibration models were developed using four of the natural samples, which were evaluated by predicting the REE content in the fifth sample. Principal component analysis was performed on the LIBS data obtained from the natural samples and it classified all the samples with high accuracy. Two-dimensional (2D) elemental mapping on coal chunk samples was also performed using both LIBS and LA-ICP-MS to study the distribution of REEs in the samples. The resulting elemental images and their correlations can be used to infer mineral distributions.

01 COAL, LIGNITE, AND PEAT↗

A call to standardize metrics for monitoring baleen whales near marine construction activities

Effective monitoring is necessary to protect marine mammal species during the construction of offshore infrastructure. The tools for detecting or monitoring marine mammals span traditional (e.g., visual observers, optical cameras), to newer (e.g., passive acoustic monitoring, infrared cameras, tags), and emerging (e.g., satellite imagery, environmental DNA, dimethyl sulfide concentration) technologies. Some are better suited for use during offshore development; however, peer-reviewed literature does not typically evaluate and report on the performance of these various technologies. We define a minimum set of metrics related to efficacy (i.e., confusion matrix, precision and recall, probability of missed mitigation), detection range (i.e., maximum and reliable detection range, spatial resolution), and data delivery (i.e., detection latency, system reliability, temporal resolution) that we recommend are needed to assess the utility of monitoring technologies for this purpose. Following a literature review of relevant studies, we highlight which publications reported these metrics and used multiple technologies to compare relative performance. We also emphasize the benefits of multi-modal approaches and recommend performance assessments through modeling or large-scale collaborative field testing. These metrics will standardize data collection, reporting, and analysis; promote consistent and comparable results; and foster collaboration among developers, regulatory agencies, and scientists. This may lead to the co-development of technology that achieves multiple goals, has greater application, and can answer research questions while collecting data to fulfill permitting requirements. These metrics may also inform decisions on what systems regulatory agencies might consider using and reduce monitoring costs, which is critical to support the marine sector's rapid growth alongside marine mammal conservation.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Analysis of Rig Parameter Data Using Drilling Process Modeling Constraints, Volume 1: Summary of Utah FORGE Wells 16A(78)-32, 56-32, 78B-32 and 16B(78)-32

Drill rig parameter measurements are routinely used during deep well construction to monitor and guide drilling conditions for improved performance and reduced costs. While insightful into the drilling process, these measurements are of reduced value without a standard to aid in data evaluation and decision making. In the main body of this work (Volume 1), a method is demonstrated whereby rock reduction model constraints are used to interpret drilling response parameters; the method could be applied in real-time to improve decision-making in the field and to further discern technology performance during post-drilling evaluations. Drilling parameters are evaluated using laboratory-validated rock reduction models for predicting the phenomenological response of drag bits (Detournay and Defourny, 1992) in computational algorithms. The method presented has applicability to development of advanced analytics on future geothermal wells using real-time electronic data recording for improved performance and reduced drilling costs. A drilling cost model is also used to show the tradeoff between rate of penetration and bit life and the influence on interval drilling costs. Details of the bit specifications and performance are cataloged in an independent volume, documented under separate cover, for each of the four wells, and include Volume 2: Utah FORGE 16A(78)-32; Volume 3: Utah FORGE 56-32; Volume 4: Utah FORGE 78B-32 and Volume 5: Utah FORGE 16B(78)-32.

15 GEOTHERMAL ENERGY↗

Physics-Informed Active Learning With Simultaneous Weak-Form Latent Space Dynamics Identification

The parametric greedy latent space dynamics identification (gLaSDI) framework has demonstrated promising potential for accurate and efficient modeling of high-dimensional nonlinear physical systems. However, it remains challenging to handle noisy data. Here, to enhance robustness against noise, we incorporate the weak-form estimation of nonlinear dynamics (WENDy) into gLaSDI. In the proposed weak-form gLaSDI (WgLaSDI) framework, an autoencoder and WENDy are trained simultaneously to discover intrinsic nonlinear latent-space dynamics of high-dimensional data. Compared with the standard sparse identification of nonlinear dynamics (SINDy) employed in gLaSDI, WENDy enables variance reduction and robust latent space discovery, therefore leading to more accurate and efficient reduced-order modeling. Furthermore, the greedy physics-informed active learning in WgLaSDI enables adaptive sampling of optimal training data on the fly for enhanced modeling accuracy. The effectiveness of the proposed framework is demonstrated by modeling various nonlinear dynamical problems, including viscous and inviscid Burgers' equations, time-dependent radial advection, and the Vlasov equation for plasma physics. With data that contains 5%–10% Gaussian white noise, WgLaSDI outperforms gLaSDI by orders of magnitude, achieving 1%–7% relative errors. Compared with the high-fidelity models, WgLaSDI achieves 121 to 1779x speed-up.

97 MATHEMATICS AND COMPUTING↗

NuMaCo

The Nuclear Material and Composition (NuMaCo) server is a web-based platform for storing, managing, and serving material composition data. It includes the SCALE Standard Composition Library and the PNNL Compendium, along with their associated references and uncertainty data. NuMaCo also provides tools to generate, compare, and validate material input definitions for both SCALE and MCNP.

Skutnik, Steve (000000016441135X)↗

Systems Innovation: Modernization & Efficiencies for ESH&Q Reviews

Environmental compliance reviews at INL have traditionally been managed through fragmented systems, relying on multiple spreadsheets and manual processes. This inefficiency led to time-consuming status updates and redundant tasks, such as manually sending reminder emails and transferring data from Excel to the Environmental Review Process (ERP). Initial attempts to streamline these processes using Power Automate and Excel revealed significant limitations, necessitating a more comprehensive solution. To address these immediate inefficiencies, automated workflows were developed using Power Automate. These workflows were designed to send scheduled status update reminders and capture responses through standardized forms, with submitted data flowing directly into centralized Excel trackers. This automation reduced the administrative burden, improved data accuracy, and enabled faster, more consistent reporting. Specifically, email automation achieved a 65% efficiency gain, while data integration saw a 48% improvement, resulting in 91% of project statuses being updated within two months. Despite the improvements brought by Power Automate, the fragmented nature of the review processes persisted. To further enhance efficiency and accuracy, the Integrated Review Tool (IRT) was developed. The IRT aims to centralize review initiation and connect team systems, creating an interconnected data infrastructure that preserves team autonomy while enhancing overall efficiency. This tool automates email reminders, centralizes reviews, and streamlines data integration, significantly improving the accuracy and efficiency of environmental compliance reviews. The design and development of the IRT involved advanced systems methodology, process mapping, project management, and collaboration with subject matter experts. The minimum viable product design is 100% complete, and system development is currently underway, with expected outcomes including a centralized entry point for all ESH&Q reviews, automated routing, real-time tracking and analytics, AI integration, and a user-friendly interface. This project demonstrates the potential of leveraging automation and integrated systems to enhance efficiency, accuracy, and decision-making in environmental reporting and compliance processes at INL.

99 - GENERAL AND MISCELLANEOUS↗

A Performant, Scalable Processing Pipeline for High‐Quality and FAIR Environmental Sensor Data

High-resolution environmental monitoring is necessary to record, understand, and predict biogeochemical and ecological changes particularly in coastal systems but brings significant challenges in processing and making rapidly available the resulting data. The COMPASS-FME project established a network of coastal observational sites across the Chesapeake Bay and western Lake Erie regions extensively instrumented with soil, vegetation, and weather sensors logging data every 15 min. Our data processing framework, written in R and completely open source, prioritizes rapid model-experiment iteration and makes biogeochemical data rapidly available for quality assurance/quality control, analysis, and model ingestion. This pipeline is distinguished by a standardized and modular approach to data curation, extensive metadata and documentation, and its high performance. These attributes combine to make biogeochemical data rapidly accessible across COMPASS-FME and the broader community. Flexible, powerful, and reproducible approaches to handling high-volume environmental data are crucial for accelerating biogeosciences research.

Pennington, Stephanie C. [Pacific Northwest Nation↗

Distributed Wind Monitoring Best Practices

Accessible performance and operational data have been identified as a key enabler for distributed wind energy industry advancement. While utility-scale wind turbines benefit from reliable and continuous supervisory control and data acquisition (SCADA)-based monitoring platforms, monitoring of the U.S. fleet of distributed wind (DW) turbines has been more inconsistent, unreliable, and sometime difficult to access. Without fleet monitoring data, the industry will never understand and thus work to improve turbine under-performance and reliability issues. For the DW industry to scale up, attract investors, and boost credibility, fleetwide monitoring must be robust and reliable, select data must be made accessible to stakeholders, and the data must be in a format useful to users. To help move the industry toward a more standardized, accessible stream of monitoring data, this distributed wind monitoring best practices report attempts to cover topics including key monitoring channels, hardware, communication strategies, and accessibility. Strategic engagement with DW original equipment manufacturers (OEMs), service providers, lab and university researchers, testing organization, certification bodies, end users and solar photovoltaic (PV) monitoring experts has enabled a better understanding of the current state-of-the-art of monitoring and aided in articulating this set of best practices that will guide OEMs toward harmonized monitoring strategies, aimed at a future goal of achieving accessible performance and operational data for the entire fleet of U.S. distributed wind turbines.

17 WIND ENERGY↗

Mist

Determining the appropriate material data is often a bottleneck for performing calculations/simulations of industrial/experimental processes and resulting material structures and properties. Beyond the time it takes to find the appropriate values in the literature, many judgement calls are involved in choosing the values. These judgement calls can lead to inconsistencies between steps in research workflow, where different material parameter values are used. Mist solves this problem by providing a mechanism to store, share, and use material information in convenient human-readable and machine-readable formats. Mist has an extensible ontology for defining a wide variety of material information, currently focused on metal alloy applications. Examples include: alloy composition, density, liquidus temperature, and the coefficient of thermal expansion. Mist converts between standardized machine-readable data formats (e.g. JSON), specialized input format for simulation tools, and human-readable documents (e.g. LaTeX, Markdown). For parameters defined by an equation (e.g. a polynomial function) or a list of tabulated values, Mist can evaluate parameter values at requested conditions. Mist also provides an API for direct usage of the Mist data structures in calculations, if supported.

DeWitt, Stephen [Oak Ridge National Laboratory (OR↗

Identifying Sample Provenance From SEM/EDS Automated Particle Analysis via Few-Shot Learning Coupled With Similarity Graph Clustering

Automated particle analysis (APA) provides a vast amount of compositional data via energy-dispersive X-ray spectroscopy along with size and shape data via scanning electron microscopy for individual particles in a sample. In many instances, APA data are leveraged to support identification of the source of a sample based on the detection of particles of a specific composition. Often, the particles that provide context make up a minuscule portion of the sample. Additionally, the interpretation of complex samples can be difficult due to the diversity of compositions both in the mixture and within a particle. In this work, we demonstrate a method to compute and cluster similarity graphs that describe inter-particle relationships within a sample using a multi-modal few-shot learning neural network. Here, as a proof-of-concept, we show that samples known to have been exposed to gunshot residue can be distinguished from samples occasionally mistaken for gunshot residue. Our workflow builds upon standard APA techniques and data processing methods to unveil additional information in a readily interpretable and quantitatively comparable format.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Divide and conquer: using RhizoVision Explorer to aggregate data from multiple root scans using image concatenation and statistical methods

Roots are important in agricultural and natural systems for determining plant productivity and soil carbon inputs. Sometimes, the amount of roots in a sample is too much to fit into a single scanned image, so the sample is divided among several scans, and there is no standard method to aggregate the data. Here, we describe and validate two methods for standardizing measurements across multiple scans: image concatenation and statistical aggregation. We developed a Python script that identifies which images belong to the same sample and returns a single, larger concatenated image. These concatenated images and the original images were processed with RhizoVision Explorer, a free and open-source software. An R script was developed, which identifies rows of data belonging to the same sample and applies correct statistical methods to return a single data row for each sample. These two methods were compared using example images from switchgrass, poplar, and various tree and ericaceous shrub species from a northern peatland and the Arctic. Most root measurements were nearly identical between the two methods except median diameter, which cannot be accurately computed by statistical aggregation. We believe the availability of these methods will be useful to the root biology community.

59 BASIC BIOLOGICAL SCIENCES↗

Developing a Database of Bio-based Materials for Building Envelope Applications

Oak Ridge National Laboratory (ORNL) has been funded by the Department of Energy (DOE) to help accelerate the introduction of building envelope materials that would reduce the carbon footprint of the buildings sector. The DOE’s Building Technologies Office has historically sought to resolve the knowledge gaps regarding the energy efficiency and moisture durability of building envelope systems and to develop the data, guidance, and tools needed to facilitate rapid industry adoption of high-performance, moisture-managed envelope systems. This project will help accelerate the widespread acceptance of a new generation of building materials developed specifically with the intent of reducing the carbon footprint of buildings. We have produced a database of hygrothermal transport properties on low embodied carbon building materials that can be added to energy and durability simulation tools. Properties that were measured include density, heat capacity, thermal conductivity as a function of temperature and relative humidity, moisture dependent permeance, and sorption isotherms as a function of relative humidity. These data sets were measured following consensus national standards using state-of-the-art facilities. The data has been compiled and is being made available to building designers who require these data to assess these new materials in their designs. We will publish the data and seek its addition to reference databases such as the ASHRAE Handbook of Fundamentals.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

In Situ Data Analysis Through Physics-informed Tensor Decompositions (LDRD Final Report)

We introduce a new low-dimensional model of high-dimensional numerical simulation data based on low-rank tensor decompositions. Our new model aims to minimize differences between the model data and simulation data as well as functions of the model data and functions of the simulation data. This novel approach to dimensionality reduction of simulation data provides a means of directly incorporating quantities of interests and invariants associated with conservation principles associated with the simulation data into the low-dimensional model, thus enabling more accurate analysis of the simulation without requiring access to the full set of high-dimensional data. Computational results of applying this approach to two standard low-rank tensor decompositions of data arising from simulation of combustion and plasma physics are presented.

97 MATHEMATICS AND COMPUTING↗

Ultra-Fast Non-Volatile Resistive Switching Devices with Over 512 Distinct and Stable Levels for Memory and Neuromorphic Computing

Low-current multilevel programmability with inherent non-volatility and high stability of resistance states is required for both multi-bit memory storage and deep learning accelerators but is difficult to achieve. Here, in a resistive switching system, this work realizes >512 (>9 bits) distinct non-volatile conductance levels with stable retention for each state with current levels down to the nanoampere range, highly promising for potential integration with small processing nodes with ultra-low power consumption requirements. This is achieved by demonstrating a new thin film design concept that encompasses three key features: an ultra-thin epitaxial oxygen ionic switching layer that provides a tunable energy barrier at the bottom electrode, an overcoat amorphous layer that acts as an ion migration barrier for stable state retention, and a partial conductive filament as a localized electronic transport channel to the epitaxial switching layer. A large dynamic resistance range of up to seven orders of magnitude is achieved with reset-free transitions among intermediate states, and programmability is demonstrated with ultra-fast (20 ns) pulses. Artificial neural network (ANN) simulations, based on the experimental performance and its non-idealities, demonstrate close-to-ideal inference accuracies for various Modified National Institute of Standards and Technology (MNIST) data sets.

36 MATERIALS SCIENCE↗

Predictive analytics of selections of russet potatoes

We explore the application of machine learning algorithms specifically to enhance the selection process of Russet potato (Solanum tuberosum L.) clones in breeding trials by predicting their suitability for advancement. This study addresses the challenge of efficiently identifying high-yield, disease-resistant, and climate-resilient potato varieties that meet processing industry standards. Leveraging manually collected data from trials in the state of Oregon, we investigate the potential of a wide variety of state-of-the-art binary classification models. The dataset includes 1086 clones, with data on 38 attributes recorded for each clone, focusing on yield, size, appearance, and frying characteristics, with several control varieties planted consistently across four Oregon regions from 2013 to 2021. We conduct a comprehensive analysis of the dataset that includes preprocessing, feature engineering, and imputation to address missing values. We focus on several key metrics such as accuracy, F1-score, and Matthews correlation coefficient (MCC) for model evaluation. The top-performing models, namely a feedforward neural network classifier (Neural Net), a histogram-based gradient boosting classifier (HGBC), and a support vector machine classifier (SVM), demonstrate consistent and significant results. To further validate our findings, we conducted a simulation study using the aims, data-generating mechanisms, estimands, methods, and performance measures (ADEMP) framework, simulating different data-generating scenarios to assess model robustness and performance through true positive, true negative, false positive, and false negative distributions, area under the receiver operating characteristic curve (AUC-ROC) and MCC. The simulation results highlight that non-linear models like SVM and HGBC consistently show higher AUC-ROC and MCC than logistic regression, thus outperforming the traditional linear model across various distributions, and emphasizing the importance of model selection and tuning in agricultural trials. Variable selection further enhances model performance and identifies influential features in predicting trial outcomes. The findings emphasize the potential of machine learning in streamlining the selection process for potato varieties, offering benefits such as increased efficiency, substantial cost savings, and judicious resource utilization. Our study contributes insights into precision agriculture and showcases the relevance of advanced technologies for informed decision-making in breeding programs.

60 APPLIED LIFE SCIENCES↗

Search for high-mass resonances in a final state comprising a gluon and two hadronically decaying W bosons in proton-proton collisions at $\sqrt{s}$ = 13 TeV

A search for high-mass resonances decaying into a gluon, g, and two W bosons is presented. A Kaluza-Klein gluon, g$_{KK}$, decaying in cascade via a scalar radion R, g$_{KK}$ → gR → gWW, is considered. The final state studied consists of three large-radius jets, two of which contain the products of hadronically decaying W bosons, and the third one the hadronization products of the gluon. The analysis is performed using proton-proton collision data at $\sqrt{s}$ = 13 TeV collected by the CMS experiment at the CERN LHC during 2016–2018, corresponding to an integrated luminosity of 138 fb$^{−1}$. The masses of the g$_{KK}$ and R candidates are reconstructed as trijet and dijet masses, respectively. These are used for event categorization and signal extraction. No excess of data events above the standard model background expectation is observed. Upper limits are set on the product of the g$_{KK}$ production cross section and its branching fraction via a radion R to gWW. This is the first analysis examining the resonant WW+jet signature and setting limits on the two resonance masses in an extended warped extra-dimensional model.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Equivariant, safe and sensitive — graph networks for new physics

This study introduces a novel Graph Neural Network (GNN) architecture that leverages infrared and collinear (IRC) safety and equivariance to enhance the analysis of collider data for Beyond the Standard Model (BSM) discoveries. By integrating equivariance in the rapidity-azimuth plane with IRC-safe principles, our model significantly reduces computational overhead while ensuring theoretical consistency in identifying BSM scenarios amidst Quantum Chromodynamics backgrounds. The proposed GNN architecture demonstrates superior performance in tagging semi-visible jets, highlighting its potential as a robust tool for advancing BSM search strategies at high-energy colliders.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Search for exotic Higgs boson decays H → 𝒜𝒜 with 𝒜 → γγ in events with a semi-merged topology in proton-proton collisions at $\sqrt{s}=13$ TeV

A search for exotic Higgs boson decays H → 𝒜𝒜, with 𝒜 → γγ is presented, using events with a semi-merged topology. One of the hypothetical particles, 𝒜, is assumed to decay promptly into a semi-merged diphoton system reconstructed as a single photon-like object, while the other 𝒜 decays into two resolved photons. The search is performed using proton-proton collision data collected by the CMS experiment at $\sqrt{s}=13$ TeV, corresponding to an integrated luminosity of 138 fb −1 . The data agree with the standard model background expectation. Upper limits are set on the product of the Higgs boson production cross section and the branching fraction, σ(pp → H) ℬ(H → 𝒜𝒜 → 4γ), which range from 0.264 to 0.005 pb at 95% confidence level, for 𝒜 masses in the range 1 < m 𝒜 < 15 GeV. These limits are the most stringent to date in the 1–5 GeV m 𝒜 range.

Beyond Standard Model↗