Search NASA⌕ Search

SEARCH · Search NASA

Results for “machine data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Fortran Program for X-Ray Photoelectron Spectroscopy Data Reformatting

A FORTRAN program has been written for use on an IBM PC/XT or AT or compatible microcomputer (personal computer, PC) that converts a column of ASCII-format numbers into a binary-format file suitable for interactive analysis on a Digital Equipment Corporation (DEC) computer running the VGS-5000 Enhanced Data Processing (EDP) software package. The incompatible floating-point number representations of the two computers were compared, and a subroutine was created to correctly store floating-point numbers on the IBM PC, which can be directly read by the DEC computer. Any file transfer protocol having provision for binary data can be used to transmit the resulting file from the PC to the DEC machine. The data file header required by the EDP programs for an x ray photoelectron spectrum is also written to the file. The user is prompted for the relevant experimental parameters, which are then properly coded into the format used internally by all of the VGS-5000 series EDP packages.

Abel, Phillip B.↗

Predicting Dynamic-to-Static Correction Factor from Petrophysical Data and Chemostratigraphy using Unsupervised Machine Learning

Estimating static mechanical properties of stratigraphic layers is critical for optimizing subsurface engineering applications. To estimate dynamic-to-static correction factor F ds (static-to-dynamic Young’s modulus ratio) across the Caney shale interval in Oklahoma, USA, we integrated triaxial test measurements and petrophysical data, including well logs and X-ray fluorescence (XRF) using unsupervised machine learning (ML). We used a novel workflow that includes principal component analysis (PCA) to reduce data set dimensionality of well logs and XRF data sets—both separately and combined—creating three scenarios, and later applied inverse distance weighting (IDW) to derive F ds profiles for these scenarios. Furthermore, we applied K-means clustering on each scenario to predict depositional facies, and built a stiffness zonation profile through chemostratigraphic analysis of the terrigenous elements to validate the predicted F ds . The predicted F ds profile from each scenario using the PCA-IDW method was compared with the constant F ds approach from our previous study by calculating the root mean square error (RMSE). The combined data sets scenario yielded the lowest RMSE value of 0.113, while the RMSE values for the well logs and XRF scenarios were 0.131 and 0.129, respectively. In addition, the predicted F ds from the XRF scenario well-matched the stiffness zonation from the chemostratigraphic analysis that was built using the optimized K-means clustering of nine clusters for that scenario. These methods and findings offer a valuable tool for refining lithological classification and improving the F ds profile, potentially enhancing drilling and stimulation strategies for subsurface energy engineering applications.

clastic rock↗

Evaluation of surface water resources from machine-processing of ERTS multispectral data

The surface water resources of a large metropolitan area, Marion County (Indianapolis), Indiana, are studied in order to assess the potential value of ERTS spectral analysis to water resources problems. The results of the research indicate that all surface water bodies over 0.5 ha were identified accurately from ERTS multispectral analysis. Five distinct classes of water were identified and correlated with parameters which included: degree of water siltiness; depth of water; presence of macro and micro biotic forms in the water; and presence of various chemical concentrations in the water. The machine processing of ERTS spectral data used alone or in conjunction with conventional sources of hydrological information can lead to the monitoring of area of surface water bodies; estimated volume of selected surface water bodies; differences in degree of silt and clay suspended in water and degree of water eutrophication related to chemical concentrations.

Mausel, P. W.↗

A Data-Driven Method for Modeling Creep-Fatigue Stress- Strain Behavior Using Neural ODEs

In this paper, we introduce a data-driven machine learning approach for modeling one-dimensional stress–strain behavior under cyclic loading, utilizing experimental data from the nickel-based Alloy 617. The study employs uniaxial creep–fatigue test data acquired under various loading histories and compares two distinct neural network-based ODE models. The first model, known as the black-box model, comprehensively describes the strain–stress relationship using a Neural ODE equation. To interpret this black-box model, we apply the Sparse Identification of Nonlinear Dynamical Systems (SINDy) technique, transforming the black-box model into an equation-based model using symbolic regression. The second model, the Neural flow rule model, incorporates Hooke’s Law for the linear elastic component, with the nonlinear part characterized by a Neural ODE. Both models are trained with experimental data to accurately reflect the observed stress–strain behavior. We conduct a detailed comparison with the standard Chaboche model, which includes three back stresses. Our results demonstrate that the neural network-based ODE models precisely capture the experimental creep–fatigue mechanical behavior, exceeding the standard Chaboche model’s accuracy. Furthermore, an interpretable model derived from the black-box neural ODE model through symbolic regression achieves accuracy comparable to the Chaboche model, enhancing its interpretability. The results highlight the potential of neural network-based ODE models to depict complex creep–fatigue behavior, eliminating the necessity for experts to define a specific, material-focused model form.

creep-fatigue↗

Evaluation of Drilling Performance at The Geysers with Machine Learning Methods Using Geologic Data

A recent well, GDC-36, was drilled in The Geysers Geothermal Field served in a Department of Energy-industry to demonstrate improved drilling performance with polycrystalline diamond compact (PDC) bits. Both PDC and roller cone drill bits were used to drill this well. Key challenges encountered during drilling included lost circulation in the mud-drilled section, and bit damage interfacial severity in the deeper, air-drilled section. The objective of this study is to evaluate the drilling performance in relation to the local geological characteristics using machine learning methods. By applying K-clustering to the sonic log data, we were able to identify areas correlated with measured lost circulation. Also, the boundaries defined by clustering of the mineralogical and lithological data from the mud logs correlate well with interfacial severity during drilling. A random forest model was employed to build correlation between drilling data and rock strength. The confined compressive strength (CCS) of the rock in the training of the machine learning model was inferred from the dipole sonic log. The R-squared of the testing data is 0.78, and the RMSE (Root Mean Squared Error) is 0.06. The trained model was used to forecast rock strength for the section where sonic log data are not available. CCS could also be inferred from mud logs provided the relationship between mineralogy and rock strength is established through core testing data.

15 GEOTHERMAL ENERGY↗

Machine Learned Empirical Numerical Integrator from Simulated Data

Recently, a number of state-of-the-art surrogate machine learning (ML) models have been designed for global weather and climate prediction, which have been trained using reanalysis data products. Reanalysis data products are constructed using numerical model simulations that combine numerical integration of partial differential equations and parameterization schemes. These products are typically only archived and made available using coarsened spatial and temporal resolutions. This study explores the impact of the numerical generation methods used to produce the training datasets and the temporal resolution of those datasets on machine learning surrogate models. Using the nonlinear vector autoregression (NVAR) machine as an explainable ML technique, simple dynamical systems are emulated with ML models trained on data produced by three classical numerical integration schemes. NVAR is validated as a skillful ML method, capable of producing accurate predictions and, more importantly, reconstructing both the underlying dynamics and the numerical integration scheme used to generate the training data. However, the machine fails to generalize predictions on unseen test data generated by different numerical integration schemes, despite the underlying dynamical system being the same. This result provides a word of caution for the growing field of machine learning emulation of weather and climate dynamics. Furthermore, we illustrate using NVAR that training on temporally coarsened data may increase the required complexity of ML models and potentially introduce new numerical challenges. Finally, we discover that empirical integration schemes with arbitrary time-stepping sizes can be constructed directly from the data, which implies a potential for the development of empirical numerical integration schemes.

54 ENVIRONMENTAL SCIENCES↗

Machine learning for seismic low-frequency extrapolation

The cycle-skipping problem that plagues full waveform inversion (FWI) can be at least partially mitigated if low frequencies (which encode the kinematics of wave propagation in seismic data) are recorded. However, seismic sources and receivers are band-limited, so seismic data does not generally include signals down to 0 Hz. To improve our ability to solve the seismic inverse problem, one can synthesize this missing low-frequency (LF) content from the recorded high-frequency (HF) data using machine learning (ML) models. Deep learning models such as convolutional neural networks (CNNs) demonstrate impressive ability to perform low frequency extrapolation. However, such models require powerful hardware (GPU machines) and careful training. We assess the extrapolation capabilities of three different ML models that do not require GPU machines, namely, random forest, Gaussian process regression and gradient boosting, on both synthetic and real data. Experimental results on two synthetic data sets (generated from a low velocity lens embedded in a homogeneous medium, and the Marmousi model) demonstrate that FWI applied to the extrapolated data consistently improves inversion accuracy relative to FWI applied to the original data sets that do not contain low frequencies. Application of low-frequency extrapolation to real data from the Northwest Shelf of Australia demonstrates that tree-based ML models such as gradient boosting can outperform CNNs in terms of both accuracy and computational cost on non-GPU architectures.

58 GEOSCIENCES↗

Evaluation of Machine Learning Models for Automated Data Analysis in In-Service Nuclear Power Plant Inspections

The commercial nuclear power industry is facing a potential shortage of certified nondestructive evaluation (NDE) analysts to meet future in-service inspection demands. Automated data analysis (ADA) currently supports human inspectors in tasks such as eddy current evaluations for steam generator examinations. Machine learning (ML) systems are nearing the capability to pass performance demonstration tests for ultrasonic testing (UT) inspections of reactor pressure vessel upper head penetrations in nuclear power plants (NPPs). Current research and development is focused on assisted analysis (AA) of ADA versus fully automated examinations. This presentation will cover assessment of ML flaw detection on dissimilar metal weld (DMW) piping joints.

36 MATERIALS SCIENCE↗

The generation of infrared and ultraviolet astronomical data bases and retrieval systems

Observations with the Infrared Astronomy Satellite (IRAS) and with the International Ultraviolet Explorer (IUE) satellite have stimulated the need for machine-readable data bases at infrared and ultraviolet wavelengths along with associated software. This paper describes the generation of three such data sets at the Astronomical Data Center (ADC) of the NASA-Goddard Space Flight Center (GSFC): the Catalog of Infrared Observations, the Combined List of Astronomical Sources, and the Bibliographical Index of Objects Observed by IUE 1978-82. The discussion is divided by spectral regime and includes summaries of the data products developed in each category.

Mead, J. M.↗

Multi‐Decadal Dynamics of Wetland Methane Emissions Revealed by Knowledge‐Guided Machine Learning

Measurement of methane fluxes (FCH 4 ) from natural systems, such as wetlands, has lagged far behind carbon dioxide fluxes. Short and fragmented wetland FCH 4 data limit our ability to assess its long-term dynamics and potential climate feedbacks. Extrapolating short-term FCH 4 records to recent decades remains challenging for both process-based models and data-driven machine learning (ML) approaches. Here, we develop a knowledge-guided ML framework that integrates eddy covariance (EC) FCH 4 observations, field warming experiments, and biogeochemical knowledge to reconstruct the long-term FCH 4 budgets and trends. Focusing on the 11 longest EC monitoring sites in the AmeriFlux network, we found considerable variability in multi-decadal trends of wetland FCH 4 , with increases up to 14% per decade from 2000 to 2024. We also found that the strength of these increasing trends declines from high to low latitudes, highlighting the vulnerability of northern wetlands. This work presents novel and robust reconstructions of long-term wetland FCH 4 , offering critical benchmark datasets for bottom-up ecosystem models and advancing fundamental understanding of wetland biogeochemistry.

AmeriFlux site↗

Ice Phase Classification Made Easy with Score-Based Denoising

Accurate identification of ice phases is essential for understanding various physicochemical phenomena. However, such classification for structures simulated with molecular dynamics is complicated by the complex symmetries of ice polymorphs and thermal fluctuations. For this purpose, both traditional order parameters and data-driven machine learning approaches have been employed, but they often rely on expert intuition, specific geometric information, or large training data sets. In this work, we present an unsupervised phase classification framework that combines a score-based denoiser model with a subsequent model-free classification method to accurately identify ice phases. Further, the denoiser model is trained on perturbed synthetic data of ideal reference structures, eliminating the need for large data sets and labeling efforts. The classification step utilizes the smooth overlap of atomic position (SOAP) descriptors as the atomic fingerprint, ensuring Euclidean symmetries and transferability to various structural systems. Our approach achieves a remarkable 100% accuracy in distinguishing ice phases of test trajectories using only seven ideal reference structures of ice phases as model inputs. This demonstrates the generalizability of the score-based denoiser model in facilitating phase identification for complex molecular systems. The proposed classification strategy can be broadly applied to investigate structural evolution and phase identification for a wide range of materials, offering new insights into the fundamental understanding of water and other complex systems.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Introduction to the Asteroids II data base

This paper describes the Asteroids II data base, which is a compilation of asteroid data published, or in press, as of March 1988 with some updates in early 1989. The Asteroids II machine-readable data base includes asteroid names and discovery circumstances; proper elements and family identifications; asteroid light-curve parameters; asteroid pole determinations; taxonomic classes; and absolute magnitudes and slope parameters, UBV colors, albedos, and diameters.

Tedesco, Edward F.↗

Application of Machine Learning Techniques in Calibration and Data Reduction of Multi-Hole Probes

This work presents procedures for implementing machine learning methods into existing algorithms for multi-hole probe calibration and data reduction. It demonstrates that using artificial neural networks (ANNs) can decrease the amount of calibration data needed to achieve a specific calibration uncertainty by over 50%, while also significantly reducing data reduction times. Instead of surface fitting methods, ANNs are employed. Initially, directional calibration coefficients related to flow angles are computed based on pressure measurements, and then these flow angles serve as input parameters for subsequent ANNs to iteratively define Mach number, static pressure, and total pressure. In an alternative approach, new calibration coefficients directly relate pressure measurements from the five-hole probe to the quantities of interest, thereby eliminating the need for iterative algorithms used in conventional surface fitting methods. This method offers several advantages: an average increase of less than 1%in calibration uncertainty for flow angles and a significant reduction in data reduction times to a few seconds on average. Additionally, the methodology is confirmed to avoid both over- and under-fitting.

Machine Learning↗

Selection of Global Climate Model Data for Downscaling With Generative Machine Learning and Use in the Power Planning for Alignment of Climate and Energy Systems Project

The range of results from climate models and scenarios is important to the understanding of uncertainty in power planning analysis. A U.S. Department of Energy-funded analytic project called Power Planning for Alignment of Climate and Energy Systems is developing data and analytic methods to reflect the effects of climate change on key variables for power system planning, as part of the Grid Modernization Lab Consortium. This project will select and prepare global climate model results for use in power system planning models. A related report (Evaluation of Global Climate Models for Use in Energy Analysis) assesses the performance of various global climate models from the Coupled Model Intercomparison Project Phase 6 data archive for their historical skill with respect to energy system performance and for their future projections under multiple climate change scenarios. Building from that report, we describe the selection of a climate scenario (Shared Socioeconomic Pathway [SSP] 2-4.5) and five climate models: TaiESM1, EC-Earth3-CC, GFDL-CM4, EC-Earth3-Veg, and MPI-ESM1-2-HR. We describe the model selection criteria, which were based on the quality of the match between model results under historical conditions and on the representation of the range of future values for several variables. These results will be downscaled via an open-source generative machine learning method called Super-Resolution for Renewable Energy Resource Data with Climate Change Impacts.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Application of Machine Learning Techniques in Calibration and Data Reduction of Multi-Hole Probes

This work presents procedures to implement machine learning methods in the existing algorithms for multi-hole probe calibrations and data reduction. It is shown here, that utilizing artificial neural networks (ANNs) can reduce the amount of calibration data that needs to be acquired in order to obtain a specific calibration uncertainty, by more than 50% while simultaneously reducing data reduction times significantly. ANNs were used instead of the surface fitting methods, where first, the directional calibration coefficients related to the flow angles are calculated based on the pressure measurements, and then the flow angles are used as a set of the input parameters for the following ANNs to define Mach number and static and total pressure iteratively. In a second approach, novel calibration coefficients were used to directly relate the pressure measurements from five-hole probe to the quantities of interest thus, eliminating the need for iterative algorithms used in the conventional surface fitting methods. The advantageous features of this method are an average increase of less than 1% in the calibration uncertainty for flow angles and significant reduction of the data reduction times (few seconds). In addition, we confirmed the methodology to avoid over-fitting and under-fitting.

Machine Learning↗

Data management, chapter 5, part C

The data management for a spacecraft radar was defined in terms of an end-to-end data system, which performs the following three functions: (1) sampling and compaction of data onboard the spacecraft, (2) manipulation of radar data on the ground and (3) conversion of radar measurements to geophysical quantities by means of pattern recognition and other machine techniques. Data processing for imaging radar onboard the spacecraft was examined with the conclusion that several techniques can be used to compact the data before storage. It is recommended that compaction techniques be studied further and that existing aircraft radars be modified to provide digital data so that these compaction techniques can be tested.

Source record↗

Using Machine Learning to Predict Core Sizes of High-Efficiency Turbofan Engines

With the rise in big data and analytics, machine learning is transforming many industries. It is being increasingly employed to solve a wide range of complex problems, producing autonomous systems that support human decision-making. For the aircraft engine industry, machine learning of historical and existing engine data could provide insights that help drive for better engine design. This work explored the application of machine learning to engine preliminary design. Engine core-size prediction was chosen for the first study because of its relative simplicity in terms of number of input variables required (only three). Specifically, machine-learning predictive tools were developed for turbofan engine core-size prediction, using publicly available data of two hundred manufactured engines and engines that were studied previously in NASA aeronautics projects. The prediction results of these models show that, by bringing together big data, robust machine-learning algorithms and data science, a machine learning-based predictive model can be an effective tool for turbofan engine core-size prediction. The promising results of this first study paves the way for further exploration of the use of machine learning for aircraft engine preliminary design.

Core Size↗