Search NASA⌕ Search

SEARCH · Search NASA

Results for “Memory Optimization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

514 records · Page 29

Hardenability and microstructural evolution of a precipitation strengthened Ni 50 Ti 21 Hf 25 Al 4 alloy

NiTi-based quaternary alloys are used in a variety of mechanical components, such as bearings, actuators, and dampers, owing to their good hardenability, wear resistance, and corrosion resistance. Additionally, one of the most notable characteristics of NiTi-based alloys is their shape memory effect and pseudoelastic properties. Connecting the macroscopic processing parameters employed in the design of new intermetallic alloys to the nanoscale structural characteristics dictating their behavior is crucial for improving their mechanical properties and expanding the spectrum of potential applications. Here, in this work, an arc melted Ni 50 Ti 21 Hf 25 Al 4 (at%) alloy was solution treated at 1050 °C followed by quenching and aging at 600 °C to investigate the effect of aging time on the microstructure and mechanical properties. Two types of nano-sized precipitates were observed and determined as face-centered orthorhombic H-phase (TiHf)Ni and L2 1 Heusler precipitates Ni 2 TiAl. The morphology and orientation of the H-phase were investigated using scanning and transmission electron microscopy (SEM and TEM), elucidating the coarsening kinetics and strengthening contribution of that phase to the intermetallic mechanical behavior. Following coarsening, the presence of Heusler nanoprecipitates was detected under overaged conditions through TEM imaging and nanobeam electron diffraction patterns. A peak hardness condition of 756 HV was achieved after 70 h of aging, indicating that the co-precipitation of H-phase and Heusler precipitates through a well-designed aging treatment can lead to optimal mechanical performance, thus elevating the alloy’s potential as a viable material for industrial applications.

36 MATERIALS SCIENCE↗

Machine learning-based interatomic potential development and phase transition analysis of ferroelectric hafnium dioxide

The ferroelectric phase (𝑃⁢𝑐⁢𝑎⁢2 1 , which is in orthorhombic symmetry) of hafnium dioxide (HfO 2 ) has gained much attention due to its potential applications in nanoelectronics and advanced memory devices. However, its complex phase behavior under external stimuli, such as pressure and temperature, remains a subject of intense investigation. This study focuses on developing a machine learning-based interatomic potential (MLIP) that is trained with data from density-functional theory (DFT) calculations to simulate phase transitions and mechanical properties of HfO 2 . The developed MLIP predicts lattice parameters, equations of state, bulk and shear moduli, and elastic constants that closely align with DFT predictions for several phases and at various pressures. Once validated, the MLIP is used to investigate the phase transitions of ferroelectric HfO 2 (𝑃⁢𝑐⁢𝑎⁢2 1 ) under both isobaric and constant stress conditions at elevated temperatures ranging from 200 to 2500 K. We used several complementary methods, including local symmetry identification, radial distribution function, and x-ray diffraction characterization, to identify interesting phase transitions among several competitive hafnia phases predicted from our simulations. The suggested methods uniformly reveal that under pure deviatoric condition, the system favors a transition from the orthorhombic 𝑃⁢𝑐⁢𝑎⁢2 1 phase to a tetragonal (𝑃⁢4 2 /𝑛⁢𝑚⁢𝑐) phase, whereas a zero stress condition drives the system from the 𝑃⁢𝑐⁢𝑎⁢2 1 phase to another orthorhombic (𝑃⁢𝑏⁢𝑐⁢𝑛) phase. These findings provide crucial insights into stress and temperature-induced phase behavior of hafnia, guiding future experimental and theoretical studies for optimizing hafnia-based ferroelectric devices.

Ferroelectric HfO2↗

Effects of input gradient regularization on neural networks time-series forecasting of thermal power systems

This study proposes using neural networks, specifically gated recurrent unit (GRU), long-short-term memory (LSTM), and transformer networks, to improve control strategies in a 450 MW coal-fired power plant. However, neural networks face issues of becoming overly dependent on just a few variables to make predictions, which negatively impacts control decisions that rely on the model to determine the value of all manipulated variables. The paper introduces regularization techniques, including noise injection and input gradient regularization, during the training phase. Here, the work presents novel contributions in adapting neural networks to control industrial systems and applying regularization techniques from computer vision to industrial process control. Results demonstrate the effectiveness of input gradient regularization in reducing model dependence on subsets of variables, emphasizing the balance between fidelity and controllability. Further exploration is recommended, including the development of recurrent transformers, closed-loop control testing, and a sensitivity analysis on computer models to provide further insight.

20 FOSSIL-FUELED POWER PLANTS↗

Development of a Computer Architecture to Support the Optical Plume Anomaly Detection (OPAD) System

The NASA OPAD spectrometer system relies heavily on extensive software which repetitively extracts spectral information from the engine plume and reports the amounts of metals which are present in the plume. The development of this software is at a sufficiently advanced stage where it can be used in actual engine tests to provide valuable data on engine operation and health. This activity will continue and, in addition, the OPAD system is planned to be used in flight aboard space vehicles. The two implementations, test-stand and in-flight, may have some differing requirements. For example, the data stored during a test-stand experiment are much more extensive than in the in-flight case. In both cases though, the majority of the requirements are similar. New data from the spectrograph is generated at a rate of once every 0.5 sec or faster. All processing must be completed within this period of time to maintain real-time performance. Every 0.5 sec, the OPAD system must report the amounts of specific metals within the engine plume, given the spectral data. At present, the software in the OPAD system performs this function by solving the inverse problem. It uses powerful physics-based computational models (the SPECTRA code), which receive amounts of metals as inputs to produce the spectral data that would have been observed, had the same metal amounts been present in the engine plume. During the experiment, for every spectrum that is observed, an initial approximation is performed using neural networks to establish an initial metal composition which approximates as accurately as possible the real one. Then, using optimization techniques, the SPECTRA code is repetitively used to produce a fit to the data, by adjusting the metal input amounts until the produced spectrum matches the observed one to within a given level of tolerance. This iterative solution to the original problem of determining the metal composition in the plume requires a relatively long period of time to execute the software in a modern single-processor workstation, and therefore real-time operation is currently not possible. A different number of iterations may be required to perform spectral data fitting per spectral sample. Yet, the OPAD system must be designed to maintain real-time performance in all cases. Although faster single-processor workstations are available for execution of the fitting and SPECTRA software, this option is unattractive due to the excessive cost associated with very fast workstations and also due to the fact that such hardware is not easily expandable to accommodate future versions of the software which may require more processing power. Initial research has already demonstrated that the OPAD software can take advantage of a parallel computer architecture to achieve the necessary speedup. Current work has improved the software by converting it into a form which is easily parallelizable. Timing experiments have been performed to establish the computational complexity and execution speed of major components of the software. This work provides the foundation of future work which will create a fully parallel version of the software executing in a shared-memory multiprocessor system.

Katsinis, Constantine↗

Coarse-Grain Bandwidth Estimation Scheme for Large-Scale Network

A large-scale network that supports a large number of users can have an aggregate data rate of hundreds of Mbps at any time. High-fidelity simulation of a large-scale network might be too complicated and memory-intensive for typical commercial-off-the-shelf (COTS) tools. Unlike a large commercial wide-area-network (WAN) that shares diverse network resources among diverse users and has a complex topology that requires routing mechanism and flow control, the ground communication links of a space network operate under the assumption of a guaranteed dedicated bandwidth allocation between specific sparse endpoints in a star-like topology. This work solved the network design problem of estimating the bandwidths of a ground network architecture option that offer different service classes to meet the latency requirements of different user data types. In this work, a top-down analysis and simulation approach was created to size the bandwidths of a store-and-forward network for a given network topology, a mission traffic scenario, and a set of data types with different latency requirements. These techniques were used to estimate the WAN bandwidths of the ground links for different architecture options of the proposed Integrated Space Communication and Navigation (SCaN) Network. A new analytical approach, called the "leveling scheme," was developed to model the store-and-forward mechanism of the network data flow. The term "leveling" refers to the spreading of data across a longer time horizon without violating the corresponding latency requirement of the data type. Two versions of the leveling scheme were developed: 1. A straightforward version that simply spreads the data of each data type across the time horizon and doesn't take into account the interactions among data types within a pass, or between data types across overlapping passes at a network node, and is inherently sub-optimal. 2. Two-state Markov leveling scheme that takes into account the second order behavior of the store-and-forward mechanism, and the interactions among data types within a pass. The novelty of this approach lies in the modeling of the store-and-forward mechanism of each network node. The term store-and-forward refers to the data traffic regulation technique in which data is sent to an intermediate network node where they are temporarily stored and sent at a later time to the destination node or to another intermediate node. Store-and-forward can be applied to both space-based networks that have intermittent connectivity, and ground-based networks with deterministic connectivity. For groundbased networks, the store-and-forward mechanism is used to regulate the network data flow and link resource utilization such that the user data types can be delivered to their destination nodes without violating their respective latency requirements.

Cheung, Kar-Ming↗

FPGA Implementation of Stereo Disparity with High Throughput for Mobility Applications

High speed stereo vision can allow unmanned robotic systems to navigate safely in unstructured terrain, but the computational cost can exceed the capacity of typical embedded CPUs. In this paper, we describe an end-to-end stereo computation co-processing system optimized for fast throughput that has been implemented on a single Virtex 4 LX160 FPGA. This system is capable of operating on images from a 1024 x 768 3CCD (true RGB) camera pair at 15 Hz. Data enters the FPGA directly from the cameras via Camera Link and is rectified, pre-filtered and converted into a disparity image all within the FPGA, incurring no CPU load. Once complete, a rectified image and the final disparity image are read out over the PCI bus, for a bandwidth cost of 68 MB/sec. Within the FPGA there are 4 distinct algorithms: Camera Link capture, Bilinear rectification, Bilateral subtraction pre-filtering and the Sum of Absolute Difference (SAD) disparity. Each module will be described in brief along with the data flow and control logic for the system. The system has been successfully fielded upon the Carnegie Mellon University's National Robotics Engineering Center (NREC) Crusher system during extensive field trials in 2007 and 2008 and is being implemented for other surface mobility systems at JPL.

Random access memory↗

Water foraging with dynamic roots in E3SM; The role of roots in terrestrial ecosystem memory on intermediate timescales (Final Technical Report)

Terrestrial ecosystems can show sustained responses to stress events that may last for years after the initial perturbation. This phenomenon is referred to as legacy or memory and it emerges from a set of ecosystem processes that are poorly captured by Earth System and Land Surface Models. Consequently, these models struggle to predict the impact that extreme climate events have on surface energy, carbon and hydrological exchange once a stressor such as drought has relaxed or in response to repeated exposure to stress. In this project, we set out to test how the addition of dynamic root profiles in land surface models impact the capacity for Earth System Models, such as the Department of Energy’s E3SM, to capture realistic legacy effects. Observations have shown that vegetation shifts their root profiles during stress events to forage for water and these altered root profiles may take years to relax back to the initial state. We hypothesized that the alteration of the belowground root structures may be a source of terrestrial ecosystem legacy that is missing from models. To test this idea we undertook three core activities. (1) We developed a global-scale analysis of the impact of dynamic roots on ecosystem legacy building from a recently developed dynamic root module. (2) We developed new root dynamics in the DOE’s Energy Land Model (ELM) that allowed not only root profiles to shift their profile but also simultaneously alter carbon allocation to fine root pools. We then tested the impact of these new dynamics with intensive sensitivity analysis at four long-term AmeriFlux sites. (3) We undertook detailed isotopic analysis of tree rings from these AmeriFlux sites to assess ecophysiological and ecohydrological legacy to stress events that can be used to benchmark the sensitivity experiments with ELM. From these core activities, we report the following key findings. Firstly, on a global scale, the addition of dynamic roots led to chronically water-stressed ecosystems recovering faster to climate stress while wetter ecosystem showed enhanced legacy. This is because across ecosystems, stress events were almost universally associated with water shortages that led to the development of deeper root profiles. These deeper root profiles proved beneficial for recovery from drought stress. While the root dynamics did not universally improve the modeled representation of legacy it showed complex transient dynamics that emerge from the addition of dynamic roots. Secondly, the sensitivity analysis illustrated long term shifts in rooting depths away from the prescribed default profile suggesting that initializing of root profiles could benefit from spin-up simulations that converge on locally optimized root profiles. In addition, by enabling dynamic allocation some of the sites predicted unrealistically low allocation to roots (and high allocation to leaves). While these changes did not dramatically alter modeled gross primary productivity, they illustrated that without more sophisticated root processes in models, there is little penalty to dramatically disinvest in roots. Thirdly, the isotopic analysis of tree rings showed highly distinct legacy responses across species and sites. For example, T. canadensis showed reduced transpiration the year after stress events illustrated by sustained elevated $\delta^{18}$O. In contrast, A saccharum displayed elevated $\delta^{13}$C associated with reduced stomatal conductance in response to the previous years’ stress event. These geochemical signatures of legacy were present despite tree growth returning to normal the year after the stress. The results show the importance of species-level dynamics in legacy that are absent in modeling that assumes common traits within plant functional types. In summary, the work here established new avenues to explore root dynamics in models while illustrating how additional root processes are needed before dynamic carbon allocation can be implemented. Lastly, this project provided mentorship to a postdoctoral fellow, training for an early career scientist and multiple undergraduate students recruited from a minority serving institution.

54 ENVIRONMENTAL SCIENCES↗

Efficient Implementation of an Optimal Interpolator for Large Spatial Data Sets

Scattered data interpolation is a problem of interest in numerous areas such as electronic imaging, smooth surface modeling, and computational geometry. Our motivation arises from applications in geology and mining, which often involve large scattered data sets and a demand for high accuracy. The method of choice is ordinary kriging. This is because it is a best unbiased estimator. Unfortunately, this interpolant is computationally very expensive to compute exactly. For n scattered data points, computing the value of a single interpolant involves solving a dense linear system of size roughly n x n. This is infeasible for large n. In practice, kriging is solved approximately by local approaches that are based on considering only a relatively small'number of points that lie close to the query point. There are many problems with this local approach, however. The first is that determining the proper neighborhood size is tricky, and is usually solved by ad hoc methods such as selecting a fixed number of nearest neighbors or all the points lying within a fixed radius. Such fixed neighborhood sizes may not work well for all query points, depending on local density of the point distribution. Local methods also suffer from the problem that the resulting interpolant is not continuous. Meyer showed that while kriging produces smooth continues surfaces, it has zero order continuity along its borders. Thus, at interface boundaries where the neighborhood changes, the interpolant behaves discontinuously. Therefore, it is important to consider and solve the global system for each interpolant. However, solving such large dense systems for each query point is impractical. Recently a more principled approach to approximating kriging has been proposed based on a technique called covariance tapering. The problems arise from the fact that the covariance functions that are used in kriging have global support. Our implementations combine, utilize, and enhance a number of different approaches that have been introduced in literature for solving large linear systems for interpolation of scattered data points. For very large systems, exact methods such as Gaussian elimination are impractical since they require 0(n(exp 3)) time and 0(n(exp 2)) storage. As Billings et al. suggested, we use an iterative approach. In particular, we use the SYMMLQ method, for solving the large but sparse ordinary kriging systems that result from tapering. The main technical issue that need to be overcome in our algorithmic solution is that the points' covariance matrix for kriging should be symmetric positive definite. The goal of tapering is to obtain a sparse approximate representation of the covariance matrix while maintaining its positive definiteness. Furrer et al. used tapering to obtain a sparse linear system of the form Ax = b, where A is the tapered symmetric positive definite covariance matrix. Thus, Cholesky factorization could be used to solve their linear systems. They implemented an efficient sparse Cholesky decomposition method. They also showed if these tapers are used for a limited class of covariance models, the solution of the system converges to the solution of the original system. Matrix A in the ordinary kriging system, while symmetric, is not positive definite. Thus, their approach is not applicable to the ordinary kriging system. Therefore, we use tapering only to obtain a sparse linear system. Then, we use SYMMLQ to solve the ordinary kriging system. We show that solving large kriging systems becomes practical via tapering and iterative methods, and results in lower estimation errors compared to traditional local approaches, and significant memory savings compared to the original global system. We also developed a more efficient variant of the sparse SYMMLQ method for large ordinary kriging systems. This approach adaptively finds the correct local neighborhood for each query point in the interpolation process.

Memarsadeghi, Nargess↗

Rao-Blackwellization for Adaptive Gaussian Sum Nonlinear Model Propagation

When dealing with imperfect data and general models of dynamic systems, the best estimate is always sought in the presence of uncertainty or unknown parameters. In many cases, as the first attempt, the Extended Kalman filter (EKF) provides sufficient solutions to handling issues arising from nonlinear and non-Gaussian estimation problems. But these issues may lead unacceptable performance and even divergence. In order to accurately capture the nonlinearities of most real-world dynamic systems, advanced filtering methods have been created to reduce filter divergence while enhancing performance. Approaches, such as Gaussian sum filtering, grid based Bayesian methods and particle filters are well-known examples of advanced methods used to represent and recursively reproduce an approximation to the state probability density function (pdf). Some of these filtering methods were conceptually developed years before their widespread uses were realized. Advanced nonlinear filtering methods currently benefit from the computing advancements in computational speeds, memory, and parallel processing. Grid based methods, multiple-model approaches and Gaussian sum filtering are numerical solutions that take advantage of different state coordinates or multiple-model methods that reduced the amount of approximations used. Choosing an efficient grid is very difficult for multi-dimensional state spaces, and oftentimes expensive computations must be done at each point. For the original Gaussian sum filter, a weighted sum of Gaussian density functions approximates the pdf but suffers at the update step for the individual component weight selections. In order to improve upon the original Gaussian sum filter, Ref. [2] introduces a weight update approach at the filter propagation stage instead of the measurement update stage. This weight update is performed by minimizing the integral square difference between the true forecast pdf and its Gaussian sum approximation. By adaptively updating each component weight during the nonlinear propagation stage an approximation of the true pdf can be successfully reconstructed. Particle filtering (PF) methods have gained popularity recently for solving nonlinear estimation problems due to their straightforward approach and the processing capabilities mentioned above. The basic concept behind PF is to represent any pdf as a set of random samples. As the number of samples increases, they will theoretically converge to the exact, equivalent representation of the desired pdf. When the estimated qth moment is needed, the samples are used for its construction allowing further analysis of the pdf characteristics. However, filter performance deteriorates as the dimension of the state vector increases. To overcome this problem Ref. [5] applies a marginalization technique for PF methods, decreasing complexity of the system to one linear and another nonlinear state estimation problem. The marginalization theory was originally developed by Rao and Blackwell independently. According to Ref. [6] it improves any given estimator under every convex loss function. The improvement comes from calculating a conditional expected value, often involving integrating out a supportive statistic. In other words, Rao-Blackwellization allows for smaller but separate computations to be carried out while reaching the main objective of the estimator. In the case of improving an estimator's variance, any supporting statistic can be removed and its variance determined. Next, any other information that dependents on the supporting statistic is found along with its respective variance. A new approach is developed here by utilizing the strengths of the adaptive Gaussian sum propagation in Ref. [2] and a marginalization approach used for PF methods found in Ref. [7]. In the following sections a modified filtering approach is presented based on a special state-space model within nonlinear systems to reduce the dimensionality of the optimization problem in Ref. [2]. First, the adaptive Gaussian sum propagation is explained and then the new marginalized adaptive Gaussian sum propagation is derived. Finally, an example simulation is presented.

state estimation↗

Chemically Enabled CO 2 -Enhanced Oil Recovery in Multi-Porosity, Hydrothermally Altered Carbonates in the Southern Michigan Basin (Final Technical Report)

This is the Final Technical Report for the project "Chemically Enabled CO 2 -Enhanced Oil Recovery in Multi-Porosity, Hydrothermally Altered Carbonates in the Southern Michigan Basin." Over the course of six years of collaboration between Battelle and project partners, all stated objectives of the program have been completed, including full geological characterization of the TBR trend (See companion report for Task2), laboratory and modeling experiments to determine the optimum composition and design of CO 2 -EOR operations in the TBR trend, execution of a field test of chemically-enhanced CO 2 in a TBR well, and integration of the data and learnings gathered during these efforts into a full-trend development plan. Detailed reporting on these activities, their outcomes, and implications for trend-wide development is provided in the report. This report and encompassed data will provide TBR field operators with detailed information on what worked, what did not work, and how to proceed with production optimization of their TBR assets using chemically-enhanced CO 2 -EOR. CO 2 -EOR is a relatively well understood and broadly implemented strategy for increasing incremental production across the oil and gas industry, but its application has been primarily focused on reservoirs with limited heterogeneity. The intention of this project was show first that the same physical mechanisms that improve recovery factors in homogeneous reservoirs (namely wettability alteration, viscosity alteration, oil swelling, and mobility control) are at play in heterogeneous reservoirs. This was proven by the project’s laboratory studies and dynamic simulations, with the potential exception of mobility control, which needs further study. The second intention was to demonstrate via direct field testing that CO 2 -EOR can work in a strongly heterogeneous reservoir. While the field test strategy implemented during this project did not succeed in producing oil, data gathered during the test sheds light on what may work for field operators who try chemically-enhanced CO 2 -EOR within their own reservoirs, significantly reducing the level of uncertainty carried by first-of-a-kind commercial efforts that could (and should) follow this test. Simultaneously, the project has identified several large-volume ethanol plants and other sources of CO 2 emissions in the region and provided a handrail that CO 2 emitters and field operators can leverage to capture, transport, and inject that CO 2 into their fields. This project has also shown that, in many cases, the economics of CO 2 -EOR in the TBR are attractive. And finally, by completing a project of this scope in the southern Michigan Basin, the project has contributed to the knowledge base and operational experience of field operators, state regulatory agencies, local service companies, and state universities, with CO 2 -EOR projects which should allow follow-on projects to proceed safely and efficiently.

02 PETROLEUM↗