Search NASA⌕ Search

SEARCH · Search NASA

Results for “Performance Tuning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

On the Efficacy of Source Code Optimizations for Cache-Based Systems

Obtaining high performance without machine-specific tuning is an important goal of scientific application programmers. Since most scientific processing is done on commodity microprocessors with hierarchical memory systems, this goal of "portable performance" can be achieved if a common set of optimization principles is effective for all such systems. It is widely believed, or at least hoped, that portable performance can be realized. The rule of thumb for optimization on hierarchical memory systems is to maximize temporal and spatial locality of memory references by reusing data and minimizing memory access stride. We investigate the effects of a number of optimizations on the performance of three related kernels taken from a computational fluid dynamics application. Timing the kernels on a range of processors, we observe an inconsistent and often counterintuitive impact of the optimizations on performance. In particular, code variations that have a positive impact on one architecture can have a negative impact on another, and variations expected to be unimportant can produce large effects. Moreover, we find that cache miss rates - as reported by a cache simulation tool, and confirmed by hardware counters - only partially explain the results. By contrast, the compiler-generated assembly code provides more insight by revealing the importance of processor-specific instructions and of compiler maturity, both of which strongly, and sometimes unexpectedly, influence performance. We conclude that it is difficult to obtain performance portability on modern cache-based computers, and comment on the implications of this result.

VanderWijngaart, Rob F.↗

On the Efficacy of Source Code Optimizations for Cache-Based Systems

Obtaining high performance without machine-specific tuning is an important goal of scientific application programmers. Since most scientific processing is done on commodity microprocessors with hierarchical memory systems, this goal of "portable performance" can be achieved if a common set of optimization principles is effective for all such systems. It is widely believed, or at least hoped, that portable performance can be realized. The rule of thumb for optimization on hierarchical memory systems is to maximize temporal and spatial locality of memory references by reusing data and minimizing memory access stride. We investigate the effects of a number of optimizations on the performance of three related kernels taken from a computational fluid dynamics application. Timing the kernels on a range of processors, we observe an inconsistent and often counterintuitive impact of the optimizations on performance. In particular, code variations that have a positive impact on one architecture can have a negative impact on another, and variations expected to be unimportant can produce large effects. Moreover, we find that cache miss rates-as reported by a cache simulation tool, and confirmed by hardware counters-only partially explain the results. By contrast, the compiler-generated assembly code provides more insight by revealing the importance of processor-specific instructions and of compiler maturity, both of which strongly, and sometimes unexpectedly, influence performance. We conclude that it is difficult to obtain performance portability on modern cache-based computers, and comment on the implications of this result.

VanderWijngaart, Rob F.↗

Tuned Chamber Core Panel Acoustic Test Results

This report documents acoustic testing of tuned chamber core panels, which can be used to supplement the low-frequency performance of conventional acoustic treatment. The tuned chamber core concept incorporates low-frequency noise control directly within the primary structure and is applicable to sandwich constructions with a directional core, including corrugated-, truss-, and fluted-core designs. These types of sandwich structures have long, hollow channels (or chambers) in the core. By adding small holes through one of the facesheets, the hollow chambers can be utilized as an array of low-frequency acoustic resonators. These resonators can then be used to attenuate low-frequency noise (below 400 Hz) inside a vehicle compartment without increasing the weight or size of the structure. The results of this test program demonstrate that the tuned chamber core concept is effective when used in isolation or combined with acoustic foam treatments. Specifically, an array of acoustic resonators integrated within the core of the panels was shown to improve both the low-frequency absorption and transmission loss of the structure in targeted one-third octave bands.

Schiller, Noah H.↗

Networks for image acquisition, processing and display

The human visual system comprises layers of networks which sample, process, and code images. Understanding these networks is a valuable means of understanding human vision and of designing autonomous vision systems based on network processing. Ames Research Center has an ongoing program to develop computational models of such networks. The models predict human performance in detection of targets and in discrimination of displayed information. In addition, the models are artificial vision systems sharing properties with biological vision that has been tuned by evolution for high performance. Properties include variable density sampling, noise immunity, multi-resolution coding, and fault-tolerance. The research stresses analysis of noise in visual networks, including sampling, photon, and processing unit noises. Specific accomplishments include: models of sampling array growth with variable density and irregularity comparable to that of the retinal cone mosaic; noise models of networks with signal-dependent and independent noise; models of network connection development for preserving spatial registration and interpolation; multi-resolution encoding models based on hexagonal arrays (HOP transform); and mathematical procedures for simplifying analysis of large networks.

Ahumada, Albert J., Jr.↗

Controller design for the improvement of feedback properties: A method for tuning the weight in the evaluation function of the LQG theory

In the controller design of a linear time-invariant system, it is important to improve the feedback properties such as robust stability and sensitivity. In the multi-input multi-output case, these properties can be estimated by using the singular values of return difference matrix. The design method to obtain better singular-value-points is desired. How to tune the weight matrix of performance index and/or covariance matrix of noise in the Linear Quadratic Gaussian (LQC) theory to get a desired singular-value-plot was studied. First, the property of singular-value plots of return difference matrix of a system designed by LQG theory is examined from the viewpoint of tuning weight. Second, a distance between real singular-value-plots and a desired plot is defined, and the weight of performance index is numerically determined by quasi-Newton method so that the distance is minimized.

Saeki, Masami↗

A multiple objective optimization approach to quality control

The use of product quality as the performance criteria for manufacturing system control is explored. The goal in manufacturing, for economic reasons, is to optimize product quality. The problem is that since quality is a rather nebulous product characteristic, there is seldom an analytic function that can be used as a measure. Therefore standard control approaches, such as optimal control, cannot readily be applied. A second problem with optimizing product quality is that it is typically measured along many dimensions: there are many apsects of quality which must be optimized simultaneously. Very often these different aspects are incommensurate and competing. The concept of optimality must now include accepting tradeoffs among the different quality characteristics. These problems are addressed using multiple objective optimization. It is shown that the quality control problem can be defined as a multiple objective optimization problem. A controller structure is defined using this as the basis. Then, an algorithm is presented which can be used by an operator to interactively find the best operating point. Essentially, the algorithm uses process data to provide the operator with two pieces of information: (1) if it is possible to simultaneously improve all quality criteria, then determine what changes to the process input or controller parameters should be made to do this; and (2) if it is not possible to improve all criteria, and the current operating point is not a desirable one, select a criteria in which a tradeoff should be made, and make input changes to improve all other criteria. The process is not operating at an optimal point in any sense if no tradeoff has to be made to move to a new operating point. This algorithm ensures that operating points are optimal in some sense and provides the operator with information about tradeoffs when seeking the best operating point. The multiobjective algorithm was implemented in two different injection molding scenarios: tuning of process controllers to meet specified performance objectives and tuning of process inputs to meet specified quality objectives. Five case studies are presented.

Seaman, Christopher Michael↗

CW laser strategies for multi-parameter measurements of high-speed flows containing either NO or O2

Measurements of gasdynamic quantities were performed using a rapid-tuning CW dye laser to resolve Doppler-shifted spectral features in either the O2 Schumann-Runge bands or the NO gamma band near 225 nm. With the rapid-tuning capability, spectral features were acquired at a repetition rate of 4 kHz. Monitoring O2 transitions provided estimates of velocities while monitoring collision-broadened NO line pairs provided simultaneous measurements of velocity, temperature and pressure. Experiments were first performed in absorption within the transient one-dimensional flows generated in a shock tube. Agreement between measured and theoretical values, as calculated from one-dimensional shock relations, was typically better than 5 percent. The method was extended to fluorescence detection of NO in a static cell. Temperature and pressure were extracted from recorded profiles, and the results agreed well with expected values.

Dirosa, M. D.↗

Modeling-Error-Driven Performance-Seeking Direct Adaptive Control

This paper presents a stable discrete-time adaptive law that targets modeling errors in a direct adaptive control framework. The update law was developed in our previous work for the adaptive disturbance rejection application. The approach is based on the philosophy that without modeling errors, the original control design has been tuned to achieve the desired performance. The adaptive control should, therefore, work towards getting this performance even in the face of modeling uncertainties/errors. In this work, the baseline controller uses dynamic inversion with proportional-integral augmentation. Dynamic inversion is carried out using the assumed system model. On-line adaptation of this control law is achieved by providing a parameterized augmentation signal to the dynamic inversion block. The parameters of this augmentation signal are updated to achieve the nominal desired error dynamics. Contrary to the typical Lyapunov-based adaptive approaches that guarantee only stability, the current approach investigates conditions for stability as well as performance. A high-fidelity F-15 model is used to illustrate the overall approach.

Kulkarni, Nilesh V.↗

Artemis I Optical Navigation System Performance

This paper summarizes the assessment of the Optical Navigation Flight Test Objective (FTO) during the flight of Artemis I. The Optical Navigation (OpNav) System was tested under a variety of range, target, and lighting conditions to evaluate the performance compared to the pre-flight predicted error models. In general, OpNav performed very well – successfully processing over a thousand images of starfields, Earth, and Moon. The performance of the algorithm when processing Moon images matched the pre-flight expected error models. The errors when processing Earth images were notably higher than the pre-flight models predicted, however this was found to be due to an over-estimation of the atmosphere bias used in the tuning of the algorithm. After the bias was re-tuned and the images reprocessed, performance significantly improved.

GNC↗

An Overview of Trajectory Design Operations for the Microwave Anisotropy Probe (MAP) Mission

The main science objective of the Microwave Anisotropy Probe (MAP) mission is to produce an accurate full-sky map of the cosmic microwave background temperature fluctuations - anisotropy. MAP will collect these measurements from a lissajous orbit about the Sun-Earth/Moon L2 Lagrange Point. The NASA Goddard Space Flight Center (GSFC) Flight Dynamics Analysis Branch provided mission analysis, maneuver planning and maneuver calibration for the MAP spacecraft. This paper will provide an overview of the MAP trajectory design, a summary of the maneuvers executed. Differences from the pre-launch nominal plan will also be discussed. During the MAP phasing loops, MAP performed three calibration maneuvers in order to characterize the performance of the primary sets of thrusters - +X, +Z, and -Z. The calibration maneuvers were designed to minimize their impact on the trajectory. Four maneuvers were performed to set up the gravity assist of the Moon - required to propel MAP out to its orbit about L2. These maneuvers were performed at the three phasing loop perigees and at 18 hours after the final perigee. It became necessary to alter some of the perigee maneuvers in order to shape the gravity assist. This shaping was done to help meet some mission goals. In particular, the gravity assist was changed slightly in order to remove lunar shadows in both the cruise out to L2 and in the first revolution about L2. This amounted to a change in the phasing loop AV of less than 1 m/s. After the gravity assist, two mid-course correction (MCC) maneuvers were performed in order to fine-tune the trajectory. MCC1 was used to clean up and errors which resulted from the gravity assist. MCC2 was performed in order to mitigate a large stationkeeping maneuver following a crucial instrument calibration period during the cruise phase. MAP executed it's first stationkeeping maneuver in January 16th and is ready for a second calibration period during late Winter / early Spring. Further information concerning subsequent stationkeeping maneuver will be added as they become available.

Cuevas, Osvaldo O.↗

Algorithms for Spectral Decomposition with Applications to Optical Plume Anomaly Detection

The analysis of spectral signals for features that represent physical phenomenon is ubiquitous in the science and engineering communities. There are two main approaches that can be taken to extract relevant features from these high-dimensional data streams. The first set of approaches relies on extracting features using a physics-based paradigm where the underlying physical mechanism that generates the spectra is used to infer the most important features in the data stream. We focus on a complementary methodology that uses a data-driven technique that is informed by the underlying physics but also has the ability to adapt to unmodeled system attributes and dynamics. We discuss the following four algorithms: Spectral Decomposition Algorithm (SDA), Non-Negative Matrix Factorization (NMF), Independent Component Analysis (ICA) and Principal Components Analysis (PCA) and compare their performance on a spectral emulator which we use to generate artificial data with known statistical properties. This spectral emulator mimics the real-world phenomena arising from the plume of the space shuttle main engine and can be used to validate the results that arise from various spectral decomposition algorithms and is very useful for situations where real-world systems have very low probabilities of fault or failure. Our results indicate that methods like SDA and NMF provide a straightforward way of incorporating prior physical knowledge while NMF with a tuning mechanism can give superior performance on some tests. We demonstrate these algorithms to detect potential system-health issues on data from a spectral emulator with tunable health parameters.

Srivastava, Askok N.↗

Application of a Constant Gain Extended Kalman Filter for In-Flight Estimation of Aircraft Engine Performance Parameters

An approach based on the Constant Gain Extended Kalman Filter (CGEKF) technique is investigated for the in-flight estimation of non-measurable performance parameters of aircraft engines. Performance parameters, such as thrust and stall margins, provide crucial information for operating an aircraft engine in a safe and efficient manner, but they cannot be directly measured during flight. A technique to accurately estimate these parameters is, therefore, essential for further enhancement of engine operation. In this paper, a CGEKF is developed by combining an on-board engine model and a single Kalman gain matrix. In order to make the on-board engine model adaptive to the real engine s performance variations due to degradation or anomalies, the CGEKF is designed with the ability to adjust its performance through the adjustment of artificial parameters called tuning parameters. With this design approach, the CGEKF can maintain accurate estimation performance when it is applied to aircraft engines at offnominal conditions. The performance of the CGEKF is evaluated in a simulation environment using numerous component degradation and fault scenarios at multiple operating conditions.

Kobayashi, Takahisa↗

Enhanced Self Tuning On-Board Real-Time Model (eSTORM) for Aircraft Engine Performance Health Tracking

A key technological concept for producing reliable engine diagnostics and prognostics exploits the benefits of fusing sensor data, information, and/or processing algorithms. This report describes the development of a hybrid engine model for a propulsion gas turbine engine, which is the result of fusing two diverse modeling methodologies: a physics-based model approach and an empirical model approach. The report describes the process and methods involved in deriving and implementing a hybrid model configuration for a commercial turbofan engine. Among the intended uses for such a model is to enable real-time, on-board tracking of engine module performance changes and engine parameter synthesis for fault detection and accommodation.

Volponi, Al↗

Solid-state lasers for coherent communication and remote sensing

Semiconductor-diode laser-pumped solid-state lasers have properties that are superior to other lasers for the applications of coherent communication and remote sensing. These properties include efficiency, reliability, stability, and capability to be scaled to higher powers. We have demonstrated that an optical phase-locked loop can be used to lock the frequency of two diode-pumped 1.06 micron Nd:YAG lasers to levels required for coherent communication. Monolithic nonplanar ring oscillators constructed from solid pieces of the laser material provide better than 10 kHz frequency stability over 0.1 sec intervals. We have used active feedback stabilization of the cavity length of these lasers to demonstrate 0.3 Hz frequency stabilization relative to a reference cavity. We have performed experiments and analysis to show that optical parametric oscillators (OPO's) reproduce the frequency stability of the pump laser in outputs that can be tuned to arbitrary wavelengths. Another measurement performed in this program has demonstrated the sub-shot-noise character of correlations of the fluctuations in the twin output of OPO's. Measurements of nonlinear optical coefficients by phase-matched second harmonic generation are helping to resolve inconsistency in these important parameters.

Byer, Robert L.↗

Implementing Access to Data Distributed on Many Processors

A reference architecture is defined for an object-oriented implementation of domains, arrays, and distributions written in the programming language Chapel. This technology primarily addresses domains that contain arrays that have regular index sets with the low-level implementation details being beyond the scope of this discussion. What is defined is a complete set of object-oriented operators that allows one to perform data distributions for domain arrays involving regular arithmetic index sets. What is unique is that these operators allow for the arbitrary regions of the arrays to be fragmented and distributed across multiple processors with a single point of access giving the programmer the illusion that all the elements are collocated on a single processor. Today's massively parallel High Productivity Computing Systems (HPCS) are characterized by a modular structure, with a large number of processing and memory units connected by a high-speed network. Locality of access as well as load balancing are primary concerns in these systems that are typically used for high-performance scientific computation. Data distributions address these issues by providing a range of methods for spreading large data sets across the components of a system. Over the past two decades, many languages, systems, tools, and libraries have been developed for the support of distributions. Since the performance of data parallel applications is directly influenced by the distribution strategy, users often resort to low-level programming models that allow fine-tuning of the distribution aspects affecting performance, but, at the same time, are tedious and error-prone. This technology presents a reusable design of a data-distribution framework for data parallel high-performance applications. Distributions are a means to express locality in systems composed of large numbers of processor and memory components connected by a network. Since distributions have a great effect on the performance of applications, it is important that the distribution strategy is flexible, so its behavior can change depending on the needs of the application. At the same time, high productivity concerns require that the user be shielded from error-prone, tedious details such as communication and synchronization.

James, Mark↗

Scramjet Performance Assessment Using Water Absorption Diagnostics (U)

Simultaneous multiple path measurements of temperature and H2O concentration will be presented for the AIMHYE test entries in the NASA Ames 16-Inch Shock Tunnel. Monitoring the progress of high temperature chemical reactions that define scramjet combustor efficiencies is a task uniquely suited to nonintrusive optical diagnostics. One application strategy to overcome the many challenges and limitations of nonintrusive measurements is to use laser absorption spectroscopy coupled with optical fibers. Absorption spectroscopic techniques with rapidly tunable lasers are capable of making simultaneous measurements of mole fraction, temperature, pressure, and velocity. The scramjet water absorption diagnostic was used to measure combustor efficiency and was compared to thrust measurements using a nozzle force balance and integrated nozzle pressures to develop a direct technique for evaluating integrated scramjet performance. Tests were initially performed with a diode laser tuning over a water absorption feature at 1391.7 nm. A second diode laser later became available at a wavelength near 1343.3 nm covering an additional water absorption feature and was incorporated in the system for a two-wavelength technique. Both temperature and mole fraction can be inferred from the lineshape analysis using this approach. Additional high temperature spectroscopy research was conducted to reduce uncertainties in the scramjet application. The lasers are optical fiber coupled to ports at the combustor exit and in the nozzle region. The output from the two diode lasers were combined in a single fiber, and the resultant two-wavelength beam was subsequently split into four legs. Each leg was directed through 60 meters of optical fiber to four combustor exit locations for measurement of beam intensity after absorption by the water within the flow. Absorption results will be compared to 1D combustor analysis using RJPA and nozzle CFD computations as well as to data from a nozzle metric balance measuring thrust and integrated pressure measurements along the length of the nozzle. Assessment of its value as a combustor performance evaluation tool will be conducted.

Cavolowsky, John A.↗

Scheduling Operations for Massive Heterogeneous Clusters

High-performance computing (HPC) programming has become increasingly difficult with the advent of hybrid supercomputers consisting of multicore CPUs and accelerator boards such as the GPU. Manual tuning of software to achieve high performance on this type of machine has been performed by programmers. This is needlessly difficult and prone to being invalidated by new hardware, new software, or changes in the underlying code. A system was developed for task-based representation of programs, which when coupled with a scheduler and runtime system, allows for many benefits, including higher performance and utilization of computational resources, easier programming and porting, and adaptations of code during runtime. The system consists of a method of representing computer algorithms as a series of data-dependent tasks. The series forms a graph, which can be scheduled for execution on many nodes of a supercomputer efficiently by a computer algorithm. The schedule is executed by a dispatch component, which is tailored to understand all of the hardware types that may be available within the system. The scheduler is informed by a cluster mapping tool, which generates a topology of available resources and their strengths and communication costs. Software is decoupled from its hardware, which aids in porting to future architectures. A computer algorithm schedules all operations, which for systems of high complexity (i.e., most NASA codes), cannot be performed optimally by a human. The system aids in reducing repetitive code, such as communication code, and aids in the reduction of redundant code across projects. It adds new features to code automatically, such as recovering from a lost node or the ability to modify the code while running. In this project, the innovators at the time of this reporting intend to develop two distinct technologies that build upon each other and both of which serve as building blocks for more efficient HPC usage. First is the scheduling and dynamic execution framework, and the second is scalable linear algebra libraries that are built directly on the former.

Humphrey, John↗