Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallel estimation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Correlation filters for orientation estimation

An important task in many vision applications is that of rapidly estimating the orientation of an object with respect to some frame of reference. Because of their speed and parallel processing capabilities, optical correlators should prove valuable in this application. This paper considers two algorithms for object orientation estimation based on optical correlations and presents some initial simulation results.

Kumar, B. V. K. Vijaya↗

Time estimation as a secondary task to measure workload

Variation in the length of time productions and verbal estimates of duration was investigated to determine the influence of concurrent activity on operator time perception. The length of 10-, 20-, and 30-sec intervals produced while performing six different compensatory tracking tasks was significantly longer, 23% on the average, than those produced while performing no other task. Verbal estimates of session duration, taken at the end of each of 27 experimental sessions, reflected a parallel increase in subjective underestimation of the passage of time as the difficulty of the task performed increased. These data suggest that estimates of duration made while performing a manual control task provide stable and sensitive measures of the workload imposed by the primary task, with minimal interference.

Sandra G. Hart↗

Performance Evaluation and Modeling Techniques for Parallel Processors

In practice, the performance evaluation of supercomputers is still substantially driven by singlepoint estimates of metrics (e.g., MFLOPS) obtained by running characteristic benchmarks or workloads. With the rapid increase in the use of time-shared multiprogramming in these systems, such measurements are clearly inadequate. This is because multiprogramming and system overhead, as well as other degradations in performance due to time varying characteristics of workloads, are not taken into account. In multiprogrammed environments, multiple jobs and users can dramatically increase the amount of system overhead and degrade the performance of the machine. Performance techniques, such as benchmarking, which characterize performance on a dedicated machine ignore this major component of true computer performance. Due to the complexity of analysis, there has been little work done in analyzing, modeling, and predicting the performance of applications in multiprogrammed environments. This is especially true for parallel processors, where the costs and benefits of multi-user workloads are exacerbated. While some may claim that the issue of multiprogramming is not a viable one in the supercomputer market, experience shows otherwise. Even in recent massively parallel machines, multiprogramming is a key component. It has even been claimed that a partial cause of the demise of the CM2 was the fact that it did not efficiently support time-sharing. In the same paper, Gordon Bell postulates that, multicomputers will evolve to multiprocessors in order to support efficient multiprogramming. Therefore, it is clear that parallel processors of the future will be required to offer the user a time-shared environment with reasonable response times for the applications. In this type of environment, the most important performance metric is the completion of response time of a given application. However, there are a few evaluation efforts addressing this issue.

Dimpsey, Robert Tod↗

Stress modeling of microdiaphragm pressure sensors

A finite element program analysis was used to model the stress distribution of two monocrystalline silicon diaphragm pressure sensors. One configuration consists of an anisotropically backside etched diaphragm into a 250 micron thick, (100) oriented, silicon wafer. The diaphragm and total chip dimensions are given. The device is rigidly clamped on the back to a support substrate. Another configuration consists of a monocrystalline, (100), microdiaphragm which is formed on top of the wafer and whose area is reduced by a factor of 25 over the first configuration. The diaphragm is rigidly clamped to the silicon wafer. The stresses were calculated at a gauge pressure of 300 mm Hg and used to estimate the piezoresistive responses of resistor elements which were placed parallel and perpendicular near the diaphragm edges.

Tack, P. C.↗

NASA Multidimensional Stirling Convertor Code Developed

A high-efficiency Stirling Radioisotope Generator (SRG) for use on potential NASA Space Science missions is being developed by the Department of Energy, Lockheed Martin, Stirling Technology Company, and the NASA Glenn Research Center. These missions may include providing spacecraft onboard electric power for deep space missions or power for unmanned Mars rovers. Glenn is also developing advanced technology for Stirling convertors, aimed at substantially improving the specific power and efficiency of the convertor and the overall power system. Performance and mass improvement goals have been established for second- and third-generation Stirling radioisotope power systems. Multiple efforts are underway to achieve these goals, both in house at Glenn and under various grants and contracts. These efforts include the development of a multidimensional Stirling computational fluid dynamics code, high-temperature materials, advanced controllers, an end-to-end system dynamics model, low-vibration techniques, advanced regenerators, and a lightweight convertor. Under a NASA grant, Cleveland State University (CSU) and its subcontractors, the University of Minnesota (UMN) and Gedeon Associates, have developed a twodimensional computer simulation of a CSUmod Stirling convertor. The CFD-ACE commercial software developed by CFD Research Corp. of Huntsville, Alabama, is being used. The CSUmod is a scaled version of the Stirling Technology Demonstrator Convertor (TDC), which was designed and fabricated by the Stirling Technology Company and is being tested by NASA. The schematic illustrates the structure of this model. Modeled are the fluid-flow and heat-transfer phenomena that occur in the expansion space, the heater, the regenerator, the cooler, the compression space, the surrounding walls, and the moving piston and displacer. In addition, the overall heat transfer, the indicated power, and the efficiency can be calculated. The CSUmod model is being converted to a two-dimensional model of the TDC at NASA Glenn. Validation of the multidimensional Stirling code is an important part of the grant effort. UMN has been generating data in an oscillating-flow test facility using two different test sections: a 90 turn and a cooler/regenerator/heater test section. CSU has created computational fluid dynamics models of both these test sections and has been making comparisons with the data, then improving their models to improve the agreement with the test data. CSU has also been using data available in the literature for code validation. UMN is now preparing to begin fabrication of a new 180 turn test section that will be more representative of certain portions of the Stirling engine geometry. Simulations to almost periodic steady state with the two-dimensional CSUmod model indicate that, to reach periodic steady state on a single 2-GHz desktop computer, 75 to 100 complete simulation cycles would be required and between 1 and 2 months of computer time. Therefore, Glenn has purchased the first 8 computers, of a 64-computer cluster, to be run in parallel to accelerate the simulation. On the basis of CFD Research Corp.'s experience with running the parallelized version of CFD-ACE on their clusters, we estimate that the complete 64-computer cluster will reduce simulation computing time by a factor of about 40. Plans are to continue development of these multidimensional Stirling codes and to use them to study the fluid-flow and heat-transfer phenomena that occur inside Stirling convertors. This is expected to lead to improved thermodynamic loss understanding, onedimensional design and performance codes, and engine performance.

Tew, Roy C.↗

Comparison of Precision of Biomass Estimates in Regional Field Sample Surveys and Airborne LiDAR-Assisted Surveys in Hedmark County, Norway

Airborne scanning LiDAR (Light Detection and Ranging) has emerged as a promising tool to provide auxiliary data for sample surveys aiming at estimation of above-ground tree biomass (AGB), with potential applications in REDD forest monitoring. For larger geographical regions such as counties, states or nations, it is not feasible to collect airborne LiDAR data continuously ("wall-to-wall") over the entire area of interest. Two-stage cluster survey designs have therefore been demonstrated by which LiDAR data are collected along selected individual flight-lines treated as clusters and with ground plots sampled along these LiDAR swaths. Recently, analytical AGB estimators and associated variance estimators that quantify the sampling variability have been proposed. Empirical studies employing these estimators have shown a seemingly equal or even larger uncertainty of the AGB estimates obtained with extensive use of LiDAR data to support the estimation as compared to pure field-based estimates employing estimators appropriate under simple random sampling (SRS). However, comparison of uncertainty estimates under SRS and sophisticated two-stage designs is complicated by large differences in the designs and assumptions. In this study, probability-based principles to estimation and inference were followed. We assumed designs of a field sample and a LiDAR-assisted survey of Hedmark County (HC) (27,390 km2), Norway, considered to be more comparable than those assumed in previous studies. The field sample consisted of 659 systematically distributed National Forest Inventory (NFI) plots and the airborne scanning LiDAR data were collected along 53 parallel flight-lines flown over the NFI plots. We compared AGB estimates based on the field survey only assuming SRS against corresponding estimates assuming two-phase (double) sampling with LiDAR and employing model-assisted estimators. We also compared AGB estimates based on the field survey only assuming two-stage sampling (the NFI plots being grouped in clusters) against corresponding estimates assuming two-stage sampling with the LiDAR and employing model-assisted estimators. For each of the two comparisons, the standard errors of the AGB estimates were consistently lower for the LiDAR-assisted designs. The overall reduction of the standard errors in the LiDAR-assisted estimation was around 40-60% compared to the pure field survey. We conclude that the previously proposed two-stage model-assisted estimators are inappropriate for surveys with unequal lengths of the LiDAR flight-lines and new estimators are needed. Some options for design of LiDAR-assisted sample surveys under REDD are also discussed, which capitalize on the flexibility offered when the field survey is designed as an integrated part of the overall survey design as opposed to previous LiDAR-assisted sample surveys in the boreal and temperate zones which have been restricted by the current design of an existing NFI.

Comparison↗

Exploring the Connection Between Sampling Problems in Bayesian Inference and Statistical Mechanics

The Bayesian and statistical mechanical communities often share the same objective in their work - estimating and integrating probability distribution functions (pdfs) describing stochastic systems, models or processes. Frequently, these pdfs are complex functions of random variables exhibiting multiple, well separated local minima. Conventional strategies for sampling such pdfs are inefficient, sometimes leading to an apparent non-ergodic behavior. Several recently developed techniques for handling this problem have been successfully applied in statistical mechanics. In the multicanonical and Wang-Landau Monte Carlo (MC) methods, the correct pdfs are recovered from uniform sampling of the parameter space by iteratively establishing proper weighting factors connecting these distributions. Trivial generalizations allow for sampling from any chosen pdf. The closely related transition matrix method relies on estimating transition probabilities between different states. All these methods proved to generate estimates of pdfs with high statistical accuracy. In another MC technique, parallel tempering, several random walks, each corresponding to a different value of a parameter (e.g. "temperature"), are generated and occasionally exchanged using the Metropolis criterion. This method can be considered as a statistically correct version of simulated annealing. An alternative approach is to represent the set of independent variables as a Hamiltonian system. Considerab!e progress has been made in understanding how to ensure that the system obeys the equipartition theorem or, equivalently, that coupling between the variables is correctly described. Then a host of techniques developed for dynamical systems can be used. Among them, probably the most powerful is the Adaptive Biasing Force method, in which thermodynamic integration and biased sampling are combined to yield very efficient estimates of pdfs. The third class of methods deals with transitions between states described by rate constants. These problems are isomorphic with chemical kinetics problems. Recently, several efficient techniques for this purpose have been developed based on the approach originally proposed by Gillespie. Although the utility of the techniques mentioned above for Bayesian problems has not been determined, further research along these lines is warranted

Pohorille, Andrew↗

Thin Film Thermal Conductivity Measurements using Superconducting Nanowires

We present a simple experimental scheme for estimating the cryogenic thermal transport properties of thin films using superconducting nanowires. In a parallel array of nanowires, the heat from one nanowire in the normal state changes the local temperature around adjacent nanowires, reducing their switching current. Calibration of this change in switching current as a function of bath temperature provides an estimate of the temperature as a function of displacement from the heater. This provides a method of determining the contribution of substrate heat transport to the cooling time of superconducting nanowire single photon detectors. Understanding this process is necessary for successful electrothermal modeling of superconducting nanowire systems.

Shaw, M. D.↗

Landau damping of auroral hiss

Auroral hiss is observed to propagate over distances comparable to an Earth radius from its source in the auroral oval. The role of Landau damping is investigated for upward propagating auroral hiss. By using a ray tracing code and a simplified model of the distribution function, the effect of Landau damping is calculated for auroral hiss propagation through the environment around the auroral oval. Landau damping is found to be the likely mechanism for explaining some of the one-sided auroral hiss funnels observed by Dynamics Explorer 1. It is also found that Landau damping puts a lower limit on the wavelength of auroral hiss. Poleward of the auroral oval, Landau damping is found in a typical case to limit omega/k(sub parallel) to values of 3.4 x 10(exp 4) km/s or greater, corresponding to resonance energies of 3.2 keV or greater and wavelengths of 2 km or greater. For equatorward propagation, omega/k(sub parallel) is limited to values greater than 6.8 x 10(exp 4) km/s, corresponding to resonance energies greater than 13 keV and wavelengths greater than 3 km. Independent estimates based on measured ratios of the magnetic to electric field intensity also show that omega/k(sub parallel) corresponds to resonance energies greater than 1 keV and wavelengths greater than 1 km. These results lead to the difficulty that upgoing electron beams sufficiently energetic to directly generate auroral hiss of the inferred wavelength are not usually observed. A partial transmission mechanism utilizing density discontinuities oblique to the magnetic field is proposed for converting auroral hiss to wavelengths long enough to avoid damping of the wave over long distances. Numerous reflections of the wave in an upwardly flared density cavity could convert waves to significantly increased wavelengths and resonance velocities.

Morgan, D. D.↗

Interval Management with Spacing to Parallel Dependent Runways (IMSPIDR) Experiment and Results

An area in aviation operations that may offer an increase in efficiency is the use of continuous descent arrivals (CDA), especially during dependent parallel runway operations. However, variations in aircraft descent angle and speed can cause inaccuracies in estimated time of arrival calculations, requiring an increase in the size of the buffer between aircraft. This in turn reduces airport throughput and limits the use of CDAs during high-density operations, particularly to dependent parallel runways. The Interval Management with Spacing to Parallel Dependent Runways (IMSPiDR) concept uses a trajectory-based spacing tool onboard the aircraft to achieve by the runway an air traffic control assigned spacing interval behind the previous aircraft. This paper describes the first ever experiment and results of this concept at NASA Langley. Pilots flew CDAs to the Dallas Fort-Worth airport using airspeed calculations from the spacing tool to achieve either a Required Time of Arrival (RTA) or Interval Management (IM) spacing interval at the runway threshold. Results indicate flight crews were able to land aircraft on the runway with a mean of 2 seconds and less than 4 seconds standard deviation of the air traffic control assigned time, even in the presence of forecast wind error and large time delay. Statistically significant differences in delivery precision and number of speed changes as a function of stream position were observed, however, there was no trend to the difference and the error did not increase during the operation. Two areas the flight crew indicated as not acceptable included the additional number of speed changes required during the wind shear event, and issuing an IM clearance via data link while at low altitude. A number of refinements and future spacing algorithm capabilities were also identified.

Baxley, Brian T.↗

Performance degradation due to multiprogramming and system overheads in real workloads - Case study on a shared memory multiprocessor

The performance degradation due to the multiprogramming (MP) overhead in a parallel execution environment is quantified. In addition, total system overhead is also measured. A methodology, which estimates the MP overhead present in real workloads, is illustrated with real measurents. It is found that MP overhead usually consumes between 10 and 23 percent of the processing power available to parallel programs. The mean MP overhead is determined to be 16 percent which is well over half the total system overhead executed on the system (the mean system overhead is determined to be 24 percent of the processing power). It is found that MP overhead, total system overhead, and application completion time are all moderately correlated.

Dimpsey, R. T.↗

Performance of a neutralizer for electron bombardment thruster

Results of the SERT II flight indicate that the hollow cathode neutralizer not only represents a power and propellant weight penalty but can be a contributing cause to accelerator grid erosion. Tests with a 30-cm diameter thruster show that a neutralizer position of approximately 9 cm axially downstream of the accelerator grid and approximately 9 cm radially away from the outer edge of the accelerator grid and pointing parallel to the thruster axis provides the best overall performance. The estimated grid wear rate was less than 0.08 mm in 10,000 hours. The coupling voltage was approximately 17 volts at a neutralizer flow rate of 22 equivalent milliamperes of mercury and a beam current of 1.5 amperes. Neutralizer power was 31 watts and the effect of neutralizer flow on overall propellant utilization efficiency is a 1.2 percentage point reduction at a thruster utilization efficiency of 90 percent. The neutralizer position defined in tests with a 30 cm thruster was tested with a 15 cm SERT II thruster. When the neutralizer was relocated further downstream with this different orientation, accelerator impingement current due to neutralizer operation was reduced by approximately a factor of seven and was nearly independent of neutralizer operation.

Bechtel, R. T.↗

Performance of a neutralizer for electron bombardment thruster.

Results of the SERT II flight indicate that the hollow cathode neutralizer not only represents a power and propellant weight penalty but can be a contributing cause to accelerator grid erosion. Tests with a 30-cm diameter thruster have shown that a neutralizer position of approximately 9 cm axially downstream of the accelerator grid and approximately 9 cm radially away from the outer edge of the accelerator grid and pointing parallel to the thruster axis provides the best overall performance. The estimated grid wear rate was less than 0.08 mm in 10,000 hr. The coupling voltage (neutralizer to beam voltage) was approximately 17 volts at a neutralizer flow rate of 22 equivalent milliamperes of mercury and a beam current of 1.5 amperes. Neutralizer power (excluding heaters) was 31 watts and the effect of neutralizer flow on overall propellant utilization efficiency is a 1.2 percentage point reduction at a thruster utilization efficiency of 90 percent.

Bechtel, R. T.↗

Geometric-optical modeling of a conifer forest canopy

A geometric-optical model of a conifer forest canopy was constructed to describe the variance of a remotely-sensed image of a forest stand. The model is driven by interpixel variance generated from three sources: (1) the number of crowns in the pixel; (2) the size of the individual crowns; and (3) the overlapping of crowns and shadows. The illumination of a tree and its shadow is described as a cone using parallel-ray geometry. The model can also be inverted to provide estimates of the size, shape and spacing of the trees using remotely-sensed imagery and a minimum of ground measurements. The results of field tests using 10-meter and 80-meter multispectral imagery of two test conifer stands in northeastern California are presented. It is shown that the model produces reasonable estimates for the geometric parameters of the stand and appears to be sufficiently robust for application to other geometric shapes corresponding to different types of vegetation.

Li, X.↗

A Simple GPU-Accelerated Two-Dimensional MUSCL-Hancock Solver for Ideal Magnetohydrodynamics

We describe our experience using NVIDIA's CUDA (Compute Unified Device Architecture) C programming environment to implement a two-dimensional second-order MUSCL-Hancock ideal magnetohydrodynamics (MHD) solver on a GTX 480 Graphics Processing Unit (GPU). Taking a simple approach in which the MHD variables are stored exclusively in the global memory of the GTX 480 and accessed in a cache-friendly manner (without further optimizing memory access by, for example, staging data in the GPU's faster shared memory), we achieved a maximum speed-up of approx. = 126 for a sq 1024 grid relative to the sequential C code running on a single Intel Nehalem (2.8 GHz) core. This speedup is consistent with simple estimates based on the known floating point performance, memory throughput and parallel processing capacity of the GTX 480.

graphics processing units↗

Autonomous attitude estimation via star sensing and pattern recognition

Results are reported on the development of an autonomous, onboard, near real time spacecraft attitude estimation technique. The approach uses CCD based star sensors to digitize relative star positions. Three microcomputers are envisioned, configured in parallel, to: (1) determine star image centroids and delete spurious images; (2) identify measured stars with stars in an onboard catalog and determine discrete attitude estimates; (3) integrate gyro rate measurements and determine optimal real time attitude estimates for use in the control system and for feedback to the star identification algorithm. Algorithms for the star identification are presented. The discrete attitude estimation algorithm recovers thermally varying interlock angles between two star sensors. The optimal state estimation process recovers rate gyro biases in addition to real time attitude estimates.

Junkins, J. L.↗

Implementation and Characterization of Three-Dimensional Particle-in-Cell Codes on Multiple-Instruction-Multiple-Data Massively Parallel Supercomputers

A three-dimensional electrostatic particle-in-cell (PIC) plasma simulation code has been developed on coarse-grain distributed-memory massively parallel computers with message passing communications. Our implementation is the generalization to three-dimensions of the general concurrent particle-in-cell (GCPIC) algorithm. In the GCPIC algorithm, the particle computation is divided among the processors using a domain decomposition of the simulation domain. In a three-dimensional simulation, the domain can be partitioned into one-, two-, or three-dimensional subdomains ("slabs," "rods," or "cubes") and we investigate the efficiency of the parallel implementation of the push for all three choices. The present implementation runs on the Intel Touchstone Delta machine at Caltech; a multiple-instruction-multiple-data (MIMD) parallel computer with 512 nodes. We find that the parallel efficiency of the push is very high, with the ratio of communication to computation time in the range 0.3%-10.0%. The highest efficiency (> 99%) occurs for a large, scaled problem with 64(sup 3) particles per processing node (approximately 134 million particles of 512 nodes) which has a push time of about 250 ns per particle per time step. We have also developed expressions for the timing of the code which are a function of both code parameters (number of grid points, particles, etc.) and machine-dependent parameters (effective FLOP rate, and the effective interprocessor bandwidths for the communication of particles and grid points). These expressions can be used to estimate the performance of scaled problems--including those with inhomogeneous plasmas--to other parallel machines once the machine-dependent parameters are known.

Lyster, P. M.↗

Intelligent flight control systems

The capabilities of flight control systems can be enhanced by designing them to emulate functions of natural intelligence. Intelligent control functions fall in three categories. Declarative actions involve decision-making, providing models for system monitoring, goal planning, and system/scenario identification. Procedural actions concern skilled behavior and have parallels in guidance, navigation, and adaptation. Reflexive actions are spontaneous, inner-loop responses for control and estimation. Intelligent flight control systems learn knowledge of the aircraft and its mission and adapt to changes in the flight environment. Cognitive models form an efficient basis for integrating 'outer-loop/inner-loop' control functions and for developing robust parallel-processing algorithms.

Stengel, Robert F.↗