Search NASA⌕ Search

SEARCH · Search NASA

Results for “Arithmetic”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 415 records · Page 23

Applications Performance on NAS Intel Paragon XP/S - 15#

The Numerical Aerodynamic Simulation (NAS) Systems Division received an Intel Touchstone Sigma prototype model Paragon XP/S- 15 in February, 1993. The i860 XP microprocessor with an integrated floating point unit and operating in dual -instruction mode gives peak performance of 75 million floating point operations (NIFLOPS) per second for 64 bit floating point arithmetic. It is used in the Paragon XP/S-15 which has been installed at NAS, NASA Ames Research Center. The NAS Paragon has 208 nodes and its peak performance is 15.6 GFLOPS. Here, we will report on early experience using the Paragon XP/S- 15. We have tested its performance using both kernels and applications of interest to NAS. We have measured the performance of BLAS 1, 2 and 3 both assembly-coded and Fortran coded on NAS Paragon XP/S- 15. Furthermore, we have investigated the performance of a single node one-dimensional FFT, a distributed two-dimensional FFT and a distributed three-dimensional FFT Finally, we measured the performance of NAS Parallel Benchmarks (NPB) on the Paragon and compare it with the performance obtained on other highly parallel machines, such as CM-5, CRAY T3D, IBM SP I, etc. In particular, we investigated the following issues, which can strongly affect the performance of the Paragon: a. Impact of the operating system: Intel currently uses as a default an operating system OSF/1 AD from the Open Software Foundation. The paging of Open Software Foundation (OSF) server at 22 MB to make more memory available for the application degrades the performance. We found that when the limit of 26 NIB per node out of 32 MB available is reached, the application is paged out of main memory using virtual memory. When the application starts paging, the performance is considerably reduced. We found that dynamic memory allocation can help applications performance under certain circumstances. b. Impact of data cache on the i860/XP: We measured the performance of the BLAS both assembly coded and Fortran coded. We found that the measured performance of assembly-coded BLAS is much less than what memory bandwidth limitation would predict. The influence of data cache on different sizes of vectors is also investigated using one-dimensional FFTs. c. Impact of processor layout: There are several different ways processors can be laid out within the two-dimensional grid of processors on the Paragon. We have used the FFT example to investigate performance differences based on processors layout.

Saini, Subhash↗

Real Automation in the Field

We provide a package of strategies for automation of non-linear arithmetic in PVS. In particular, we describe a simplication procedure for the field of real numbers and a strategy for cancellation of common terms.

Munoz, Cesar↗

Finding New Math Identities by Computer

Recently a number of interesting new mathematical identities have been discovered by means of numerical searches on high performance computers, using some newly discovered algorithms. These include the following: pi = ((sup oo)(sub k=0))(Sigma) (1 / 16) (sup k) ((4 / 8k+1) - (2 / 8k+4) - (1 / 8k+5) - (1 / 8k+6)) and ((17 pi(exp 4)) / 360) = ((sup oo)(sub k=1))(Sigma) (1 + (1/2) + (1/3) + ... + (1/k))(exp 2) k(exp -2), zeta(3, 1, 3, 1, ..., 3, 1) = (2 pi(exp 4m) / (4m+2)! where m = number of (3,1) pairs. and where zeta(n1,n2,...,nr) = (sub k1 (is greater than) k2 (is greater than) ... (is greater than) kr)(Sigma) (1 / (k1 (sup n1) k2 (sup n2) ... kr (sup nr). The first identity is remarkable in that it permits one to compute the n-th binary or hexadecimal digit of pu directly, without computing any of the previous digits, and without using multiple precision arithmetic. Recently the ten billionth hexadecimal digit of pi was computed using this formula. The third identity has connections to quantum field theory. (The first and second of these been formally established; the third is affirmed by numerical evidence only.) The background and results of this work will be described, including an overview of the algorithms and computer techniques used in these studies.

Bailey, David H.↗

Efficacy of Code Optimization on Cache-based Processors

The current common wisdom in the U.S. is that the powerful, cost-effective supercomputers of tomorrow will be based on commodity (RISC) micro-processors with cache memories. Already, most distributed systems in the world use such hardware as building blocks. This shift away from vector supercomputers and towards cache-based systems has brought about a change in programming paradigm, even when ignoring issues of parallelism. Vector machines require inner-loop independence and regular, non-pathological memory strides (usually this means: non-power-of-two strides) to allow efficient vectorization of array operations. Cache-based systems require spatial and temporal locality of data, so that data once read from main memory and stored in high-speed cache memory is used optimally before being written back to main memory. This means that the most cache-friendly array operations are those that feature zero or unit stride, so that each unit of data read from main memory (a cache line) contains information for the next iteration in the loop. Moreover, loops ought to be 'fat', meaning that as many operations as possible are performed on cache data-provided instruction caches do not overflow and enough registers are available. If unit stride is not possible, for example because of some data dependency, then care must be taken to avoid pathological strides, just ads on vector computers. For cache-based systems the issues are more complex, due to the effects of associativity and of non-unit block (cache line) size. But there is more to the story. Most modern micro-processors are superscalar, which means that they can issue several (arithmetic) instructions per clock cycle, provided that there are enough independent instructions in the loop body. This is another argument for providing fat loop bodies. With these restrictions, it appears fairly straightforward to produce code that will run efficiently on any cache-based system. It can be argued that although some of the important computational algorithms employed at NASA Ames require different programming styles on vector machines and cache-based machines, respectively, neither architecture class appeared to be favored by particular algorithms in principle. Practice tells us that the situation is more complicated. This report presents observations and some analysis of performance tuning for cache-based systems. We point out several counterintuitive results that serve as a cautionary reminder that memory accesses are not the only factors that determine performance, and that within the class of cache-based systems, significant differences exist.

VanderWijngaart, Rob F.↗

On the Floating Point Performance of the i860 Microprocessor

The i860 microprocessor is a pipelined processor that can deliver two double precision floating point results every clock. It is being used in the Touchstone project to develop a teraflop computer by the year 2000. With such high computational capabilities it was expected that memory bandwidth would limit performance on many kernels. Measured performance of three kernels showed performance is less than what memory bandwidth limitations would predict. This paper develops a model that explains the discrepancy in terms of memory latencies and points to some problems involved in moving data from memory to the arithmetic pipelines.

Lee, King↗

Triangle Geometry Processing for Surface Modeling and Cartesian Grid Generation

Cartesian mesh generation is accomplished for component based geometries, by intersecting components subject to mesh generation to extract wetted surfaces with a geometry engine using adaptive precision arithmetic in a system which automatically breaks ties with respect to geometric degeneracies. During volume mesh generation, intersected surface triangulations are received to enable mesh generation with cell division of an initially coarse grid. The hexagonal cells are resolved, preserving the ability to directionally divide cells which are locally well aligned.

Aftosmis, Michael J.↗

Analysis of the Effect of Surface Modification on Polyimide Composites Coated with Erosion Resistant Materials

The aim of this research is to enhance performance of composite coatings through modification of graphite-reinforced polyimide composite surfaces prior to metal bond coat/ hard topcoat application for use in the erosive and/or oxidative environments of advanced engines. Graphite reinforced polyimide composites, PMR-15 and PMR-II-50, formed by sheet molding and pre-pregging will be surface treated, overlaid with a bond coat and then coated with WC-Co. The surface treatment will include cleaning, RF plasma or ultraviolet light- ozone etching, and deposition of SiO(x) groups. These surface treatments will be studied in order to investigate and improve adhesion and oxidation resistance. The following panels were provided by NASA-Glenn Research Center(NASA-GRC): Eight compression molded PMR-II-50; 6 x 6 x 0.125 in. Two vacuum-bagged PMR-II-50; 12 x 12 x 0.125 in. Eight compression molded PMR-15; 6 x 6 x 0.125 in. One vacuum-bagged PMR-15; 12 x 12 x 0.125 in. All panels were made using a 12 x 12 in. T650-35 8HS (3K-tow) graphite fabric. A diamond-wafering blade, with deionized water as a cutting fluid, was used to cut PMR-II-50 and PMR-15 panels into 1 x 1 in. pieces for surface tests. The panel edges exhibiting delamination were used for the preliminary surface preparation tests as these would be unsuitable for strength and erosion testing. PMR-15 neat resin samples were also provided by NASA GRC. Surface profiles of the as-received samples were determined using a Dektak III Surface profile measuring system. Two samples of compression molded PMR-II-50 and PMR-15, vacuum-bagged PMR-II-50 and PMR-15 were randomly chosen for surface profile measurement according to ANSI/ASME B46.1. Prior to each measurement, the samples were blasted with compressed air to remove any artifacts. Five 10 mm-long scans were made on each sample. The short and long wavelength cutoff filter values were set at 100 and 1000 m, diamond stylus radius was 12.5 microns. Table 1 is a summary of the arithmetic average roughness (Ra) and waviness (Wa) for the composite surfaces.

Ndalama, Tchinga↗

Multi-Decadal Pathfinder Data Sets of Global Land Biophysical Variables from AVHRR and MODIS and their Use in GCM Studies of Biogeophysics and Biogeochemistry

The problem of how the scale, or spatial resolution, of reflectance data impacts retrievals of vegetation leaf area index (LAI) and fraction absorbed photosynthetically active radiation (PAR) has been investigated. We define the goal of scaling as the process by which it is established that LAI and FPAR values derived from coarse resolution sensor data equal the arithmetic average of values derived independently from fine resolution sensor data. The increasing probability of land cover mixtures with decreasing resolution is defined as heterogeneity, which is a key concept in scaling studies. The effect of pixel heterogeneity on spectral reflectances and LAI/FPAR retrievals is investigated with 1 km Advanced Very High Resolution Radiometer (AVHRR) data aggregated to different coarse spatial resolutions. It is shown that LAI retrieval errors at coarse resolution are inversely related to the proportion of the dominant land cover in such pixel. Further, large errors in LAI retrievals are incurred when forests are minority biomes in non-forest pixels compared to when forest biomes are mixed with one another, and vice-versa. A physically based technique for scaling with explicit spatial resolution dependent radiative transfer formulation is developed. The successful application of this theory to scaling LAI retrievals from AVHRR data of different resolutions is demonstrated

Myneni, Ranga↗

Design and Application of Strategies/Tactics in Higher Order Logics

This Proceedings includes both a paper from the implementors of PVS providing guidance for PVS strategy writers and a tutorial on PVS strategy writing distilled from the experience of three PVS users who have written extensive sets of PVS user strategies. Following these are three full papers from the higher-order logic theorem proving community that discuss PVS strategies to enhance arithmetic and other interactive reasoning in PVS; implementing first-order tactics in higher-order provers; and a proposed technique for specifying small step semantics that can be used in multiple higher order logic theorem provers, with illustrations from both Coq and PVS. The Proceedings concludes with three position papers for a panel session that discuss three settings in which development of PVS strategies is worth while.

Archer, Myla↗

Rapid Prototyping in PVS

PVSio is a conservative extension to the PVS prelude library that provides basic input/output capabilities to the PVS ground evaluator. It supports rapid prototyping in PVS by enhancing the specification language with built-in constructs for string manipulation, floating point arithmetic, and input/output operations.

Munoz, Cesar A.↗

Vestibulosympathetic reflex during mental stress

Increases in sympathetic neural activity occur independently with either vestibular or mental stimulation, but it is unknown whether sympathetic activation is additive or inhibitive when both stressors are combined. The purpose of the present study was to investigate the combined effects of vestibular and mental stimulation on sympathetic neural activation and arterial pressure in humans. Muscle sympathetic nerve activity (MSNA), arterial pressure, and heart rate were recorded in 10 healthy volunteers in the prone position during 1) head-down rotation (HDR), 2) mental stress (MS; using arithmetic), and 3) combined HDR and MS. HDR significantly (P < 0.05) increased MSNA (9 +/- 2 to 13 +/- 2 bursts/min). MS significantly increased MSNA (8 +/- 2 to 13 +/- 2 bursts/min) and mean arterial pressure (87 +/- 2 to 101 +/- 2 mmHg). Combined HDR and MS significantly increased MSNA (9 +/- 1 to 16 +/- 2 bursts/min) and mean arterial pressure (89 +/- 2 to 100 +/- 3 mmHg). Increases in MSNA (7 +/- 1 bursts/min) during the combination trial were not different from the algebraic sum of each trial performed alone (8 +/- 2 bursts/min). We conclude that the interaction for MSNA and arterial pressure is additive during combined vestibular and mental stimulation. Therefore, vestibular- and stress-mediated increases of MSNA appear to occur independently in humans.

Non-NASA Center↗

VLSI processors for signal detection in SETI

The objective of the Search for Extraterrestrial Intelligence (SETI) is to locate an artificially created signal coming from a distant star. This is done in two steps: (1) spectral analysis of an incoming radio frequency band, and (2) pattern detection for narrow-band signals. Both steps are computationally expensive and require the development of specially designed computer architectures. To reduce the size and cost of the SETI signal detection machine, two custom VLSI chips are under development. The first chip, the SETI DSP Engine, is used in the spectrum analyzer and is specially designed to compute Discrete Fourier Transforms (DFTs). It is a high-speed arithmetic processor that has two adders, one multiplier-accumulator, and three four-port memories. The second chip is a new type of Content-Addressable Memory. It is the heart of an associative processor that is used for pattern detection. Both chips incorporate many innovative circuits and architectural features.

NASA Program Exobiology↗

Verification of IEEE Compliant Subtractive Division Algorithms

A parameterized definition of subtractive floating point division algorithms is presented and verified using PVS. The general algorithm is proven to satisfy a formal definition of an IEEE standard for floating point arithmetic. The utility of the general specification is illustrated using a number of different instances of the general algorithm.

Miner, Paul S.↗

Analysis of cell mechanics in single vinculin-deficient cells using a magnetic tweezer

A magnetic tweezer was constructed to apply controlled tensional forces (10 pN to greater than 1 nN) to transmembrane receptors via bound ligand-coated microbeadswhile optically measuring lateral bead displacements within individual cells. Use of this system with wild-type F9 embryonic carcinoma cells and cells from a vinculin knockout mouse F9 Vin (-/-) revealed much larger differences in the stiffness of the transmembrane integrin linkages to the cytoskeleton than previously reported using related techniques that measured average mechanical properties of large cell populations. The mechanical properties measured varied widely among cells, exhibiting an approximately log-normal distribution. The median lateral bead displacement was 2-fold larger in F9 Vin (-/-) cells compared to wild-type cells whereas the arithmetic mean displacement only increased by 37%. We conclude that vinculin serves a greater mechanical role in cells than previously reported and that this magnetic tweezer device may be useful for probing the molecular basis of cell mechanics within single cells. Copyright 2000 Academic Press.

Non-NASA Center↗

Effect of vergence on the gain of the linear vestibulo-ocular reflex

We measured the linear vestibulo-ocular reflex (LVOR) and vergence, using binocular search coils, in 3 humans. The subjects were accelerated sinusoidally at 0.5 Hz and 0.2 g peak acceleration, in complete darkness, while performing three different tasks: i) mental arithmetic; ii) tracking a remembered target at either 0.34 m or 0.14 m distance; and iii) maintaining vergence at either of these distances by means of audio biofeedback based on vergence. Subjects could control vergence using the audio feedback; there was greater convergence with the near audio target. However, there was no significant difference in vergence between the near and far remembered target conditions. With audio feedback, the amplitude of smooth tracking was not consistently different for the near and the far conditions. However, the amplitude of tracking (saccades and smooth component) in the remembered target conditions was greater for near than for far targets. These results suggest that linear VOR amplitude is not determined by vergence alone.

NASA Program Space Physiology and Countermeasures↗

Use of promethazine to hasten adaptation to provocative motion

In an earlier study, the authors found that severely motion sick individuals could be greatly relieved of their symptoms by intramuscular injections of promethazine (50 mg) or scopolamine (.5 mg). Comparable 50-mg injections of promethazine also have been found effective in alleviating symptoms of space motion sickness. The concern has risen, however, that such drugs may delay or retard the acquisition of adaptation to stressful environments. In the current study, we controlled arousal using a mental arithmetic task and precisely equated the exposure history (number of head movements during rotation) of a placebo, control group and an experimental group who had received promethazine. No differences in total adaptation or in rates of adaptation were present between the two groups. Another experimental group also received promethazine and was allowed to make as many head movements as they could, before reaching nausea, up to 800. This group showed a greater level of adaptation than the placebo group. These results suggest a strategy for dealing with space motion sickness that is described.

Non-NASA Center↗

Formal Consistency Verification of Deliberative Agents with Respect to Communication Protocols

The aim of this paper is to show a method that is able to detect inconsistencies in the reasoning carried out by a deliberative agent. The agent is supposed to be provided with a hybrid Knowledge Base expressed in a language called CCR-2, based on production rules and hierarchies of frames, which permits the representation of non-monotonic reasoning, uncertain reasoning and arithmetic constraints in the rules. The method can give a specification of the scenarios in which the agent would deduce an inconsistency. We define a scenario to be a description of the initial agent s state (in the agent life cycle), a deductive tree of rule firings, and a partially ordered set of messages and/or stimuli that the agent must receive from other agents and/or the environment. Moreover, the method will make sure that the scenarios will be valid w.r.t. the communication protocols in which the agent is involved.

Ramirez, Jaime↗

Statistical Considerations of Data Processing in Giovanni Online Tool

The GES DISC Interactive Online Visualization and Analysis Infrastructure (Giovanni) is a web-based interface for the rapid visualization and analysis of gridded data from a number of remote sensing instruments. The GES DISC currently employs several Giovanni instances to analyze various products, such as Ocean-Giovanni for ocean products from SeaWiFS and MODIS-Aqua; TOMS & OM1 Giovanni for atmospheric chemical trace gases from TOMS and OMI, and MOVAS for aerosols from MODIS, etc. (http://giovanni.gsfc.nasa.gov) Foremost among the Giovanni statistical functions is data averaging. Two aspects of this function are addressed here. The first deals with the accuracy of averaging gridded mapped products vs. averaging from the ungridded Level 2 data. Some mapped products contain mean values only; others contain additional statistics, such as number of pixels (NP) for each grid, standard deviation, etc. Since NP varies spatially and temporally, averaging with or without weighting by NP will be different. In this paper, we address differences of various weighting algorithms for some datasets utilized in Giovanni. The second aspect is related to different averaging methods affecting data quality and interpretation for data with non-normal distribution. The present study demonstrates results of different spatial averaging methods using gridded SeaWiFS Level 3 mapped monthly chlorophyll a data. Spatial averages were calculated using three different methods: arithmetic mean (AVG), geometric mean (GEO), and maximum likelihood estimator (MLE). Biogeochemical data, such as chlorophyll a, are usually considered to have a log-normal distribution. The study determined that differences between methods tend to increase with increasing size of a selected coastal area, with no significant differences in most open oceans. The GEO method consistently produces values lower than AVG and MLE. The AVG method produces values larger than MLE in some cases, but smaller in other cases. Further studies indicated that significant differences between AVG and MLE methods occurred in coastal areas where data have large spatial variations and a log-bimodal distribution instead of log-normal distribution.

Suhung, Shen↗