Search NASASearch

SEARCH · Search NASA

Results for “Dependency analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Benchmarking Variables for Checkpointing in HPC Applications

Checkpoint/Restart (C/R) is a widely used fault tolerance mechanism in converged systems of cloud, edge, and HPC. However, users often rely on their experience to determine which variables to checkpoint, as there is currently no benchmark that can provide a reference. This can result in checkpointing redundant or even incorrect variables. To address this issue, we propose a benchmark suite that includes critical variables for checkpointing, which have been manually identified, and a method for identifying those critical variables, with 20 representative HPC applications. Our method involves analyzing data dependency between variables to identify critical variables analytically. We verify the identified variables' correctness with a widely used C/R library FTI by an ablation study. With our benchmark suite and data dependency analysis, HPC practitioners now have a reference for identifying checkpointing variables and better knowledge of what kind of variables to checkpoint.

Fu, Xiang

Ambipolar Transport in Polycrystalline GeSn Transistors for Complementary Metal-Oxide-Semiconductor Applications

Group-IV alloy GeSn is a promising material for electronic and optoelectronic applications due to its compatibility with both Si substrates and established Si fabrication processes. This study focuses on polycrystalline GeSn (10% Sn), which offers a cost-effective, large-area, and versatile alternative to epitaxial GeSn. We demonstrate ambipolar transport behavior in polycrystalline GeSn thin film transistors, achieving electron and hole field-effect mobilities reaching up to 0.05 cm 2 /Vs and 2.05 cm 2 /Vs, respectively. Through temperature-dependent analysis, we elucidate the underlying mechanism of this phenomenon, which we attribute to quantum tunneling between the Schottky barrier contact and the channel, as well as potential barriers between the grain boundaries of this polycrystalline film, thereby advancing the understanding of polycrystalline GeSn's electrical properties. Furthermore, this work highlights the potential of ambipolar transport as a technique to employ towards the development of GeSn complementary metal-oxide-semiconductor field-effect transistors, promising to simplify and reduce the cost of GeSn manufacturing processes for edge computing and sensing applications.

42 ENGINEERING

Concerning the origin of novae and U Geminorum stars

The dynamical stability of the contact component in a semi-detached binary system is investigated by a linear time dependent analysis. The boundary condition imposed on the photosphere by the Roche lobe is assumed to be one of constant pressure over the whole surface of the lobe, which is treated as being spherically symmetric. The eigenvalues of the linear equation of motion are evaluated for sequences of stellar envelope models under the constraint of this boundary condition. An extensive region in the Hertzprung-Russell diagram is found to be dynamically unstable. This region coincides almost precisely with that predicted by a previous quasi-static study. Time scales for the growth of the instability are also evaluated. Color observations of several novae and U Geminorum stars are discussed in an attempt to provide semi-empirical tests of the proposed theoretical model of the outbursts. In particular the track in the H-R diagram of the U Geminorum star, EM Cygni, is shown to be consistent with an outburst which is purely a temperature change at constant radius, as is demanded by the theory.

Bath, G. T.

Time-dependent difference theory for noise propagation in a two-dimensional duct

A time dependent numerical formulation was derived for sound propagation in a two dimensional straight soft-walled duct in the absence of mean flow. The time dependent governing acoustic-difference equations and boundary conditions were developed along with the maximum stable time increment. Example calculations were presented for sound attenuation in hard and soft wall ducts. The time dependent analysis were found to be superior to the conventional steady numerical analysis because of much shorter solution times and the elimination of matrix storage requirements.

Baumeister, K. J.

Time-dependent difference theory for noise propagation in a two-dimensional duct

A time-dependent numerical formulation is derived for sound propagation in a two-dimensional straight soft-walled duct in the absence of mean flow. The time-dependent governing acoustic-difference equations and boundary conditions are developed along with the maximum stable time increment. Example calculations are presented for sound attenuation in hard- and soft-wall ducts. The time-dependent analysis has been found to be superior to the conventional steady numerical analysis because of much shorter solution times and the elimination of matrix storage requirements.

Baumeister, K. J.

A search for large-scale convection cells in the solar atmosphere

Mount Wilson magnetograph velocity observations are used to search for east-west motions resulting from hypothetical cellular patterns extending over one or two hemispheres in the latitude direction. No such solar patterns were found. Upper limits established by this analysis depend on the cell lifetime and the pattern stability, but in all cases they are no more than about 10 m/s.

Howard, R.

The life cycle of thunderstorm gust fronts as viewed with Doppler radar and rawinsonde data

This paper presents the time-dependent analysis of the thunderstorm gust front with the use of Project NIMROD data. RHI cross sections of reflectivity and Doppler velocity are constructed to determine the entire vertical structure. The life cycle of the gust front is divided into four stages: (1) the formative stage; (2) the early mature stage; (3) the late mature stage; and (4) the dissipation stage. A new finding is a horizontal roll detected in the reflectivity pattern resulting from airflow that is deflected upward by the ground, while carrying some of the smaller precipitation ahead of the main echo core of the squall line. This feature is called a 'precipitation roll'. As determined from rawinsonde data, the cold air behind the gust front accounts for the observed surface pressure rise. Calculations confirm that the collision of two fluids produce a nonhydrostatic pressure at the leading edge of the outflow. The equation governing the propagation speed of a density current accurately predicts the movement of the gust front.

Wakimoto, R. M.

On numerical integration and computer implementation of viscoplastic models

Due to the stringent design requirement for aerospace or nuclear structural components, considerable research interests have been generated on the development of constitutive models for representing the inelastic behavior of metals at elevated temperatures. In particular, a class of unified theories (or viscoplastic constitutive models) have been proposed to simulate material responses such as cyclic plasticity, rate sensitivity, creep deformations, strain hardening or softening, etc. This approach differs from the conventional creep and plasticity theory in that both the creep and plastic deformations are treated as unified time-dependent quantities. Although most of viscoplastic models give better material behavior representation, the associated constitutive differential equations have stiff regimes which present numerical difficulties in time-dependent analysis. In this connection, appropriate solution algorithm must be developed for viscoplastic analysis via finite element method.

Chang, T. Y.

Visual slant misperception and the 'black-hole' landing situation

A theory which explains the tendency for dangerously low approaches during night landing situations is presented. The two-dimensional information at the pilot's eye contains sufficient information for the visual system to extract the angle of slant of the runway relative to the approach path. The analysis depends upon perspective information which is available at a certain distance out from the aimpoint, to either side of the runway edgelights. Under black hole landing conditions, however, this information is not available, and it is proposed that the visual system use instead the only available information, the perspective gradient of the runway edgelights. An equation is developed which predicts the perceived approach angle when this incorrect parameter is used. The predictions are in close agreement with existing experimental data.

Perrone, J. A.

Transonic cascade flow calculations using non-periodic C-type grids

A new kind of C-type grid is proposed for turbomachinery flow calculations. This grid is nonperiodic on the wake and results in minimum skewness for cascades with high turning and large camber. Euler and Reynolds averaged Navier-Stokes equations are discretized on this type of grid using a finite volume approach. The Baldwin-Lomax eddy-viscosity model is used for turbulence closure. Jameson's explicit Runge-Kutta scheme is adopted for the integration in time, and computational efficiency is achieved through accelerating strategies such as multigriding and residual smoothing. A detailed numerical study was performed for a turbine rotor and for a vane. A grid dependence analysis is presented and the effect of artificial dissipation is also investigated. Comparison of calculations with experiments clearly demonstrates the advantage of the proposed grid.

Arnone, Andrea

Run-time parallelization and scheduling of loops

Run-time methods are studied to automatically parallelize and schedule iterations of a do loop in certain cases where compile-time information is inadequate. The methods presented involve execution time preprocessing of the loop. At compile-time, these methods set up the framework for performing a loop dependency analysis. At run-time, wavefronts of concurrently executable loop iterations are identified. Using this wavefront information, loop iterations are reordered for increased parallelism. Symbolic transformation rules are used to produce: inspector procedures that perform execution time preprocessing, and executors or transformed versions of source code loop structures. These transformed loop structures carry out the calculations planned in the inspector procedures. Performance results are presented from experiments conducted on the Encore Multimax. These results illustrate that run-time reordering of loop indexes can have a significant impact on performance.

Saltz, Joel H.

The thermal and plasma-physical evolution of laminar current sheets formed in the solar atmosphere by emerging flux

A time-dependent analysis of emerging flux is carried out, and the time evolution of both the current sheet energetics and the plasma state is calculated. This evolution is determined in two different regimes. In the first case the width of the current sheet is assumed to be independent of the sheet thermodynamics and is fixed by the initial conditions. In the second, the width of the current sheet is a function of the resistivity and is allowed to decrease to its minimum given by the electron gyroradius. In both cases the resistivity is computed according to the marginal stability hypothesis. In each case the thermodynamic evolution is found to be quite rapid, with the temperature increasing from 10,000 to 1,000,000 K in a second or less. In contrast to previous studies, it is found that the resistivity is not significantly enhanced by the current-driven plasma wave turbulence. It is concluded that a laminar current sheet cannot be responsible for the activity associated with emerging flux.

Larosa, T. N.

Working Notes from the 1992 AAAI Spring Symposium on Practical Approaches to Scheduling and Planning

The symposium presented issues involved in the development of scheduling systems that can deal with resource and time limitations. To qualify, a system must be implemented and tested to some degree on non-trivial problems (ideally, on real-world problems). However, a system need not be fully deployed to qualify. Systems that schedule actions in terms of metric time constraints typically represent and reason about an external numeric clock or calendar and can be contrasted with those systems that represent time purely symbolically. The following topics are discussed: integrating planning and scheduling; integrating symbolic goals and numerical utilities; managing uncertainty; incremental rescheduling; managing limited computation time; anytime scheduling and planning algorithms, systems; dependency analysis and schedule reuse; management of schedule and plan execution; and incorporation of discrete event techniques.

Drummond, Mark

On the application of multifrequency polarimetric radar observations for sea-ice classification

The use of multifrequency polarimetric radar imagery to enhance the ability to separate different sea-ice types using single-frequency, single-polarization synthetic aperture radar (SAR) data is investigated. Backscatter characteristics of six radiometrically and polarimetrically distinct sea-ice types are selected in an unsupervised range-dependent analysis of multifrequency polarimetric SAR data using the maximum a posteriori (MAP) polarimetric classifier. Maximum ice discrimination is achieved with combined C- and L-band full polarimetry, and collocated passive microwave imagery suggests greater than 90 percent classification accuracy. C-band VV-pol alone achieves only 68 percent relative accuracy because it confuses multiyear and rough compressed first year ice. L-band, relative classification accuracy is 75 percent, 83 percent, and 85 percent, using HH-pol, HH- and VV-combined, or the full polarimetry, respectively. P-band is less accurate. Combinations of two frequencies at a single polarization show the greatest improvement over a single channel.

Rignot, Eric

Efficient Parallel Kernel Solvers for Computational Fluid Dynamics Applications

Distributed-memory parallel computers dominate today's parallel computing arena. These machines, such as Intel Paragon, IBM SP2, and Cray Origin2OO, have successfully delivered high performance computing power for solving some of the so-called "grand-challenge" problems. Despite initial success, parallel machines have not been widely accepted in production engineering environments due to the complexity of parallel programming. On a parallel computing system, a task has to be partitioned and distributed appropriately among processors to reduce communication cost and to attain load balance. More importantly, even with careful partitioning and mapping, the performance of an algorithm may still be unsatisfactory, since conventional sequential algorithms may be serial in nature and may not be implemented efficiently on parallel machines. In many cases, new algorithms have to be introduced to increase parallel performance. In order to achieve optimal performance, in addition to partitioning and mapping, a careful performance study should be conducted for a given application to find a good algorithm-machine combination. This process, however, is usually painful and elusive. The goal of this project is to design and develop efficient parallel algorithms for highly accurate Computational Fluid Dynamics (CFD) simulations and other engineering applications. The work plan is 1) developing highly accurate parallel numerical algorithms, 2) conduct preliminary testing to verify the effectiveness and potential of these algorithms, 3) incorporate newly developed algorithms into actual simulation packages. The work plan has well achieved. Two highly accurate, efficient Poisson solvers have been developed and tested based on two different approaches: (1) Adopting a mathematical geometry which has a better capacity to describe the fluid, (2) Using compact scheme to gain high order accuracy in numerical discretization. The previously developed Parallel Diagonal Dominant (PDD) algorithm and Reduced Parallel Diagonal Dominant (RPDD) algorithm have been carefully studied on different parallel platforms for different applications, and a NASA simulation code developed by Man M. Rai and his colleagues has been parallelized and implemented based on data dependency analysis. These achievements are addressed in detail in the paper.

Sun, Xian-He

By Hand or Not By-Hand: A Case Study of Alternative Approaches to Parallelize CFD Applications

While parallel processing promises to speed up applications by several orders of magnitude, the performance achieved still depends upon several factors, including the multiprocessor architecture, system software, data distribution and alignment, as well as the methods used for partitioning the application and mapping its components onto the architecture. The existence of the Gorden Bell Prize given out at Supercomputing every year suggests that while good performance can be attained for real applications on general purpose multiprocessors, the large investment in man-power and time still has to be repeated for each application-machine combination. As applications and machine architectures become more complex, the cost and time-delays for obtaining performance by hand will become prohibitive. Computer users today can turn to three possible avenues for help: parallel libraries, parallel languages and compilers, interactive parallelization tools. The success of these methodologies, in turn, depends on proper application of data dependency analysis, program structure recognition and transformation, performance prediction as well as exploitation of user supplied knowledge. NASA has been developing multidisciplinary applications on highly parallel architectures under the High Performance Computing and Communications Program. Over the past six years, the transition of underlying hardware and system software have forced the scientists to spend a large effort to migrate and recede their applications. Various attempts to exploit software tools to automate the parallelization process have not produced favorable results. In this paper, we report our most recent experience with CAPTOOL, a package developed at Greenwich University. We have chosen CAPTOOL for three reasons: 1. CAPTOOL accepts a FORTRAN 77 program as input. This suggests its potential applicability to a large collection of legacy codes currently in use. 2. CAPTOOL employs domain decomposition to obtain parallelism. Although the fact that not all kinds of parallelism are handled may seem unappealing, many NASA applications in computational aerosciences as well as earth and space sciences are amenable to domain decomposition. 3. CAPTOOL generates code for a large variety of environments employed across NASA centers: MPI/PVM on network of workstations to the IBS/SP2 and CRAY/T3D.

Yan, Jerry C.

The Design and Evaluation of "CAPTools"--A Computer Aided Parallelization Toolkit

Writing applications for high performance computers is a challenging task. Although writing code by hand still offers the best performance, it is extremely costly and often not very portable. The Computer Aided Parallelization Tools (CAPTools) are a toolkit designed to help automate the mapping of sequential FORTRAN scientific applications onto multiprocessors. CAPTools consists of the following major components: an inter-procedural dependence analysis module that incorporates user knowledge; a 'self-propagating' data partitioning module driven via user guidance; an execution control mask generation and optimization module for the user to fine tune parallel processing of individual partitions; a program transformation/restructuring facility for source code clean up and optimization; a set of browsers through which the user interacts with CAPTools at each stage of the parallelization process; and a code generator supporting multiple programming paradigms on various multiprocessors. Besides describing the rationale behind the architecture of CAPTools, the parallelization process is illustrated via case studies involving structured and unstructured meshes. The programming process and the performance of the generated parallel programs are compared against other programming alternatives based on the NAS Parallel Benchmarks, ARC3D and other scientific applications. Based on these results, a discussion on the feasibility of constructing architectural independent parallel applications is presented.

Yan, Jerry