Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallel processing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,027 records · Page 57

Hardware acceleration for HPS algorithms in two and three dimensions

We provide a flexible, open-source framework for hardware acceleration, namely massively-parallel execution on general-purpose graphics processing units (GPUs), applied to the hierarchical Poincaré–Steklov (HPS) family of algorithms for building fast direct solvers for linear elliptic partial differential equations. To take full advantage of the power of hardware acceleration, we propose two variants of HPS algorithms to improve performance on two- and three-dimensional problems. In the two-dimensional setting, we introduce a novel recomputation strategy that minimizes costly data transfers to and from the GPU; in three dimensions, we modify and extend the adaptive discretization technique of Geldermans and Gillman [1] to greatly reduce peak memory usage. We provide an open-source implementation of these methods written in JAX, a high-level accelerated linear algebra package, which allows for the first integration of a high-order fast direct solver with automatic differentiation tools. We conclude with extensive numerical examples showing our methods are fast and accurate on two- and three-dimensional problems.

Fast direct solvers↗

Collisionless cooling of perpendicular electron temperature in the thermal quench of a magnetized plasma

Thermal quench of a nearly collisionless plasma against an isolated cooling boundary or region is an undesirable off-normal event in magnetic fusion experiments, but an ubiquitous process of cosmological importance in astrophysical plasmas. Parallel transport theory of ambipolar-constrained tail electron loss is known to predict rapid cooling of the parallel electron temperature $T_{e\Vert}$ although $T_{e\Vert}$ is difficult to diagnose in actual experiments. Instead direct experimental measurements can readily track the perpendicular electron temperature $T_{e\bot}$ via electron cyclotron emission. The physics underlying the observed fast drop in $T_{e\bot}$ requires a resolution. Here two collisionless mechanisms, dilutional cooling by infalling cold electrons and wave-particle interaction by two families of whistler instabilities, are shown to enable fast $T_{e\bot}$ cooling that closely tracks the mostly collisionless crash of $T_{e\Vert}$. These findings motivate both experimental validation and reexamination of a broad class of plasma cooling problems in laboratory, space, and astrophysical settings.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Arrayed in vivo barcoding for multiplexed sequence verification of plasmid DNA and demultiplexing of pooled libraries

Sequence verification of plasmid DNA is critical for many cloning and molecular biology workflows. To leverage high-throughput sequencing, several methods have been developed that add a unique DNA barcode to individual samples prior to pooling and sequencing. However, these methods require an individual plasmid extraction and/or in vitro barcoding reaction for each sample processed, limiting throughput and adding cost. Here, we develop an arrayed in vivo plasmid barcoding platform that enables pooled plasmid extraction and library preparation for Oxford Nanopore sequencing. This method has a high accuracy and recovery rate, and greatly increases throughput and reduces cost relative to other plasmid barcoding methods or Sanger sequencing. We use in vivo barcoding to sequence verify >45 000 plasmids and show that the method can be used to transform error-containing dispersed plasmid pools into sequence-perfect arrays or well-balanced pools. In vivo barcoding does not require any specialized equipment beyond a low-overhead Oxford Nanopore sequencer, enabling most labs to flexibly process hundreds to thousands of plasmids in parallel.

59 BASIC BIOLOGICAL SCIENCES↗

Sparse Linear Solvers for Large-scale Electromagnetic Transient Simulations

Linear solvers form the basis for electromagnetic transient (EMT) simulations. There is a need to speed up EMT simulations as larger regions are analyzed using EMT simulations. For the same, the performance of linear solvers plays an important role. Exploiting the sparsity of the matrices generated in EMT simulations could assist with speed-up. Scalability is also crucial as power grids expand, demanding solutions capable of accommodating the increasing system size. Recent studies from the North American Electric Reliability Corporation (NERC) increasingly emphasize that EMT simulation models of the power grid will grow larger with the inclusion of power electronics components. Parallelisms in sparsity patterns exploit modern central processing units (CPUs), multi-core CPUs, and graphics processing units (GPUs) architectures in sparse solver designs. Therefore, this paper explores publicly available existing linear solvers and investigates their efficiency in large-scale power grid simulations. A large-scale power grid is developed by increasing the size of the IEEE 39 bus test system to up to 39000 bus systems.

Hsu, Kuan-Chieh↗

Airborne LiDAR to Improve Canopy Fuels Mapping for Wildfire Modeling

Increasing conflict between wildfire and the built environment has increased the need for more up-to-date and finer resolution canopy fuels data to improve wildfire modeling and associated risk forecasts. The US Forest Service and US Department of the Interior’s LANDFIRE product, which provides 30-m resolution canopy fuels data for the entire US, is one of the most widely used sources of fuels data. However, the last complete mapping effort for LANDFIRE is based on 2016 conditions, and subsequent updates reflect disturbances 1-2 years behind the release year. Airborne systems equipped with Light Detection and Ranging (LiDAR) sensors can be deployed to actively sense canopy structure and estimate canopy fuels data (cover, height, base height, bulk density) at finer resolutions. Canopy base height (CBH) and canopy bulk density (CBD) are difficult to measure both in the field and in LiDAR point clouds. Still, they are important for accurately modeling crown fires, which are often intense and difficult to contain. Additionally, point cloud datasets are large, and calculations require efficient utilization of computational resources. To address these challenges, we are working on an approach that uses openly available National Ecological Observatory Network (NEON) airborne LiDAR data, with calculations processed in the R programming language and parallelized through the lidR package. CBH and CBD are often derived from tree height, diameter at breast height, and species-specific allometries using the Fire and Fuels Extension of the Forest Vegetation Simulator (FFE-FVS). We aim to test if airborne LiDAR can estimate CBH and CBD without the use of empirical equations. Reliable estimates of canopy fuels data directly from airborne LiDAR could streamline quick, fine-resolution updates for use in wildfire behavior models.

54 ENVIRONMENTAL SCIENCES↗

Predicting Selective Laser Printing Print Quality of Polymer Powders through Melt Flow Index

Selective Laser Sintering (SLS) uses a precisely controlled laser to fuse polymer powder to build complex 3D shapes. While SLS covers a wide application space, the processing knowledge of polymer powder is limited, restricting the number of commercial powders available. This study serves to expand the processing knowledge of polypropylene and polyethylene in parallel with a “mature” SLS feedstock, nylon, through melt flow index (MFI) characterization. Differential Scanning Calorimetry (DSC) was used to explore the sintering window of the polymers. Polyethylene and polypropylene exhibited a relatively narrow sintering window between 4 – 5 °C, whereas the sintering window of nylon was much wider, between 23 – 24 °C. This suggests that the print quality between polyethylene and polypropylene would be similar; however, X-ray computed tomography revealed a higher volume of print defects, voids, and de lamination in polyethylene than in polypropylene. MFI analysis provided additional insight into the difference in print quality, as the MFI of polyethylene was 7.16 g/10 min, 9.95 g/10 min for polypropylene, and 17.23 g/10 min for nylon. MFI is inversely correlated to melt viscosity, and a low melt viscosity is desired for proper coalescing between the layers and particles. These results suggest that MFI is a promising tool, in conjunction with traditional thermal analysis, for screening candidate powder feedstocks for SLS and optimization of print parameters for novel powders.

36 MATERIALS SCIENCE↗

Initial results from the NASA Lewis Bumpy Torus experiment

Initial results were obtained from low power operation of the NASA Lewis Bumpy Torus experiment, in which a steady-state ion heating method based on the modified Penning discharge is applied in a bumpy torus confinement geometry. The magnet facility consists of 12 superconducting coils, each 19 cm i.d. and capable of 3.0 T, equally spaced in a toroidal array 1.52 m in major diameter. A 18 cm i.d. anode ring is located at each of the 12 midplanes and is maintained at high positive potentials by a dc power supply. Initial observations indicate electron temperatures from 10 to 150 eV, and ion kinetic temperatures from 200 eV to 1200 eV. Two modes of operation were observed, which depend on background pressure, and have different radial density profiles. Steady state neutron production was observed. The ion heating process in the bumpy torus appears to parallel closely the mechanism observed when the modified Penning discharge was operated in a simple magnetic mirror field.

Roth, J. R.↗

The current status of super computers

In this paper, commercially available super computers are surveyed. Computer performance in general is limited by circuit speeds and physical size. Assuming the use of the fastest technology, super computers typically use parallelism in the form of either vector processing or array processing to obtain performance. The Burroughs Scientific Processor is an array computer with 16 separate processors, the Cray-1 and CDC STAR-100 are vector processors, the Goodyear Aerospace STARAN is an array processor with up to 8192 single bit processors, and the Systems Development Corporation PEPE is a collection of up to 288 separate processors.

Knight, J. C.↗

Design of a massively parallel processor

The massively parallel processor (MPP) system is designed to process satellite imagery at high rates. A large number (16,384) of processing elements (PE's) are configured in a square array. For optimum performance on operands of arbitrary length, processing is performed in a bit-serial manner. On 8-bit integer data, addition can occur at 6553 million operations per second (MOPS) and multiplication at 1861 MOPS. On 32-bit floating-point data, addition can occur at 430 MOPS and multiplication at 216 MOPS.

Batcher, K. E.↗

Avenues and incentives for commercial use of a low-gravity environment

The scientific and commercial utilization of the low-g environments for materials research and for process and product development is considered. Any products of commercial interest which necessitate processing in space will probably be low volume, high value items. To encourage the commercialization of materials processing in low-g, NASA, in parallel with establishing and demonstrating the scientific/technological precepts for analyzing and using a low-g environment, is establishing the legal and management mechanisms to share in the cost and risk of early commercial ventures, and is now working with commercial firms on a case-by basis to explore applications of this new technology to specific needs of the company.

Brown, R. L.↗

Bulk processing techniques for very large areas - Landsat classification of California

In 1977, California Law AB452 was passed to provide a mandate for the California Department of Forestry (CDF) to design and implement an information system to assess the forest land base for multiple uses and values. In connection with this mandate, a land-cover map of the entire state, emphasizing forest types, was produced. In producing this map, the latest techniques in digital image mosaicking were combined with the highspeed processing capability available on the ILLIAC IV parallel processor and other computer systems at the Ames Research Center (ARC). An operational and very responsive analysis method was developed at ARC that permitted on-time response to weekly workshops conducted with CDF field personnel to identify all 1,200 spectral classes and to produce final products. Over 100,000,000 acres were classified in the period between December 1, 1978, and April 15, 1979. All analyses were conducted using existing software.

Newland, W.↗

U.S. welding technology - Constraints to space implementation

U.S. and European efforts to develop welding techniques for satellite solar arrays are described. Soldering practices in the U.S. have benefitted from Mo solders, which are well fitted to Si solar cell material thermal characteristics. Analyses have indicated that welds are the preferred method for interconnections and bonds. Extensive work has been done with the parallel gap resistance method (RW), which involves process heat generated by passing a current through resistive layers. Confining the primary heat input to the interconnector/cell contact interface results in welds being formed beneath both electrodes. Pulse welding has become the dominant RW technique in Europe, while ultrasonic welding is used in the U.S.; silver is employed as the interconnect material on both continents. The bonding techniques have been developed empirically instead of theoretically. An IR inspection technique has been produced for monitoring the weld temperature.

Stella, P. M.↗

Nuclear Science Symposium, 30th, and Symposium on Nuclear Power Systems, 15th, San Francisco, CA, October 19-21, 1983, Proceedings

The range of disciplines covered includes physics instrumentation, data acquisition, FASTBUS, radiation detectors, scintillators, photomultipliers and optical detectors, subnanosecond 2-D imaging, medical and health instrumentation, reactor instrumentation, and space instrumentation. Cerenkov detectors, the prevention and rate capability of breakdown processes in wire chambers, and an interactive parallel processor for data analysis are studied. Attention is also given to bismuth germanate's role in gamma-ray spectroscopy, and the construction of a broadband universal sampling head is described. Gamma-ray imaging with a rotating hexagonal uniform redundant array, and broadband X-ray astronomical spectrometry are considered.

Nakamura, M.↗

A numerical study of the vertical transport of momentum in a tropical rainband

The vertical transport of horizontal momentum in a convective tropical rainband is studied using a two-dimensional cloud ensemble model. Twelve simulations are made under the same large-scale conditions. The vertical transports of v momentum (parallel to the rainband) are essentially the same in all of the simulations, even though the structure of the clouds is different in each of the runs. The magnitude of the v-momentum transport by clouds is fairly large. It takes only half of a day to smooth out the tropical low-level easterly jet parallel to the rainband if no other processes are operating. The vertical transports of u momentum (perpendicular to the rainband) are quite different in all of the simulations. This difference can be explained by the dissimilarities in the distributions of horizontal momentum associated with various cloud configurations. The simulated vertical transports of horizontal momentum are compared with those computed with the Schneider and Lindzen scheme. The results suggest that their scheme is basically correct and usable if some improvements are made.

Soong, S.-T.↗

National full-scale aerodynamic complex integrated systems test data system

The data acquisition system of the 80 by 120 foot wind tunnel of the National Full-Scale Aerodynamic Facility (NFAC) is described. How the various satellite data stations are connected to the data acquisition system is shown. As an illustrative example, a strain gage signal is traced from one of the satellite data locations to its final destination in the data system where the signal is processed, observed in real time on various parallel graphic displays, and stored on magnetic disks for postrun data reduction.

Jung, Oscar↗

Applications of a transonic wing design method

A method for designing wings and airfoils at transonic speeds using a predictor/corrector approach was developed. The procedure iterates between an aerodynamic code, which predicts the flow about a given geometry, and the design module, which compares the calculated and target pressure distributions and modifies the geometry using an algorithm that relates differences in pressure to a change in surface curvature. The modular nature of the design method makes it relatively simple to couple it to any analysis method. The iterative approach allows the design process and aerodynamic analysis to converge in parallel, significantly reducing the time required to reach a final design. Viscous and static aeroelastic effects can also be accounted for during the design or as a post-design correction. Results from several pilot design codes indicated that the method accurately reproduced pressure distributions as well as the coordinates of a given airfoil or wing by modifying an initial contour. The codes were applied to supercritical as well as conventional airfoils, forward- and aft-swept transport wings, and moderate-to-highly swept fighter wings. The design method was found to be robust and efficient, even for cases having fairly strong shocks.

Campbell, Richard L.↗

Feed-forward volume rendering algorithm for moderately parallel MIMD machines

Algorithms for direct volume rendering on parallel and vector processors are investigated. Volumes are transformed efficiently on parallel processors by dividing the data into slices and beams of voxels. Equal sized sets of slices along one axis are distributed to processors. Parallelism is achieved at two levels. Because each slice can be transformed independently of others, processors transform their assigned slices with no communication, thus providing maximum possible parallelism at the first level. Within each slice, consecutive beams are incrementally transformed using coherency in the transformation computation. Also, coherency across slices can be exploited to further enhance performance. This coherency yields the second level of parallelism through the use of the vector processing or pipelining. Other ongoing efforts include investigations into image reconstruction techniques, load balancing strategies, and improving performance.

Yagel, Roni↗