Search NASA⌕ Search

SEARCH · Search NASA

Results for “Portability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

IRIS-MEMFLOW: Data Flow-Enabled Portable Memory Orchestration in IRIS Runtime for Diverse Heterogeneity

Task-based programming models and execution paradigms provide a means to decompose a computation by expressing it as a graph in which each node represents a specific computation operating on memory objects and the edges define the dependencies in the execution flow. In this execution model, independent nodes in the graph can be executed concurrently in different computing devices, making it suitable for heterogeneous systems in which computing devices with different architectures coexist. However, careful memory orchestration across heterogeneous devices is needed because copies of the same memory object may reside in multiple devices during execution. Manually ensuring such an orchestration is quite challenging. Not only must an application developer guard against race conditions, but they must also optimize data movement between the host and devices because unnecessary data movement significantly impacts performance. To mitigate these challenges, we enhance the IRIS heterogeneous runtime and introduce IRIS-MEMFLOW–a data flow–enabled portable memory abstraction for seamlessly orchestrating memory in diverse heterogeneous computing environments. By using data-flow analysis, IRIS-MEMFLOW guards against race conditions while multiple heterogeneous devices access memory objects. IRIS-MEMFLOW also optimizes data movement between the host and devices without manual intervention. As a result, IRIS provides improved programming productivity, performance, and portability for multidevice heterogeneous executions in high-performance computing and cloud systems that run diverse architectures from different vendors. The efficacy of IRIS-MEMFLOW is evaluated through experiments that show its capability in terms of programming productivity, multidevice heterogeneity, portability, and low overhead versus the state of the art.

Monil, M. A. H. [ORNL] (ORCID:0000000334194037)↗

Toward Automated Detection of Portability Bugs in Kokkos Parallel Programs

Performance-portable programming frameworks provide abstractions for parallel execution to allow easily porting an application to multiple backend programming models, such as CUDA, HIP, and OpenMP. However, programs may still have portability bugs that manifest only on specific backends. Traditional testing is ineffective in discovering these bugs, as it would require concrete execution on all supported hardware configurations for a potentially infinite set of inputs. To mitigate this issue, we focused on a specific programming framework, Kokkos, and identified several categories of common portability bugs. We then developed Klokkos, a static analysis approach based on symbolic execution that can run on commodity hardware, before execution on supercomputers. As a proof-of-concept, we ran Klokkos on examples encoding the identified bugs. Our results show that Klokkos is effective, efficient, and precise: it detected all the considered bugs, quickly, and without any false positives. Although preliminary, our results motivate further research and development in this direction.

Kale, Vivek↗

Integrating ORNL’s HPC and Neutron Facilities with a Performance-Portable CPU/GPU Ecosystem

We explore the development of a performance-portable CPU/GPU ecosystem to integrate two of the US Department of Energy’s (DOE’s) largest scientific instruments, the Oak Ridge Leadership Computing facility and the Spallation Neutron Source (SNS), both of which are housed at Oak Ridge National Laboratory. We select a relevant data reduction workflow use-case to obtain the differential scattering cross-section from data collected by SNS’s CORELLI and TOPAZ instruments. We compare the current CPU-only production implementation using the Garnet Python multiprocess package based on the Mantid C++ framework against our proposed CPU/GPU implementation that uses the LLVM-based, just-in-time Julia scientific language and the JACC.jl performance-portable package. Two proxy apps were developed: (i) an app for extracting relevant Mantid kernels (MDNorm) in C++ and (ii) the Julia MiniVATES.jl miniapp. We present performance results for NVIDIA A100 and AMD MI100 GPUs and AMD EPYC 7513 and 7662 CPUs. The results provide insights for future generations of data reduction software that can embrace performance portability for an integrated research infrastructure across DOE’s experimental and computational facilities.

Hahn, Steven↗

A Study of Performance Portability of Low-bit Fused Matrix-Vector Multiplication Kernels in SYCL

Understanding the causes of performance gaps between a portable programming model and a vendor-specific programming model is important for improving performance portability. This paper studies performance portability of low-bit fused general matrix-vector multiplication kernels in SYCL on vendors’ graphics processing units (GPUs). This work introduces the use case, explains the kernel implementations in detail, evaluates the performance of the CUDA, HIP, and SYCL kernels on datacenter, desktop, and laptop GPUs, and investigates the causes of performance gaps. The results show that loop unrolling, kernel dispatch overhead, and sum reduction contribute to the gaps.

Jin, Zheming [ORNL] (ORCID:000000027197780X)↗

Portable fiber optic sensor for rare earth elements and other critical metals using photoluminescence methods

Rare earth elements and other metals are vital to a range of technologies that are used in the energy and defense sectors. However, monopolistic market conditions have caused significant concern over the stability of the critical metal supply chain, and this has spurred extensive efforts in many nations to produce these metals domestically, both from conventional sources such as mining and well as from unconventional sources such as coal and its utilization byproducts. Slow and expensive characterization methods pose a significant barrier for both metals prospecting and process monitoring. A promising solution to this challenge is the development of highly sensitive luminescent sensors for metals, which can offer low costs, portability, and sensitivity. Anionic zinc adeninate metal-organic frameworks (BioMOFs) are known to distinguishing and detect part-per-billion levels of terbium, europium, samarium, and dysprosium in water by sensitizing the narrow, element-specific emission bands from these lanthanides. Here, a BioMOF material is immobilized onto a large diameter, solarization-resistant fiber optic tip integrated with a portable, low-cost spectrometer for rare earth element sensing. Immobilizing the sensing material on fiber instead of dispersing the sensing material in solution offers several advantages: it facilitates solvent removal, which enhances luminescent signal from the sensitized lanthanides, and it also allows the BioMOF to be recycled for multiple uses. The sensing system was deployed on a simulated process stream and exhibited qualitative agreement with inductively-coupled plasma mass spectrometry for terbium and europium detection, highlighting the potential for the sensing system to be deployed for real-world applications. By using different sensing materials, the same portable sensor may be deployed to detect other energy relevant metals such as cobalt, providing a cost-effective and sensitive platform for critical metal characterization.

Crawford, Scott↗

Oak Ridge National Laboratory Modernizing the Kokkos Build System: Using CMake to Encapsulate the Complexity of Build Instructions for Performance Portable Libraries

Kokkos, a C++ library focused on performance portability, requires a build system that can work with a variety of compilers and hardware. Ideally, users need only select the compiler and architecture and should not have to know or specify how programs using Kokkos are built. CMake can be used to create a flexible, robust build system and automatically configures compilers and settings based on the user’s inputs. Nevertheless, Kokkos’ requirements as a performance portability library for the build system exceed CMake’s current capabilities. This report describes the requirements, solutions, and testing of various implementations to create a CMake-based build system suitable for Kokkos. It compares the strengths and shortcomings of the approaches and evaluates the implementations with respect to the requirements. Because no solution was found to meet all of the requirements, the Kokkos team engaged with the CMake development team to discuss and plan a path toward support for performance-portable build systems in CMake in the future.

97 MATHEMATICS AND COMPUTING↗

Low-Cost and Portable Biosensor Based on Monitoring Impedance Changes in Aptamer-Functionalized Nanoporous Anodized Aluminum Oxide Membrane

We report a low-cost, portable biosensor composed of an aptamer-functionalized nanoporous anodic aluminum oxide (NAAO) membrane and a commercial microcontroller chip-based impedance reader suitable for electrochemical impedance spectroscopy (EIS)-based sensing. The biosensor consists of two chambers separated by an aptamer-functionalized NAAO membrane, and the impedance reader is utilized to monitor transmembrane impedance changes. The biosensor is utilized to detect amodiaquine molecules using an amodiaquine-binding aptamer (OR7)-functionalized membrane. The aptamer-functionalized membrane is exposed to different concentrations of amodiaquine molecules to characterize the sensitivity of the sensor response. The specificity of the sensor response is characterized by exposure to varying concentrations of chloroquine, which is similar in structure to amodiaquine but does not bind to the OR7 aptamer. A commercial potentiostat is also used to measure the sensor response for amodiaquine and chloroquine. The sensing response measured using both the portable impedance reader and the commercial potentiostat showed a similar dynamic response and detection threshold. The specific and sensitive sensing results for amodiaquine demonstrate the efficacy of the low-cost and portable biosensor.

60 APPLIED LIFE SCIENCES↗

TChem-atm (v2.0.0): scalable performance-portable multiphase atmospheric chemistry

We present TChem-atm, a performance-portable approach that enables efficient simulation of chemically detailed and multiphase atmospheric chemistry on modern heterogeneous computing architectures. Unlike previous efforts that rely on architecture-specific code or focus exclusively on gas-phase chemistry, TChem-atm supports fully coupled gas–aerosol systems with execution across CPUs, NVIDIA GPUs, and AMD GPUs through the Kokkos programming model. It integrates the flexible multiphase capabilities of the Community Atmospheric Model Chemistry Package (CAMP) with the high-performance kinetic routines of TChem, and includes automatic Jacobian construction with support for a range of stiff ODE solvers. In a proof-of-concept integration with the particle-resolved model PartMC, TChem-atm reproduces the existing PartMC–CAMP implementation within solver tolerances and delivers substantial GPU speedups, especially for large particle populations. Performance benchmarks reveal substantial speedups on GPU platforms, particularly for large particle populations, with consistent results across hardware backends. TChem-atm enables performance-portable execution across CPUs and GPUs, though optimal efficiency may require modest architecture-specific tuning (e.g., team and vector sizes), with up to a twofold improvement on the NVIDIA H100. It directly supports sectional and particle-resolved host models, while modal aerosol schemes require minor adaptation to provide particle-scale quantities such as representative diameters. By enabling chemically detailed, multiphase simulations with performance portability and host-model flexibility, TChem-atm facilitates the incorporation of advanced chemistry into atmospheric models.

Díaz-Ibarra, Oscar Homero [Sandia National Laborat↗

Technology assessment of portable energy RDT and P, phase 1

A technological assessment of portable energy research, development, technology, and production was undertaken to assess the technical, economic, environmental, and sociopolitical issues associated with portable energy options. Those courses of action are discussed which would impact aviation and air transportation research and technology. Technology assessment workshops were held to develop problem statements. The eighteen portable energy problem statements are discussed in detail along with each program's objective, approach, task description, and estimates of time and costs.

Spraul, J. R.↗

A portable telescope, photometer, and data-recording system

A description is presented of a portable telescope, two-channel photometer and high-speed data-recording system that meet the unique requirements of occultation astronomy. The instrumentation has been designed for portability, high time resolution, accurate absolute timing, and ease of use. The instrumentation includes a light-weight 35.5-cm (14-inch) objective portable telescope, a two-channel high-speed photometer and a data-recording system built into a suitcase. The equipment has been successfully used in field operations and at fixed observatories in India, Australia, and Hawaii.

Baron, R. L.↗

Development of a portable precision landing system

A portable, tactical approach guidance (PTAG) system, based on a novel, X-band, precision approach concept, was developed and flight tested as a part of NASA's Rotorcraft All-Weather Operations Research Program. The system is based on state-of-the-art X-band technology and digital processing techniques. The PTAG airborne hardware consists of an X-band receiver and a small microprocessor installed in conjunction with the aircraft instrument landing system (ILS) receiver. The microprocessor analyzes the X-band, PTAG pulses and outputs ILS compatible localizer and glide slope signals. The ground stations are inexpensive, portable units, each weighing less than 85 lb, including battery, that can be quickly deployed at a landing site. Results from the flight test program show that PTAG has a significant potential for providing tactical aircraft with low cost, portable, precision instrument approach capability.

Davis, T. J.↗

ProperCAD: A portable object-oriented parallel environment for VLSI CAD

Most parallel algorithms for VLSI CAD proposed to date have one important drawback: they work efficiently only on machines that they were designed for. As a result, algorithms designed to date are dependent on the architecture for which they are developed and do not port easily to other parallel architectures. A new project under way to address this problem is described. A Portable object-oriented parallel environment for CAD algorithms (ProperCAD) is being developed. The objectives of this research are (1) to develop new parallel algorithms that run in a portable object-oriented environment (CAD algorithms using a general purpose platform for portable parallel programming called CARM is being developed and a C++ environment that is truly object-oriented and specialized for CAD applications is also being developed); and (2) to design the parallel algorithms around a good sequential algorithm with a well-defined parallel-sequential interface (permitting the parallel algorithm to benefit from future developments in sequential algorithms). One CAD application that has been implemented as part of the ProperCAD project, flat VLSI circuit extraction, is described. The algorithm, its implementation, and its performance on a range of parallel machines are discussed in detail. It currently runs on an Encore Multimax, a Sequent Symmetry, Intel iPSC/2 and i860 hypercubes, a NCUBE 2 hypercube, and a network of Sun Sparc workstations. Performance data for other applications that were developed are provided: namely test pattern generation for sequential circuits, parallel logic synthesis, and standard cell placement.

Ramkumar, Balkrishna↗

Portable computer system architecture for the Space Station Freedom program

This paper outlines various mission requirements and technical approaches that support the potential use of portable computers in several defined activities within the Space Station Freedom (SSF) program. Specifically, the use of portable computers as consoles for both spacecraft control and payload applications is presented. Various issues and proposed solutions regarding the incorporation of portable computers within the program are presented. The primary issues presented regard architecture (standard interface for expansion, advanced processors and displays), integration (methods of high-speed data communication, peripheral interfaces, and interconnectivity within various support networks), and evolution (wireless communications and multimedia data interface methods).

Alena, Richard↗

Portability and Cross-Platform Performance of an MPI-Based Parallel Polygon Renderer

Visualizing the results of computations performed on large-scale parallel computers is a challenging problem, due to the size of the datasets involved. One approach is to perform the visualization and graphics operations in place, exploiting the available parallelism to obtain the necessary rendering performance. Over the past several years, we have been developing algorithms and software to support visualization applications on NASA's parallel supercomputers. Our results have been incorporated into a parallel polygon rendering system called PGL. PGL was initially developed on tightly-coupled distributed-memory message-passing systems, including Intel's iPSC/860 and Paragon, and IBM's SP2. Over the past year, we have ported it to a variety of additional platforms, including the HP Exemplar, SGI Origin2OOO, Cray T3E, and clusters of Sun workstations. In implementing PGL, we have had two primary goals: cross-platform portability and high performance. Portability is important because (1) our manpower resources are limited, making it difficult to develop and maintain multiple versions of the code, and (2) NASA's complement of parallel computing platforms is diverse and subject to frequent change. Performance is important in delivering adequate rendering rates for complex scenes and ensuring that parallel computing resources are used effectively. Unfortunately, these two goals are often at odds. In this paper we report on our experiences with portability and performance of the PGL polygon renderer across a range of parallel computing platforms.

Crockett, Thomas W.↗

Portable Parallel Programming for the Dynamic Load Balancing of Unstructured Grid Applications

The ability to dynamically adapt an unstructured -rid (or mesh) is a powerful tool for solving computational problems with evolving physical features; however, an efficient parallel implementation is rather difficult, particularly from the view point of portability on various multiprocessor platforms We address this problem by developing PLUM, tin automatic anti architecture-independent framework for adaptive numerical computations in a message-passing environment. Portability is demonstrated by comparing performance on an SP2, an Origin2000, and a T3E, without any code modifications. We also present a general-purpose load balancer that utilizes symmetric broadcast networks (SBN) as the underlying communication pattern, with a goal to providing a global view of system loads across processors. Experiments on, an SP2 and an Origin2000 demonstrate the portability of our approach which achieves superb load balance at the cost of minimal extra overhead.

Biswas, Rupak↗

MPIRUN: A Portable Loader for Multidisciplinary and Multi-Zonal Applications

Multidisciplinary and multi-zonal applications are an important class of applications in the area of Computational Aerosciences. In these codes, two or more distinct parallel programs or copies of a single program are utilized to model a single problem. To support such applications, it is common to use a programming model where a program is divided into several single program multiple data stream (SPMD) applications, each of which solves the equations for a single physical discipline or grid zone. These SPMD applications are then bound together to form a single multidisciplinary or multi-zonal program in which the constituent parts communicate via point-to-point message passing routines. One method for implementing the message passing portion of these codes is with the new Message Passing Interface (MPI) standard. Unfortunately, this standard only specifies the message passing portion of an application, but does not specify any portable mechanisms for loading an application. MPIRUN was developed to provide a portable means for loading MPI programs, and was specifically targeted at multidisciplinary and multi-zonal applications. Programs using MPIRUN for loading and MPI for message passing are then portable between all machines supported by MPIRUN. MPIRUN is currently implemented for the Intel iPSC/860, TMC CM5, IBM SP-1 and SP-2, Intel Paragon, and workstation clusters. Further, MPIRUN is designed to be simple enough to port easily to any system supporting MPI.

Fineberg, Samuel A.↗

Portable Welder

A low cost, low power, self-contained portable welding gun designed for joining thermoplastics which become soft when heated and harden when cooled was developed originally by NASA's Langley Research Center for repairing helicopter windshields. Welder has a broad range of applications for joining both thermoplastic materials in the aerospace, automotive, appliance, and construction industries. Welders portability and low power requirement allow its use on-site in any type of climate, with power supplied by a variety of portable sources.

Source record↗

Evaluation of portable air samplers for monitoring airborne culturable bacteria

Airborne culturable bacteria were monitored at five locations (three in an office/laboratory building and two in a private residence) in a series of experiments designed to compare the efficiency of four air samplers: the Andersen two-stage, Burkard portable, RCS Plus, and SAS Super 90 samplers. A total of 280 samples was collected. The four samplers were operated simultaneously, each sampling 100 L of air with collection on trypticase soy agar. The data were corrected by applying positive hole conversion factors for the Burkard portable, Andersen two-stage, and SAS Super 90 air samplers, and were expressed as log10 values prior to statistical analysis by analysis of variance. The Burkard portable air sampler retrieved the highest number of airborne culturable bacteria at four of the five sampling sites, followed by the SAS Super 90 and the Andersen two-stage impactor. The number of bacteria retrieved by the RCS Plus was significantly less than those retrieved by the other samplers. Among the predominant bacterial genera retrieved by all samplers were Staphylococcus, Bacillus, Corynebacterium, Micrococcus, and Streptococcus.

NASA Center JSC↗