Search NASA⌕ Search

SEARCH · Search NASA

Results for “Hardware acceleration”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

International Space Station Microgravity Analytical Model Correlation And Update

The acceleration environment aboard the completed International Space Station (ISS) is a key resource for scientific and technological endeavors. Hardware verification activIties and early measurements indicate that the ISS is well on the way of meeting these "Assembly Complete" "microgravity" provisions, however, the simulation models that compute these accelerations have, to date, lacked the high degree of empirical validation typical of standard aerospace industry practices. Assembly stage, on-orbit measurements are used to address this shortcoming and to develop higher confidence in the simulation models. The Phase I correlation results show the analyses to be consistently conservative, producing higher than measured levels. The 25 to 30% greater quasi-steady computations are deemed acceptable for verification. Updates are made to localized structural dynamic and vibroacoustic parameters that reduce responses in selected one-third octave bands by almost 50%. These models are then used for the Assembly Complete verification analysis which concludes that the ISS vehicle meets the ISS microgravity requirements with minor reservations. Two of the sixteen rack are marginally non-compliant in the quasi-steady regime, and operational constraints are needed on the U. S. Lab and ESA APM vacuum resource vents, and the Russian Resistive Exercise Device in the structural dynamic regime.

DelBasso, Steve↗

Integrating quantum computing resources into scientific HPC ecosystems

Quantum Computing (QC) offers significant potential to enhance scientific discovery in fields such as quantum chemistry, optimization, and artificial intelligence. Yet QC faces challenges due to the noisy intermediate-scale quantum era’s inherent external noise issues. Here, this paper discusses the integration of QC as a computational accelerator within classical scientific high-performance computing (HPC) systems. By leveraging a broad spectrum of simulators and hardware technologies, we propose a hardware-agnostic framework for augmenting classical HPC with QC capabilities. Drawing on the HPC expertise of the Oak Ridge National Laboratory (ORNL) and the HPC lifecycle management of the Department of Energy (DOE), our approach focuses on the strategic incorporation of QC capabilities and acceleration into existing scientific HPC workflows. This includes detailed analyses, benchmarks, and code optimization driven by the needs of the DOE and ORNL missions. Our comprehensive framework integrates hardware, software, workflows, and user interfaces to foster a synergistic environment for quantum and classical computing research. This paper outlines plans to unlock new computational possibilities, driving forward scientific inquiry and innovation in a wide array of research domains.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Space Acceleration Measurement System-II: Microgravity Instrumentation for the International Space Station Research Community

The International Space Station opens for business in the year 2000, and with the opening, science investigations will take advantage of the unique conditions it provides as an on-orbit laboratory for research. With initiation of scientific studies comes a need to understand the environment present during research. The Space Acceleration Measurement System-II provides researchers a consistent means to understand the vibratory conditions present during experimentation on the International Space Station. The Space Acceleration Measurement System-II, or SAMS-II, detects vibrations present while the space station is operating. SAMS-II on-orbit hardware is comprised of two basic building block elements: a centralized control unit and multiple Remote Triaxial Sensors deployed to measure the acceleration environment at the point of scientific research, generally within a research rack. Ground Operations Equipment is deployed to complete the command, control and data telemetry elements of the SAMS-II implementation. Initially, operations consist of user requirements development, measurement sensor deployment and use, and data recovery on the ground. Future system enhancements will provide additional user functionality and support more simultaneous users.

Sutliff, Thomas J.↗

Motion and force controlled vibration testing

A technique for controlling both the input acceleration and force in vibration tests is proposed to alleviate the overtesting risks and the problems associated with response limiting in conventional vibration tests of aerospace hardware. Previous research on impedance and force controlled vibration tests is reviewed and a simple equation governing the dual control of acceleration and force is derived. A practical method for implementing the dual control technique in random vibration tests has been demonstrated in JPL's environmental test facility using a conventional digital controller operating in the extremal mode. The dual control technique provides appropriate real-time notching of the input acceleration and a corresponding reduction of the test item response at resonances. Issues concerning the need for force and acceleration phase information, the adequacy of specifying the blocked force, and the derivation of the total force for multipoint supports are discussed.

Scharton, Terry D.↗

Accelerating high-order continuum kinetic plasma simulations using multiple GPUs

Kinetic plasma simulations solve the Vlasov-Poisson or Vlasov-Maxwell equations to evolve scalar-variable distribution functions in position-velocity phase space and vector-variable electromagnetic fields in configuration space. The immense computational cost of evolving high-dimensional variables, and their large number of degrees of freedom, often limits the utility of continuum kinetic simulations and presents a challenge when it comes to accurately simulating real-world physical phenomena. To address this challenge, we present techniques that accelerate and minimize the computational work required for a scalable Vlasov-Poisson solver. We show theoretical hardware compute and communication bounds for solving a fourth-order finite-volume Vlasov-Poisson system. These bounds are then used to inform and evaluate the design of performance portable algorithms for a multiple graphics processing unit (GPU) accelerated version of the Vlasov-Poisson solver VCK-CPU [1]. We demonstrate that the multi-GPU Vlasov solver implementation, VCK-GPU, simultaneously minimizes required inter-process data transfer while also being bounded by the machine network performance limits. This results in an overall strong scaling speedup per timestep of up to 40x in three-dimensional phase space (one position, two velocity coordinates) and 54x in four dimensional phase space (two position, two velocity coordinates) and a 341x increase in simulation throughput of the GPU accelerated code over the existing CPU code. The GPU code is also able to weak scale up to 256 compute nodes and 1024 GPUs. In conclusion, we demonstrate that the improved compute performance enables exploring configurations which were previously computationally infeasible, including resolving fine-scale distribution function filamentation and multi-species dynamics with realistic electron-proton mass ratios.

Continuum kinetics↗

Acceleration of the particle-in-cell code Osiris with graphics processing units

Fully relativistic particle-in-cell (PIC) simulations are crucial for advancing our knowledge of plasma physics. Modern supercomputers based on graphics processing units (GPUs) offer the potential to perform PIC simulations of unprecedented scale, but require robust and feature-rich codes that can fully leverage their computational resources. In this work, this demand is addressed by adding GPU acceleration to the PIC code Osiris. An overview of the algorithm, which features a CUDA extension to the underlying Fortran architecture, is given. Detailed performance benchmarks for thermal plasmas are presented, which demonstrate excellent weak scaling on NERSC's Perlmutter supercomputer and high levels of absolute performance. The robustness of the code to model a variety of physical systems is demonstrated via simulations of Weibel filamentation and laser-wakefield acceleration run with dynamic load balancing. Finally, measurements and analysis of energy consumption are provided that indicate that the GPU algorithm is up to ~14 times faster and ~7 times more energy efficient than the optimized CPU algorithm on a node-to-node basis. The described development addresses the PIC simulation community's computational demands both by contributing a robust and performant GPU-accelerated PIC code and by providing insight into efficient use of GPU hardware.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Quantum Circuit Partitioning for Scalable Noise-Aware Quantum Circuit Re-Synthesis

Re-synthesis techniques are utilized to optimize the quantum circuit. To enable scalable re-synthesis a divide-and-conquer approach is adopted that partitions the circuit into smaller blocks, which are optimized independently. Several algorithms have been proposed to minimize the block number while maximizing the gate count of each block. However, they vary in their performance and may not yield the highest output fidelity. We propose a reinforcement learning-based quantum circuit partitioning framework that incorporates the physical properties of the quantum hardware to maximize the output fidelity post-quantum circuit optimization. To accelerate the training, we also propose a noise injection method that enables on-the-fly optimization in the reinforcement learning environment, independent of the adopted optimization/re-synthesis method at the block level. We evaluate our approach compared to different partitioning techniques using various quantum benchmarks executed on IBM Q Hanoi quantum computer.

Charrwi, Mohammad Walid↗

Dry film lubricants

The Space Age has necessitated accelerated development of many technologies which have direct application to commercial, earthbound hardware. One striking example is dry film lubrication. The stimulus for the development of dry film lubricants was provided by the unsatisfactory condition created when commercial liquid lubricants were exposed to the multiple environs of space. In the search for a satisfactory solution to lubrication in space, dry films were investigated. The dry film lubrication concept has proven most desirable for space, and a natural fallout is the potential shown by such films for consumer item application. On the acceptance that the coefficient of friction of dry films is not equivalent to that of a liquid lubricant under ideal conditions, this paper discusses how the ideal condition for liquid lubrication is rarely achieved in practical applications. Comparisons are made between the performance of dry film and liquid lubricants as a function of temperature, environmental pressure, and life. Furthermore, consideration is given to the comparative costs of liquid versus dry film lubricants to show the practicality and desirability of dry film lubrication to consumer items. Although no data on the performance of consumer items lubricated with dry films are presented, the comparisons clearly establish the potential of dry film lubricants in a competitive consumer market.

SOLID LUBRICANT↗

Spacecraft-spacecraft radio-metric tracking: Signal acquisition requirements and application to Mars approach navigation

Doppler and ranging measurements between spacecraft can be obtained only when the ratio of the total received signal power to noise power density (P(sub t)/N(sub 0)) at the receiving spacecraft is sufficiently large that reliable signal detection can be achieved within a reasonable time period. In this article, the requirement on P(sub t)/N(sub 0) for reliable carrier signal detection is calculated as a function of various system parameters, including characteristics of the spacecraft computing hardware and a priori uncertainty in spacecraft-spacecraft relative velocity and acceleration. Also calculated is the P(sub t)/N(sub 0) requirements for reliable detection of a ranging signal, consisting of a carrier with pseudonoise (PN) phase modulation. Once the P(sub t)/N(sub 0) requirement is determined, then for a given set of assumed spacecraft telecommunication characteristics (transmitted signal power, antenna gains, and receiver noise temperatures) it is possible to calculate the maximum range at which a carrier signal or ranging signal may be acquired. For example, if a Mars lander and a spacecraft approaching Mars are each equipped with 1-m-diameter antennas, the transmitted power is 5 W, and the receiver noise temperatures are 350 K, then S-band carrier signal acquisition can be achieved at ranges exceeding 10 million km. An error covariance analysis illustrates the utility of in situ Doppler and ranging measurements for Mars approach navigation. Covariance analysis results indicate that navigation accuracies of a few km can be achieved with either data type. The analysis also illustrates dependency of the achievable accuracy on the approach trajectory velocity.

Kahn, R. D.↗

The cardiovascular response to the AGS

This paper reports the preliminary results of experiments on human subjects conducted to study the cardiovascular response to various g-levels and exposure times using an artificial gravity simulator (AGS). The AGS is a short arm centrifuge consisting of a turntable, a traction system, a platform and four beds. Data collection hardware is part of the communication system. The AGS provides a steep acceleration gradient in subjects in the supine position.

Cardus, David↗

In-situ radio-metric tracking to support navigation for interplanetary missions with multiple spacecraft

Doppler and ranging measurements between spacecraft can be obtained only when the ratio of the total received signal power to noise power density (P(sub t/N(sub 0)) at the receiving spacecraft is sufficiently large that reliable signal detection can be achieved within a reasonable time period. In this paper, the requirements on P(sub t)/N(sub 0) for reliable carrier signal detection is calculated as a function of various system parameters, including characteristics of the spacecraft computing hardware and a priori uncertainty in spacecraft-spacecraft relative velocity and acceleration. Also calculated is the P(sub t)/N(sub 0) requirement for relaible detection of a ranging signal, consistting of a carrier with pseudo-noise phase modulation. Once the P(sub t)/N(sub 0) requirement is determined, then for a given set of assumed spacecraft telecommunication characteristics (transmitted signal power, antenna gains, receiver noise temperatures) it is possible to calculate the maximum range at which a carrier signal or ranging signal may be acquired. A brief error covariance analysis has been conducted to illustrate the utility of in situ Doppler and ranging measurements for Mars approach navigation. The results indicate that navigation accuracies of a few kilometers can be achieved with either data type. The analysis also illustrates dependency of the achievable accuracy on the approach trajectory velocity.

Kahn, Robert D.↗

Demonstration of a Pyrotechnic Bolt-Retractor System

A paper describes a demonstration of the X-38 bolt-retractor system (BRS) on a spacecraft-simulating apparatus, called the Large Mobility Base, in NASA's Flight Robotics Laboratory (FRL). The BRS design was proven safe by testing in NASA's Pyrotechnic Shock Facility (PSF) before being demonstrated in the FRL. The paper describes the BRS, FRL, PSF, and interface hardware. Information on the bolt-retraction time and spacecraft-simulator acceleration, and an analysis of forces, are presented. The purpose of the demonstration was to show the capability of the FRL for testing of the use of pyrotechnics to separate stages of a spacecraft. Although a formal test was not performed because of schedule and budget constraints, the data in the report show that the BRS is a successful design concept and the FRL is suitable for future separation tests.

Johnston, Nick↗

Observations on Dynamic Qualification Testing of a Component with Nonlinear Deadband Interfaces

Typical flight hardware dynamic qualification tests exhibit nearly linear structural response when exposed to acceleration inputs from low to full qualification levels. Even in these “linear” cases, there is typically a trend of increasing damping with test levels. The nonlinear mechanism behind this increase in energy dissipation well understood – typically stick/slip hysteresis at joint connections. This paper is concerned with a different type of nonlinearity: deadbands at structural interfaces. Deadband nonlinearities can have a significant influence on structural response and modal/spectral characteristics which can present difficulties in test, analysis, and structural certification. This subject nonlinear behaviour is observed during flight qualification testing of the AQUARIUS instrument and discussed here. Simple physical reasoning and analytical model is utilized to explain the behaviour which is consistent with the test findings.

Majed, Arya↗

wa-hls4ml and lui-gnn: A benchmark and GNN-based surrogate model for hls4ml resource and latency estimation

As machine learning (ML) increasingly serves as a tool for addressing real-time challenges in scientific applications, the development of advanced tooling has significantly reduced the time required to iterate on various designs. These advancements have solved major obstacles, but also exposed new challenges. For example, processes that were not previously considered bottlenecks, such as model synthesis, are now becoming limiting factors in the rapid iteration of designs. To reduce these emerging constraints, multiple efforts are being launched toward designing an ML-based surrogate model that estimates resource usage of synthesized accelerator architectures. This model would reduce the design iteration time, especially when designing within a set of given hardware constraints. This approach shows considerable potential, but as it stands, the effort is early and would benefit from coordination and standardization to assist future work as it emerges. We introduce wa-hls4ml, a benchmark for ML accelerator resource and latency estimation, and its corresponding initial dataset of more than 100,000 fully connected neural networks, all synthesized using hls4ml and targeting Xilinx FPGAs. In addition to the resource utilization and latency data provided, the dataset includes generated artifacts and log files for many of the synthesized neural networks, in order to support future research in ML-based code generation. The benchmark evaluates the performance of resource and latency predictors against several common ML model architectures, primarily originating from scientific domains, as exemplar models, as well as the average performance across a subset of the dataset. We measure the performance of a given predictor model through multiple metrics, including $R^2$ score and SMAPE on regression tasks, as well as inference time to further characterize the estimator under test. Additionally, we introduce the latency/utilization inference graph neural network (lui-gnn), a surrogate model that uses a graph neural network to represent input architectures in the form of a directed graph. This graph representation allows for a diverse set of model architectures to all be effectively handled by a surrogate model. We present the architecture and performance of the model, as evaluated by the new proposed benchmark, including SMAPE, $R^2$ score, and inference times, and find that lui-gnn generally predicts latency and utilization for the 75\% quantile within several percent of the synthesized resources on the synthetic test dataset, indicating that this approach of estimating resource and latency via a surrogate models has promise and warrants further research.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

High-power test of a C-band linear accelerating structure with an RFSoC-based LLRF system

Normal conducting linear particle accelerators consist of multiple rf stations with accelerating structure cavities. Low-level rf (LLRF) systems are employed to set the phase and amplitude of the field in the accelerating structure and to compensate for the pulse-to-pulse fluctuation of the rf field in the accelerating structures with a feedback loop. The LLRF systems are typically implemented with analog rf mixers, heterodyne-based architectures, and discrete data converters. There are multiple rf signals from each of the rf stations, so the number of rf channels required increases rapidly with multiple rf stations. With a large number of rf channels, the footprint, component cost, and system complexity of the LLRF hardware will increase significantly. To meet the design goals of being compact and affordable for future accelerators, we have designed the next-generation LLRF (NG-LLRF) with a higher integration level based on RFSoC technology. The NG-LLRF system samples rf signals directly and performs rf mixing digitally. Further, the NG-LLRF has been characterized in loopback mode to evaluate the performance of the system and has also been tested with a standing-wave accelerating structure, a prototype for the Cool Copper Collider (C 3 ) with a peak rf power level up to 16.45 MW. The loopback test demonstrated amplitude fluctuation below 0.15% and phase fluctuation below 0.15°, which are considerably better than the requirements of C 3 . The rf signals from the different stages of the accelerating structure at different power levels are measured by the NG-LLRF, which will be critical references for the control algorithm designs. The NG-LLRF also offers flexibility in waveform modulation, so we have used rf pulses with various modulation schemes, which could be useful for controlling some of the rf stations in accelerators. In this paper, the high-power test results at different stages of the test setup will be summarized, analyzed, and discussed.

47 OTHER INSTRUMENTATION↗

A Relief Zone Architecture for Enhanced Durability of Ultrathin Membranes in Electrochemical Conversion Devices

Premature cell failures in electrochemical conversion systems often result when membrane electrode assemblies (MEAs) are prepared with ultra-thin (= 15 micrometer-thick) polymer electrolyte membranes (PEMs). These high-performance PEMs are susceptible to mechanical degradation from stress concentrations arising from component and device-level integration. Herein, a relief zone was developed to mitigate mechanical degradation by alleviating excess and non-uniform compression across the active area. Relief zones, created through the ablation of carbonaceous diffusion material achieves a seamless adaptation across a range of MEA dimensions and electrochemical applications without need for costly hardware modifications. This concept was demonstrated using fuel cells as a case study. Accelerated stress test (AST) validations yielded a 6-fold improvement for MEAs prepared with relief zones (i.e., lifetime approximately 1500 h) compared to those fabricated using conventional edge-protection techniques, showcasing the utility of this technique in decoupling integration and device-level engineering effects from intrinsic material limitations for durable electrochemical devices.

14 SOLAR ENERGY↗

The hypervelocity impact facilities at the University of Kent at Canterbury, UK

The University of Kent at Canterbury facilities for production of hypervelocity impacts are discussed. Meteoroids are simulated by electrostatic acceleration of small (10(exp -12) - 10(exp -17) kg) particles using a 2-MV Van de Graaff accelerator. This machine has been operational for many years; both the machine and experimental area are currently being upgraded. Larger particles (10(exp -10) - 10(exp -14) kg) are accelerated using a more recently installed light gas gun. The status of all hardware (including experimental areas) is given, along with brief details of recent, current, and future projects making use of them.

Burchell, Mark J.↗

Reliability of Electronics for Cryogenic Space Applications Being Assessed

Many future NASA missions will require electronic parts and circuits that can operate reliably and efficiently in extreme temperature environments below typical device specification temperatures. These missions include the Mars Exploration Laboratory, the James Webb Space Telescope, the Europa Orbiter, surface rovers, and deep-space probes. In addition to NASA, the aerospace and commercial sectors require cryogenic electronics in applications that include advanced satellites, military hardware, medical instrumentation, magnetic levitation, superconducting energy management and distribution, particle confinement and acceleration, and arctic missions. Besides surviving hostile space environments, electronics capable of low-temperature operation would enhance circuit performance, improve system reliability, extend lifetime, and reduce development and launch costs. In addition, cryogenic electronics are expected to result in more efficient systems than those at room temperature.

Patterson, Richard L.↗