Search NASASearch

SEARCH · Search NASA

Results for “Parallel in time”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Three-wave coupling observed between a shear Alfvén wave and a kink-unstable magnetic flux rope

Results from a laboratory experiment are presented in which, for the first time, a shear Alfvén wave is launched using an antenna in a current-carrying plasma column that is tailored to be either stable or unstable to the kink oscillation. As the plasma is driven kink unstable, the frequency power spectrum of the Alfvén wave evolves from a single peak to a peak with multiple sidebands separated by integer multiples of the kink frequency. The main sidebands (one on either side of the launched wave peak in the power spectrum) are analyzed using azimuthal wavenumber matching, perpendicular and parallel wavenumber decomposition, and bispectral time series analysis. The dispersion relation and three-wave matching conditions are satisfied, given each sideband is a propagating Alfvén wave that results from the interaction of the pump Alfvén wave and the co-propagating component of a half-wavelength, standing kink mode. The interaction is shown to generate smaller perpendicular wavelength Alfvén waves that drive energy transport to scales that will approach the dissipation scale of k⊥ρs=1, with k⊥ being the perpendicular wavenumber and ρs being the ion gyroradius at the electron temperature.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Hydrotreatment of Nylon 66 and Amide Model Compounds Over Sulfided NiMo Catalysts

Molybdenum sulfide-based catalysts, such as nickel–molybdenum on alumina (NiMoS x /Al 2 O 3 ), are widely used in hydrotreating and have potential for catalyzing waste plastic conversion via hydrogenolysis, yet their performance, such as reaction kinetics and network, for amide-rich polymer feeds is poorly defined. Here we combine Nylon 66 with the amide model compound, N,N-dibutylhexanediamide (DBDAD), to quantify hydrodeoxygenation (HDO) and hydrodenitrogenation (HDN) chemistry in a stirred batch reactor (53 bar H 2 , 280–320°C). DBDAD conversion is near-linear with time, indicating strong adsorption of the substrates on the active sites. Time-resolved product identification indicates parallel C─O first-cleavagedeoxygenation (DO) and C─N first-cleavagedenitrogenation (DN) sequences proceeding through amine and diol intermediates, respectively, to C 4 ─C 6 alkanes. Increasing temperature shifts selectivity toward DN, decreasing the initial r(DO)/r(DN) from 1.38 (280°C) to 0.69 (320°C), with an apparent activation energy of 173 kJ mol −1 for DBDAD conversion. At 300°C, nylon 66 converts faster than DBDAD, producing a complex mixture of oxygen- and nitrogen-containing species and an initial rate ratio r(DO)/r(DN) of 1.6. No heteroaromatic nitrogen products are detected by the method used. These results provide reaction pathways and product signatures relevant to hydro-processing catalysts exposed to polyamide-derived streams.

Nylon 66

Interfacial Inversion of Stealth Surfactants

Amphiphilic macromolecular surfactants segregate to liquid–liquid interfaces, thereby reducing the interfacial tension and free energy. Here, we investigated “stealth surfactants” in the form of core–shell bottlebrush polymers comprised of pH-responsive diblock copolymer side chains forming a hydrophilic core and a hydrophobic shell, enabling solubility in oil. At liquid–liquid interfaces, these polymers undergo a structural “inversion”, with hydrophilic blocks segregating into the aqueous phase and hydrophobic blocks residing in the oil phase. The reconfiguration kinetics and surfactant properties are influenced by multiple factors, including the molecular weights of the backbone and side chain components, the hydrophilic-to-hydrophobic balance of the side chains, and the pH of the aqueous phase. An observed nonmonotonic dependence of interfacial tension with time is attributed to a progressive structural inversion, where the projected area of the macromolecule onto the interface decreases. To validate this inversion hypothesis, interfacial properties were characterized by sum-frequency generation vibrational spectroscopy, which revealed configurational changes of the core–shell bottlebrush polymers at the fluid interface and revealed a pH-dependent interfacial coverage. Coarse-grained molecular dynamics simulations supported these experimental findings, showing that the pH-responsive core and hydrophobic shell assume a time-averaged configuration with orientations parallel and perpendicular to the plane of the interface, respectively. These findings open routes to design multistimuli-responsive polymeric surfactants and compatibilizers, expanding their potential applications in advanced interfacial systems.

Stealth surfactants

Low-Temperature Plasma-Based Metrology of Lithium-Ion Battery Electrode Materials (CRADA Final Report)

As part of the Cyclotron Road program, SirenOpt Inc. evaluated its low-temperature plasma-based metrology sensor prototype for measuring multiple critical properties of lithium-ion battery electrode materials in parallel and in real-time. Cost-effective, minimal-waste manufacturing of high-performance battery electrode materials will be vital for achieving society’s net-zero carbon emission goals. Because existing electrode metrology sensors cannot operate within most sections of manufacturing lines, manufacturers often complete hundreds of processing steps before they can test their products and detect problems. When manufacturers perform these offline tests, they typically only test a small portion of the manufactured products. Current electrode manufacturing thus often yields many low-quality products, or off-spec products that must be thrown away all together. For example, at least 6% of the total lithium-ion battery manufacturing cost (i.e., over $250 million/year for the average gigafactory) is devoted to processing defective electrodes that are not scrapped until performance tests are failed during late-stage quality control checks. Electrode variability also leads manufacturers to build extra cells into battery packs to reduce the risk of poor performance. For example, many electric vehicle (EV) manufacturers include up to 10% more cells than needed, which substantially increases the cost and weight of the final EV product. The SirenOpt sensor can potentially enable early detection of poorly manufactured electrodes and allow them to be removed earlier from manufacturing lines, which can save battery manufacturers (hundreds of) millions of dollars per year. The sensor can further be used to improve product quality by accelerating R&D and process optimization, improving quality control, and enabling real-time process control. Overall, a real-time, in-situ metrology strategy can create unprecedented opportunities for implementation of smart manufacturing practices and advanced quality and process control solutions to realize higher battery electrode throughput and performance.

25 ENERGY STORAGE

Comparative Analysis of Radial and Random Microstructures of Mesophase Pitch Carbon Fibers

Carbon fibers (CF) with radial and random microstructures are produced. Here, these fibers are subjected to identical treatment before being mechanically tested and analyzed with Weibull analysis, with the results revealing a statistically significant difference in tensile strengths of 2.23 GPa for random CF and 1.69 GPa for radial CF. Raman mapping probed the crystalline structure perpendicular to the fiber axis and found a uniform structure, while wide‐angle X‐ray diffraction showed a significant difference of 7.5 Å in the crystallites’ basal lengths parallel to the fiber. Small‐angle X‐ray scattering is completed parallel to the fiber for the first time. A cross‐section Guinier plot of the 1D azimuthal integration is generated assuming symmetric scattering, and the parallel scatterers are found to have a similar length scale to the crystallite's length, validating the testing method. Finally, transmission electron microscopy is completed on the longitudinal cross‐section of each fiber. The radial carbon fiber is found to have a core–shell structure, as evidenced further by fast Fourier transform images. Through all studies, it is shown that the structure developed during mesophase pitch spinning altered the microstructure, thus impacting the mechanical properties, confirming a direct relationship between processing, structure, and properties.

Scherschel, Alexander [Univ. of Virginia, Charlott

Using Hardware-In-The-Loop Methodology to Develop Test Systems

Hardware in the Loop (HIL) testing methodologies have become widespread in industry. Typically, they focus on developing control algorithms for systems such as autonomous vehicles or aircraft. An oft overlooked aspect of product development is the design and fabrication of a test system for validating that the product meets requirements. Abstractly, a test system differs little from a control system—testers provide signals to the unit, monitor feedback, and base decisions on the results. While the time scales may differ, the functionalities are conceptually similar. Viewed in this light, it becomes natural to extend HIL approaches to tester development. By replacing a physical unit with a proxy model deployed to a real-time or pseudo real-time target, test systems can be developed in parallel with the design and fabrication of a first production unit. This saves considerable time in the life cycle from conceptual design to realized product. This manuscript demonstrates the process flow using a capacitive discharge unit as an exemplar.

42 ENGINEERING

Continental rift evolution and drainage reorganization along the Dead Sea rift since the Miocene

The Dead Sea fault is a section of the Arabian-African plate boundary. Widespread field relations indicate that three major drainage systems (stages) occupied the landscape west of the Dead Sea fault since its initiation at ca. 20 Ma. Specifically, (1) an early to middle Miocene drainage system, only minorly reconfigured by the fault (all sediments of this system belong to the Hazeva Formation); (2) a late Miocene to early Pleistocene fault-parallel drainage system named Paran-Neqarot (all sediments of this system belong to the Arava and Zehiha Formations; and (3) the early Pleistocene to present drainage configuration. The temporal and spatial frameworks of drainage stage 1 are generally constrained by radiometric dating of interfingering volcanic units, and the onset and temporal and spatial frameworks of drainage stage 3 are well constrained by cosmogenic 10 Be surface exposure ages. The timing and longevity of the rift-parallel drainage (stage 2) have until now been evasive to direct dating. The overall time gap between stages 1 and 3 is ~12–13 million years. Thus, an early age (within this time gap) of stage 2 would imply an immediate response of drainage reorganization to rift tectonics, while a later age of this drainage system would imply a delayed response. We present 11 10 Be- 26 Al cosmogenic burial ages of alluvial and colluvial units related to the fault-parallel drainage system (stage 2), which collectively constrain the time of deposition of the Arava Formation sediments in the central Negev to ca. 8 Ma. The general lack of stratigraphic order, together with the large dispersion of ages both across and within the sampling sites, attests to significant recycling of sediments from drainage stage 1 into the Arava Formation deposits. The termination of Arava and Zehiha Formation sediment deposition at ca. 1.8 Ma was determined previously using cosmogenic exposure ages of desert pavements that cover the formations. Combining the previously published data with our new data, we established the longevity and character of the Paran-Neqarot drainage system. In conlcusion, this framework highlights the temporal aspect of drainage system build-up and collapse as it responded to transform and extensional plate boundary tectonics during the Neogene.

58 GEOSCIENCES

Real-Time Bayesian Inference at Extreme Scale: A Digital Twin for Tsunami Early Warning Applied to the Cascadia Subduction Zone

We present a Bayesian inversion-based digital twin that employs acoustic pressure data from seafloor sensors, along with 3D coupled acoustic–gravity wave equations, to infer earthquake-induced spatiotemporal seafloor motion in real time and forecast tsunami propagation toward coastlines for early warning with quantified uncertainties. Our target is the Cascadia subduction zone, with one billion parameters. Computing the posterior mean alone would require 50 years on a 512 GPU machine. Instead, exploiting the shift invariance of the parameter-to-observable map and devising novel parallel algorithms, we induce a fast offline–online decomposition. The offline component requires just one adjoint wave propagation per sensor; using MFEM, we scale this part of the computation to the full El Capitan system (43,520 GPUs) with 92% weak parallel efficiency. Moreover, given real-time data, the online component exactly solves the Bayesian inverse and forecasting problems in 0.2 seconds on a modest GPU system, a ten-billion-fold speedup.

97 MATHEMATICS AND COMPUTING

Enabling Parallel Performance and Portability of Solid Mechanics Simulations Across CPU and GPU Architectures

Efficiently simulating solid mechanics is vital across various engineering applications. As constitutive models grow more complex and simulations scale up in size, harnessing the capabilities of modern computer architectures has become essential for achieving timely results. This paper presents advancements in running parallel simulations of solid mechanics on multi-core CPUs and GPUs using a single-code implementation. This portability is made possible by the C++ matrix and array (MATAR) library, which interfaces with the C++ Kokkos library, enabling the selection of fine-grained parallelism backends (e.g., CUDA, HIP, OpenMP, pthreads, etc.) at compile time. MATAR simplifies the transition from Fortran to C++ and Kokkos, making it easier to modernize legacy solid mechanics codes. We applied this approach to modernize a suite of constitutive models and to demonstrate substantial performance improvements across different computer architectures. This paper includes comparative performance studies using multi-core CPUs along with AMD and NVIDIA GPUs. Results are presented using a hypoelastic–plastic model, a crystal plasticity model, and the viscoplastic self-consistent generalized material model (VPSC-GMM). The results underscore the potential of using the MATAR library and modern computer architectures to accelerate solid mechanics simulations.

Morgan, Nathaniel (ORCID:0000000276118449)

Cardinal: Seismic and Geoacoustic Array Processing

Data collected via seismic and infrasound array deployments are leveraged in the geosciences to detect and characterize a myriad of natural and anthropogenic sources. These deployments consist of numerous sensors placed in a predetermined configuration to amplify signal strength and improve the efficacy of array processing techniques used to measure signal directionality and waveform coherence. High‐fidelity feature extraction is often predicated on interstation distance as well as the frequency content and wavelength of an incident signal. Numerous array processing softwares analyze data in sequential frequency bands to obtain a more detailed characterization of a signal. However, current algorithms are limited in their ability to determine optimal array configuration for each band. We introduce an open‐source Python code, called Cardinal, to process seismic and infrasound array data in discretized time–frequency space with the option of applying an adaptive array design to determine optimal subarray configuration for each frequency band. To reduce computational time, the array processing step can be run in parallel using multithreading. Furthermore, the software has the capability to aggregate array processing results from different time–frequency pixels to produce separate sets of detections, or families, with added utility via the application of an adaptive semblance threshold, which aids in isolating signals‐of‐interest from coherent background noise. Upon appropriate configuration, Cardinal exhibits the potential to combine distinct seismic and infrasound phases into separate families.

Adaptive Array

A worldwide climatology of extreme air masses

Extreme temperature events are among the most damaging weather phenomena. In a warming world, more heat extremes and fewer cold extremes are expected in most regions in the future, a trade-off that warrants further understanding of such events. Here, we track and analyze large, persistent areas of hot and cold extreme temperatures in parallel, relative to the location and time of year, to quantify overall regional exposure to extreme temperatures. To accomplish this, we compare the frequencies, movements, trends, and sources/sinks in each type of extreme air mass, calling them extreme cold or extreme hot air masses (ECAMs/EHAMs). For most land regions, ECAMs occur more often than EHAMs, and ECAMs are more common in each hemisphere’s winter, when cold-air advection is strongest and most widespread. Average movement of ECAMs has a stronger equatorward component in winter than in summer, while movement of EHAMs is eastward all year, with less meridional movement than ECAMs. EHAMs have become more common almost everywhere on land, and the reverse is true for ECAMs, with the strongest trends in the northern hemisphere occurring in autumn, especially in the Arctic. The number of EHAMs around the world is increasing at a higher rate (+ 2.32 events/year) than ECAM numbers are decreasing (-1.27 events/year), causing a net increase in extremes, especially in some midlatitude regions. These results help to document the different processes driving hot and cold extremes, and thus the asymmetric and regionally varying trends in frequency for extreme air masses.

Ryan, James M

MFC 5.0: An exascale many-physics flow solver

Many problems of interest in engineering, medicine, and the fundamental sciences rely on high-fidelity flow simulation, making performant computational fluid dynamics solvers a mainstay of the open-source software community. Previous work MFC 3.0 was made a published, documented, and open-source solver via Bryngelson et al. Comp. Phys. Comm. (2021) with numerous physical features, numerical methods, and scalable infrastructure. MFC 5.0 is a significant update to MFC 3.0, featuring a broad set of well-established and novel physical models and numerical methods, as well as the introduction of GPU and APU (or superchip) acceleration. Here, we exhibit state-of-the-art performance and ideal scaling on the first two exascale supercomputers, OLCF Frontier and LLNL El Capitan. Combined with MFC’s single-accelerator performance, MFC achieves exascale computation in practice, and achieved the largest-to-date public CFD simulation at 200 trillion grid points as a 2025 ACM Gordon Bell Prize finalist. New physical features include the immersed boundary method, N-fluid phase change, Euler–Euler and Euler–Lagrange sub-grid bubble models, fluid-structure interaction, hypo- and hyper-elastic materials, chemically reacting flow, two-material surface tension, magnetohydrodynamics (MHD), and more. Numerical techniques now represent the current state-of-the-art, including general relaxation characteristic boundary conditions, WENO variants, Strang splitting for stiff sub-grid flow features, and low Mach number treatments. Weak scaling to tens of thousands of GPUs on OLCF Summit and Frontier and LLNL El Capitan achieves efficiencies within 5% of ideal to over 90% of their respective system sizes. Strong scaling results for a 16-times increase in device count show parallel efficiencies over 90% on OLCF Frontier. MFC’s software stack has undergone further improvements, including continuous integration, which ensures code resilience and correctness through over 300 regression tests; metaprogramming, which reduces code length while maintaining performance portability; and code generation for computing chemical reactions

Computational fluid dynamics

PyHydroGeophysX: An extensible open-source platform for integrating hydrological models with geophysical measurements

Hydrological models and geophysical measurements are widely used tools for understanding subsurface hydrological processes relevant to water resource management, yet they typically remain disconnected due to technical barriers. We present PyHydroGeophysX, an open-source Python platform bridging this gap by providing standardized interfaces between hydrological modeling software (MODFLOW, ParFlow) and geophysical simulation tools (PyGIMLi, SimPEG). The platform implements bidirectional workflows: translating hydrological outputs into simulated geophysical responses through petrophysical models, and extracting hydrological information from geophysical inversions. Key features include bidirectional workflow modules, configurable petrophysical models, time-lapse inversion with temporal regularization, parallel computing, and mesh utilities for property transfer between geophysical and hydrological grids. The modular architecture of PyHydroGeophysX enables researchers to incorporate additional models and methods, fostering broader adoption of integrated hydrogeophysical approaches. The software is freely available on GitHub and is intended for researchers and practitioners working at the intersection of hydrology and geophysics.

Hydrogeophysics

Fast and Accurate Greenberger-Horne-Zeilinger Encoding Using All-to-All Interactions

The 𝑁-qubit Greenberger-Horne-Zeilinger (GHZ) state is an important resource for quantum technologies. Here, we consider the task of GHZ encoding using all-to-all interactions, which prepares the GHZ state in a special case, and is furthermore useful for quantum error correction, interaction-rate enhancement, and transmitting information using power-law interactions. The naive protocol based on parallelizing CNOT gates takes O(1)-time of Hamiltonian evolution. In this work, we propose a fast protocol that achieves GHZ encoding with high accuracy. The evolution time O⁡(log 2 ⁡𝑁/𝑁) almost saturates the theoretical limit Ω⁡(log⁡𝑁/𝑁). Moreover, the final state is close to the ideal encoded one with high fidelity >1–10 −3 , up to large system sizes 𝑁 ≲ 2000. The protocol only requires a few stages of time-independent Hamiltonian evolution; the key idea is to use the data qubit as control, and to use fast spin-squeezing dynamics generated by e.g., two-axis twisting.

quantum computation

Photochemically Induced Acousto-optics Fluid Simulations

PIAFS is a finite-difference code to solve the compressible Navier-Stokes equations with chemical heating on Cartesian grids. It models chemical reactions of air (oxygen and carbon dioxide) with ozone subject to radiation. It uses a high-order WENO spatial discretization and explicit Runge-Kutta time integration. It is capable of parallel simulations using MPI. The code is written in C/C++.

Oudin, AlbertineN [Lawrence Livermore National Lab

On a Simplified Approach to Achieve Parallel Performance and Portability Across CPU and GPU Architectures

This paper presents software advances to easily exploit computer architectures consisting of a multi-core CPU and CPU+GPU to accelerate diverse types of high-performance computing (HPC) applications using a single code implementation. The paper describes and demonstrates the performance of the open-source C++ matrix and array (MATAR) library that uniquely offers: (1) a straightforward syntax for programming productivity, (2) usable data structures for data-oriented programming (DOP) for performance, and (3) a simple interface to the open-source C++ Kokkos library for portability and memory management across CPUs and GPUs. The portability across architectures with a single code implementation is achieved by automatically switching between diverse fine-grained parallelism backends (e.g., CUDA, HIP, OpenMP, pthreads, etc.) at compile time. The MATAR library solves many longstanding challenges associated with easily writing software that can run in parallel on any computer architecture. This work benefits projects seeking to write new C++ codes while also addressing the challenges of quickly making existing Fortran codes performant and portable over modern computer architectures with minimal syntactical changes from Fortran to C++. We demonstrate the feasibility of readily writing new C++ codes and modernizing existing codes with MATAR to be performant, parallel, and portable across diverse computer architectures.

97 MATHEMATICS AND COMPUTING

Electron Influence on the Parallel Proton Firehose Instability in 10-moment, Multifluid Simulations

Instabilities driven by pressure anisotropy play a critical role in modulating the energy transfer in space and astrophysical plasmas. For the first time, we simulate the evolution and saturation of the parallel proton firehose instability using a multifluid model without adding artificial viscosity. These simulations are performed using a 10-moment, multifluid model with local and gradient relaxation heat-flux closures in high-β proton–electron plasmas. When these higher-order moments are included and pressure anisotropy is permitted to develop in all species, we find that the electrons have a significant impact on the saturation of the parallel proton firehose instability, modulating the proton pressure anisotropy as the instability saturates. Even for lower β's more relevant to heliospheric plasmas, we observe a pronounced electron energization in simulations using the gradient relaxation closure. Our results indicate that resolving the electron pressure anisotropy is important to correctly describe the behavior of multispecies plasma systems.

79 ASTRONOMY AND ASTROPHYSICS

Characterization and Optimization of the Fitting of Quantum Correlation Functions

This case study presents a characterization and optimization of an application code for extracting parton distribution functions from high energy electron-proton scattering data. Profiling this application code reveals that the phase-space density computation accounts for 93% of the overall execution time for a single iteration on a single core. When executing multiple iterations in parallel on a multicore system, the application spends 78% of its overall execution time idling due to load imbalance. We address these issues by first transforming the application code from Python to C++ and then tackling the application load imbalance via a hybrid scheduling strategy that combines dynamic and static scheduling. These techniques result in a 62% reduction in CPU idle time and a 2.46x speedup in overall execution time per node. In addition, the typically enabled power-management mechanisms in supercomputers (e.g., AMD Turbo Core, Intel Turbo Boost, and RAPL) can significantly impact intra-node scalability when more than 50% of the CPU cores are used. This finding underscores the importance of understanding system interactions with power management, as they can adversely impact application performance, and highlights the necessity of intra-node scaling tests to identify performance degradation that inter-node scaling tests might otherwise overlook.

Chuang, Pi-Yueh [Virginia Tech,Dept. of Computer S