Search NASA⌕ Search

SEARCH · Search NASA

Results for “Parallelization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 451 records · Page 25

Impact of Iron Species Dispersion on Fe/ZSM–5 Catalyst Performance for Methane Dehydroaromatization (MDA)

Methane dehydroaromatization (MDA) is one of the most promising technologies for directly transforming methane into aromatics. Unlike the extensively investigated Mo/ZSM-5 catalysts, the structure and, consequently, the catalytic activity of Fe/ZSM-5 are markedly influenced by the method of preparation, as shown here. In this study, we prepared 2 % and 4 % Fe/ZSM-5 catalysts via wet impregnation (WI) and incipient wetness impregnation (IWI). Characterizations (XRD, STEM, UV-Vis, NH 3 -TPD and H 2 -TPR) reveal that 2 %Fe-WI mainly possesses isolated or low-polymerized Fe species within zeolite channels, leading to a rapid activation and a higher benzene yield due to the faster reduction to iron suboxides under MDA conditions. In contrast, 2 %Fe-IWI contains bulk iron oxide aggregates, resulting in a slower activation as these aggregates transform into iron carbide through successive reduction and carbonization. Here, a deactivation kinetic study applied to the 2 % catalysts further demonstrates the quantitative relation between Fe site isolation and catalytic activity. Although both 4 % catalysts inevitably form sizable iron oxide clusters and particles due to the high Fe/Al ratio, similar trends are noted, with the WI catalysts exhibiting a shorter induction/activation period and a higher yield of benzene, paralleling observations made with 2 % catalysts.

03 NATURAL GAS↗

Hydrotreatment of Nylon 66 and Amide Model Compounds Over Sulfided NiMo Catalysts

Molybdenum sulfide-based catalysts, such as nickel–molybdenum on alumina (NiMoS x /Al 2 O 3 ), are widely used in hydrotreating and have potential for catalyzing waste plastic conversion via hydrogenolysis, yet their performance, such as reaction kinetics and network, for amide-rich polymer feeds is poorly defined. Here we combine Nylon 66 with the amide model compound, N,N-dibutylhexanediamide (DBDAD), to quantify hydrodeoxygenation (HDO) and hydrodenitrogenation (HDN) chemistry in a stirred batch reactor (53 bar H 2 , 280–320°C). DBDAD conversion is near-linear with time, indicating strong adsorption of the substrates on the active sites. Time-resolved product identification indicates parallel C─O first-cleavagedeoxygenation (DO) and C─N first-cleavagedenitrogenation (DN) sequences proceeding through amine and diol intermediates, respectively, to C 4 ─C 6 alkanes. Increasing temperature shifts selectivity toward DN, decreasing the initial r(DO)/r(DN) from 1.38 (280°C) to 0.69 (320°C), with an apparent activation energy of 173 kJ mol −1 for DBDAD conversion. At 300°C, nylon 66 converts faster than DBDAD, producing a complex mixture of oxygen- and nitrogen-containing species and an initial rate ratio r(DO)/r(DN) of 1.6. No heteroaromatic nitrogen products are detected by the method used. These results provide reaction pathways and product signatures relevant to hydro-processing catalysts exposed to polyamide-derived streams.

Nylon 66↗

A GPU ‐Accelerated 3D Unstructured Mesh Based Particle Tracking Code for Multi‐Species Impurity Transport Simulation in Fusion Tokamaks

ABSTRACT This paper presents the multi‐species global impurity transport capability developed in a GPU‐accelerated fully 3D unstructured mesh‐based code, GITRm, to simultaneously track multiple impurity species and handle interactions of these impurities with mixed‐material surfaces. Different computational approaches to model particle‐surface interaction or surface response have been developed and compared. Sheath electric field is taken into account by employing a fast distance‐to‐boundary calculation, which is carried out in parallel on distributed or partitioned meshes on multiple GPUs without the need for any inter‐process communication during the simulation. Several example cases, including two for the DIII‐D tokamak, that is, one with the SAS‐V divertor and the other with the collector probes, are used to demonstrate the utility of the current multi‐species capability. For the DIII‐D probe case, the capability of GITRm to resolve the spatial distribution of particles in localized regions, such as diagnostic probes, within non‐axisymmetric tokamak geometries is demonstrated. These simulations involve up to 320 million particles and utilize up to 48 GPUs.

Nath, Dhyanjyoti D. [Scientific Computation Resear↗

A GPU Accelerated Mixed‐Precision Finite Difference Informed Random Walker (FDiRW) Solver for Strongly Inhomogeneous Diffusion Problems

In nature, many complex multi‐physics coupling problems exhibit significant diffusivity inhomogeneity, where one process occurs several orders of magnitude faster than others temporally. Simulating rapid diffusion alongside slower processes demands intensive computational resources due to the necessity for small time steps. To address these computational challenges, we have developed an efficient numerical solver named Finite Difference informed Random Walker (FDiRW). In this study, we propose a GPU‐accelerated, mixed‐precision configuration for the FDiRW solver to maximize efficiency through GPU multi‐threaded parallel computation and lower precision computation. Numerical evaluation results reveal that the proposed GPU‐accelerated mixed‐precision FDiRW solver can achieve a 117× speedup over the CPU baseline, while an additional 1.75× speedup is achieved by employing lower precision GPU computation. Notably, for large model sizes, the GPU‐accelerated mixed‐precision FDiRW solver demonstrates strong scaling with the number of nodes used in simulation. When simulating radionuclide absorption processes by porous wasteform particles with a medium‐sized model of 192 × 192 × 192, this approach reduces the total computational time to 10 min, enabling the simulation of larger systems with strongly inhomogeneous diffusivity.

97 MATHEMATICS AND COMPUTING↗

ReactionMechanismSimulator.jl: A modern approach to chemical kinetic mechanism simulation and analysis

Abstract We present ReactionMechanismSimulator.jl (RMS), a modern differentiable software for the simulation and analysis of chemical kinetic mechanisms, including multiphase systems. RMS has already been applied to problems in combustion, pyrolysis, polymers, pharmaceuticals, catalysis, and electrocatalysis. RMS is written in Julia, making it easy to develop and allowing it to take advantage of Julia's extensive numerical computing ecosystem. In addition to its extensive library of optimized analytic Jacobians, RMS can generate and use Jacobians computed using automatic differentiation and symbolically generated analytic Jacobians. RMS is demonstrated to be faster than Cantera and Chemkin in several benchmarks. RMS also implements an extensive set of features for analyzing chemical mechanisms, including a library of easy‐to‐call plotting functions, molecular structure resolved flux diagram generation, crash analysis, traditional sensitivity analysis, transitory sensitivity analysis, and an automatic mechanism analysis toolkit. RMS implements efficient adjoint and parallel forward sensitivity analyses. We also demonstrate the ease of adding new features to RMS.

Johnson, Matthew S.↗

Synergistic Alignment of Low Aspect‐Ratio π‐Conjugated Molecules Enables Exceptional UV–vis–NIR Polarization Detection

Abstract Polarization detection enhances signal contrast and is widely utilized in diverse advanced applications. An ongoing challenge is the development of high‐performance polarization‐sensitive photodetectors based on optically anisotropic organic semiconductors, particularly in the near‐infrared (NIR) region. While uniaxially aligned π‐conjugated polymers with high aspect ratios exhibit strong linear dichroism and have shown promise, their limited NIR performance and heavy reliance on polymer material now represent critical limitations. Here, a breakthrough is reported in achieving giant linear dichroism and exceptional polarization detection with low aspect‐ratios (AR) non‐fullerene small‐molecule (NFSM) acceptors, extending polarization sensitivity from the UV–vis to the NIR range. An impressive dichroic ratio of 27.1 at 605 nm and 12.0 at 780 nm is demonstrated. The maximum polarization photocurrent ratio is 11.2 at 780 nm under parallel versus perpendicular polarized light. This unprecedented performance originates from synergistic molecular alignment, wherein NFSMs significantly enhance the uniaxial orientation of both the polymer matrix and the NFSMs themselves during self‐assembly and thermal annealing. Besides, such a linear‐polarization‐sensitive photodetectors (LPS‐PDs) are showcased in generating degree‐of‐linear‐polarization imaging. The work establishes NFSMs as a viable material system for next‐generation of organic LPS‐PDs and provides fundamental insights into structural origins of polarization sensitivity in low AR organic semiconductors.

Xue, Yingying↗

Nitrogen Status Rewires Transcriptional Regulation of Dhurrin, a Dual‐Purpose Defense Metabolite in Sorghum bicolor

Dhurrin, a cyanogenic glucoside, plays an important role in Sorghum bicolor physiology and defense. The concentration of dhurrin in sorghum is influenced by both nitrogen status and stage of plant organ development. While nitrogen resupply activates the expression of genes for dhurrin biosynthesis, the molecular mechanisms underlying this regulation remain unclear. In this study, we investigated the transcriptional response of sorghum to nitrogen resupply following growth under nitrogen-limiting conditions. Using a time-course design, we measured hydrogen cyanide potential (HCNp), growth, and nitrate content at 0-, 2-, 6-, 12-, 24-, 36-, 48-, and 60-h after resupply and collected tissue for RNAseq analysis in parallel for analysis of gene expression and construction of gene regulatory networks (GRNs). HCNp (mg g −1 DW) increased significantly in leaf and stem tissues following nitrogen resupply, with increases in the leaf partially driven by continued declines in controls under ongoing nitrogen stress. Expression of the dhurrin pathway genes was upregulated in leaves from 24 h after nitrogen resupply, with diel expression patterns observable over the remaining time points. No upregulation was observed in roots or stems, suggesting that developmental context overrides environmental cues. GRN analysis identified candidate transcription factors regulating dhurrin biosynthesis genes, including members of the MYB, bZIP, and GARP-type transcription factor families. Some of these candidate transcription factors may be involved in relieving senescence-associated suppression of dhurrin biosynthesis and link nitrogen signaling to pathway activation. These findings provide new insight into the nitrogen-responsive regulation of dhurrin in sorghum, highlighting candidate regulators for future functional characterization.

S. bicolor↗

Direct Exfoliation of Nanoribbons from Bulk van der Waals Crystals

Confinement of monolayers into quasi-1D atomically thin nanoribbons could lead to novel quantum phenomena beyond those achieved in their bulk and monolayer counterparts. However, current experimental availability of nanoribbon species beyond graphene is limited to bottom-up synthesis or lithographic patterning. Here, in this study, a versatile and direct approach is introduced to exfoliate bulk van der Waals crystals as nanoribbons. Akin to the Scotch tape exfoliation method for producing monolayers, this technique provides convenient access to a wide range of nanoribbons derived from their corresponding bulk crystals, including MoS 2 , WS 2 , MoSe 2 , WSe 2 , MoTe 2 , WTe 2 , ReS 2 , and hBN. The nanoribbons are predominantly monolayer, single-crystalline, parallel-aligned, flat, and exhibit high aspect ratios. The role of confinement, strain, and edge configuration of these nanoribbons is observed in their electrical, magnetic, and optical properties. This versatile exfoliation technique provides a universal route for producing a variety of nanoribbon materials and supports the study of their fundamental properties and potential applications.

2D materials↗

Sustainable Production of Biomass‐Derived Graphite and Graphene Conductive Inks from Biochar

Abstract Graphite is a commonly used raw material across many industries and the demand for high‐quality graphite has been increasing in recent years, especially as a primary component for lithium‐ion batteries. However, graphite production is currently limited by production shortages, uneven geographical distribution, and significant environmental impacts incurred from conventional processing. Here, an efficient method of synthesizing biomass‐derived graphite from biochar is presented as a sustainable alternative to natural and synthetic graphite. The resulting bio‐graphite equals or exceeds quantitative quality metrics of spheroidized natural graphite, achieving a RamanI D /I G ratio of 0.051 and crystallite size parallel to the graphene layers (L a ) of 2.08 µm. This bio‐graphite is directly applied as a raw input to liquid‐phase exfoliation of graphene for the scalable production of conductive inks. The spin‐coated films from the bio‐graphene ink exhibit the highest conductivity among all biomass‐derived graphene or carbon materials, reaching 3.58 ± 0.16 × 10 4 S m −1 . Life cycle assessment demonstrates that this bio‐graphite requires less fossil fuel and produces reduced greenhouse gas emissions compared to incumbent methods for natural, synthesized, and other bio‐derived graphitic materials. This work thus offers a sustainable, locally adaptable solution for producing state‐of‐the‐art graphite that is suitable for bio‐graphene and other high‐value products.

Chemistry↗

Discovery of a Ferromagnetic Nickel Chalcogenide Nanocluster Ni 3 S 3 H(PEt 3 ) 5

Atomically precise ligated nanoclusters (NC) are promising cluster-based materials with novel molecular architectures and tunable magnetic properties. Herein, the synthesis and characterization of a nickel sulfide NC Ni 3 S 3 H(PEt 3 ) 5 (PEt 3 = triethylphosphine) with distinct magnetic properties are reported. Magnetization measurements reveal its magnetic moment of 1.5 µ B in the solid phase, consistent with the existence of one unpaired electron predicted by density functional theory (DFT) calculations. Additionally, experimental measurements indicate the presence of ferromagnetic ordering within each Ni 3 S 3 H(PEt 3 ) 5 NC and strong coercivity at temperatures below 20 K. Ion mobility-mass spectrometry is employed in conjunction with DFT calculations and collision cross-section simulations to investigate the structure of the isolated Ni 3 S 3 H(PEt 3 ) 5 . Theoretical studies show that [Ni 3 S 3 H(PEt 3 ) 5 ] + has a planar Ni 3 S 3 core where three Ni atoms are arranged in a triangle with three bridging S atoms residing in the same plane. This structure is preserved in both solution and solid phases, which is confirmed by spectroscopic studies of Ni 3 S 3 H(PEt 3 ) 5 . Additionally, DFT calculations indicate that all spins at the Ni sites are aligned parallel, confirming the presence of ferromagnetic coupling. Overall, this study provides key insights into the structure and magnetic properties of Ni 3 S 3 H(PEt 3 ) 5 , which will facilitate the design of new NC-based magnetic materials.

Nickel sulfide nanocluster↗

Microscale Metal Additive Manufacturing by Solid‐State Impact Bonding of Shaped Thin Films

The deposition of device-grade inorganic materials is one key challenge toward the implementation of additive manufacturing (AM) in microfabrication, and to that end, a broad range of physico-chemical principles has been explored for 3D fabrication with micro- and nanoscale resolution. Yet, for metals, a process that achieves material quality rivalling that of established thin-film deposition methods, and at the same time, has the potential to combine high throughput production with a broad palette of processable materials, is still lacking. Here, the kinetic, solid-state bonding of metal thin films for the additive assembly of high-purity, high-density metals with micrometer-scale precision is introduced. Indirect laser ablation accelerates micrometer-thick gold films to hundreds of meters per second without their heating or ablation. Their subsequent impact on the substrate above a critical velocity forms a permanent, metallic bond in the solid state. Stacked layers are of high density (>99%). By defining thin-film layers with established lithographic methods prior to launch, a variable feature size (2–50 µm), arbitrary shape of bonded layers, and parallel transfer of up to 36 independent film units in a single shot, is demonstrated. Thus, the solid-state kinetic bonding principle as a viable and potentially versatile route for micro-scale AM of metals is established.

3D printing↗

Understanding the Effects of Inhomogeneities at the Back Interface of CdTe‐Based Solar Cells Using 2D Modeling

One-dimensional modeling cannot capture lateral inhomogeneities in CdTe-based devices. Here, we use 2D modeling to investigate the role of varying energetics at the back interface. We consider improvements in the back interface layer (BIL) through either reducing back surface recombination velocity (BSRV) or decreasing the downward band bending near the back interface. We show that when the BSRV is reduced, but strong downward band bending remains, there is no change in the device performance until the BSRV of 90% of the back interface is improved by the BIL. On the other hand, any coverage with a BIL that improves band bending results in device improvements. We use band bending, back interface recombination current densities, and voltage dependent current flow through the device to understand these improvements. The modeling shows that lateral flow of carriers greatly affects device performance, which is not captured in parallel diode modeling, and demonstrates improved understanding with 2D modeling.

2D modelling↗

AMR-Wind: A Performance-Portable, High-Fidelity Flow Solver for Wind Farm Simulations

We present AMR-Wind, a verified and validated high-fidelity computational-fluid-dynamics code for wind farm flows. AMR-Wind is a block-structured, adaptive-mesh, incompressible-flow solver that enables predictive simulations of the atmospheric boundary layer and wind plants. It is a highly scalable code designed for parallel high-performance computing with a specific focus on performance portability for current and future computing architectures, including graphical processing units (GPUs). In this paper, we detail the governing equations, the numerical methods, and the turbine models. Establishing a foundation for the correctness of the code, we present the results of formal verification and validation. The verification studies, which include a novel actuator line test case, indicate that AMR-Wind is spatially and temporally second-order accurate. The validation studies demonstrate that the key physics capabilities implemented in the code, including actuator disk models, actuator line models, turbulence models, and large eddy simulation (LES) models for atmospheric boundary layers, perform well in comparison to reference data from established computational tools and theory. We conclude with a demonstration simulation of a 12-turbine wind farm operating in a turbulent atmospheric boundary layer, detailing computational performance and realistic wake interactions.

17 WIND ENERGY↗

Lessons Learned and Scalability Achieved When Porting Uintah to DOE Exascale Systems

A key challenge faced when preparing codes for Department of Energy (DOE) exascale systems was designing scalable applications for systems featuring hardware and software not yet available at leadership-class scale. With such systems now available, it is important to evaluate scalability of the resulting software solutions on these target systems. One such code designed with the exascale DOE Aurora and DOE Frontier systems in mind is the Uintah Computational Framework, an open-source asynchronous many-task (AMT) runtime system. To prepare for exascale, Uintah adopted a portable MPI+X hybrid parallelism approach using the Kokkos performance portability library (i.e., MPI+Kokkos). This paper complements recent work with additional details and an evaluation of the resulting approach on Aurora and Frontier. Results are shown for a challenging benchmark demonstrating interoperability of 3 portable codes essential to Uintah-related combustion research. These results demonstrate single-source portability across Aurora and Frontier with scaling characteristics shown to 3,072 Aurora nodes and 9,216 Frontier nodes. In addition to showing results run to new scales on new systems, this paper also discusses lessons learned through efforts preparing Uintah for exascale systems.

Holmen, John [ORNL] (ORCID:0000000259342641)↗

Flexible User-Defined Domain Decomposition in Kilometer-Scale E3SM Land Model Simulation

The Energy Exascale Earth System Model (E3SM) Land Model (ELM) has been extended to kilometer-scale (km-ELM) resolutions, enabling high-fidelity simulations of terrestrial processes at 1 km x 1 km grid spacing. In ELM, domain decomposition partitions the computational domain across processors, ensuring efficient parallel execution. Currently, round-robin decomposition is applied, providing a straightforward way to distribute computational workload. As ELM continues evolving at the kilometer-scale (km-scale), particularly with integrating lateral flow modeling, decomposition strategies must also account for the increased workload and data movement. This paper introduces a flexible user-defined domain decomposition framework, allowing users to customize domain partitioning based on application requirements. The impact of different decomposition strategies is evaluated across various applications concerning computation, communication, and I/O. Results demonstrate that while 1D partitioning yields superior I/O performance, k-nearest neighbors (KNN) clustering effectively reduces inter-process communication overhead. This study lays the groundwork for scalable partitioning in large-scale land surface simulations, enhancing next-generation Earth system modeling.

Wang, Dali [ORNL] (ORCID:0000000168065108)↗

Priority-BF: A Task Manager for Priority-Based Scheduling

The increasing demand for computational resources, particularly in High-Performance Computing environments, necessitates to rethink how we handle job scheduling strategies. This work addresses the challenge of managing concurrent jobs with differing priorities on overloaded parallel systems, where strict QoS constraints are often difficult for users to define. Our solution relies on a qualitative description of priorities and pulls from two key approaches: the Easy-BF algorithm and the Conservative Backfilling algorithms. This solution improves the response time for high-priority jobs by 50% without affecting the overall system utilization. We show its applicability in several critical scenarios such as High-Performance Computing (HPC) resource management and in-situ computing.

Gainaru, Ana [ORNL]↗

Symbol alphabets in QCD and flag cluster algebras

The full 245-letter symbol alphabet for all planar massless two-loop six-point Feynman integrals was recently determined in arXiv:2412.19884 and arXiv:2501.01847. In a parallel mathematical development, it was shown in arXiv:2408.14956 that there is an embedding of the cluster algebra associated to the partial flag variety $\mathcal{Fl}$ $2,n-2;n$ , which describes the kinematics of n massless particles, into that of the Grassmannian Gr(n–2, 2n–4). In this paper we connect these developments by showing that most of the rational symbol letters can be expressed in terms of flag cluster variables, and that all of the algebraic symbol letters arise from infinite mutation sequences.

97 MATHEMATICS AND COMPUTING↗

Trigonometric continuous-variable gates and hybrid quantum simulations of the sine-Gordon model

Hybrid qubit-qumode quantum computing platforms provide a natural setting for simulating interacting bosonic quantum field theories. However, existing continuous-variable gate constructions rely predominantly on polynomial functions of canonical quadratures. In this work, we introduce a complementary universality paradigm based on trigonometric continuous-variable gates, which enable a Fourier-like representation of bosonic operators and are particularly well suited for periodic and non-perturbative interactions. We present an ancilla-based framework for implementing trigonometric gates with arguments given by arbitrary Hermitian functions of qumode quadratures. The protocol yields unitary gates deterministically, and non-unitary gates through probabilistic post-selection. As a concrete application, we develop a hybrid qubit-qumode quantum simulation of the lattice sine-Gordon model. Using these gates, we prepare ground states via quantum imaginary-time evolution, simulate real-time dynamics, compute time-dependent vertex two-point correlation functions, and extract quantum kink profiles under topological boundary conditions. Our results demonstrate that trigonometric continuous-variable gates provide a physically natural framework for simulating interacting field theories on near-term hybrid quantum hardware, while establishing a parallel route to universality beyond polynomial gate constructions. We expect that the trigonometric gates introduced here to find broader applications, including quantum simulations of condensed matter systems, quantum chemistry, and biological models.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗