Search NASA⌕ Search

SEARCH · Search NASA

Results for “Parallel in time”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 361 records · Page 20

Microscale Metal Additive Manufacturing by Solid‐State Impact Bonding of Shaped Thin Films

The deposition of device-grade inorganic materials is one key challenge toward the implementation of additive manufacturing (AM) in microfabrication, and to that end, a broad range of physico-chemical principles has been explored for 3D fabrication with micro- and nanoscale resolution. Yet, for metals, a process that achieves material quality rivalling that of established thin-film deposition methods, and at the same time, has the potential to combine high throughput production with a broad palette of processable materials, is still lacking. Here, the kinetic, solid-state bonding of metal thin films for the additive assembly of high-purity, high-density metals with micrometer-scale precision is introduced. Indirect laser ablation accelerates micrometer-thick gold films to hundreds of meters per second without their heating or ablation. Their subsequent impact on the substrate above a critical velocity forms a permanent, metallic bond in the solid state. Stacked layers are of high density (>99%). By defining thin-film layers with established lithographic methods prior to launch, a variable feature size (2–50 µm), arbitrary shape of bonded layers, and parallel transfer of up to 36 independent film units in a single shot, is demonstrated. Thus, the solid-state kinetic bonding principle as a viable and potentially versatile route for micro-scale AM of metals is established.

3D printing↗

PyJMAK: An Open-Source Python Toolkit for Modeling Solid-State Metallurgical Phase Transformations

Accurate prediction of metallurgical phase transformations is an essential basis for autonomous optimization and rapid part qualification. Several methods can be used to estimate the evolution of phase fractions such as JMAK kinetics-based models, phase-field models, thermodynamic models, and data-driven machine learning models. Thermodynamic and phase-field-based methodologies solve multiphysics equations requiring numerous calibration parameters and significant computational resources. As a result, the computation domain is limited to a point or on order of micron-meters. The data-driven models rely on large datasets from experiments and simulations. While the JMAK model only provides information about phase fraction evolution, it can predict this evolution in near real-time using thermal history and thermodynamic data without restriction on the domain. JMAK models have been popularly used by researchers to model phase transformations occuring during additive manufacturing or over arbitrary temperature profiles. Commercial proprietary software such as Abaqus and Ansys or closed-source in-house implementations offer the ability to model JMAK based kinetics to predict phase transformation. However, these software packages are not open-source or freely available for use and development in conjunction with manufacturing machines, sensors, and machine learning algorithms. In addition, the use of the model is restricted by a license token. In contrast, given temperature profiles at multiple points in the domain, this Python-based PyJMAK model can compute phase evolution in parallel due to its stand-alone modular, voxel-based structure, and it can be executed on high-performance computing resources without any license restrictions.

Prabhune, Bhagya [Oak Ridge National Laboratory (O↗

Designing a Framework for Solving Multiobjective Simulation Optimization Problems

Multiobjective simulation optimization (MOSO) problems are optimization problems with multiple conflicting objectives, where evaluation of at least one of the objectives depends on a black-box numerical code or real-world experiment, which we refer to as a simulation. Whereas an extensive body of research is dedicated to developing new algorithms and methods for solving these and related problems, it is challenging and time-consuming to integrate these techniques into real-world production-ready solvers. This is partly because of the diversity and complexity of modern state-of-the-art MOSO algorithms and methods and partly because of the complexity and specificity of many real-world problems and their corresponding computing environments. The complexity of this problem is only compounded when introducing potentially complex and/or domain-specific surrogate-modeling techniques, problem formulations, design spaces, and data acquisition functions. Here, this paper carefully surveys the current state of the art in MOSO algorithms, techniques, and solvers, as well as problem types and computational environments where MOSO is commonly applied. We then present several key challenges in the design of a parallel multiobjective simulation optimization framework (ParMOO) and how they have been addressed. Finally, we provide two case studies demonstrating how customized ParMOO solvers can be quickly built and deployed to solve real-world MOSO problems.

engineering design optimization↗

Characterization of DESI fiber assignment incompleteness effect on 2-point clustering and mitigation methods for DR1 analysis

We present an in-depth analysis of the fiber assignment incompleteness in the Dark Energy Spectroscopic Instrument (DESI) Data Release 1 (DR1). This incompleteness is caused by the restricted mobility of the robotic fiber positioner in the DESI focal plane, which limits the number of galaxies that can be observed at the same time, especially at small angular separations. As a result, the observed clustering amplitude is suppressed in a scale-dependent manner, which, if not addressed, can severely impact the inference of cosmological parameters. We discuss the methods adopted for simulating fiber assignment on mocks and data. In particular, we introduce the fast fiber assignment (FFA) emulator, which was employed to obtain the power spectrum covariance adopted for the DR1 full-shape analysis. We present the mitigation techniques, organised in two classes: measurement stage and model stage. We then use high fidelity mocks as a reference to quantify both the accuracy of the FFA emulator and the effectiveness of the different measurement-stage mitigation techniques. This complements the studies conducted in a parallel paper for the model-stage techniques, namely the θ-cut approach. We find that pairwise inverse probability (PIP) weights with angular upweighting recover the “true” clustering in all the cases considered, in both Fourier and configuration space. Notably, we present the first ever power spectrum measurement with PIP weights from real data.

cosmological simulations↗

A Performance Portable, Fully Implicit Landau Collision Operator with Batched Linear Solvers

Modern accelerators use hierarchical parallel programming models that enable massive multithreading within a processing element (PE), with multiple PEs per device driven by traditional processes. Batching is a technique for exposing PE-level parallelism in algorithms that have traditionally run on MPI processes or multiple threads within a single process. Opportunities for batching arise in, for example, kinetic discretizations of magnetized plasmas where collisions are advanced in velocity space at each spatial point independently. This paper builds on previous work on a high-performance, fully nonlinear, Landau collision operator by batching the linear solver, as well as batching the spatial point problems and adding new support for multiple grids for multiscale, multispecies problems. An anisotropic relaxation verification test that agrees well with previously published results and analytical models is presented. The performance results from NVIDIA A100 and AMD MI250X nodes are presented with hardware utilization analysis for each architecture. Finally, the entire implicit Landau operator time advance is implemented in Kokkos for performance portability, running entirely on the device and is available in the PETSc numerical library.

97 MATHEMATICS AND COMPUTING↗

Modeling of hepatitis B virus infection spread in primary human hepatocytes

ABSTRACT Chronic hepatitis B virus (HBV) infection poses a significant global health threat, causing severe liver diseases including cirrhosis and hepatocellular carcinoma. We characterized HBV DNA kinetics in primary human hepatocytes (PHHs) over 32 days post-inoculation (p.i.) and modified ourin-vivoagent-based modeling (ABM) to gain insights into the HBV lifecycle and spreadin vitro. Parallel PHH cultures were mock-treated or treated with HBV entry inhibitor Myr-preS1 (6.25 µg/mL) was initiated 24 h p.i. In untreated PHH, three viral DNA kinetic patterns were identified: (i) an initial decline, followed by (ii) rapid amplification and (iii) slower amplification/accumulation. In the presence of Myr-preS1, viral DNA and infected cell numbers in phase 3 were effectively blocked, with minimal to no increase. This suggests that phase 2 represents viral amplification in initially infected cells, while phase 3 corresponds to viral spread to naïve cells. The ABM reproduced well the HBV kinetic patterns observed and predicted that the viral eclipse phase lasts between 18 and 38 h. After the eclipse phase, the viral production rate increased over time, starting with a slow production cycle of 1 virion per day, which gradually accelerated to 1 virion per hour after 3 days. Approximately 4 days later, virion production reached a steady state production rate of 4 virions/h. The estimated median efficacy of Myr-preS1 in blocking HBV spread was 91% (range: 90–92%). The HBV kinetics and the predicted estimates of the HBV eclipse phase duration and HBV production cycles in PHH are similar to those predicted in uPA/SCID mice with human livers. IMPORTANCE While primary human hepatocytes (PHHs) are the most physiologically relevant culture system for studying HBV infectionin vitro, a comprehensive understanding of HBV infection kinetics and spread in PHH is lacking. In this study, we characterize HBV viral kinetics and modify ourin vivoagent-based modeling (ABM) to provide quantitative insights into the HBV production cycle and viral spread in PHH. The ABM provides an estimate of the HBV eclipse phase duration, HBV production cycles, and Myr-preS1 efficacy in blocking HBV spread in PHH. The results resemble those predicted in uPA/SCID mice with human livers, demonstrating that estimated HBV infection kinetic parameters in PHHin vitromirror those observed in thein vivoHBV infection chimeric mouse model.

Virology↗

Particle-based modelling of axisymmetric tandem mirror devices

In this work, we describe the use of a 1D-2V quasi-neutral hybrid electrostatic PIC with Monte-Carlo Coulomb collisions and non-uniform magnetic field to model the parallel transport and confinement in an axisymmetric tandem mirror device. End-plugs, based on simple-mirrors, are positioned at each end of the device and fueled with neutral beams (25 and 100 keV) to produce a sloshing ion population and increase the density of the end-plugs relative to the central cell. Results show the formation of a potential difference barrier between the central cell and the end-plugs. This potential confines a large fraction of the low energy thermal ions in the central cell which would otherwise be lost in a simple mirror, demonstrating the advantage of the beam-driven tandem mirror configuration relative to simple mirrors. In addition, we explore the effect of end-plug electron temperature on the confinement time of the device and compare it with theoretical estimates. Finally, we discuss the limitations of the code in its present form and describe the next logical steps to improve its predictive capability such as a fully nonlinear Fokker–Planck collision operator, multiply nested flux surface solutions and modeling the exhaust region up to the wall.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Thermodynamics and collisionality in firehose-susceptible high- β plasmas

We study the evolution of collisionless plasmas that, due to their macroscopic evolution, are susceptible to the firehose instability, using both analytic theory and hybrid-kinetic particle-in-cell simulations. We establish that, depending on the relative magnitude of the plasma β, the characteristic time scale of macroscopic evolution and the ion-Larmor frequency, the saturation of the firehose instability in high-β plasmas can result in three qualitatively distinct thermodynamic (and electromagnetic) states. By contrast with the previously identified ‘ultra-high-beta’ and ‘Alfvén-inhibiting’ states, the newly identified ‘Alfvén-enabling’ state, which is realised when the macroscopic evolution time τ exceeds the ion-Larmor frequency by a β-dependent critical parameter, can support linear Alfvén waves and Alfvénic turbulence because the magnetic tension associated with the plasma’s macroscopic magnetic field is never completely negated by anisotropic pressure forces. We characterise these states in detail, including their saturated magnetic-energy spectra. The effective collision operator associated with the firehose fluctuations is also described; we find it to be well approximated in the Alfvén-enabling state by a simple quasi-linear pitch-angle scattering operator. The box-averaged collision frequency is ν eff ∼ β/τ, in agreement with previous results, but certain subpopulations of particles scatter at a much larger (or smaller) rate depending on their velocity in the direction parallel to the magnetic field. Our findings are essential for understanding low-collisionality astrophysical plasmas including the solar wind, the intracluster medium of galaxy clusters and black hole accretion flows. We show that all three of these plasmas are in the Alfvén-enabling regime of firehose saturation and discuss the implications of this result.

astrophysical plasmas↗

QCD–Gravity Double Copy in Regge Asymptotics: From \(2\rightarrow n\) Amplitudes to Radiation in Shockwave Collisions

This paper discusses multi-particle production in QCD and in gravity at ultrarelativistic energies, their double-copy relations, and strong parallels in emergent shockwave dynamics. Dispersive techniques are applied to derive the BFKL equation for multi-gluon production in Regge asymptotics. Identical methods apply in gravity and are captured by a gravitational Lipatov equation. The building blocks in both cases are Lipatov vertices and reggeized propagators satisfying double-copy relations; in gravity, Weinberg’s soft theorem is recovered as a limit of the Lipatov framework. BFKL evolution in QCD generates wee parton states of maximal occupancy characterized by an emergent semi-hard saturation scale. Renormalization group equations in the Color Glass Condensate (CGC) EFT describe wee parton correlations and their rapidity evolution. A shockwave picture of deeply inelastic scattering and hadron–hadron collisions follows, with multi-particle production described by Cutkosky’s rules in strong time-dependent fields. Gluon radiation in the CGC EFT has a double copy in gravitational shockwave collisions, with a similar correspondence applicable between gluon and graviton shockwave propagators. Possible extensions of this semi-classical double copy are outlined for computing multi-particle production in gravitational shockwave collisions, self-force and tidal contributions, and classical and quantum noise in the focusing of geodesics.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Adaptive Protection and Validated Models to Enable Deployment of High Penetrations of Solar PV (PV-MOD)

The availability and validation of various PV models in commercial tools differ, with some models not yet thoroughly validated for advanced inverter functionalities and reliable performance under weak system conditions. Many existing models do not fully incorporate new inverter control functions, which can affect system stability. The increasing deployment of solar PV and other inverter-based resources (IBRs), including distributed energy resources (DERs), is influencing the reliable operation of protection schemes in distribution systems and microgrids. Emerging adaptive protection schemes (APS) offer new opportunities for protecting these systems during varying configurations and DER operating conditions, though their demonstration and validation remain limited. Adaptive protection schemes face similar challenges, as they are typically designed for specific configurations. There is a growing need for tools and methodologies to streamline the deployment of adaptive protection for safe and reliable DER integration. The project main objective was to develop and validate high-fidelity generic models of solar PV facilities for stability, protection, EMT, and QSTS analyses. This objective was achieved, and these models can now be integrated into commercial software tools, enabling utilities, vendors, and developers to study high-penetration PV systems more confidently. The project also demonstrated advanced applications of these models, including the design and deployment of adaptive protection schemes in high-penetration field applications and microgrids, supporting grid safety and reliability. Several milestones were reached by the end of the project. A sophisticated inverter test plan was developed, and inverters representative of the North American marketplace were selected. EPRI and NREL tested various inverters, conforming to IEEE standards. Improvements were made to existing generic models of IBR units, IBR plants, and aggregated feeders for various analyses. The first generic electromagnetic transient (EMT) model for a solar PV plant was developed, conforming to IEEE Std 2800™-2022 and validated against laboratory measurements of a 2.2 MVA large-scale battery energy storage system (BESS) inverter. That model was then used to produce reference responses illustrating examples of validated and verified IBR plant models that pass or fail tests for technical minimum capability and performance as specified in the IEEE standard. The developed, tested, and validated generic models can be used for transmission planning, stability assessments, expansion planning, and evaluating potential future IBR interconnection requirements. They can also support interconnection screens and conformity assessments of IBR plants, including solar PV. The project significantly contributed to the ongoing standardization and model-based representation and verification of IBR responses. The project further addressed challenges of common distribution protection schemes with increasing deployment of DER by developing, validating, and demonstrating adaptive protection schemes (APS) that can improve the reliable and safe integration of DER into distribution systems. New APS were designed using improved DER models for three common distribution systems: a radial feeder, a meshed network, and a microgrid. Modeling and hardware-in-the-loop (HIL) testing of the APS were conducted, successfully showing their effectiveness and selectivity. Proof-of-concept field demonstration was achieved for two APS, i.e., one on a radial feeder and another one in a microgrid. Field demonstration could not be achieved for the APS on a meshed network, primarily due apprehension of one utility partner and also due to limited access to the protective algorithms in the network protectors. Guidelines developed from the lessons learned in the project lay out the general process followed in the design, installation, and commissioning of APS for various distribution systems. Distribution utility partners’ apprehension about field demonstration of the new APS were addressed—with varying success—by taking a stepped risk-management approach of modeling of a wide range of sensitivities first, performing in-depth proof-of-concept testing in the laboratory including HIL next, and finally deliberately implementing and commissioning the actual protection equipment and algorithms into parts of—or in parallel operation to—the three real distribution systems. Future work should include pilot projects that further show the acceptable performance of the developed APS before these schemes be rolled out more widely. Inclusion of both utility and original equipment manufacturers (OEMs) in future projects could increase chances of successful field demonstration. Despite challenges in achieving the field demonstration goal of the project for all three APS, the research significantly contributed to the innovation of adaptive protection solutions for scalable and reliable DER integration into distribution systems. This project significantly enhances the understanding of the impact of using appropriate inverter models on distribution and transmission (T&D) systems. By addressing the limitations of existing generic models, the project introduces high-fidelity models for stability, protection, electromagnetic transient (EMT), and quasi-static time series (QSTS) analyses. These models, integrated into commercial software tools, enable utilities, vendors, and developers to confidently study high-penetration PV systems. The project also demonstrates advanced applications, including adaptive protection schemes (APS) for distribution systems and microgrids, ensuring grid safety and reliability. The technical effectiveness and economic feasibility of the methods are evident through the development and validation of sophisticated inverter test plans and the selection of representative inverters. Testing by EPRI and NREL on retail, commercial, and utility-scale inverters, conforming to IEEE standards, underscores the robustness of the models. Improvements to existing generic models for various analyses further enhance their validity and applicability. The project also identifies gaps in common distribution protection schemes and designed new APS using improved DER models, demonstrating their effectiveness through modeling and hardware-in-the-loop (HIL) testing. The project’s benefits to the public are manifold. By advancing the standardization and model-based representation of IBR response, it supports transmission planning, stability assessments, and future IBR interconnection requirements. The generic models can facilitate better communication between transmission planners and developers, supporting expected IBR plant capability and performance. Additionally, the development of APS for radial feeders, meshed networks, and microgrids supports the integration of distributed energy resources (DERs) into distribution systems, enhancing grid reliability and safety. The project’s emphasis on thorough testing and simplicity in design ensures practical and scalable solutions for DER integration.

14 SOLAR ENERGY↗

Evaluation of DED and LPBF Fe-based Alloys Process Application Envelopes based on Performance, Process Economics, Supply Chain Risks, and Reactor-specific Targeted Components

The U.S. Department of Energy (DOE), Office of Nuclear Energy (NE), Advanced Materials and Manufacturing Technologies (AMMT) program aims to develop extreme-environment materials solutions for use in the deployment of advanced nuclear reactors and the sustainment of the current fleet. To achieve this objective, a combination of experiment, a computational tool, and machine learning (ML) for the design of materials is adopted for the maturation of materials for nuclear technology. Through advanced manufacturing techniques such as laser powder bed fusion (LPBF) and laser powder direct energy deposition (LP-DED), components with complex geometries can be fabricated with reduced time and effort. Such advanced manufacturing methods can also provide the opportunity to improve materials performance through optimized microstructures and mechanical properties. However, existing engineering alloys are not always well suited for fabrication with additive manufacturing (AM), as their compositions have been tuned to optimize fabrication via conventional methods. Thus, similar alloys with modified compositions that are better suited for AM can be studied for improved performance. Over the past three years, the AMMT teams from Argonne National Laboratory (ANL) and Pacific Northwest National Laboratory (PNNL) studied various known Fe-based alloys by evaluating their initial printability using LPBF, and an AMMT-developed down-selection and decision matrix reduced the number of alloys to be studied from six to three in fiscal year (FY) 2024. Additionally, in FY 2024, for parallel evaluation, these three alloys were studied using LPDED. While LPBF is better for small- to medium-sized components with high detail and internal features, LP-DED combines a material feed system to place the powder onto the exact spot where the laser will melt the material. This AM method can be easily scaled to extremely large components and provides high build rate speeds compared to those of conventional LPBF systems. Additionally, DED is a better choice for complex geometries and compositional gradients.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Wehnelt photoemission in an ultrafast electron microscope: Stability and usability

We tested and compared the stability and usability of three different cathode materials and configurations in a thermionic-based ultrafast electron microscope: (1) on-axis thermionic and photoemission from a custom 100 μm diameter LaB 6 source with a graphite guard ring, (2) off-axis photoemission from the Ni aperture surface of the Wehnelt electrode, and (3) on-axis thermionic and photoemission from a custom 200 μm diameter polycrystalline Ta source. For each cathode type and configuration, including the Ni Wehnelt aperture, we illustrate how the photoelectron beam-current stability is deleteriously impacted by simultaneous cooling of the source following thermionic heating. Furthermore, we demonstrate usability via collection of parallel- and convergent-beam electron diffraction patterns and by formation of the optimum probe size. We find that usability of the off-axis Ni Wehnelt-aperture photoemission is at least comparable to on-axis LaB 6 thermionic emission, as well as to on-axis photoemission [the heretofore conventional approach to ultrafast electron microscopy (UEM) in thermionic-based instruments]. However, the stability and achievable beam currents for off-axis photoemission from the Wehnelt aperture were superior to that of the other cathode types and configurations, regardless of the electron-emission mechanism. Beam-current stability for this configuration was found to be ±1% (one standard deviation from the mean) for 70 min (longest duration tested), and steady-state beam current was reached within the sampling-time resolution used here (~1 s) for 15 pA beam currents (i.e., 460 electrons per packet for a 200 kHz repetition rate). Repeatability and robustness of the steady-state condition were also found to be within ±1% of the mean. We discuss the implications of these findings for UEM imaging and diffraction experiments, for pulsed-beam damage measurements, and for practical switching between optimum conventional TEM and UEM operation within the same instrument.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Streamlined spatial and environmental expression signatures characterize the minimalist duckweed Wolffia australiana

Single-cell genomics permits a new resolution in the examination of molecular and cellular dynamics, allowing global, parallel assessments of cell types and cellular behaviors through development and in response to environmental circumstances, such as interaction with water and the light–dark cycle of the Earth. Here, we leverage the smallest, and possibly most structurally reduced, plant, the semiaquaticWolffia australiana, to understand dynamics of cell expression in these contexts at the whole-plant level. We examined single-cell-resolution RNA-sequencing data and foundWolffiacells divide into four principal clusters representing the above- and below-water-situated parenchyma and epidermis. Although these tissues share transcriptomic similarity with model plants, they display distinct adaptations thatWolffiahas made for the aquatic environment. Within this broad classification, discrete subspecializations are evident, with select cells showing unique transcriptomic signatures associated with developmental maturation and specialized physiologies. Assessing this simplified biological system temporally at two key time-of-day (TOD) transitions, we identify additional TOD-responsive genes previously overlooked in whole-plant transcriptomic approaches and demonstrate that the core circadian clock machinery and its downstream responses can vary in cell-specific manners, even in this simplified system. Distinctions between cell types and their responses to submergence and/or TOD are driven by expression changes of unexpectedly few genes, characterizingWolffiaas a highly streamlined organism with the majority of genes dedicated to fundamental cellular processes.Wolffiaprovides a unique opportunity to apply reductionist biology to elucidate signaling functions at the organismal level, for which this work provides a powerful resource.

Biochemistry & Molecular Biology↗

DIF3D-VARIANT 12.0: Updates and New Features

The DIF3D code has been a workhorse of fast reactor analysis work at Argonne National Laboratory for over 40 years. In 1995, a transport option called VARIANT was added to DIF3D to improve the flux solutions for fast reactor problems which we term DIF3D-VARIANT today. DIF3D-VARIANT performs nodal neutron transport calculations using P N or SP N theory in Cartesian and hexagonal two- and three-dimensional geometries. The limited computing capabilities of the time restricted DIF3D-VARIANT to use at most a 6 th order spatial approximation combined with a P3 flux approximation and P1 scattering kernel for a 33 group structure on most studied reactor problems. Computer capabilities have increased steadily since 1995 and today much larger space-angle-energy approximations are possible. This manuscript serves as an update to the theory section of the original DIF3D-VARIANT manual and details more than twenty years of changes made to DIF3D to make version 12 which was released on November 1 st , 2024. The primary focus of the initial work was to extend the space-angle approximations available in DIF3D-VARIANT such that the error due to transport approximations could be better understood. This work was started and completed in 2002 and marked the official version 10. Unfortunately, those higher order approximations could not be used at that time due to the memory constraints of the BPOINTER part of DIF3D (limited to 2 GB). In version 11, completed in 2012, BPOINTER was circumvented in DIF3D-VARIANT for the largest arrays by introducing a Fortran 90 module called LMA (Large Memory Array). This seamlessly replaces all of the functionality of the BPOINTER concept, but it allows 64 bit addressing for every array such that they can be larger than 2 GB. It is now common for DIF3D-VARIANT jobs to consume 50 GB of memory on modern workstations when using high order space-angle approximations and a large number of groups. Many improvements were made to version 11 from 2012 to 2022 when work to create version 12 started. For version 12, several parts of DIF3D were updated to improve performance and thread parallelism was introduced to further reduce the runtime. Numerous minor bugs were discovered in DIF3D-VARIANT as part of the process of creating the perturbation and sensitivity code PERSENT. All of these algorithmic problems were identified in the transition from version 10 to version 11 which prevented DIF3D-VARIANT from running efficiently and reliably. Firstly, the coarse mesh rebalance scheme would routinely diverge and a study detailed in this report demonstrates how it was also typically not effective. This is not a failure of the coarse mesh rebalance methodology, but a failure of its implementation in DIF3D-VARIANT for hexagonal geometries. The fission source extrapolation algorithm was also found to be unreliable on larger group structure problems, leading to divergence in some cases and a negligible improvement in performance overall. Finally, the “Omega” acceleration applied to the partial current solver routine of DIF3D-VARIANT was found to cause DIF3D-VARIANT to converge to the wrong answer. To resolve these issues, both the coarse mesh rebalance and fission source extrapolation were permanently disabled in version 11. The Tchebychev acceleration was put in as a temporary reliable alternative but it is generally inferior to coarse mesh rebalance or coarse mesh finite difference. For the Omega acceleration, the factor was restricted to guarantee that it would not cause follow-on errors in PERSENT. Due to limited funding to support maintenance and development of DIF3D in the last 10 years, no effort was spent since to resolve the outer iteration acceleration. Except for the threading work, all of the changes discussed in this manuscript refer to changes made between version 10 and version 11. Performance comparisons are done to demonstrate the improvements from version 9 to version 12. As will be demonstrated, the updated versi

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Coriolis forces modify magnetostatic ponderomotive potentials

It is possible to produce a ponderomotive effect in a plasma system without time-varying fields, if the plasma flows over spatial oscillations in the field. This can be achieved by superimposing a spatially oscillatory perturbation on a guide field, then setting up an electric field perpendicular to the guide field to drive flow over the perturbation. However, subtle distinctions in the structure of the resulting electric field can entirely change the behavior of the resulting ponderomotive force. Previous work has shown that, in slab models, these distinctions can be explained in terms of the polarization of the effective wave that appears in the co-moving frame. Here, we consider what happens to this picture in a cylindrical system, where the transformation to the co-moving (rotating) frame is not inertial. It turns out that the non-inertial nature of this frame transformation can lead to counterintuitive behavior, partly due to the appearance of parallel (magnetic-field-aligned) electric fields in the rotating frame even in cases where none existed in the laboratory frame. Apart from the academic interest of this study, the practical impact lies in being better able to anticipate the antenna configuration on the plasma periphery of a cylindrical plasma that will lead to optimal ponderomotive barrier formation in the interior plasma.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Stage-local partitioned two-step runge-kutta methods for large systems of ordinary differential equations

We introduce stage-local partitioned two-step Runge-Kutta methods are an extension of standard two-step Runge-Kutta methods, which are an alternative to the standard additive two-step Runge-Kutta methods currently existing in the literature. Furthermore, these new schemes are designed with an eye towards truly N-partitioned systems and leverage local stage approximations to make several computationally interesting approximations viable. Specifically, the focus on local stage approximations makes possible the construction of truly asynchronous schemes, in the parallel sense, possible. In addition, we show that an implicit-explicit approach to these schemes can lead to methods that require the inversion of only local nonlinear systems.

Applied Dynamical Systems↗

Two-Tower Quantum Matrix Chain Multiplication: Trading Qubits for Depth

Matrix chain multiplication -- computing $\mathcal{W} = M^{(0)}\cdots M^{(K-1)}$ where $M^{(k)} \in \mathbb{R}^{P_k \times P_{k+1}}$-- arises in scientific computing, machine learning, and graph analysis. Despite the importance of this problem, for chains of distinct matrices, the classical number of operations grows linearly with the chain length $K$ and polynomially in the matrix dimensions. We present \emph{Two-Tower Matrix Multiplication}, a quantum subroutine that encodes the product $\mathcal{W}$ of the $K$ matrices into a quantum state in circuit depth $\mathcal{O}(\max_{k} \mathrm{polylog} (P_k P_{k+1}))$, which is independent of~$K$ within the QRAM-based state-preparation model, whereas the qubit count is $\mathcal{O}\bigl(\sum_{k} \log P_k \bigr)$; the total gate count remains linear in $K$, so the gain is in the circuit depth. The construction interleaves state-preparation operators across two layers; within each layer, all operators act on disjoint registers and execute in parallel. This subroutine can be specialized for the chain-vector case, which computes the product of $K-1$ matrices applied to a vector. We prove the correctness of the subroutine for all $K$ and provide two implementations using the Qiskit and QCLAB frameworks. The subroutine is applicable to any downstream quantum algorithm that operates on a matrix encoded in the statevector, including norm estimation, graph-matrix powers, linear system solving, and quantum machine learning kernels.

Antonioli, Giacomo [Pisa U.] (ORCID:00090000668703↗

Sensor Co-design for $\textit{smartpixels}$

Pixel tracking detectors at upcoming collider experiments will see unprecedented charged-particle densities. Real-time data reduction on the detector will enable higher granularity and faster readout, possibly enabling the use of the pixel detector in the first level of the trigger for a hadron collider. This data reduction can be accomplished with a neural network (NN) in the readout chip bonded with the sensor that recognizes and rejects tracks with low transverse momentum (p$_T$) based on the geometrical shape of the charge deposition (``cluster''). To design a viable detector for deployment at an experiment, the dependence of the NN as a function of the sensor geometry, external magnetic field, and irradiation must be understood. In this paper, we present first studies of the efficiency and data reduction for planar pixel sensors exploring these parameters. A smaller sensor pitch in the bending direction improves the p$_T$ discrimination, but a larger pitch can be partially compensated with detector depth. An external magnetic field parallel to the sensor plane induces Lorentz drift of the electron-hole pairs produced by the charged particle, broadening the cluster and improving the network performance. The absence of the external field diminishes the background rejection compared to the baseline by $\mathcal{O}$(10%). Any accumulated radiation damage also changes the cluster shape, reducing the signal efficiency compared to the baseline by $\sim$ 30 - 60%, but nearly all of the performance can be recovered through retraining of the network and updating the weights. Finally, the impact of noise was investigated, and retraining the network on noise-injected datasets was found to maintain performance within 6% of the baseline network trained and evaluated on noiseless data.

Shekar, Danush [Illinois U., Chicago]↗