Search NASA⌕ Search

SEARCH · Search NASA

Results for “Adjoint”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Development of the tangent linear and adjoint models of the global online chemical transport model MPAS-CO 2 v7.3

We describe the development of the tangent linear (TL) and adjoint models of the Model for Prediction Across Scales (MPAS)-CO 2 transport model, which is a global online chemical transport model developed upon the non-hydrostatic Model for Prediction Across Scales – Atmosphere (MPAS-A). The primary goal is to make the model system a valuable research tool for investigating atmospheric carbon transport and inverse modeling. First, we develop the TL code, encompassing all CO 2 transport processes within the MPAS-CO 2 forward model. Then, we construct the adjoint model using a combined strategy involving re-calculation and storage of the essential meteorological variables needed for CO 2 transport. This strategy allows the adjoint model to undertake a long-period integration with moderate memory demands. To ensure accuracy, the TL and adjoint models undergo vigorous verifications through a series of standard tests. The adjoint model, through backward-in-time integration, calculates the sensitivity of atmospheric CO 2 observations to surface CO 2 fluxes and the initial atmospheric CO 2 mixing ratio. To demonstrate the utility of the newly developed adjoint model, we conduct simulations for two types of atmospheric CO 2 observations, namely the tower-based in situ CO 2 mixing ratio and satellite-derived column-averaged CO 2 mixing ratio (X CO 2 ). A comparison between the sensitivity to surface flux calculated by the MPAS-CO 2 adjoint model with its counterpart from CarbonTracker–Lagrange (CT-L) reveals a spatial agreement but notable magnitude differences. These differences, particularly evident for X CO 2 , might be attributed to the two model systems' differences in the simulation configuration, spatial resolution, and treatment of vertical mixing processes. Moreover, this comparison highlights the substantial loss of information in the atmospheric CO 2 observations due to CT-L's spatial domain limitation. Furthermore, the adjoint sensitivity analysis demonstrates that the sensitivities to both surface flux and initial CO 2 conditions spread out throughout the entire Northern Hemisphere within a month. MPAS-CO 2 forward, TL, and adjoint models stand out for their calculation efficiency and variable-resolution capability, making them competitive in computational cost. In conclusion, the successful development of the MPAS-CO 2 TL and adjoint models, and their integration into the MPAS-CO 2 system, establish the possibility of using MPAS's unique features in atmospheric CO 2 transport sensitivity studies and in inverse modeling with advanced methods such as variational data assimilation.

54 ENVIRONMENTAL SCIENCES↗

An Optimization-Based Coupling of Reduced Order Models with an Efficient Reduced Adjoint Basis Generation Approach

Optimization-based coupling (OBC) is an attractive alternative to traditional Lagrange multiplier approaches in multiple modeling and simulation contexts. However, application of OBC to time-dependent problems has been hindered by the computational cost of finding the stationary points of the associated Lagrangian, which requires primal and adjoint solves. This issue can be mitigated by using OBC in conjunction with computationally efficient reduced order models (ROMs). To demonstrate the potential of this combination, in this paper, we develop an optimization-based ROM-ROM coupling for a transient advection-diffusion transmission problem. We pursue the “optimize-then-reduce” path toward solving the minimization problem at each time step and solve reduced space adjoint system of equations, where the main challenge in this formulation is the generation of adjoint snapshots and reduced bases for the adjoint systems required by the optimizer. One of the main contributions of the paper is a new technique for an efficient adjoint snapshot collection for gradient-based optimizers in the context of optimization-based ROM-ROM couplings. In conclusion, we present numerical studies demonstrating the accuracy of the approach along with comparison between various approaches for selecting a reduced order basis for the adjoint systems, including decay of snapshot energy, average iteration counts, and timings.

coupled problems↗

Small circle expansion for adjoint QCD 2 with periodic boundary conditions

We study 1 + 1-dimensional SU(N) gauge theory coupled to one adjoint multiplet of Majorana fermions on a small spatial circle of circumference L. Using periodic boundary conditions, we derive the effective action for the quantum mechanics of the holonomy and the fermion zero modes in perturbation theory up to order (gL) 3 . When the adjoint fermion mass-squared is tuned to g 2 N/(2π), the effective action is found to be an example of supersymmetric quantum mechanics with a nontrivial superpotential. We separate the states into the ℤN center symmetry sectors (universes) labeled by p = 0, . . . , N – 1 and show that in one of the sectors the supersymmetry is unbroken, while in the others it is broken spontaneously. These results give us new insights into the (1, 1) supersymmetry of adjoint QCD 2 , which has previously been established using light-cone quantization. When the adjoint mass is set to zero, our effective Hamiltonian does not depend on the fermions at all, so that there are 2 N−1 degenerate sectors of the Hilbert space. This construction appears to provide an explicit realization of the extended symmetry of the massless model, where there are 2 2N−2 operators that commute with the Hamiltonian. We also generalize our results to other gauge groups G, for which supersymmetry is found at the adjoint mass-squared g 2 h ∨ /(2π), where h ∨ is the dual Coxeter number of G.

effective field theories↗

When ancient numerical demons meet physics-informed machine learning: adjoint-based gradients for implicit differentiable modeling

Recent advances in differentiable modeling, a genre of physics-informed machine learning that trains neural networks (NNs) together with process-based equations, have shown promise in enhancing hydrological models' accuracy, interpretability, and knowledge-discovery potential. Current differentiable models are efficient for NN-based parameter regionalization, but the simple explicit numerical schemes paired with sequential calculations (operator splitting) can incur numerical errors whose impacts on models' representation power and learned parameters are not clear. Implicit schemes, however, cannot rely on automatic differentiation to calculate gradients due to potential issues of gradient vanishing and memory demand. Here we propose a “discretize-then-optimize” adjoint method to enable differentiable implicit numerical schemes for the first time for large-scale hydrological modeling. The adjoint model demonstrates comprehensively improved performance, with Kling–Gupta efficiency coefficients, peak-flow and low-flow metrics, and evapotranspiration that moderately surpass the already-competitive explicit model. Therefore, the previous sequential-calculation approach had a detrimental impact on the model's ability to represent hydrological dynamics. Furthermore, with a structural update that describes capillary rise, the adjoint model can better describe baseflow in arid regions and also produce low flows that outperform even pure machine learning methods such as long short-term memory networks. The adjoint model rectified some parameter distortions but did not alter spatial parameter distributions, demonstrating the robustness of regionalized parameterization. Despite higher computational expenses and modest improvements, the adjoint model's success removes the barrier for complex implicit schemes to enrich differentiable modeling in hydrology.

58 GEOSCIENCES↗

Supercurrents and (partial) supersymmetry in adjoint QCD 2 and its generalizations

1 + 1-dimensional SU(N) gauge theory coupled to an adjoint Majorana fermion, also known as adjoint QCD 2 , has the surprising feature that at fermion mass $\sqrt{\frac{g^2N}{2\pi }}$ it exhibits supersymmetry. In this paper, we obtain a deeper insight into how the supersymmetry works by constructing the gauge invariant, Lorentz covariant supercurrent j μA . Its conservation relies crucially on the presence of a quantum anomaly. We generalize this construction to a class of models where, in addition to an adjoint Majorana fermion of an appropriate mass, the gauge theory is coupled to some collection of massless fermions (SU(N) may be replaced by a more general gauge group). In general, these models have a supersymmetric massive sector and a non-supersymmetric CFT sector [1], but there are cases in which both sectors are supersymmetric. An example of such a gapless, fully supersymmetric model is SU(N) gauge theory coupled to three adjoint Majorana fermions, of which two are massless and the third has mass $\sqrt{\frac{3{g}^2N}{2\pi }}$.

anomalies in field and string theories↗

Adjoint DSMC for nonlinear spatially-homogeneous Boltzmann equation with a general collision model

We derive an adjoint method for the Direct Simulation Monte Carlo (DSMC) method for the spatially homogeneous Boltzmann equation with a general collision law. This generalizes our previous results in Caflisch et al., which was restricted to the case of Maxwell molecules, for which the collision rate is constant. The main difficulty in generalizing the previous results is that a rejection sampling step is required in the DSMC algorithm in order to handle the variable collision rate. We find a new term corresponding to the so-called score function in the adjoint equation and a new adjoint Jacobian matrix capturing the dependence of the collision parameter on the velocities. The new formula works for a much more general class of collision models.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Comparing Adjoint Waveform Tomography Models of California Using Different Starting Models

Abstract Adjoint waveform tomography (AWT) sits at the cutting edge of seismic tomography on local, regional, and global scales. However, the choice in starting model may have a significant impact on the final inversion results. In this paper, we present 3 AWT models of California that are based on different starting models. We chose three models that were inverted at different scales: SPiRaL, a global travel‐time tomography model (Simmons et al., 2021, 10.1093/gji/ggab277 ), CSEM_NA, a regional adjoint tomography model of North America and the North Atlantic (Krischer et al., 2018, 10.1029/2017JB015289 ), and WUS256, a regional adjoint tomography model of the western US (Rodgers et al., 2022, https://doi.org/10.1029/2022JB024549 ). We then inverted three AWT models using the same source and receiver set. We ran each model over three period bands: 30–100 s, 25–100 s, and 20–80 s. Once the iterations were finalized, we used five methods of testing model similarity in both the model and data space. We conclude that the choice of starting model has a minimal impact on long wavelength models if an appropriate multi‐scale inversion approach is used.

58 GEOSCIENCES↗

On Properties of Adjoint Systems for Evolutionary PDEs

We investigate the geometric structure of adjoint systems associated with evolutionary partial differential equations at the fully continuous, semi-discrete, and fully discrete levels and the relations between these levels. We show that the adjoint system associated with an evolutionary partial differential equation has an infinite-dimensional Hamiltonian structure, which is useful for connecting the fully continuous, semi-discrete, and fully discrete levels. We subsequently address the question of discretize-then-optimize versus optimize-then-discrete for both semi-discretization and time integration, by characterizing the commutativity of discretize-then-optimize methods versus optimize-then-discretize methods uniquely in terms of an adjoint-variational quadratic conservation law. For Galerkin semi-discretizations and one-step time integration methods in particular, we explicitly construct these commuting methods by using structure-preserving discretization techniques.

97 MATHEMATICS AND COMPUTING↗

Cascading from $\mathscr{N}$ = 2 supersymmetric Yang–Mills theory to confinement and chiral symmetry breaking in adjoint QCD

We argue that adjoint QCD in 3 + 1 dimensions, with any SU(N) gauge group and two Weyl fermion flavors (i.e. one adjoint Dirac fermion), confines and spontaneously breaks its chiral symmetries via the condensation of a fermion bilinear. We flow to this theory from pure $\mathscr{N}$ = 2 SUSY Yang–Mills theory with the same gauge group, by giving a SUSY-breaking mass M to the scalars in the $\mathscr{N}$ = 2 vector multiplet. This flow can be analyzed rigorously at small M, where it leads to a deconfined vacuum at the origin of the $\mathscr{N}$ = 2 Coulomb branch. The analysis can be extended to all M using an Abelian dual description that arises from the N multi-monopole points of the $\mathscr{N}$ = 2 theory. At each such point, there are N −1 hypermultiplet Higgs fields h$^{i=1,2}_m$, which are SU(2) R doublets. We provide a detailed study of the phase diagram as a function of M, by analyzing the semi-classical phases of the dual using a combination of analytic and numerical techniques. The result is a cascade of first-order phase transitions, along which the Higgs fields h i m successively turn on, and which interpolates between the Coulomb branch at small M, where all h$^{i}_m$ = 0, and a maximal Higgs branch, where all h$^{i}_m$ ≠ 0, at sufficiently large M. We show that this maximal Higgs branch precisely matches the confining and chiral symmetry breaking phase of two-flavor adjoint QCD, including its broken and unbroken symmetries, its massless spectrum, and the expected large-N scaling of various observables. The spontaneous breaking pattern SU(2) R → U(1) R , consistent with the Vafa–Witten theorem, is ensured by an intricate alignment mechanism for the h$^{i}_m$ in the dual, and leads to a CP 1 sigma model of increasing radius along the cascade.

D’Hoker, Eric [Univ. of California, Los Angeles, C↗

An adjoint-based method for optimising MHD equilibria against the infinite- n , ideal ballooning mode

We demonstrate a fast adjoint-based method to optimise tokamak and stellarator equilibria against a pressure-driven instability known as the infinite-n ideal ballooning mode. We present three finite-β (the ratio of thermal to magnetic pressure) equilibria: one tokamak equilibrium and two stellarator equilibria that are unstable against the ballooning mode. Using the self-adjoint property of ideal magnetohydrodynamics, we construct a technique to rapidly calculate the change in the eigenvalue, a measure of ideal ballooning instability. Using the SIMSOPT optimisation framework, we then implement our fast adjoint gradient-based optimiser to minimise the eigenvalue and find stable equilibria for each of the three originally unstable equilibria.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Candidate phases for SU(2) adjoint QCD$_4$ with two flavors from $\mathcal{N}=2$ supersymmetric Yang-Mills theory

We study four-dimensional adjoint QCD with gauge group SU(2) and two Weyl fermion flavors, which has an SU(2) R chiral symmetry. The infrared behavior of this theory is not firmly established. We explore candidate infrared phases by embedding adjoint QCD into N = 2 supersymmetric Yang-Mills theory deformed by a supersymmetry-breaking scalar mass M that preserves all global symmetries and 't Hooft anomalies. This includes 't Hooft anomalies that are only visible when the theory is placed on manifolds that do not admit a spin structure. The consistency of this procedure is guaranteed by a nonabelian spin-charge relation involving the SU(2) R symmetry that is familiar from topologically twisted N = 2 theories. Since every vacuum on the Coulomb branch of the N = 2 theory necessarily matches all 't Hooft anomalies, we can generate candidate phases for adjoint QCD by deforming the theories in these vacua while preserving all symmetries and 't Hooft anomalies. One such deformation is the supersymmetry-breaking scalar mass M itself, which can be reliably analyzed when M is small. In this regime it gives rise to an exotic Coulomb phase without chiral symmetry breaking. By contrast, the theory near the monopole and dyon points can be deformed to realize a candidate phase with monopole-induced confinement and chiral symmetry breaking. The low-energy theory consists of two copies of a CP 1 sigma model, which we analyze in detail. Certain topological couplings that are likely to be present in this CP 1 model turn the confining solitonic string of the model into a topological insulator. We also examine the behavior of various candidate phases under fermion mass deformations. We speculate on the possible large-M behavior of the deformed N = 2 theory and conjecture that the CP 1 phase eventually becomes dominant.

Córdova, Clay↗

AdjointBackMapV2: Precise reconstruction of arbitrary CNN unit’s activation via adjoint operators

Adjoint operators have been found to be effective in the exploration of CNN’s inner workings (Wan and Choe, 2022). However, the previous no-bias assumption restricted its generalization. We overcome the restriction via embedding input images into an extended normed space that includes bias in all CNN layers as part of the extended space and propose an adjoint-operator-based algorithm that maps high-level weights back to the extended input space for reconstructing an effective hypersurface. Such hypersurface can be computed for an arbitrary unit in the CNN, and here we prove that this reconstructed hypersurface, when multiplied by the original input (through an inner product), will precisely replicate the output value of each unit. We show experimental results based on the CIFAR-10 and CIFAR-100 data sets where the proposed approach achieves near 0 activation value reconstruction error.

97 MATHEMATICS AND COMPUTING↗

More about the lattice Hamiltonian for Adjoint QCD 2

In our earlier work [1], we introduced a lattice Hamiltonian for Adjoint QCD 2 using staggered Majorana fermions. We found the gauge invariant space of states explicitly for the gauge group SU(2) and used them for numerical calculations of observables, such as the spectrum and the expectation value of the fermion bilinear. In this paper, we carry out a more in-depth study of our lattice model, extending it to any compact and simply-connected gauge group G. We show how to find the gauge invariant space of states and use it to study various observables. We also use the lattice model to calculate the mixed ’t Hooft anomalies of Adjoint QCD 2 for arbitrary G. We show that the matrix elements of the lattice Hamiltonian can be expressed in terms of the Wigner 6j-symbols of G. For G = SU(3), we perform exact diagonalization for lattices of up to six sites and study the low-lying spectrum, the fermion bilinear condensate, and the string tension. We also show how to write the lattice strong coupling expansion for ground state energies and operator expectation values in terms of the Wigner 6j-symbols. For SU(3) we carry this out explicitly and find good agreement with the exact diagonalizations, and for SU(4) we give expansions that can be compared with future numerical studies.

confinement↗

A Type II Hamiltonian Variational Principle and Adjoint Systems for Lie Groups

We present a novel Type II variational principle on the cotangent bundle of a Lie group which enforces Type II boundary conditions, i.e., fixed initial position and final momentum. In general, such Type II variational principles are only globally defined on vector spaces or locally defined on general manifolds; however, by left translation, we are able to define this variational principle globally on cotangent bundles of Lie groups. Type II boundary conditions are particularly important for adjoint sensitivity analysis, which is our motivating application. As such, we additionally discuss adjoint systems on Lie groups, their properties, and how they can be used to solve optimization problems subject to dynamics on Lie groups.

97 MATHEMATICS AND COMPUTING↗

CANVAS: An Adjoint Waveform Tomography Model of California and Nevada

Abstract We present the California‐Nevada Adjoint Simulations (CANVAS) model, an adjoint waveform tomography model of the crust and uppermost mantle of the states of California and Nevada. We used WUS256 (Rodgers et al., 2022, https://doi.org/10.1029/2022jb024549 ) as the starting model and iteratively decreased the minimum period of CANVAS from 30 to 12 s. CANVAS was iterated in two distinct stages: the first stage with source mechanisms from the Global Centroid Moment Tensor (GCMT) catalog and the second stage with inverted moment tensors (MT) using the CANV_WUS model (Doody et al., 2023, https://doi.org/10.1029/2023jb026463 ). We show that updating the MTs with 3D Green's functions improved waveform fits and azimuthal coverage of windowed data used to calculate the gradients. As for the model itself, we improved waveform fits over WUS256, particularly in the dispersed surface waves. CANVAS resolved tectonic features seen in other models and accurately defined the depth to basement of major basins, including the Central Valley and the Ventura Basin. We propose CANVAS as a starting model for crustal tomography models on smaller scales.

58 GEOSCIENCES↗

Adjoint Waveform Tomography for Next Generation Seismic Analyses and Monitoring

The development of methods and capabilities to compute complete waveform simulations in three-dimensional (3D) Earth models along with adjoint methods for computing the fully 3D sensitivity kernels in the 2000's set the stage for new advances in seismic imaging. I believe that the full benefits of adjoint waveform tomography (AWT) are not yet fully realized and this will be an important direction for the future of seismic tomography.

58 GEOSCIENCES↗

A direct-adjoint approach for material point model calibration with application to plasticity

Here, this paper proposes a new approach for the calibration of material parameters in local elastoplastic constitutive models. The calibration is posed as a constrained optimization problem, where the constitutive model evolution equations for a single material point serve as constraints. The objective function quantifies the mismatch between the stress predicted by the model and corresponding experimental measurements. To improve calibration efficiency, a novel direct-adjoint approach is presented to compute the Hessian of the objective function, which enables the use of second-order optimization algorithms. Automatic differentiation is used for gradient and Hessian computations. Two numerical examples are employed to validate the Hessian matrices and to demonstrate that the Newton–Raphson algorithm consistently outperforms gradient-based algorithms such as L-BFGS-B.

36 MATERIALS SCIENCE↗

MITgcm-AD v2: Open source tangent linear and adjoint modeling framework for the oceans and atmosphere enabled by the Automatic Differentiation tool Tapenade

The Massachusetts Institute of Technology General Circulation Model (MITgcm) is widely used by the climate science community to simulate planetary atmosphere and ocean circulations. A defining feature of the MITgcm is that it has been developed to be compatible with an algorithmic differentiation (AD) tool, TAF, enabling the generation of tangent-linear and adjoint models. These provide gradient information which enables dynamics-based sensitivity and attribution studies, state and parameter estimation, and rigorous uncertainty quantification. Importantly, gradient information is essential for computing comprehensive sensitivities and performing efficient large-scale data assimilation, ensuring that observations collected from satellites and in-situ measuring instruments can be effectively used to optimize a large uncertain control space. As a result, the MITgcm forms the dynamical core of a key data assimilation product employed by the physical oceanography research community: Estimating the Circulation and Climate of the Ocean (ECCO) state estimate. Although MITgcm and ECCO are used extensively within the research community, the AD tool TAF is proprietary and hence inaccessible to a large proportion of these users. The new version 2 (MITgcm-AD v2) framework introduced here is based on the source-to-source AD tool Tapenade, which has recently been open-sourced. Another feature of Tapenade is that it stores required variables by default (instead of recomputing them) which simplifies the implementation of efficient, AD-compatible code. The framework has been integrated with the MITgcm model’s main branch and is now freely available.

Adjoints↗