Search NASA⌕ Search

SEARCH · Search NASA

Results for “Iterative”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 685 records · Page 38

Improved Guarantees for Optimal Nash Equilibrium Seeking and Bilevel Variational Inequalities

We consider a class of hierarchical variational inequality (VI) problems that subsumes VI-constrained optimization and several other problem classes, including the optimal solution selection problem and the optimal Nash equilibrium (NE) seeking problem. Our main contribution is threefold. (i) We consider bilevel VIs with monotone and Lipschitz continuous mappings and devise a single-timescale iteratively regularized extragradient method, named IR-EG 𝚖,𝚖 . We improve the existing iteration complexity results for addressing both bilevel VI and VI-constrained convex optimization problems. (ii) Under the strong monotonicity of the outer-level mapping, we develop a method named IR-EG 𝚜,𝚖 and derive faster guarantees than those in (i). We also study the iteration complexity of this method under a constant regularization parameter. These results appear to be new for both bilevel VIs and VI-constrained optimization. (iii) To our knowledge, complexity guarantees for computing the optimal NE in nonconvex settings do not exist. Motivated by this lacuna, we consider VI-constrained nonconvex optimization problems and devise an inexactly projected gradient method, named IPR-EG, where the projection onto the unknown set of equilibria is performed using IR-EG 𝚜,𝚖 with a prescribed termination criterion and an adaptive regularization parameter. We obtain new complexity guarantees in terms of a residual map and an infeasibility metric for computing a stationary point. Here, we validate the theoretical findings using preliminary numerical experiments for computing the best and the worst NEs.

bilevel optimization↗

Using parallel banded linear system solvers in generalized eigenvalue problems

Subspace iteration is a reliable and cost effective method for solving positive definite banded symmetric generalized eigenproblems, especially in the case of large scale problems. This paper discusses an algorithm that makes use of two parallel banded solvers in subspace iteration. A shift is introduced to decompose the banded linear systems into relatively independent subsystems and to accelerate the iterations. With this shift, an eigenproblem is mapped efficiently into the memories of a multiprocessor and a high speedup is obtained for parallel implementations. An optimal shift is a shift that balances total computation and communication costs. Under certain conditions, we show how to estimate an optimal shift analytically using the decay rate for the inverse of a banded matrix, and how to improve this estimate. Computational results on iPSC/2 and iPSC/860 multiprocessors are presented.

DISTRIBUTED MEMORY MULTIPROCES↗

Survey of tungsten gross erosion from main plasma facing components in WEST during a L-mode high fluence campaign

An initial high fluence campaign was performed in WEST, in 2023, on the newly installed actively cooled tungsten divertor composed of ITER-grade monoblocks. The campaign consisted in the repetition of a 60 s long Deuterium L-mode pulse in attached divertor conditions, cumulating over 10000s of plasma exposure. A maximum deuterium fluence of approximately 5⋅1⁢026 m−2 was reached in the outer strike point region, representative of a few high performance ITER pulses. Gross tungsten erosion inferred from visible spectroscopy shows that the most eroded plasma facing component is the inner divertor target with rates ten times larger than on the outer divertor target. The outer midplane tungsten bumpers, located a few centimeters from the plasma, show gross erosion rates two times lower than at the outer divertor. We conclude that the outer midplane bumpers have a negligible contribution to the long range tungsten migration and deposition onto the lower divertor. The cumulated gross erosion rate on the inner divertor translates in an effective gross erosion thickness of about 20μ⁢m, while it is about 2μ⁢m for the outer divertor. Strikingly, these orderings coincide with the thickness of deposits found locally on the divertor: the exposed surfaces of high field side monoblocks are covered with several tens of μ⁢m tungsten deposits, while on the lower field side, few μ⁢m thin tungsten deposits are only found on the magnetically shadowed parts of monoblocks. The strong impact of those deposits on WEST operation, namely perturbation of surface temperature measurement with infra-red thermography, and the emission of flakes causing radiative perturbation of the confined plasma, calls for anticipating similar issues in ITER. In particular, the start of research operation shall consider the definition of a divertor erosion budget in order to anticipate the formation of deleterious deposits.

Fedorczak, N.↗

Experiment-modeling studies comparing energy dissipation in the DIII-D SAS and SAS-VW divertors

Recent DIII-D experiments on Small Angle Slot (SAS) divertors have confirmed that a combination of divertor closure and target shaping can enhance cooling across the divertor target and increase energy dissipation, but with significant dependence on B T (toroidal magnetic field) direction. In these novel divertors, the roles of closure, target shaping, drifts, and scale lengths are all interconnected in optimizing dissipation, with the separatrix electron density n eSEP being the key parameter associated with the level of dissipation/detachment. After modifying the original flat-targeted graphite SAS to include a V shape with a tungsten coating on the outer side of the divertor (SAS-VW), matched series of discharges were run to compare to detailed SOLPS-ITER modeling. Experimentally, when run as designed with the outer strike point at the slot vertex, SAS-VW requires nearly identical n eSEP for detachment as the original SAS, with little difference in dissipation for the new geometry. This is in contrast to (1) earlier modeling predictions that a small change of the SAS geometry to a V shape should enhance dissipation at the same n eSEP for magnetic configurations having better H-mode access (ion B × ∇B drift directed into the divertor), and (2) despite the achievement of significantly higher (2-7x) neutral pressures and compression in the SAS-VW slot. Comparisons of experimental density scans to the most recent SOLPS-ITER modeling with ExB drifts show reasonable agreement for dissipation/detachment onset when using separatrix density as the independent parameter. In order to help understand the discrepancy in modeled vs actual performance for the new configuration, additional measurements varying gas injection location and impurity injection were undertaken. In-slot D 2 gas fueling is more effective (5–22 %) in promoting detachment, in accord with modeling. In-slot impurity injection (N 2 or Ne) can yield 30 % lower core Z eff and 15 % less confinement degradation after detachment compared to main chamber puffing, as well as relatively lower tungsten leakage from the divertor. Modeling can also reproduce the improved detachment seen as the strike point moves inboard of the slot vertex. While we can explain the effects of the most important parameters causing energy dissipation in these slot divertors, it remains that many aspects of their behavior cannot be accurately modeled using state-of-art codes such as SOLPS-ITER. This is of concern for future model-driven designs utilizing similar V-shaped geometries.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Uncertainty estimation of bifurcated solutions in the Rayleigh–Bénard problem for advanced nuclear reactors applications

Multiphysics models of nuclear reactors frequently comprise nonlinear systems of equations. The nonlinear nature of these models could lead to solution bifurcations, where a small change in a certain parameter, e.g., the thermophysical properties of the coolant, can lead to a sudden change in the system’s behavior. At the point in parameter space where this happens, called a critical point, the Jacobian matrix of the model’s nonlinear operator becomes singular potentially permitting multiple solutions to coexist. In this paper, we perform uncertainty estimation (UE) in a parameter range that includes bifurcated solutions within the context of Rayleigh–Bénard problem. We perform this analysis assuming uncertain temperature difference, and tilt angle for the iterative solution algorithm with a unit Prandtl number (Pr = 1). Also, we perform this analysis under uncertain thermophysical properties for both FLiBe molten salt and liquid sodium as working fluid. We deploy two approaches to compute statistical moments for the resulting distributions of selected flow-field variables. The first approach is the blind computation of the mean and the standard deviation without any consideration of solution bifurcation, while the second approach utilizes k-means clustering to cluster each branch’s solutions together and compute separate statistical moments for each branch. The statistical distributions are obtained by perturbing the selected parameters about nominal values that correspond to a solution on one of the valid branches, and that solution is used as initial guess for the iterative solution algorithm. We found that perturbation of any parameter when its nominal value is close to its critical point always leads to branch jumping, i.e., the iterations converge to a solution on a branch different from the branch of the initial guess. This produces a statistical ensemble comprised of fundamentally different solutions leading to wrong mean values and uncertainty estimates, whereas clustering provides an efficient way to deal with this type of computation. This work is important for developing Gen IV nuclear systems because many of these systems rely on natural convection for cooling especially in accident conditions.

97 - MATHEMATICS AND COMPUTING↗

Active learning enables generation of molecules that advance the known Pareto front

Although generative models hold promise for discovering molecules with optimized desired properties, they often fail to suggest synthesizable molecules that improve upon the properties of the structures represented in the training distribution. We find that this limitation arises not only from the molecule generation process itself, but also from the poor generalization capabilities of molecular property predictors. We address this challenge by creating a closed-loop molecule generation pipeline with iterative retraining on new quantum chemical simulation data. Compared against static, single-pass generative modeling approaches, only our closed-loop iterative workflow generates molecules with properties extending beyond the training distribution (up to 0.44 standard deviations beyond the original range) and achieves a 79% improvement in out-of-distribution molecule classification accuracy. Furthermore, by conditioning molecular generation on thermodynamic stability data obtained during the iterative loop, the proportion of stable and hence potentially synthesizable molecules generated is 3.5x higher than the next-best model.

Chemistry↗

Thermal-Fluid and Thermal-Structural Response of the T-Tube Modular Divertor to Spatiotemporally Varying Heat Loads

Tungsten (W) is the leading candidate for divertor target plates because of its high melting point (>3000°C), thermal conductivity, and ultimate tensile stress. While W and its alloys are the only solid materials that can survive the high heat fluxes incident on the divertor, W’s low-ductility high ductile-to-brittle transition temperature of ~600°C and relatively low recrystallization temperature (RT) of ~1300°C pose structural (among other) challenges. The objective of this work is to estimate the thermal-fluid and thermal-structural performance of the helium (He)-cooled T-tube divertor, which was originally developed by the Advanced Reactor Innovation and Evaluation Study (ARIES) using numerical simulations. Here, predictions of temperature distributions across the plasma-facing structural component and surface pressures from computational fluid dynamics simulations are used to determine stress distributions using commercial structural finite element modeling software over a range of fusion-relevant conditions. The maximum allowable incident heat fluxes are determined based on the temperature limits imposed by the ITER elastic Structural Design Criteria for In-vessel Components (SDC-IC) and the maximum RT over a range of He mass flow rates and presented in the form of performance design charts. Our recent work found that thermal- structural criteria accounting for the low ductility of W in a finger-type modular divertor constrain the maximum incident heat fluxes to values well below the ITER specifications, and those based on considering only the RT demonstrate that integrated thermal-fluid and elastic structural performance evaluation are required for accurate assessment of divertor performance. This novel analysis of the T-tube considers how nonuniform and transient incident heat fluxes affect its thermal-fluid and thermal-structural performance, as well as the effect of volumetric heating, which can be as great as 27% of the power incident on the divertor surface. The W tile of the T-tube, with its relatively large plasma-facing area of ~15 cm 2 , will likely experience significant spatial variations in incident heat flux. This work therefore assesses whether steady-state incident heat flux profiles with a peak of 10 MW/m 2 and maximum heat flux gradients of 200 MW/m 2 per m exceed the structural limits imposed by the ITER elastic SDC-IC and the maximum RT over a range of fusion-relevant conditions. The effect of transient heat fluxes typical of plasma detachment and reattachment from the target plate due, for example, to gas injection are also evaluated

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Implementation of the D1S Methodology for Shutdown Dose Rate Calculations in the OpenMC Monte Carlo Particle Transport Code

We present an implementation of the direct one-step (D1S) methodology for shutdown dose rate (SDR) calculations in the OpenMC Monte Carlo particle transport code. In addition to being the first fully open-source D1S implementation, it is also the first to require no ad hoc source code or nuclear data library modifications. The code can seamlessly switch between production of prompt and decay photons based on a user input parameter, and the decay data needed for decay photon generation are made available through a depletion chain file, which is already used for OpenMC’s built-in depletion/activation solver. A set of Python functions significantly eases the burden of computing and applying time correction factors needed to properly account for the time dependence of radionuclide activity. To assess the accuracy of the D1S implementation, SDR calculations have been carried out for three problems: a prism of iron irradiated by 14-MeV neutrons, the ITER port plug computational benchmark, and the Frascati Neutron Generator (FNG) ITER dose rate benchmark problem from the Shielding INtegral Benchmark Archive and Database (SINBAD). For each of these problems, comparisons were made to calculations using the rigorous two-step (R2S) method. The results on the iron prism problem illustrate how the D1S method achieves superior spatial resolution compared to the R2S method without the need for spatial discretization of the activation regions. The D1S and R2S results for the ITER port plug benchmark agree well with previously reported results in the literature. While the D1S results are 10% to 15% lower than the R2S results, this may be due to stochastic uncertainty and/or spatial discretization in the R2S calculations. On the FNG dose rate benchmark problem, the D1S method produces dose rate estimates that are within 4% of the dose rates predicted using a cell-based R2S workflow. The D1S estimates of the SDR are also in reasonable agreement with the experimental measurements and show the same basic trends that have been observed in previous works. A qualitative analysis of the execution time and uncertainty for the R2S and D1S workflows suggests that the D1S method would attain a higher figure of merit.

D1S method↗

Impact of ionization and transport on pedestal density structure in DIII-D and Alcator C-Mod

Abstract This paper investigates the role of ionization on the pedestal structure using both measurements and modeling for H-mode plasmas on DIII-D and Alcator C-Mod to enhance our ability to predict pedestal behavior in future pilot plants. The impact of the neutral penetration depth on the pedestal density is investigated using dimensionally matching hydrogen and deuterium DIII-D H-mode discharges at low and high electron density. The DIII-D Lyman- α diagnostic measurements show that hydrogen neutrals penetrate deeper inside the plasma on both the high field and low field side, while the pedestal electron density structure is similar for both isotopes. However, as the opaqueness increases we observe that the pedestal density gradient becomes stiff, similar to prior observations on DIII-D and C-Mod (Mordijck 2020 Nuclear Fusion 60 082006). In addition, these results also confirm prior measured and modeled poloidal asymmetries in neutral densities, indicating that to make transport predictions, 2D neutral modeling is necessary. The first direct validation of SOLPS-ITER for the measured brightness, emissivity and neutral densities for three different confinement regimes on C-Mod is introduced. The SOLPS-ITER model shows good agreement, within the constrains of the model for all regimes. In addition, a comparison of SOLPS-ITER modeling for DIII-D and C-Mod shows that as opaqueness increases, the role of divertor fueling and thus poloidal asymmetries in the neutral density profiles decreases. Based on these experimental and modeling results we estimate the size of a potential particle pinch using typical values for the diffusion coefficient for both DIII-D and C-Mod H-mode discharges.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Assessing the impact of alpha particles on thermal confinement in JET D-T plasmas through global GENE-Tango simulations

The capability of the global, electromagnetic gyrokinetic GENE code interfaced with the transport Tango solver is exploited to address the impact of fusion alpha particles (in their dual role of fast particles and heating source) on plasma profiles and performance at JET in the discharges with the highest quasi-stationary peak fusion power during the DTE2 experimental campaigns. Employing radially global nonlinear electromagnetic GENE-Tango simulations, we compare results with/without alpha particles and alpha heating. Our findings reveal that alpha particles have a negligible impact on turbulent transport, with GENE-Tango converging to similar plasma profiles regardless of their inclusion as a kinetic species in GENE. On the other hand, alpha heating is found to contribute to the peaking of the electron temperature profiles, leading to a 1 keV drop on the on-axis electron temperature when alpha heating is neglected in Tango. The minimal impact of alpha particles on turbulent transport in this JET discharge–despite this being the shot with the highest fusion output–is attributed to the low content of fusion alpha in this discharge. To assess the potential impact of alpha particles on turbulent transport in regimes with higher alpha particle density, as expected in ITER and fusion reactors, we artificially increased the alpha particle concentration to levels expected for ITER. By performing global nonlinear GENE standalone simulations, we found that increasing the alpha particle density beyond five times the nominal value lead to significant overall turbulence destabilization. These results demonstrate that an increased alpha particle concentration can significantly impact transport properties under simulated JET experimental conditions. However, these findings cannot be directly extrapolated to ITER due to the substantial differences in parameters such as plasma size, magnetic field, plasma current, and thermal pressure.

energetic particles↗

MHD, disruptions and control physics: Chapter 4 of the special issue: on the path to tokamak burning plasma operation

In this chapter, we review the progress in MHD stability, disruptions and control in magnetic fusion research that has occurred over the past (more than) one and a half decades since the publication by Hender et al in 2007 on the same topic as part of the update of ITER Physics Basis. During this period, remarkable progress has been achieved in the understanding of the basic physics and overall control of MHD instabilities through a wide spectrum of dedicated experiments, theory and modeling. The sawtooth activities are probably today one of the best understood of MHD events and very robust control schemes have been developed for reliable operation of tokamaks through core heating. Similarly, significant improvements have been achieved in understanding and control of neoclassical tearing modes, resistive wall modes or locked modes and their control through ECCD or error field control. The field of disruption prediction through application of artificial intelligence, machine learning or deep learning methods, which had already started at the time of the 2007 review, has progressed significantly due to general progress in these fields and application of newer, more sophisticated algorithms. However, although remarkable progress has been achieved in the field of Disruptions, their understanding, prediction, possible avoidance and mitigation still remain probably the most active fields of R&D globally in this field. This is especially because reactor grade machines like ITER and DEMO will be much less tolerant in respect of disruptions and runaway currents, and their occurrences must be either avoided altogether or minimized to an acceptable value without causing any significant hindrance to robust machine operations. This review is intended to present a broad spectrum of the R&D that has occurred in this field in support of ITER, which will also be of immense significance for all future machines, especially reactors like DEMO.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Autoregressive long-horizon prediction of plasma edge dynamics *

Accurate modeling of scrape-off layer (SOL) and divertor-edge dynamics is vital for designing plasma-facing components in fusion devices. High-fidelity edge fluid/neutral codes such as SOLPS-ITER capture SOL physics with high accuracy, but their computational cost limits broad parameter scans and long transient studies. We present transformer-based, autoregressive surrogates for efficient prediction of 2D, time-dependent plasma edge state fields. Trained on SOLPS-ITER spatiotemporal data for the KSTAR tokamak, the surrogates forecast electron temperature, electron density, and radiated power over extended horizons. We evaluate model variants trained with increasing autoregressive horizons (1–100 steps) on short- and long-horizon prediction tasks. Longer-horizon training systematically improves rollout stability and mitigates error accumulation, enabling stable predictions over hundreds to thousands of steps and reproducing key dynamical features such as the motion of high-radiation regions. Measured end-to-end wall-clock times show the surrogate is orders of magnitude faster than SOLPS-ITER, enabling rapid parameter exploration. Prediction accuracy degrades when the surrogate enters physical regimes not represented in the training dataset, motivating future work on data enrichment and physics-informed constraints. Overall, this approach provides a fast, accurate surrogate for computationally intensive plasma edge simulations, supporting rapid scenario exploration, control-oriented studies, and progress toward real-time applications in fusion devices.

autoregressive deep learning↗

Enhancing Photosynthesis Simulation Performance in ESMs with Machine Learning-Assisted Solvers

When simulating vegetation dynamics, photosynthesis accounts for a large fraction of the computational cost in most Earth System Models (ESMs). This is largely since photosynthesis is represented as a system of nonlinear equations, and the solution requires the use of an initial guess followed by many iterations of the numerical solver to obtain a solution. We use machine learning (ML) to replicate the response surface of the model’s numerical solver to improve the choice of initial guess, therefore requiring fewer iterations to obtain a final solution. We implemented this test on the leaf-level calculations as well as at the canopy scale, and for both we observed fewer iterations of the photosynthesis solver when a ML-based initial guess was implemented. The model tested here is the Energy Exascale Earth System Model - Land Model (ELM). The ML-based algorithms used here are trained on simulations from the model itself and used only to improve the initial guess for the solver; therefore, the model maintains its own set of physics to obtain the final solution. This work shows novel ways to utilize ML-based methods to improve the performance of numerical solvers in ESMs.

Massoud, Elias [ORNL] (ORCID:0000000217725361)↗

SymProp: Scaling Sparse Symmetric Tucker Decomposition via Symmetry Propagation

Sparse symmetric tensors are an important class of tensors, and their decompositions serve as powerful tools for revealing low-rank structures. This paper introduces SymProp, a novel approach for scaling sparse symmetric Tucker decomposition by propagating symmetry through intermediate computations. SymProp optimizes two key computational kernels: Sparse Symmetric Tensor Times Same Matrix chain (S3 TTMc) for Higher-Order Orthogonal Iteration (HOOI) and Sparse Symmetric Tensor Times Same Matrix chain Times Core (S3 TTMcTC) for Higher-Order QR Iteration (HOQRI). Our method employs a metaprogramming-based index iteration approach to efficiently handle the upper triangular parts of intermediate dense symmetric tensors. SymProp achieves up to 50.9× speedup over SPLATT and up to 360.8× over Compressed Sparse Symmetric (CSS) format on the S3 TTMc operation. Moreover, our S3 TTMc and S3 TTMcTC implementations support tensor orders four levels higher than state-of-the-art methods. Our HOQRI demonstrates superior scalability and up to a 33.6× speedup over optimized HOOI. By enabling more scalable Tucker decompositions for higher orders, decomposition ranks, and dimension sizes, SymProp opens new possibilities for analyzing complex hypergraph structures in fields such as network science, data mining, and machine learning.

Li, Zecheng [North Carolina State University]↗

Aluminum-Based Superconducting Tunnel Junction Sensors for Nuclear Recoil Spectroscopy

The BeEST experiment is searching for sub-MeV sterile neutrinos by measuring nuclear recoil energies from the decay of 7 Be implanted into superconducting tunnel junction (STJ) sensors. The recoil spectra are affected by interactions between the radioactive implants and the sensor materials. We are therefore developing aluminum-based STJs (Al-STJs) as an alternative to existing tantalum devices (Ta-STJs) to investigate how to separate material effects in the recoil spectrum from potential signatures of physics beyond the Standard Model. Three iterations of Al-STJs were fabricated. The first had electrode thicknesses similar to existing Ta-STJs. They had low responsivity and reduced resolution, but were used successfully to measure 7 Be nuclear recoil spectra. The second iteration had STJs suspended on thin SiN membranes by backside etching. These devices had low leakage current, but also low yield. The final iteration was not backside etched, and the Al-STJs had thinner electrodes and thinner tunnel barriers to increase signal amplitudes. These devices achieved 2.96 eV FWHM energy resolution at 50 eV using a pulsed 355 nm (~3.5 eV) laser. These results establish Al-STJs as viable detectors for systematic material studies in the BeEST experiment.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

DESIGN ISSUES AND QUALIFICATION OF HYDROFORMED DOUBLE WALLED EXPANSION JOINTS IN VACUUM SERVICE

The Vacuum Auxiliary System for ITER is devoted to pumping out, venting, and purging the vacuum volumes of the tokamak. Over 5000 clients including the cryostat and vacuum vessel, at 8500 m3 and 1400 m3 respectively, are serviced by approximately 150 pumping stations through 6 km of pipework. The piping and functions provided to the clients include component operation, routing of potentially tritiated gases, and timely leak localization. Expansion joints are used to reduce loadings on pumps and piping stresses but also to qualify in-line components having low allowable design loads. The ITER project utilizes vacuum valves that are not designed to withstand typical piping system loadings and require detailed design and evaluation. Loading scenarios prescribed by governing system specifications must be considered. To maintain allowable design loads for components and to keep piping stresses below code requirements, double walled hydroformed expansion joints have been employed in the design. These are not standard items and require custom fabrication. Installation space constraints and adjacent pipe support attachment location availability at ITER limit standard expansion joint installation guidelines provided by the Expasion Joint Manufacturers Association (EJMA). Careful consideration must be given to the analysis model to ensure proper function and life expectancy of the component. These considerations include accurate accounting of thrust forces, thermal movements, seismic accelerations, equipment and building differential displacements. In addition to displacements, the process internal, external, and interspace pressures affect the qualification and selection of the double walled expansion joints. The calculation results shall confirm that deflections, forces, and moments are reasonable for the size and type required for the system’s demands as evaluated against the manufacturer’s design.

Clark, Forrest [ORNL] (ORCID:0009000678106843)↗

Stochastic Trust-Region Algorithm in Random Subspaces with Convergence and Expected Complexity Analyses

Here, this work proposes a framework for large-scale stochastic derivative-free optimization (DFO) by introducing STARS, a trust-region method based on iterative minimization in random subspaces. This framework is both an algorithmic and theoretical extension of a random subspace derivative-free optimization (RSDFO) framework, and an algorithm for stochastic optimization with random models (STORM). Moreover, like RSDFO, STARS achieves scalability by minimizing interpolation models that approximate the objective in low-dimensional affine subspaces, thus significantly reducing per-iteration costs in terms of function evaluations and yielding strong performance on largescale stochastic DFO problems. The user-determined dimension of these subspaces, when the latter are defined, for example, by the columns of so-called Johnson-Lindenstrauss transforms, turns out to be independent of the dimension of the problem. For convergence purposes, inspired by the analyses of RSDFO and STORM, both a particular quality of the subspace and the accuracies of random function estimates and models are required to hold with sufficiently high, but fixed, probabilities. Using martingale theory under the latter assumptions, an almost sure global convergence of STARS to a first-order stationary point is shown, and the expected number of iterations required to reach a desired first-order accuracy is proved to be similar to that of STORM and other stochastic DFO algorithms, up to constants.

97 MATHEMATICS AND COMPUTING↗

Convergence Analysis of the Alternating Anderson–Picard Method for Nonlinear Fixed-Point Problems

Anderson acceleration (AA) has been widely used to solve nonlinear fixed-point problems due to its rapid convergence. This work focuses on a variant of AA in which multiple Picard iterations are performed between each AA step, referred to as the Alternating Anderson–Picard (AAP) method. Furthermore, despite introducing more “slow” Picard iterations, this method has been shown to be efficient and even more robust in both linear and nonlinear cases. However, there is a lack of theoretical analysis for AAP in the nonlinear case. In this paper, we address this gap by establishing the equivalence between AAP and a multisecant-GMRES method that uses GMRES to solve a multisecant linear system at each iteration. From this perspective, we show that AAP “converges” to the Newton-GMRES method. Specifically, as the residual approaches zero, the multisecant matrix, the approximate Jacobian inverse, the search direction, and the optimization gain of AAP converge to their counterparts in the Newton-GMRES method. These connections provide insights for analyzing the asymptotic convergence properties of AAP. Consequently, we show that AAP is locally 𝑞-linear convergent and provide an upper bound for the convergence factor of AAP. To validate the theoretical results, numerical examples are provided.

Anderson acceleration↗