Search NASASearch

SEARCH · Search NASA

Results for “parameterization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

CONCURRENT, CONDENSED STEIN VARIATIONAL GRADIENT DESCENT FOR UNCERTAINTY QUANTIFICATION OF NEURAL NETWORKS

In this work, we propose a Stein variational gradient descent (SVGD) method to concurrently sparsify, train, and provide uncertainty quantification (UQ) of a complexly parameterized model, such as a neural network (NN). It employs a graph reconciliation and condensation process to reduce complexity and increase similarity in the Stein ensemble of parameterizations. Therefore, the proposed concurrent, condensed SVGD (ccSVGD) method can provide UQ on parameters, not just outputs. Furthermore, the parameter reduction speeds up the convergence of the Stein gradient descent as it reduces the combinatorial complexity by aligning and differentiating the sensitivity to parameters. These properties are demonstrated with an illustrative example and an application to a mechanical response representation problem in solid mechanics.

42 ENGINEERING

Size-resolved particle and black carbon deposition over the cryosphere (Final Report)

Our project focused on providing observational constraints and investigation of aerosol deposition by performing the first unambiguous direct eddy flux covariance measurements of aerosol dry deposition over the cryosphere. The project aimed to make eddy covariance flux measurements of dry deposition and off-line wet deposition measurements of black carbon containing aerosol plus eddy covariance measurements of size-resolved accumulation mode scattering aerosol over the DOE ARM AMF3 Oliktok Point site in Fall 2020. We aimed to compare our measurements to model parameterizations to constrain uncertainties and systematic bias, and improve deposition parameterizations in models. Due to disruptions in field work from COVID-19, our work was delayed and moved to an alternate site on the North Slope of Alaska. This report summarizes the key findings from the project, including data analysis from previous field projects and from the North Slope of Alaska.

54 ENVIRONMENTAL SCIENCES

Advancing the Understanding of Cloud Microphysical Processes and Aerosol Indirect Effects in High-Latitude Mixed-Phase Clouds by Linking ARM Measurements with Climate Model Simulations (Final Report)

The key objectives of this project were to advance our understanding of cloud microphysical characteristics and aerosol indirect effects on mixed-phase clouds in high latitudes. To improve the representation of ice and mixed-phase clouds in Earth System Models (ESMs), we propose an integrated observation and modeling study of cloud macro- and microphysical properties, including spatial heterogeneities, mass partitioning between ice crystals and supercooled liquid water, effects of ice nucleating particles (INPs), and efficiency of secondary ice production (SIP), etc. Specifically, we took four main approaches in this project: (1) examining macro- and microphysical properties of ice and mixed-phase clouds based on in-situ and ground-based observations from multiple field campaigns funded by the U.S. Department of Energy (DOE) Atmospheric Radiation Measurement (ARM) program, including the Mixed-Phase Arctic Cloud Experiment (M-PACE), Indirect and Semi-Direct Aerosol Campaign (ISDAC), Ice Nucleating Particle Sources at Oliktok Point (INPOP), ARM West Antarctic Radiation Experiment (AWARE), Measurements of Aerosols, Radiation, and Clouds over the Southern Ocean (MARCUS), and Macquarie Island Cloud and Radiation Experiment (MICRE); (2) evaluating the DOE Energy Exascale Earth System Model (E3SM) simulations based on observations, particularly for ice and mixed-phase cloud microphysical properties; (3) examining the impacts of INPs on ice and mixed-phase clouds. Specifically, a series of comparisons were conducted using observations over the Arctic, Southern Ocean, and Antarctica, including comparisons between the lower and higher southern latitudes as well as comparisons between the northern and southern hemispheres. In addition, aerosol indirect effects from distinct sources of dust particles were examined; and (4) investigating the impacts of SIP. Ultimately, these results helped to improve cloud microphysics and aerosol-cloud interaction parameterizations in the E3SM model. Overall, the project provided improved understanding regarding various factors, including thermodynamic, dynamic, and aerosol conditions, on the micro- and macrophysical properties of ice and mixed-phase clouds in the high latitudes. Resulting analysis helped to provide an improved physical basis for refining the current cloud microphysics parameterizations related to ice and mixed-phase clouds in E3SM.

54 ENVIRONMENTAL SCIENCES

Five Year Comparison of Mixing Height Determinations at the Savannah River Site

Air quality dispersion modeling is performed for the Savannah River Site (SRS) to demonstrate compliance with applicable regulations. The AMS/EPA Regulatory Model (AERMOD) modeling system is an EPA recommended model for air quality applications with a data preprocessor (AERMET) to incorporate meteorological data collected on site. AERMET parameterizes or calculates meteorological variables that are not directly measured onsite. One of the parameters estimated by AERMET is the atmospheric mixing height. While the mixing height is not currently a measurement input into AERMET, SRS has the capability to measure the local mixing height. The Savannah River National Laboratory (SRNL) operates a Vaisala CL31 Lidar Ceilometer which estimates mixing height from aerosol backscatter. This study compares the parameterized mixing height from AERMET to the ceilometer estimated mixing height for the current regulatory period at SRS incorporating data from 2015-2019. Results from this study showed the average daily minimum values (morning) from AERMET were an order of magnitude lower than the commonly used Holzworth (1972) method and the ceilometer estimated mixing heights. Additionally, on average, the ceilometer exhibited a daily maximum mixing height value that occurred 1-3 hours later than the AERMET estimated maximum. This difference is likely due to the nighttime atmospheric mixing height assumptions and calculations used by AERMET. The AERMET algorithm cuts off mixing height growth at sunset while the ceilometer data show ongoing evening convection typical of the southeastern United States. These results suggest that the AERMET parametrization scheme assumptions may not be representative of a forested landscape and evening convection which could account for more mixing overnight. The results obtained in this study are significant for air dispersion modeling applications for regulatory purposes and worker safety. Mixing height can impact model estimated pollutant concentrations. A greater mixing height will provide more volume for pollutant dispersion. This report documents efforts to quantify the dependence of mixing height inputs toward a conservative estimated pollutant concentration.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W

Collaborative Proposal: Improving understanding of the internal structure and dynamics of deep convection using ARM observations and large eddy simulations

Recent observational and large eddy simulation (LES) modeling studies have nearly unanimously supported the view of deep cumulus convection being composed of a series of quasi-spherical bubbles of buoyant air, known as moist thermals. Despite the prevalence of moist thermals in deep convection, a comprehensive theory for the dynamics of these structures is lacking. Most current conceptual models for cumulus convection are based on canonical scaling theories for dry thermals or plumes; however, there is considerable evidence that the behavior of moist thermals differs markedly from these theories. Furthermore, the theoretical basis for most cumulus parameterizations originates from the plume conceptual model, and therefore these parameterizations are inconsistent with the real structure of moist convection. Motivated by the aforementioned knowledge gaps, this “end-to-end” research effort use theory, observations, numerical simulations, and direct improvements to the Zhang-McFarlane (ZM) convection scheme in the global climate Community Atmosphere Model (CAM) to address the following research questions: What key environmental parameters determine whether or not shallow convection will transition into deep convection, in the context of thermal-like updrafts? What factors regulate the size of thermals within cumulus updrafts? How does vertical wind shear influence thermal behavior, and as a consequence, vertical velocity and mass flux profiles and the shallow-to-deep convective transition? What are the critical processes that determine updraft vertical velocities and their connection to the vertical mass flux profile for thermal-like updrafts? Idealized LES modeling will be used in conjunction with theoretical models for the core properties of thermal-like updrafts to better understand key processes that regulate thermal ascent rates and entrainment properties. Thermal-tracking procedures will be used to characterize the behavior of thermals within the LES, and recently developed direct measures of entrainment and detrainment will be used to quantify entrainment/detrainment rates. Building from these results, we will analyze the structure of moist thermals from hemispheric range-height indicator scans taken during the Atmospheric Radiation Measurement Cloud, Aerosol, and Complex Terrain Interactions (CACTI) field campaign, and from “real case” LES of CACTI events. This combined modeling and observational analysis will provide essential validation for the existing body of research on moist thermal dynamics, which is based primarily on modeling studies. With the insight gained from the aforementioned activities, we will modify the Zhang-McFarlane convection scheme to improve its representation of updraft vertical velocity and entrainment rate profiles. These process-level changes will be tested in the Community Atmosphere Model to assess the impact on global climate simulations.

54 ENVIRONMENTAL SCIENCES

Interactions Between Clouds and Wind-Driven Surface Heat Exchanges over Land

Earth system model experiments show that increasing horizontal resolution fundamentally alters the simulated soil-moisture-precipitation feedback. Kilometer-scale simulations often produce weaker or even negative feedback compared to coarse-resolution models. A key difference of kilometer-scale models is that they resolve mesoscale secondary circulations, including boundary layer horizontal rolls and cellular structures, in addition to cold pools and downdrafts associated with convective precipitation. However, because the relevant processes occur on yet-smaller scales, these circulations are often poorly resolved. This project demonstrated that boundary layer secondary circulations significantly affect surface heat exchanges and wind gusts, and that current model parameterizations can misrepresent these processes at kilometer-scale resolution. Using DOE Atmospheric Radiation Measurement (ARM) observations and targeted experiments with the DOE Energy Exascale Earth System Model (E3SM), we identified physically unrealistic wind gust and surface flux responses to secondary circulations, diagnosed a systematic overestimation of wind shear in convective cold pools, and uncovered a multivariate relationship between land surface fluxes and the scales of updrafts that form shallow cumulus clouds. These findings provide observation-based recommendations for improving parameterizations of surface fluxes and wind gusts in high-resolution Earth system models, thereby reducing uncertainty in convective storm prediction and land-atmosphere feedbacks.

54 ENVIRONMENTAL SCIENCES

Simulation of Particulate Transport for Delivery of Solid Amendments into the Subsurface: FY24 Status Report

For particulate-based amendments to be viable for field-scale remediation at the Hanford Site (e.g., 200 DV-1 Operable Unit), particles need to be delivered a sufficient radial distance from an injection well and retained at concentrations high enough for effective treatment. An accurate description of the particle radius of influence (ROI) is critical for developing an overall remediation strategy. However, field-scale particle simulations are currently limited due to insufficient simulation capabilities and a lack of experimental data to validate and parameterize particle transport models. To help build toward field-scale deployment, this fiscal year (FY) we have (1) developed a pre screening tool to estimate particle transport, (2) implemented particle transport models within PFLOTRAN, and (3) conducted preliminary estimations of particle ROI. While field-scale numerical simulations will ultimately be necessary before remedy design and field implementation, we have developed a pre-screening tool that offers valuable estimations of expected particle injectability and ROI in a 1-D system. The advantage of the tool is that it does not require extensive laboratory experiments and instead makes predictions based solely on routine laboratory measurements. This tool can assist in down-selection and decision-making by identifying which particle amendment systems are worth pursuing in future laboratory experiments, such as 1-D column tests and beyond. With any system, scaling up from the lab to the field presents challenges. Currently, there is no field data available for model calibration or validation. However, the theoretical particle models being developed herein are the best tools available to guide progress toward field deployment. To help bridge this gap and verify model predictions, larger-scale lab experiments are being proposed. To advance simulation capabilities, six particle transport models are being integrated into the reactive transport simulator PFLOTRAN. These include colloid filtration theory (CFT) and five additional particle transport models (M1-M5). Each model, from M1 to M5, progressively incorporates additional particle transport and retention processes. Ultimately, the simplest model capable of accurately describing 1-D column data will be selected and parameterized. During FY24, the CFT and M1 model have been fully implemented within PFLOTRAN. Using an existing 1 D column experiment, the two currently implemented particle transport models (CFT and M1), and associated parameters, were fit to this experiment. While simpler model formulations are helpful for estimations, these formulations could not fully describe particle transport and retention behavior in the previous 1-D column experiment. Thus, additional complexities will need to be considered, which will be accounted for in the M2-M5 model formulations. Additionally, because a viscous, shear thinning fluid was required to keep particles in suspension, considerations for flow will also need to also be accounted for. Therefore, a new immiscible two-phase flow mode is currently being implemented in PFLOTRAN. With some modifications, this new flow module could also support simulation of non-Newtonian liquid amendments, foams, and emulsions. We also estimated the expected ROI of solid amendments using 1-D simulations. The average predicted ROI was approximately 15 ft for micron-sized zero valent iron (mZVI) suspended in xanthan gum (XG). Using the pre screening tool and ROI estimates, additional amendment-delivery laboratory characterization and experiments are proposed. The results from additional experiments can be used to validate and parametrize particulate transport model formulations, which will ultimately provide predictive capabilities for field amendment-delivery systems.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W

VIC-Global Parameter Dataset Sensitivity with the Variable Infiltration Capacity Model: Evaluating the importance of dynamic land surface parameters when using the VIC-Global parameter dataset

Accurate prediction of runoff is essential to water resources management, flood risk assessment, and ecosystem protection. However, many hydrological models still have relatively substantial limitations when representing the influence of land use and land cover (LULC) on runoff generation and routing. Changes in LULC, such as deforestation, urban expansion, agricultural intensification, and wetland loss, have been shown to alter the water balance at the land surface through fundamental hydrologic processes (e.g., interception, infiltration, evapotranspiration, and soil storage). However, it remains an open question what the exact magnitude and timing of these impacts are for the spatial and temporal scales commonly used in engineering applications. In this analysis we focus on one aspect of recent LULC change for assessing human impacts, which is urbanization. Specifically we seek to determine the impacts of urbanization on the magnitude and timing of surface runoff and baseflow in HUC-12 basins in Clark County, Nevada which has experienced rapid urbanization. We use the Variable Infiltration Capacity (VIC) hydrology model with a widely used off-the-shelf dataset of land surface parameters, VIC-Global, both of which have been commonly used in the past for water and energy balance modeling for large scale hydrologic studies. We examine two scenarios where the first scenario removes all urbanized land cover and parameterizes those areas of the basins as barren or open shrubland. The second scenario tests the opposite case where all areas of the basins are classified as urban regardless of their present classification. The results from the VIC model show there is a low sensitivity for daily surface runoff between scenarios. The daily baseflow values indicate similar low sensitivity to the classification change during specific periods, but then have substantial differences during other period when large precipitation events are occurring. This is likely due to the assumed parameter values for the urban land cover classification made by the VIC-Global dataset. Using a static land cover parameterization is reasonable for large domain hydrology models that are being used for near-term planning horizons (<30 years). However, longer planning horizons where feedbacks between the atmosphere and land surface are important, especially in transient climate situations, considerations for how to update land surface parameters should be incorporated.

42 ENGINEERING

WFIP3

The Wind Forecasting Improvement Project 3 (WFIP-3) is the first offshore-based wind resource characterization project within the WFIP construct, funded by the U.S. Department of Energy. WFIP-3 will provide a unique field study that will deliver the comprehensive suite of data needed to inform a series of modeling efforts that will develop and evaluate parameterization schemes suited to offshore environments and improved industry-targeted applications. The field study has two goals: (1) detailed sampling of the vertical structure of the Marine Atmospheric Boundary Layer (MABL) at key observational areas, creating a rich dataset that will be used to refine and validate parameterization schemes, and (2) wide-area sampling of the MABL to create a multi-scale array of observations informing and guiding models of resource characterization. We will deploy a multi-platform array of measurements that span the MABL and create a multi-scale observational array stretching south from Marth’s Vineyard across the wind energy areas.

17 WIND ENERGY

Infinite quantum signal processing

Quantum signal processing (QSP) represents a real scalar polynomial of degree d using a product of unitary matrices of size 2 × 2 , parameterized by ( d + 1 ) real numbers called the phase factors. This innovative representation of polynomials has a wide range of applications in quantum computation. When the polynomial of interest is obtained by truncating an infinite polynomial series, a natural question is whether the phase factors have a well defined limit as the degree d → ∞ . While the phase factors are generally not unique, we find that there exists a consistent choice of parameterization so that the limit is well defined in the ℓ 1 space. This generalization of QSP, called the infinite quantum signal processing, can be used to represent a large class of non-polynomial functions. Our analysis reveals a surprising connection between the regularity of the target function and the decay properties of the phase factors. Our analysis also inspires a very simple and efficient algorithm to approximately compute the phase factors in the ℓ 1 space. The algorithm uses only double precision arithmetic operations, and provably converges when the ℓ 1 norm of the Chebyshev coefficients of the target function is upper bounded by a constant that is independent of d . This is also the first numerically stable algorithm for finding phase factors with provable performance guarantees in the limit d → ∞ .

Dong, Yulong [Department of Mathematics, Universit

A Full Accounting of the Visible Mass in SDSS MaNGA Disk Galaxies

We present a study of the ratio of visible mass to total mass in spiral galaxies to better understand the relative amount of dark matter present in galaxies of different masses and evolutionary stages. Using the velocities of the Hα emission line measured in spectroscopic observations from the Sloan Digital Sky Survey (SDSS) MaNGA Data Release 17 (DR 17), we evaluate the rotational velocity of over 5500 disk galaxies at their 90% elliptical Petrosian radii, R 90 . We compare this to the velocity expected from the total visible mass, which we compute from the stellar, H i, H 2 , and heavy metals and dust masses. H 2 mass measurements are available for only a small subset of galaxies observed in SDSS MaNGA DR17, so we derive a parameterization of the H 2 mass as a function of absolute magnitude in the r band using galaxies observed as part of SDSS DR7. With these parameterizations, we calculate the fraction of visible mass within R 90 that corresponds to the observed velocity. Based on statistically analyzing the likelihood of this fraction, we conclude that the null hypothesis (no dark matter) cannot be excluded at a confidence level better than 95% within the visible extent of the disk galaxies. We also find that when all mass components are included, the ratio of visible to total mass within the visible extent of star-forming disk galaxies increases with galaxy luminosity.

79 ASTRONOMY AND ASTROPHYSICS

HSW-V v1.0: localized injections of interactive volcanic aerosols and their climate impacts in a simple general circulation model

Abstract. A new set of standalone parameterizations is presented for simulating the injection, evolution, and radiative forcing by stratospheric volcanic aerosols against an idealized Held–Suarez–Williamson (HSW) atmospheric background in the Energy Exascale Earth System Model version 2 (E3SMv2). In this model configuration (HSW with enabled volcanism, HSW-V), sulfur dioxide (SO2) and ash are injected into the atmosphere with a specified profile in the vertical, and they proceed to follow a simple exponential decay. The SO2 decay is modeled as a perfect conversion to a long-living sulfate aerosol which persists in the stratosphere. All three species are implemented as tracers in the model framework and are transported by the dynamical core's advection algorithm. The aerosols contribute simultaneously to local heating of the stratosphere and cooling of the surface by a simple plane-parallel Beer–Lambert law applied on two zonally symmetric radiation broadbands in the longwave and shortwave ranges. It is shown that the implementation parameters can be tuned to produce realistic temperature anomaly signatures of large volcanic events. In particular, results are shown for an ensemble of runs that mimic the volcanic eruption of Mt. Pinatubo in 1991. The design requires no coupling to microphysical subgrid-scale parameterizations and thus approaches the computational affordability of prescribed aerosol forcing strategies. The idealized simulations contain a single isolated volcanic event against a statistically uniform climate, where no background aerosols or other sources of externally forced variability are present. HSW-V represents a simpler-to-understand tool for the development of climate source-to-impact attribution methods.

Hollowed, Joseph P. (ORCID:0000000286581672)

NeuralMie (v1.0): an aerosol optics emulator

The direct interactions of atmospheric aerosols with radiation significantly impact the Earth's climate and weather and are important to represent accurately in simulations of the atmosphere. This work introduces two contributions to enable a more accurate representation of aerosol optics in atmosphere models: (1) NeuralMie, a neural network Mie scattering emulator that can directly compute the bulk optical properties of a diverse range of aerosol populations and is appropriate for use in atmosphere simulations where aerosol optical properties are parameterized, and (2) TAMie, a fast Python-based Mie scattering code based on the Toon and Ackerman (1981) Mie scattering algorithm that can represent both homogeneous and coated particles. TAMie achieves speed and accuracy comparable to established Fortran Mie codes and is used to produce training data for NeuralMie. NeuralMie is highly flexible and can be used for a wide range of particle types, wavelengths, and mixing assumptions. It can represent core-shell scattering and, by directly estimating bulk optical properties, is more efficient than existing Mie code and Mie code emulators while incurring negligible error compared to existing aerosol optics parameterization schemes (0.08 % mean absolute percentage error).

54 ENVIRONMENTAL SCIENCES

Large-eddy simulation of an atmospheric bore and associated gravity wave effects on wind farm performance in the southern Great Plains

Gravity waves are a common occurrence in the atmosphere, with a variety of generation mechanisms. Their impact on wind farms has only recently gained attention, with most studies focused on wind farm-induced gravity waves. In this study, the interaction between a wind farm and gravity waves generated by an atmospheric bore event is assessed using multiscale large-eddy simulations. The atmospheric bore is created by a thunderstorm downdraft from a nocturnal mesoscale convective system (MCS). The associated gravity waves impact the wind resource and power production at a nearby wind farm during the American Wake Experiment (AWAKEN) in the US southern Great Plains. A two-domain nested setup (Δx=300 and 20 m) is used in the Weather Research and Forecasting (WRF) model, forced with data from the High-Resolution Rapid Refresh model, to capture both the formation of the bore and its interaction with individual wind turbines. The MCS is resolved on the large outer domain, where the structure of the bore and the associated gravity waves are found to be especially sensitive to parameterized microphysics processes. On the finer inner domain, gravity wave interactions with individual wind turbines are resolved; wake dynamics are captured using a generalized actuator disk parameterization in WRF. The gravity waves are found to have a strong effect on the atmosphere above the wind farm; however, the effect of the waves is more nuanced closer to the surface where there is additional turbulence, both ambient and wake-generated. Notably, the gravity waves modulate the mesoscale environment by weakening and dissipating the preexisting low-level jet, which reduces hub-height wind speed and hence the simulated power output, which is confirmed by the observed supervisory control and data acquisition (SCADA) power data. Additionally, the gravity waves induce local wind direction variations correlated with fluctuations in pressure, which lead to fluctuations in the simulated power output as various turbines within the farm are subjected to waking from nearby turbines.

17 WIND ENERGY

Differences in cluster and internal wake effects from mesoscale and large-eddy simulations off the US East Coast

Mesoscale simulations are increasingly used to estimate wake effects within and between large wind farms, despite limited validation for large-scale wake effects. This study evaluates the capabilities and limitations of mesoscale simulations in capturing wake-induced impacts on wind turbine power production through a direct comparison with large-domain large-eddy simulations (LESs) for three planned offshore wind farms under realistic atmospheric conditions and a range of atmospheric stabilities. We assess mesoscale performance in replicating wake characteristics behind single and multiple turbine clusters and quantify the resulting variability in mean turbine power. Results show that mesoscale Weather Research and Forecasting simulations with the Fitch wind farm parameterization capture key features of the velocity deficit downstream of both single and multiple wind farms, with mean root-mean-square errors near 5 % and good agreement with stability-driven wake behavior. However, in these simulations, the mesoscale Fitch parameterization underestimates power losses from internal wake effects, particularly when turbines align with the prevailing wind direction or under stable stratification. In these conditions, individual wakes persist and dominate downstream power deficits. The coarse resolution of the mesoscale simulations limits their ability to resolve individual wind turbine wakes that drive power fluctuations within wind farms. Nonetheless, mesoscale simulations can yield accurate estimates of combined wake losses from internal and cluster effects across some wind direction sectors, where errors in wake representation may cancel each other out. These findings underscore the strengths of mesoscale simulations for capturing broader wake patterns while highlighting their limitations for modeling turbine-level power losses. Future work should explore hybrid modeling approaches to capture both long-range cluster wake propagation and localized internal wake dynamics.

17 WIND ENERGY

Exploring gauge-fixing conditions with gradient-based optimization

Lattice gauge fixing is required to compute gauge-variant quantities, for example those used in RI-MOM renormalization schemes or as objects of comparison for model calculations. Recently, gauge-variant quantities have also been found to be more amenable to signal-to-noise optimization using contour deformations. These applications motivate systematic parameterization and exploration of gauge-fixing schemes. This work introduces a differentiable parameterization of gauge fixing which is broad enough to cover Landau gauge, Coulomb gauge, and maximal tree gauges. The adjoint state method allows gradient-based optimization to select gauge-fixing schemes that minimize an arbitrary target loss function.

Detmold, William

A geometric framework for momentum-based optimizers for low-rank training

Low-rank pre-training and fine-tuning have recently emerged as promising techniques for reducing the computational and storage costs of large neural networks. Training low-rank parameterizations typically relies on conventional optimizers such as heavy ball momentum methods or Adam. In this work, we identify and analyze potential difficulties that these training methods encounter when used to train low-rank parameterizations of weights. In particular, we show that classical momentum methods can struggle to converge to a local optimum due to the geometry of the underlying optimization landscape. To address this, we introduce novel training strategies derived from dynamical low-rank approximation, which explicitly account for the underlying geometric structure. Our approach leverages and combines tools from dynamical low-rank approximation and momentum-based optimization to design optimizers that respect the intrinsic geometry of the parameter space. We validate our methods through numerical experiments, demonstrating faster convergence, and stronger validation metrics at given parameter budgets.

Schotthoefer, Steffen [ORNL] (ORCID:00000002156965

Benchmarking quantum trial wavefunctions for phaseless auxiliary-field quantum Monte Carlo

The phaseless auxiliary-field quantum Monte Carlo (ph-AFQMC) method is a stochastic imaginary-time projection technique for computing ground-state properties of strongly correlated quantum systems, with accuracy that depends critically on the choice of trial wavefunction. Here, we investigate ph-AFQMC with trial states prepared using parameterized quantum circuits. In this work, we present a comprehensive benchmarking study of quantum trial wavefunctions spanning unitary coupled-cluster, Hamiltonian-informed, Jastrow-inspired, and adaptively constructed ansatze. The benchmarking evaluates accuracy, expressibility, and scalability of these ansatze within the QC-AFQMC framework. We test these ansatze on linear hydrogen chains under bond stretching and find that several ansatz families produce chemically accurate ph-AFQMC energies across the dissociation curve. We have performed simulations using the CUDA-Q quantum development platform on the GPU partition of the Perlmutter supercomputer. When comparing ansatze at similar numbers of variational parameters, we find that different ansatz families yield comparable ph-AFQMC results despite exhibiting substantially different variational energies, optimization costs, and circuit depths. Our results indicate that the variational energy of an ansatz is not always a reliable indicator of its quality for ph-AFQMC and reveal instances of over-parameterization. In the strongly correlated regime, trial wavefunctions obtained from adaptive ansatze, exemplified here by ADAPT-VQE with the UCCSD operator pool, can outperform their fixed-ansatz counterparts (UCCSD) in terms of projected energies while using substantially more compact circuits, providing a flexible route to optimize quantum resources within the ph-AFQMC framework.

Rofougaran, Rod [LBNL, Berkeley; Columbia U.; PNL,