Search NASA⌕ Search

SEARCH · Search NASA

Results for “Adaptive momentum”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Improving Deep Neural Networks’ Training for Image Classification With Nonlinear Conjugate Gradient-Style Adaptive Momentum

Momentum is crucial in stochastic gradient-based optimization algorithms for accelerating or improving training deep neural networks (DNNs). In deep learning practice, the momentum is usually weighted by a well-calibrated constant. However, tuning the hyperparameter for momentum can be a significant computational burden. In this article, we propose a novel adaptive momentum for improving DNNs training; this adaptive momentum, with no momentum-related hyperparame- ter required, is motivated by the nonlinear conjugate gradient (NCG) method. Stochastic gradient descent (SGD) with this new adaptive momentum eliminates the need for the momentum hyperparameter calibration, allows using a significantly larger learning rate, accelerates DNN training, and improves the final accuracy and robustness of the trained DNNs. For example, SGD with this adaptive momentum reduces classification errors for training ResNet110 for CIFAR10 and CIFAR100 from 5.25% to 4.64% and 23.75% to 20.03%, respectively. Furthermore, SGD, with the new adaptive momentum, also benefits adversarial training and, hence, improves the adversarial robustness of the trained DNNs.

97 MATHEMATICS AND COMPUTING↗

Adaptive momentum management for large space structures

Momentum management is discussed for a Large Space Structure (LSS) with the structure selected configuration being the Initial Orbital Configuration (IOC) of the dual keel space station. The external forces considered were gravity gradient and aerodynamic torques. The goal of the momentum management scheme developed is to remove the bias components of the external torques and center the cyclic components of the stored angular momentum. The scheme investigated is adaptive to uncertainties of the inertia tensor and requires only approximate knowledge of principle moments of inertia. Computational requirements are minimal and should present no implementation problem in a flight type computer and the method proposed is shown to be effective in the presence of attitude control bandwidths as low as .01 radian/sec.

Hahn, E.↗

Adaptive momentum management for the dual keel Space Station

The report discusses momentum management for a large space structure with the structure selected configuration being the Initial Orbital Configuration of the dual-keel Space Station. The external torques considered were gravity gradient and aerodynamic torques. The goal of the momentum management scheme developed is to remove the bias components of the external torques and center the cyclic components of the stored angular momentum. The scheme investigated is adaptive to uncertainties of the inertia tensor and requires only approximate knowledge of principal moments of inertia. Computational requirements are minimal and should present no implementation problem in a flight-type computer. The method proposed is shown to be effective in the presence of attitude control bandwidths as low as 0.01 radian/sec.

Hopkins, M.↗

Accelerated Sparse Recovery via Gradient Descent with Nonlinear Conjugate Gradient Momentum

This paper applies an idea of adaptive momentum for the nonlinear conjugate gradient to accelerate optimization problems in sparse recovery. Specifically, we consider two types of minimization problems: a (single) differentiable function and the sum of a non-smooth function and a differentiable function. In the first case, we adopt a fixed step size to avoid the traditional line search and establish the convergence analysis of the proposed algorithm for a quadratic problem. This acceleration is further incorporated with an operator splitting technique to deal with the non-smooth function in the second case. As a result, we use the convex ι 1 and the nonconvex ι 1 – ι 2 functionals as two case studies to demonstrate the efficiency of the proposed approaches over traditional methods.

97 MATHEMATICS AND COMPUTING↗

Adaptive attitude control and momentum management for large-angle spacecraft maneuvers

The fully coupled equations of motion are systematically linearized around an equilibrium point of a gravity gradient stabilized spacecraft, controlled by momentum exchange devices. These equations are then used for attitude control system design of an early Space Station Freedom flight configuration, demonstrating the errors caused by the improper approximation of the spacecraft dynamics. A full state feedback controller, incorporating gain-scheduled adaptation of the attitude gains, is developed for use during spacecraft on-orbit assembly or operations characterized by significant mass properties variations. The feasibility of the gain adaptation is demonstrated via a Space Station Freedom assembly sequence case study. The attitude controller stability robustness and transient performance during gain adaptation appear satisfactory.

Parlos, Alexander G.↗

Adapting FEFF to 5f Angular Momentum Coupling

Here, it is demonstrated that the spectral simulation program FEFF can be adapted to include the effects of 5f total angular momentum coupling in the fluorite actinide dioxide systems ThO 2 , UO 2 , and PuO 2 . N 4,5 x-ray absorption spectra produced with this modified FEFF approach will be compared to the previous experimental results, obtaining a strong agreement between the two.

36 MATERIALS SCIENCE↗

A mass–momentum consistent coupling for mesh-adaptive two-phase flow simulations

Here, we present a novel mass-momentum consistent coupling between a geometric volume-of-fluid scheme and an incompressible flow solver with differing directional-splitting approaches. The advection of the volume fraction is performed using a direction-split algorithm, whereas the momentum advection algorithm uses a traditional unsplit, fractional-step approach. Both algorithms employ finite-volume discretizations based on Cartesian meshes. In solving the mass-momentum consistency problem, momentum fluxes at the cell faces are weighted by the density fluxes based on the already advected volume fraction. The success of our approach lies on introducing a Favre-averaged velocity interpolation at the liquid/gas interface along with a minmod slope limiter. Mesh-convergence studies show that when the minmod slope limiter is used, the two-phase solver retains an accuracy between first and second order, but when a purely upwind scheme is considered, its accuracy drops to first order. Finally, after considering several validation problems, the solver is shown to agree well with reference numerical and experimental data while retaining its robustness and efficiency.

97 MATHEMATICS AND COMPUTING↗

Numerical simulations of loops heated to solar flare temperatures. III - Asymmetrical heating

A numerical model is defined for asymmetric full solar flare loop heating and comparisons are made with observational data. The Dynamic Flux Tube Model is used to describe the heating process in terms of one-dimensional, two fluid conservation equations of mass, energy and momentum. An adaptive grid allows for the downward movement of the transition region caused by an advancing conduction front. A loop 20,000 km long is considered, along with a flare heating system and the hydrodynamic evolution of the loop. The model was applied to generating line profiles and spatial X-ray and UV line distributions, which were compared with SMM, P78-1 and Hintori data for Fe, Ca and Mg spectra. Little agreement was obtained, and it is suggested that flares be treated as multi-loop phenomena. Finally, it is concluded that chromospheric evaporation is not an effective mechanism for generating the soft X-ray bursts associated with flares.

Cheng, C.-C.↗

Adaptive control system for large annular momentum control device

A dual momentum vector control concept, consisting of two counterrotating rings (each designated as an annular momentum control device), was studied for pointing and slewing control of large spacecraft. In a disturbance free space environment, the concept provides for three axis pointing and slewing capabilities while requiring no expendables. The approach utilizes two large diameter counterrotating rings or wheels suspended magnetically in many race supports distributed around the antenna structure. When the magnets are energized, attracting the two wheels, the resulting gyroscopic torque produces a rate along the appropriate axis. Roll control is provided by alternating the radiative rotational velocity of the two wheels. Wheels with diameters of 500 to 800 m and with sufficient momentum storage capability require rims only a few centimeters thick. The wheels are extremely flexible; therefore, it is necessary to account for the distributed nature of the rings in the design of the bearing controllers. Also, ring behavior is unpredictably sensitive to ring temperature, spin rate, manufacturing imperfections, and other variables. An adaptive control system designed to handle these problems is described.

Montgomery, R. C.↗

Rotational dating of middle-aged stars

By the age of about 100 million years, the rotational velocity of solar mass stars becomes independent of the initial rotation rate on the main sequence, due to the strong dependence of the angular momentum loss rate on the rotation rate; once this common value is reached, the rotation velocity depends only on time and the mass of the star. Thus, the measured rotation rate of a low-mass star provides an estimate of its age when its mass is known. The observed rotation periods of low-mass stars in the Hyades are used to test this conclusion and to check the validity of the theoretical angular momentum loss rates. Adapting the magnetic braking model of Kawaler (1988), the angular momentum loss rate is integrated to derive a relation between rotation period and age, and this relation is converted to a period-age-color relation using simple stellar models. Using this method, a mean rotational age for the Hyades is obtained which is in close agreement with the age determined by isochrone fitting.

Kawaler, Steven D.↗

Review of Work Done with Professor Nussenzveig Regarding the Mie Theory

Prof. Nussenzveig has dedicated part of his career to a surprising and unusual pursuit harking back to the heyday of classical physics: Mie theory, or the scattering of electromagnetic radiation by a homogeneous sphere. M e theory was not put forward until around 1908 (nearly simultaneously by Debye in a different form) and was quickly forgotten in the rush to the then-new quantum mechanics. It remained somewhat of a backwater until Prof. Nussenzveig brought it back by adapting complex angular momentum ideas from Regge pole theory, which had originally been invented for quantum mechanics. Since 1960, he has made fundamental contributions to actually understanding (as opposed to merely calculating) Mie theory. I became involved with this work in 1978 by inviting Prof. Nussenzveig to visit me at the National Center for Atmospheric Research. Since then we have written several papers together on approximate methods for Mie cross-sections, bubble scattering, and other subjects. This lecture will review that work, ending with his recent work on Mie resonances. The emphasis will be on the applications in atmospheric sciences.

Wiscombe, W.↗

Lion Cub: Minimizing Communication Overhead in Distributed Lion

Communication overhead is a key challenge in distributed deep learning, especially on slower Ethernet intercon nects, and given current hardware trends, communication is likely to become a major bottleneck. While gradient compression techniques have been explored for SGD and Adam, the Lion optimizer has the distinct advantage that its update vectors are the output of a sign operation, enabling straightforward quantization. However, simply compressing updates for communication and using techniques like majority voting fails to lead to end-to-end speedups due to inefficient communication algorithms and reduced convergence. We analyze three factors critical to distributed learning with Lion: optimizing communication methods, identifying effective quantization methods, and assessing the necessity of momentum synchronization. Our findings show that quantization techniques adapted to Lion and selective momentum synchronization can significantly reduce communication costs while maintaining convergence. We combine these into Lion Cub, which enables up to 5x speedups in end-to-end training compared to Lion. This highlights Lion’s potential as a communication-efficient solution for distributed training.

97 MATHEMATICS AND COMPUTING↗

Self-tuning control of attitude and momentum management for the Space Station

This paper presents a hybrid state-space self-tuning design methodology using dual-rate sampling for suboptimal digital adaptive control of attitude and momentum management for the Space Station. This new hybrid adaptive control scheme combines an on-line recursive estimation algorithm for indirectly identifying the parameters of a continuous-time system from the available fast-rate sampled data of the inputs and states and a controller synthesis algorithm for indirectly finding the slow-rate suboptimal digital controller from the designed optimal analog controller. The proposed method enables the development of digitally implementable control algorithms for the robust control of Space Station Freedom with unknown environmental disturbances and slowly time-varying dynamics.

Shieh, L. S.↗

Attitude control of spacecraft using neural networks

This paper investigates the use of radial basis function neural networks for adaptive attitude control and momentum management of spacecraft. In the first part of the paper, neural networks are trained to learn from a family of open-loop optimal controls parameterized by the initial states and times-to-go. The trained is then used for closed-loop control. In the second part of the paper, neural networks are used for direct adaptive control in the presence of unmodeled effects and parameter uncertainty. The control and learning laws are derived using the method of Lyapunov.

Vadali, Srinivas R.↗

Physics-based adaptivity of a spectral method for the Vlasov–Poisson equations based on the asymmetrically-weighted Hermite expansion in velocity space

We propose a spectral method for the 1D-1V Vlasov–Poisson system where the discretization in velocity space is based on asymmetrically-weighted Hermite functions, dynamically adapted via a scaling α and shifting u of the velocity variable. Specifically, at each time instant an adaptivity criterion selects new values of α and u based on the numerical solution of the discrete Vlasov–Poisson system obtained at that time step. Once the new values of the Hermite parameters α and u are fixed, the Hermite expansion is updated and the discrete system is further evolved for the next time step. The procedure is applied iteratively over the desired temporal interval. The key aspects of the adaptive algorithm are: the map between approximation spaces associated with different values of the Hermite parameters that preserves total mass, momentum and energy; and the adaptivity criterion to update α and u based on physics considerations relating the Hermite parameters to the average velocity and temperature of each plasma species. For the discretization of the spatial coordinate, we rely on Fourier functions and use the implicit midpoint rule for time stepping. The resulting numerical method possesses intrinsically the property of fluid-kinetic coupling, where the low-order terms of the expansion are akin to the fluid moments of a macroscopic description of the plasma, while kinetic physics is retained by adding more spectral terms. Moreover, the scheme features conservation of total mass, momentum and energy associated in the discrete, for periodic boundary conditions. A set of numerical experiments confirms that the adaptive method outperforms the non-adaptive one in terms of accuracy and stability of the numerical solution.

97 MATHEMATICS AND COMPUTING↗

Exploiting stochastic locality in lattice QCD: hadronic observables and their uncertainties

Abstract Because of the mass gap, lattice QCD simulations exhibit stochastic locality: distant regions of the lattice fluctuate independently. There is a long history of exploiting this to increase statistics by obtaining multiple spatially-separated samples from each gauge field; in the extreme case, we arrive at the master-field approach in which a single gauge field is used. Here we develop techniques for studying hadronic observables using position-space correlators, which are more localized, and compare with the standard time-momentum representation. We also adapt methods for estimating the variance of an observable from autocorrelated Monte Carlo samples to the case of correlated spatially-separated samples.

Physics↗

Adaptive filtering and maximum entropy spectra with application to changes in atmospheric angular momentum

The spectral resolution and statistical significance of a harmonic analysis obtained by low-order MEM can be improved by subjecting the data to an adaptive filter. This adaptive filter consists of projecting the data onto the leading temporal empirical orthogonal functions obtained from singular spectrum analysis (SSA). The combined SSA-MEM method is applied both to a synthetic time series and a time series of AAM data. The procedure is very effective when the background noise is white and less so when the background noise is red. The latter case obtains in the AAM data. Nevertheless, reliable evidence for intraseasonal and interannual oscillations in AAM is detected. The interannual periods include a quasi-biennial one and an LF one, of 5 years, both related to the El Nino/Southern Oscillation. In the intraseasonal band, separate oscillations of about 48.5 and 51 days are ascertained.

Penland, Cecile↗

Two-Dimensional Projected-Momentum Covariance Mapping for Coulomb Explosion Imaging

We introduce projected-momentum covariance mapping, an extension of recoil-frame covariance mapping for 2D ion imaging studies. By considering the two-dimensional projection of the ion momenta as recorded by the detector, one opens the door to a complex suite of analysis tools adapted from three-dimensional momentum imaging studies. This includes the use of different frames of reference to unravel the dynamics of fragmentation and the application of fragment momentum constraints to isolate specific fragmentation channels. The technique is demonstrated on data from a two-dimensional ion imaging study of the Coulomb explosion of the cis and trans isomers of 1,2-dichloroethene, following strong-field ionization by an intense near-infrared femtosecond laser pulse. Classical simulations are used to guide the interpretation of projected-momentum covariance maps. The results offer a detailed insight into the distinct Coulomb explosion dynamics for this pair of isomers and lay the groundwork for future time-resolved studies of photoisomerization dynamics in this molecular system.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗