Search NASA⌕ Search

SEARCH · Search NASA

Results for “Gaussian Process”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 271 records · Page 15

Implementation of McMurchie–Davidson Algorithm for Gaussian AO Integrals Suited for SIMD Processors

We report an implementation of the McMurchie− Davidson evaluation scheme for 1- and 2-particle Gaussian AO integrals designed for processors with Single Instruction Multiple Data (SIMD) instruction sets. Like in our recent MD implementation for graphical processing units (GPUs) [Asadchev, A.; Valeev, E. F.. J. Chem. Phys. 2024, 160, 244109.], variable-sized batches of shellsets of integrals are evaluated at a time. By optimizing for the floating point instruction throughput rather than minimizing the number of operations, this approach achieves up to 50% of the theoretical hardware peak FP64 performance for many common SIMD-equipped platforms (AVX2, AVX512, NEON), which translates to speedups of up to 30 over the state-of-the-art one-shellset-at-a-time implementation of Obara−Saika-type schemes in Libint for a variety of primitive and contracted integrals. As with our previous work, we rely on the standard C++ programming language such as the std::simd standard library feature to be included in the 2026 ISO C++ standard without any explicit code generation to keep the code base small and portable. The implementation is part of the open source LibintX library freely available at https://github.com/ValeevGroup/libintx.

Basis sets↗

coh3

CoH3 (CoH ver.3) is an optical model, exciton pre-equilibrium, and Hauser-Feshbach statistical model code, which calculates nuclear reaction cross sections for medium to heavy targets in the keV to MeV energy region. This program is written in standard C++, divided into approximately 200 source and header files. CoH solves the Schroedinger equation for optical potentials defined in the code, and calculates differential elastic scattering, reaction, and total cross sections, for neutron, proton, deuteron, triton, 3He, and alpha-particle. Deformed optical potentials are solved with the coupled-channels method, in which the ground state rotational band members, or vibrational phonon states are coupled. The optical model gives particle transmission coefficients that are fed into the statistical model calculations. CoH includes the pre-equilibrium model (exciton model), the direct/semidirect capture model, and the multi-stage Hauser-Feshbach statistical decay with width fluctuation correction based on the Gaussian orthogonal ensemble. For weakly coupled levels, the DWBA (distorted wave Born approximation) method is used to calculate the direct inelastic scattering process to the excited states.

Kawano, Toshihiko↗

Non-Gaussian Generalized Two-Mode Squeezing: Applications to Two-Ensemble Spin Squeezing and Beyond

Bosonic two-mode squeezed states are paradigmatic entangled Gaussian states that have wide utility in quantum information and metrology. Here, in this study, we show that the basic structure of these states can be generalized to arbitrary bipartite quantum systems in a manner that allows simultaneous, Heisenberg-limited estimation of two independent parameters for finite-dimensional systems. Further, we show that these general states can always be stabilized by a relatively simple Markovian dissipative process. In the specific case where the two subsystems are ensembles of two-level atoms or spins, our generalized states define a notion of two-mode spin squeezing that is valid beyond the Gaussian limit and that enables true multiparameter estimation. We discuss how generalized Ramsey measurements allow one to reach the two-parameter quantum Cramér-Rao bound, and how the dissipative preparation scheme is compatible with current experiments.

Mamaev, Mikhail [Univ. of Chicago, IL (United Stat↗

Joint cosmic density reconstruction from photometric and spectroscopic samples

ABSTRACT We reconstruct the dark matter density field from spatially overlapping spectroscopic and photometric redshift catalogues through a field-level forward modelling approach. Instead of directly inferring the underlying density field, we find the best-fitting initial Gaussian fluctuations that will evolve into the observed cosmic volume. To account for the substantial uncertainty of photometric redshifts we employ a differentiable continuous Poisson process. As an initial test, we construct a mock based on the upcoming Prime Focus Spectrograph combined with photometric sample modelled on the Subaru Hyper Suprime-Cam. Depending on the statistic of interest, we find improvements in cosmic structure classification equivalent to 50–100 per cent more spectroscopic targets by combining relatively sparse spectroscopic with dense photometric samples.

Horowitz, B.↗

Efficient near-field ptychography reconstruction using the Hessian operator

X-ray ptychography is a powerful and robust coherent imaging method providing access to the complex object and probe (illumination). Ptychography reconstruction is typically performed using first-order methods due to their computational efficiency. Higher-order methods, while potentially more accurate, are often prohibitively expensive in terms of computation. In this study, we present a mathematical framework for reconstruction using second-order information derived from an efficient computation of the bilinear Hessian and Hessian operator. The formulation is provided for Gaussian-based models, enabling the simultaneous reconstruction of the object, probe, and object positions. Synthetic data tests, along with experimental near-field ptychography data processing, demonstrate a ten-fold reduction in computation time compared to first-order methods. The derived formulas for computing the Hessians, along with the strategies for incorporating them into optimization schemes, are well-structured and easily adaptable to various ptychography problem formulations.

Carlsson, Marcus [Lund Univ. (Sweden)] (ORCID:0000↗

CAFE AU LAIT: Compute-Aware Federated Augmented Low-Rank AI Training

Federated finetuning is crucial for unlocking the knowledge embedded in pretrained Large Language Models (LLMs) when data are geographically distributed across clients. Unlike finetuning with data from a single institution, federated finetuning allows collaboration across multiple institutions, enabling the utilization of diverse and decentralized datasets while preserving data privacy. Given the high computing costs of LLM training and the emphasis on energy efficiency in Federated Learning (FL), Low-Rank Adaptation (LoRA) has emerged as a widely adopted algorithm due to its significantly reduced number of trainable parameters. However, this assumes that all data silos have the necessary computing resources to compute local updates of LLMs. Nevertheless, in practice, the computing resources across clients are highly heterogeneous: while some may have access to hundreds of GPUs, others might have limited or no GPU access. Recently, federated finetuning using synthetic data has been proposed, allowing clients to participate in a collaborative training run without training LLMs locally. However, our experimental results reveal a performance gap between models trained using synthetic data and those trained using local updates. Motivated by the observed heterogeneity in computing resources and the performance gap, we propose a novel two-stage algorithm that leverages the storage and computing capabilities of a strong server. In the first stage, under the coordination of the strong server, clients with limited computing resources collaborate to generate synthetic data, which is transferred to and stored on the strong server. In the second stage, the strong server uses this synthetic data on behalf of the resource-constrained clients to perform federated LoRA finetuning alongside clients with sufficient computing resources. This approach ensures that all clients can participate in the finetuning process. Experimental results demonstrate that incorporating local updates from even a small fraction of clients improves performance compared to using synthetic data for all clients. Furthermore, we incorporate the Gaussian mechanism in both stages to guarantee client-level differential privacy.

Wang, Jiayi [ORNL]↗

Reconfigurable unitary transformations of optical beam arrays

Spatial transformations of light are ubiquitous in optics, with examples ranging from simple imaging with a lens to quantum and classical information processing in waveguide meshes. Multi-plane light converter (MPLC) systems have emerged as a platform that promises completely general spatial transformations, i.e., a universal unitary. However, until now, MPLC systems have demonstrated transformations that are far from general, e.g., converting from a Gaussian to Laguerre-Gauss mode. Here, we demonstrate the promise of an MLPC, the ability to impose an arbitrary unitary transformation that can be reconfigured dynamically. Specifically, we consider transformations on superpositions of parallel free-space beams arranged in an array, which is a common information encoding in photonics. We experimentally test the full gamut of unitary transformations for a system of two parallel beams and make a map of their fidelity. We obtain an average transformation fidelity of 0.85 ± 0.03. This high-fidelity suggests that MPLCs are a useful tool for implementing the unitary transformations that comprise quantum and classical information processing.

47 OTHER INSTRUMENTATION↗

Investigation of post-breakup Coulomb acceleration using a trajectory model

Intermediate mass fragments ejected during the deexcitation of excited projectilelike fragments may promptly decay following ejection; the daughter particles that are subsequently produced are subject to interactions with the residual nucleus that affect final-state observables, a process herein referred to as post-breakup Coulomb acceleration. A simple classical Coulomb interaction model was used to study modification of 8 Be (2 + ), 5 Li (3/2 – ), 7 Li (7/2 – ), 7 Be (7/2 – ), and states in 12 B emitted following heavy-ion collisions of 28 Si + 12 C at 35 MeV/nucleon. Here, in contrast to previous work studying 8 Be (2 + ), excellent agreement between simulation and experiment was obtained using only Coulomb forces when either a Lorentzian or R-matrix line shape was used to describe the initial relative energy rather than a Gaussian. In consideration of the obtained results, improvements to the model and evaluation of experimental data are discussed as future directions, but it was concluded that the effects observed in the present data can be accurately described using only elements of classical mechanics and that the process is largely understood for a wide range of state lifetimes and mass and charge (a)symmetries.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Divide and conquer: separating the two probabilities in seismic phase picking

There are two fundamental probabilities in the seismic phase picking process—the probability of the existence of a seismic phase (detection probability) and the probability associated with the phase arrival time estimation (timing probability). The nearly ubiquitous approach in developing deep learning phase picking models is to use a kernel, such as a truncated Gaussian, to mask the labelled phase arrival time and train a segmentation model. Once a model is trained, the times of the peaks in the output are taken as phase arrival times (picks), and the height of the peaks are taken as ‘probability’ of the picks. Here, we show that this ‘probability’ represents neither the detection nor the timing probability because this approach forces the output to follow the shape of the kernel. We introduce an approach using two models to estimate these two distinct probabilities. We use a binary classifier with a calibrated confidence to address the detection probability and a multiclass classifier to obtain a probability mass function to address the timing probability. This new approach can make the deep learning-based phase picking process more interpretable and provide options to logically control seismic monitoring workflows.

58 GEOSCIENCES↗

Predicting Pulsed-Laser Deposition SrTiO 3 Homoepitaxy Growth Dynamics Using High-Speed Reflection High-Energy Electron Diffraction

Pulsed-laser deposition (PLD) is a powerful technique for growing complex oxides with controlled stoichiometry. To understand growth dynamics therein, it is common to leverage in situ spectroscopies, such as reflection high-energy electron diffraction (RHEED), to monitor surface crystallinity. Most commercial systems rely on video-rate cameras operating at 60-120 Hz that lack sufficient temporal resolution to capture growth dynamics at practical deposition frequencies. Here, a high-speed platform to record in situ dynamics via RHEED at >500 Hz is implemented. An open-source analysis package is designed to fit diffraction spots to 2D Gaussians, allowing single-pulse surface reconstruction kinetics extraction. Using homoepitaxially deposited (001)-oriented SrTiO 3 as a model system, we demonstrate how high-speed RHEED can provide real-time insight into growth processes obscured by slower acquisition systems. By fitting the single-pulse intensity to a set of exponential functions, we observe changes in the characteristic decay time and mechanism correlated to the substrate step width and surface termination. We observe distinct surface effects, with diffraction intensity decaying on lower-energy TiO 2 -terminated surfaces and stabilizing on SrO- or mixed-terminated surfaces. Similarly, using an exponential model, the extracted characteristic time of adatom deposition decreases with increased density of bonding sites associated with mixed termination and narrower step widths. Ultimately, this work shows how increasing RHEED temporal resolution can uncover new insights into growth processes, with practical implications for the design and control of PLD processes. This experimental platform provides new capabilities to enable data-driven machine learning analysis and autonomous control systems to enhance the complexity and fecundity of PLD.

(SrO)↗

Stimulated emission tomography for efficient characterization of spatial entanglement

Stimulated-emission tomography (SET) is an excellent tool for characterizing the process of spontaneous parametric down conversion (SPDC), which is commonly used to create pairs of entangled photons for use in quantum information protocols. The use of stimulated emission increases the average number of detected photons by several orders of magnitude compared to the spontaneous process. In a SET measurement, the parametric down conversion is seeded by an intense signal field prepared with specified mode properties rather than by broadband multimodal vacuum fluctuations, as is the case for the spontaneous process. The SET process generates an intense idler field in a mode that is the complex conjugate to the signal mode. In this work we use SET to estimate the joint spatial mode distribution (JSMD) in the Laguerre-Gaussian (LG) basis of the two photons of an entangled photon pair. The pair is produced by parametric down conversion in a beta barium borate (BBO) crystal with type-II phase matching pumped at a wavelength of 405 nm along with a 780 nm seed signal beam prepared in a variety of LG modes to generate an 842 nm idler beam of which the spatial mode distribution is measured. We observe strong idler production and good agreement with the theoretical prediction of its spatial mode distribution. Our experimental procedure should enable the efficient determination of the photon-pair wavefunctions produced by low-brightness SPDC sources and the characterization of high-dimensional entangled-photon pairs. Published by the American Physical Society 2024

Xu, Yang (ORCID:0009000954547320)↗

Resonant Slow Extraction Simulation using Bmad

Simulations of slow extraction and transport of charged particle beams from synchrotrons requires careful modeling and experimentation. In this paper, we outline how the Bmad modeling software was adapted and developed to run third-integer resonant extraction simulations of beams at Brookhaven National Laboratory’s Booster synchrotron. Further, we show experimental comparisons of the slowly extracted beams transferred to the NASA Space Radiation Laboratory (NSRL) transport line. In this process, beam passes through a stripping foil element at the extraction point, which, along with stripping remaining electrons from the beam ions, acts as a scatterer to modify the phase space, moving to a more Gaussian-like distribution. This modification to the beam helps generate a uniform beam at the beam line’s target location; which is necessary for the variety of experiments performed at the facility. During this work, beam energy loss and multiple scattering by foil routines were built in conjunction with the Bmad code developers and are now integrated into the software.

43 PARTICLE ACCELERATORS↗

Silicon-On-Sapphire Metasurfaces Generate Arrays of Dark and Bright Traps for Neutral Atoms

We demonstrated crystalline silicon-on-sapphire (c-SOS) metasurfaces that convert a Gaussian beam into arrays of complex optical traps, including arrays of optical bottle beams that trap atoms in dark regions interleaved with bright tweezer arrays. The high refractive index and indirect band gap of crystalline silicon make it possible to design high-resolution near-infrared (λ > 700 nm) metasurfaces that can be manufactured at scale using CMOS-compatible processes. Compared with active components like spatial light modulators (SLMs) that have become widely used to generate trap arrays, metasurfaces provide an indefinitely scalable number of pixels, enabling large arrays of complex traps in a very small form factor, as well as reduced dynamic noise. To design metasurfaces that can generate three-dimensional bottle beams to serve as dark traps, we modified the Gerchberg-Saxton algorithm to enforce complex-amplitude profiles at the focal plane of the metasurface and to optimize the uniformity of the traps across the array. We fabricated and measured c-SOS metasurfaces that convert a Gaussian laser beam into arrays of bright traps, dark traps, and interleaved bright/dark traps.

Gerchberg-Saxton↗

Robust direct laser acceleration of electrons with flying-focus laser pulses

Direct laser acceleration (DLA) offers a compact source of high-charge, energetic electrons for generating secondary radiation or neutrons. While DLA in high-density plasma optimizes the energy transfer from a laser pulse to electrons, it exacerbates nonlinear propagation effects, such as filamentation, that can disrupt the acceleration process. Here, we show that superluminal flying-focus pulses (FFPs) mitigate nonlinear propagation, thereby enhancing the number of high-energy electrons and resulting x-ray yield. Three-dimensional particle-in-cell simulations show that, compared to a Gaussian pulse of equal energy (1 J) and intensity (2 × 10 20 W/cm 2 ), an FFP produces 80 × more electrons above 100 MeV, increases the electron cutoff energy by 20%, triples the high-energy x-ray yield, and improves x-ray collimation. These results illustrate the ability of spatiotemporally structured laser pulses to provide additional control in the highly nonlinear, relativistic regime of laser-plasma interactions.

Laser-produced plasmas↗

The Langdon effect in laser plasmas: Absorption and conduction

A plasma heated by inverse bremsstrahlung absorption of laser light develops a non-Maxwellian electron distribution function, called the Langdon effect [A. B. Langdon, Phys. Rev. Lett. 44, 575 (1980)]. These non-Maxwellian distributions are sufficiently long-lived to impact the absorption processes itself as well as the transport of heat by electrons. The theory of the Langdon effect in a homogeneous plasma is reviewed to clarify some aspects of Langdon's derivation as well as to confirm that the widely used super-Gaussian approximation works fairly well to describe the shape of the distribution function and reduction of the absorption rate. The Langdon effect on thermal conduction in an inhomogeneous plasma is developed by considering perturbations in a homogeneous absorbing plasma, which develops a heat flux due to both temperature and density gradients. A practical theory of the heat flux is developed by fitting the results of Vlasov–Fokker–Planck simulations, which avoids several approximations that compromised the usefulness of past theoretical predictions, most critically, the effect of electron–electron collisions on the fluxes. The present fits parameterize the coefficients of the temperature gradient (thermal conductivity) and the density gradient for a plasma of any ionization state and for any laser intensity where the theory of the Langdon effect remains locally valid. It is expected that this generalized theory of heat flow in an absorbing plasma will improve the predictive capability of radiation-hydrodynamics simulations of laser-produced plasmas, especially those formed in inertial confinement fusion experiments.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Lie-algebraic classical simulations for quantum computing

The classical simulation of quantum dynamics plays an important role in our understanding of quantum complexity and in the development of quantum technologies. Efficient techniques such as those based on the Gottesman-Knill theorem for Clifford circuits, tensor networks for low entanglement-generating circuits, or Wick's theorem for fermionic Gaussian states have become central tools in quantum computing. In this work, we contribute to this body of knowledge by presenting a framework for classical simulations, dubbed “𝔤-sim”, which is based on the underlying Lie algebraic structure of the dynamical process. When the dimension of the algebra grows at most polynomially in the system size, there exist observables for which the simulation is efficient. Indeed, we show that 𝔤-sim enables new regimes for classical simulations, is able to deal with certain forms of noise in the evolution, as well as can be used to tackle several paradigmatic variational and nonvariational quantum computing tasks. For the former, we perform Lie-algebraic simulations to train and optimize parametrized quantum circuits (thus effectively showing that some variational models can be dequantized), design enhanced parameter initialization strategies, solve tasks of quantum circuit synthesis, and train a quantum-phase classifier. For the latter, we report large-scale noiseless and noisy simulations on benchmark problems. By comparing the limitations of 𝔤-sim and certain Wick's theorem-based simulations, we find that the two methods become inefficient for different types of states or observables, hinting at the existence of distinct, nonequivalent resources for classical simulation.

97 MATHEMATICS AND COMPUTING↗

Search for electroweak-scale dijet resonances using trigger-level analysis with the ATLAS detector in 132 fb −1 of 𝑝⁢𝑝 collisions at $\sqrt{𝑠}$ = 13 TeV

This article reports on a search for dijet resonances using 132 fb −1 of 𝑝⁢𝑝 collision data recorded at $\sqrt{𝑠}$ = 13 TeV by the ATLAS detector at the Large Hadron Collider. The search is performed solely on jets reconstructed within the ATLAS trigger to overcome bandwidth limitations imposed on conventional single-jet triggers, which would otherwise reject data from decays of sub-TeV dijet resonances. Collision events with two jets satisfying transverse momentum thresholds of 𝑝 T ≥ 85 GeV and jet rapidity separation of |𝑦*| <0.6 are analysed for dijet resonances with invariant masses from 375 to 1800 GeV. A data-driven background estimate is used to model the dijet mass distribution from multijet processes. No significant excess above the expected background is observed. Upper limits are set at 95% confidence level on coupling values for a benchmark leptophobic axial-vector 𝑍′ model and on the production cross section for a new resonance contributing a Gaussian-distributed line-shape to the dijet mass distribution.

hadron colliders↗

The Poisson tensor completion non-parametric differential entropy estimator

We introduce the Poisson tensor completion (PTC) estimator, a non-parametric differential entropy estimator. The PTC estimator leverages inter-sample relationships to compute a low-rank Poisson tensor decomposition of the frequency histogram. Our crucial observation is that the histogram bins are an instance of a space partitioning of counts and thus can be identified with a spatial Poisson process. The Poisson tensor decomposition leads to a completion of the intensity measure over all bins—including those containing few to no samples—and leads to our proposed PTC differential entropy estimator. A Poisson tensor decomposition models the underlying distribution of the count data and guarantees non-negative estimated values and so can be safely used directly in entropy estimation. Our estimator is the first tensor-based estimator that exploits the underlying spatial Poisson process related to the histogram explicitly when estimating the probability density with low-rank tensor decompositions for the purpose of tensor completion. Furthermore, we demonstrate that our PTC estimator is a substantial improvement over standard histogram-based estimators for sub-Gaussian probability distributions because of the concentration of norm phenomenon.

42 ENGINEERING↗