Search NASA⌕ Search

SEARCH · Search NASA

Results for “High dimensional data,”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

High-throughput spin-bath characterization of spin defects in semiconductors

Detailed knowledge of the local environments of spin defects in semiconductors, such as nitrogenvacancy (NV) centers in diamond or divacancies in silicon carbide, is crucial for optimizing control and entanglement protocols in quantum sensing and information applications. However, at present a direct experimental characterization of individual defect environments is not scalable, as conventional spin-bath measurements are time consuming and difficult to automate. Achieving high-throughput characterization requires short experiments to probe the spin bath. However, with fewer and noisier measurements, the inverse problem of recovering spin-bath properties from measured data becomes ill posed, with multiple spin baths having a high likelihood of yielding the same data. In this work, we present a set of computational tools to resolve the ill-posed inverse problem of recovering the atomic positions and hyperfine couplings of random nuclei surrounding spin defects from sparse, noisy experimental coherence data, which can be obtained in hours. Here, we use a trans-dimensional Bayesian approach that incorporates ab initio data to yield full posterior distributions over nuclear spin environments, enabling robust recovery from limited data. We also provide practical tools and guidelines to determine the limits of detectability for hyperfine couplings under specific dynamical decoupling sequences and sampling conditions. In addition, we demonstrate how the tools developed here, in combination with ab initio simulations of spin baths, can guide the design of efficient experimental protocols for application-specific high-throughput screening. To showcase the utility of our approach, we apply it to design fast dynamical decoupling experiments to characterize the spin baths often individual NV centers in diamond. While the primary focus is on accelerating spin-bath characterization of spin defects, this Bayesian approach also lays the foundation for digital-twin studies of spin defects, where a virtual model of the spin-defect system evolves in real time with ongoing experimental measurements. Together, the set of tools we designed and applied paves the way for scalable deployment of spin defects in semiconductors for quantum sensing and information applications.

Bayesian methods↗

Dense autoencoders, clustering techniques, and semi-supervised learning for HPGe $γ$-spectra

Classifying high-resolution gamma spectra by their isotopic content is an essential task in nuclear forensics and other applications. Traditional analysis methods are often time-intensive, but machine learning (ML) may help analysts quickly process many spectra. Such methods tend to rely on abundant, well-labeled data for training. Historical gamma data exists in various fields but is not uniformly useful for supervised ML due to inconsistent labeling. Here, to address some of these challenges, we present a method to classify and organize unlabeled data from high-purity germanium detectors using an autoencoding neural network (autoencoder). We trained dense autoencoders to compress gamma data into latent representations that enable efficient data characterization. By clustering the encoded spectra or lower-dimensional mappings of them, we identified and removed portions of over-abundant data categories, resulting in a more balanced dataset and improved autoencoder performance. This encoding and clustering pipeline also enabled the organization of spectra into self-consistent categories. Finally, we found that encoded representations showed potential as inputs for semi-supervised learning of nuclide identification (NID) labels, achieving an average F1 score of 0.85 ± 0.03 when mapping encodings to a set of 65 isotope labels.

Autoencoders↗

Toward ultra-efficient high-fidelity predictions of wind turbine wakes: Augmenting the accuracy of engineering models with machine learning

This study proposes a novel machine learning (ML) methodology for the efficient and cost-effective prediction of high-fidelity three-dimensional velocity fields in the wake of utility-scale turbines. The model consists of an autoencoder convolutional neural network with U-Net skipped connections, fine-tuned using high-fidelity data from large-eddy simulations (LES). The trained model takes the low-fidelity velocity field cost-effectively generated from the analytical engineering wake model as input and produces the high-fidelity velocity fields. The accuracy of the proposed ML model is demonstrated in a utility-scale wind farm for which datasets of wake flow fields were previously generated using LES under various wind speeds, wind directions, and yaw angles. Comparing the ML model results with those of LES, the ML model was shown to reduce the error in the prediction from 20% obtained from the Gauss Curl hybrid (GCH) model to less than 5%. In addition, the ML model captured the non-symmetric wake deflection observed for opposing yaw angles for wake steering cases, demonstrating a greater accuracy than the GCH model. The computational cost of the ML model is on par with that of the analytical wake model while generating numerical outcomes nearly as accurate as those of the high-fidelity LES.

Mechanics↗

High-Fidelity Simulation Aerodynamics Dataset of NACA 0012 and 0021 Airfoils

The repository contains time series data of pressure, viscous and moment forces for the NACA 0012 and 0021 airfoils at angles of attack (AOA), 5, 17, 30, 45, 60 and 90 degrees. The data was generated during the high-fidelity simulations of the airfoils using the open-source CFD code, Nalu-Wind (https://github.com/Exawind/nalu-wind). The research objective was to study the effects of mesh resolution and turbulence model on the three-dimensional deep stall aerodynamics of the airfoils, which has been published in the Journal of Turbulence (https://doi.org/10.1080/14685248.2023.2225141). Several meshes of the O-grid type were generated using the modified Pointwise Glyph script of Carrigan (https://github.com/pointwise/AirfoilMesh) by varying the wall normal and spanwise resolutions. Note that, the geometry was extruded by four chord lengths in the spanwise direction. In the paper, we showed that Improved Delayed Detached Eddy Simulation (IDDES) hybrid RANS/LES turbulence model with a minimum of 24 cells per chord length in the spanwise direction is necessary to correctly predict the loads in the deep stall regime. Under the NACA 0012 airfoil directory, the data for several combinations of wall-normal and spanwise resolutions are presented in each of the AOA subdirectories. The corresponding experimental data for the lift and drag polars can be found in the "exp_data" subdirectory. The directory for the NACA 0021 airfoil consists of CFD force data and airfoil polars for Reynolds numbers 2.7e5 and 2.0e6.

17 WIND ENERGY↗

Enhancing Short-Range Weather Forecasts through Temporal Variation Encoding: A Multiperiod Embedding Approach

Machine learning (ML) techniques have emerged as promising approaches to improve regional weather forecast accuracy and reliability through data-driven methods. We propose a novel ML-based weather forecasting model, the Multiperiod Embed Net (MPENet). A key distinguishing feature of MPENet is its explicit utilization of the inherent cyclic nature in weather dynamics, unlike the autoregressive strategies commonly used in other ML weather forecasting approaches. Critical cyclic structures are identified via Fourier analyses of dynamic time series. Cyclicity in the convolutional representation is achieved by transforming one-dimensional time series of meteorological variables into two-dimensional tensors based on identified periods. This approach enables the model to leverage intrinsic weather patterns, enhancing regional forecast performance. To demonstrate the effectiveness of MPENet, we conduct a comparative analysis with Nvidia’s FourCastNet. Both models are trained on High-Resolution Rapid Refresh (HRRR) data from 2015 to 2022, over a 192 km × 192 km region in Tennessee. The comparisons are performed locally at two specific locations known to have different weather dynamics due to orographic effects: Crossville, on the relatively flat Cumberland Plateau with fewer topographic airflow disruptions, and Oak Ridge, in the ridge-and-valley region, where airflow is heavily influenced by surrounding valleys and mountains. Our results indicate that FourCastNet achieves strong accuracy at very short lead times, while MPENet maintains competitive skill and shows advantages in capturing temporal evolution over longer periods. Cross-correlation analyses of MPENet and FourCastNet predictions with the HRRR data suggest that encoding critical cyclicity into the network architecture leads to improvements in the forecasting skill.

Artificial intelligence↗

Hardware acceleration for HPS algorithms in two and three dimensions

We provide a flexible, open-source framework for hardware acceleration, namely massively-parallel execution on general-purpose graphics processing units (GPUs), applied to the hierarchical Poincaré–Steklov (HPS) family of algorithms for building fast direct solvers for linear elliptic partial differential equations. To take full advantage of the power of hardware acceleration, we propose two variants of HPS algorithms to improve performance on two- and three-dimensional problems. In the two-dimensional setting, we introduce a novel recomputation strategy that minimizes costly data transfers to and from the GPU; in three dimensions, we modify and extend the adaptive discretization technique of Geldermans and Gillman [1] to greatly reduce peak memory usage. We provide an open-source implementation of these methods written in JAX, a high-level accelerated linear algebra package, which allows for the first integration of a high-order fast direct solver with automatic differentiation tools. We conclude with extensive numerical examples showing our methods are fast and accurate on two- and three-dimensional problems.

Fast direct solvers↗

Two-Frequency RF Fields Induced Multipactor in Coaxial Transmission Lines

Multipactor is a nonlinear discharge phenomenon that occurs in vacuum RF systems, potentially leading to signal distortion, power loss, and even permanent damage to high-power components. This study presents a detailed investigation of two-surface multipactor in coaxial transmission lines under two-frequency excitation, using one-dimensional (1D) Monte Carlo simulations validated by three-dimensional (3D) Particle-in-Cell (PIC) results and experimental data. Introducing a second carrier mode is shown to suppress multipactor by reshaping and shrinking the susceptibility region, with the extent and location of suppression strongly dependent on the device aspect ratio and the relative phase of the second mode. Distinct suppression patterns emerge across different frequency–gap distance (fd) regimes, and in certain cases, susceptibility expansion is also observed. The study identifies and distinguishes pure and mixed multipactor modes in coaxial geometry, where analytical mode boundaries are not readily defined. Unlike planar systems, pure-mode regions in coaxial structures overlap with mixed-mode domains, complicating classification. Image charge forces are found to have minimal effect on susceptibility thresholds but do influence electron growth rates. These findings provide new insights into waveform-driven control of multipactor in high-power RF systems.

43 PARTICLE ACCELERATORS↗

Data mining and computational screening of Rashba-Dresselhaus splitting and optoelectronic properties in two-dimensional perovskite materials

Recent developments highlighting the promise of two-dimensional perovskites have vastly increased the compositional search space in the perovskite family. This presents a great opportunity for the realization of highly performant devices and practical challenges associated with the identification of candidate materials. High-fidelity computational screening offers great value in this regard. In this study, we carry out a multiscale computational workflow, generating a dataset of two-dimensional perovskites in the Dion-Jacobson and Ruddlesden-Popper phases. Our dataset comprises ten B-site cations, four halogens, and over 20 organic cations across over 2000 materials. We compute electronic properties, thermoelectric performance, and numerous geometric characteristics. Furthermore, we introduce a framework for the high-throughput computation of Rashba-Dresselhaus splitting. Finally, we use this dataset to train machine learning models for the accurate prediction of band gaps, candidate Rashba-Dresselhaus materials, and partial charges. The work presented herein can aid future investigations of two-dimensional perovskites with targeted applications in mind.

14 SOLAR ENERGY↗

DOE-ICoM/RIFT

Rapid Infrastructure Flood Tool (RIFT) is a two-dimensional hydrodynamic model based on the complete shallow water equations. RIFT has specifically been designed with rapid simulation in mind by utilizing commodity high performance computing technology and best-available nation-wide data. RIFT is used to predict the movement of water over land and resolve the spatial and temporal variability of flood depths, extent, and velocity. RIFT can be applied to many flood situations and has primarily been used to quantify flood extents from dam/levee failure or inland rainfall flooding.

Perkins, Bill [Pacific Northwest National Laborato↗

Architecture-Aware Models of AI Engines for High-Performance Matrix Matrix Multiplication

The AI Engine (AIE) architecture, available in systems from mobile SoCs to server-class FPGAs, aims to efficiently execute AI/ML tasks through a two-dimensional array of compute tiles. Previous work on AIEs has explored different approaches to mapping computation across spatial arrays, but the compute kernel running on each tile has not been the focus. Additionally, the AIE-ML architecture introduces memory tiles and omits programmable logic, requiring new approaches to staging and moving data throughout the array. In this work we update analytical models developed for CPUs to produce the design of high performance kernels while introducing new model considerations such as memory structure, throughput, and latency as required by the AIE hardware. We evaluate our models by developing AIE-ML kernels for matrix multiplication in low-precision data types showing performance up to 95% of compute peak for the kernel when data resides in local memory and above 90% of compute peak when data resides in main memory.

Binder, Elliott D. [Carnegie Mellon University, Pi↗

Properties of Electronic Materials

This final technical report summarizes the research conducted under DOE Grant DE-SC0002623, "Properties of Electronic Materials," led by Principal Investigator Shengbai Zhang at Rensselaer Polytechnic Institute. Over the 16-year period, the project employed first-principles computational methods to investigate the structural, electronic, and dynamic properties of a wide range of electronic materials, with applications in energy technologies, optoelectronics, and data storage. Key areas included topological insulators, phase-change materials, graphene and two-dimensional systems, perovskites for photovoltaics, defect engineering in semiconductors, kagome lattices, and ultrafast carrier dynamics. The research resulted in 115 peer-reviewed publications, advancing fundamental understanding of material behaviors at the atomic scale and contributing to innovations in renewable energy, memory devices, and quantum materials. Findings have implications for improving energy efficiency, developing lead-free solar cells, and enabling high-speed data processing. The work has trained numerous graduate students and postdocs, fostering the next generation of computational materials scientists. The original goals were to develop theoretical models and computational tools to predict and optimize electronic properties of materials for energy applications. All objectives were accomplished, with no major departures from planned methodologies. Challenges in computational scaling were addressed through access to high-performance computing resources.

36 MATERIALS SCIENCE↗

Nucleon structure studies: DVCS on polarised protons with the CLAS12 experiment, and development of Micromegas detectors for EIC

The quark and gluon structure of nucleons is crucial for understanding the origin of their mass and spin. This information is encoded in structure function such as Generalised Parton Distributions (GPDs), which provide a three-dimensional picture of the nucleon in terms of its constituents. Deeply Virtual Compton Scattering (DVCS) offers the most direct access to GPDs, but their extraction requires high-precision measurements of multiple observables over a wide kinematic range. In 2022 and 2023 the CLAS12 experiment at Jefferson Lab (JLab) collected data from polarised electron scattering on a longitudinally polarised proton target, enabling the first measurement of polarised DVCS asymmetries at JLab 12GeV kinematics. This thesis will present preliminary measurement of the beam, target and double spin DVCS asymmetries from the CLAS12 polarised proton target data. Nucleon structure studies will also constitute a major part of the physics program at the future Electron Ion Collider (EIC). The development of the first detector at the EIC is ongoing and it requires light Micro-Pattern Gaseous Detectors (MPGDs) for its tracking system. This thesis further reports initial tests of Micromegas MPGDs with a two-dimensional readout developed for EIC.

Polcher, Samy [Univ. Paris-Saclay, Gif-sur-Yvette ↗

Efficient lattice QCD computation of radiative-leptonic-decay form factors at multiple positive and negative photon virtualities

In previous work [D. Giusti, Methods for high-precision determinations of radiative-leptonic decay form factors using lattice QCD, Phys. Rev. D 107, 074507 (2023)], we showed that form factors for radiative leptonic decays of pseudoscalar mesons can be determined efficiently and with high precision from lattice QCD using the “three-dimensional (3D) method,” in which three-point functions are computed for all values of the current insertion time and the time integral is performed at the data-analysis stage. Here, we demonstrate another benefit of the 3D method: the form factors can be extracted for any number of nonzero photon virtualities from the same three-point functions at no extra cost. We present results for the $D_s → ℓνγ*$ vector form factor as a function of photon energy and photon virtuality, for both positive and negative virtuality, for a single ensemble with 340 MeV pion mass and 0.11 fm lattice spacing. In our analysis, we separately consider the two different time orderings and the different quark flavors in the electromagnetic current. We discuss in detail the behavior of the unwanted exponentials contributing to the three-point functions, as well as the choice of fit models and fit ranges used to remove them for various values of the virtuality. While positive photon virtuality is relevant for decays to multiple charged leptons, negative photon virtuality suppresses soft contributions and is of interest in QCD-factorization studies of the form factors.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Analysis of streaked images of x-ray self-emission in laser-driven spherical implosions

Imaging of x-ray self-emission provides a powerful in situ measurement of the spatial and temporal evolution of high-energy-density plasmas. However, interpretation of these measurements requires detailed understanding of the data-generating process. This work presents a case study in the interpretation of x-ray self-emission data for the specific application of streaked one-dimensional slit imaging of spherical laser-driven implosions. A comprehensive generative model of the streaked slit-imaging diagnostic is developed including detailed treatments of the radiation transfer, photometrics, and photostatistics associated with the measurement. The model is used to generate realistic synthetic streaked images and to analyze experimental streaked images to extract important physical quantities of interest. An example analysis of streaked images from implosion experiments on the OMEGA laser is presented, where the model developed in this work is used to constrain the trajectory and peak velocity of the implosion using Bayesian inference.

Bayesian inference↗

SO(3)-invariant PCA with application to molecular data

Principal component analysis (PCA) is a fundamental technique for dimensionality reduction and denoising; however, its application to three-dimensional data with arbitrary orientations -- common in structural biology -- presents significant challenges. A naive approach requires augmenting the dataset with many rotated copies of each sample, incurring prohibitive computational costs. In this paper, we extend PCA to 3D volumetric datasets with unknown orientations by developing an efficient and principled framework for SO(3)-invariant PCA that implicitly accounts for all rotations without explicit data augmentation. By exploiting underlying algebraic structure, we demonstrate that the computation involves only the square root of the total number of covariance entries, resulting in a substantial reduction in complexity. We validate the method on real-world molecular datasets, demonstrating its effectiveness and opening up new possibilities for large-scale, high-dimensional reconstruction problems.

Fraiman, Michael [Tel Aviv Univ., Tel Aviv (Israel↗

Conservative projection-based data-driven model order reduction of a fluid-kinetic spectral solver

Kinetic simulations are computationally intensive due to six-dimensional phase space discretization. Many kinetic spectral solvers use the asymmetrically weighted Hermite expansion due to its conservation and fluid-kinetic coupling properties, i.e., the lower-order Hermite moments capture and describe the macroscopic fluid dynamics, and higher-order Hermite moments describe the microscopic kinetic dynamics. We leverage this structure by developing a parametric data-driven reduced-order model based on the proper orthogonal decomposition, which projects the higher-order kinetic moments while retaining the fluid moments intact. We demonstrate analytically and numerically that the method ensures local and global mass, momentum, and energy conservation. The numerical results show that the proposed method effectively replicates the high-dimensional spectral simulations at a fraction of the computational cost and memory, as validated on the weak Landau damping and two-stream instability benchmark problems.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Validation of Pronghorn for Natural-Circulation Molten Salt Loops

This paper presents the development and validation of a high-fidelity thermal-hydraulic model of a molten salt natural circulation flow loop, designed for integration within a digital twin framework. The study evaluates the performance of Idaho National Laboratory’s Pronghorn against experimental data from Texas A&M University Molten Salt Flow Loop (MSFL) four Hitec-salt test benchmark data. Natural circulation of high-Prandtl-number fluids exhibits complex, counter-intuitive flow patterns that make pointwise thermocouple readings unreliable. Experimental work at TAMU’s MSFL provides benchmark data, including flow visualization at a test-section and centerline steady-state temperature measurements along the loop. Validation includes four single-phase natural circulation test cases with Hitec salt. Key metrics include flow profile agreement and steady-state temperature accuracy. Pronghorn results for two-dimensional single-phase agree qualitatively with the experimental flow profile. This paper illustrates the importance of Computational Fluid Dynamics (CFD) in elucidating the behavior of high-Prandtl-number thermal-hydraulics, along with how misleading centerline temperature measurements can be. Pronghorn reproduces the axial and radial stratification that makes single thermocouple readings unreliable. Future research will focus on reduced-order modeling techniques to enable rapid simulation suitable for real-time digital twin applications. The validated cases provide a basis for developing reduced-order surrogates aimed at real-time digital-twin applications.

22 - GENERAL STUDIES OF NUCLEAR REACTORS↗