Search NASA⌕ Search

SEARCH · Search NASA

Results for “Kernel”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

A methodology for using nonlinear aerodynamics in aeroservoelastic analysis and design

A methodology is presented for using the Volterra-Wiener theory of nonlinear systems in aeroservoelastic (ASE) analyses and design. The theory is applied to the development of nonlinear aerodynamic response models that can be defined in state-space form and are, therefore, appropriate for use in modern control theory. The theory relies on the identification of nonlinear kernels that can be used to predict the response of a nonlinear system due to an arbitrary input. A numerical kernel identification technique, based on unit impulse responses, is presented and applied to a simple bilinear, single-input-single-output system. The linear kernel (unit impulse response) and the nonlinear second-order kernel of the system are numerically-identified and compared with the exact, analytically-defined linear and second-order kernels. This kernel identification technique is then applied to the CAP-TSD code for identification of the linear and second-order kernels of a NACA64A010 rectangular wing undergoing pitch at M = 0.5, M = 0.85 (transonic), and M = 0.93 (transonic). Results presented demonstrate the feasibility of this approach for use with nonlinear, unsteady aerodynamic responses.

Silva, Walter A.↗

Reduced Dimensionality Analysis of TEMPO Ozone Profile Retrievals Using the Compact Phase Space (CPSR) Algorithm

TEMPO ozone (O 3 ) profile retrievals are expected to have fidelity in the troposphere due the sensitivities of the associated averaging kernels. However, those averaging kernels are severely rank deficiency meaning that a visual inspection of the vertical structure of the averaging kernel profile sensitivities is misleading due linear dependencies in the profile. The Compact Phase Space Retrieval (CPSR) algorithm use singular value decompositions of the averaging kernels and the ‘compressed’ retrieval solution error covariance to project the transformed averaging kernels into a space that removes the linear dependencies and accounts for the solution error uncertainties. In this oral presentation and poster, we apply the CPSR dimensional reduction analysis to TEMPO and TROPOMI O 3 profile retrievals for 13:45 UTC March 29, 2024 to study the phase space characteristics of the transformed averaging kernels as a function of latitude for North America. Our results show that TEMPO generally has more phase space vertical structure in the troposphere than TROPOMI. TEMPO has four to five dominant modes, and TROPOMI has five to six dominant modes. That means that dimensional reduction can reduce the TEMPO resource requirements by ~77% and the TROPOMI requirements by ~81%. Finally, we found that after removing linear dependences and after accounting for solution uncertainties TEMPO still has sensitivities throughout the troposphere.

TEMPO↗

DNS of ignition and flame stabilization in a simplified gas turbine premixer

With the increasing need for fuel flexibility, mitigation of auto-ignition (AI) inside gas turbine (GT) premixers becomes crucial. They must be designed to yield a sufficiently homogeneous fuel-air mixture to achieve low emissions while at the same time avoiding the occurrence of AI and subsequent flame stabilization. This challenge requires a detailed understanding of turbulent mixing and chemistry interactions. In the present work, a direct numerical simulation (DNS) of an array of jets in crossflow (JICF), representative of an industrial GT premixer, is reported to shed light on these complex phenomena. It is found that AI kernels form in the aft part of the premixer and coalesce into a flame front that then propagates upstream, mainly through the boundary layer, and successively engulfs the jets. This, therefore, suggests a significant role of the jet array pattern on the flame stabilization. It is noted that AI kernels continue to form independently during the whole time of the simulation. To clarify the contribution of AI and diffusion in the ignition kernels and the main flame, chemical explosive mode analysis (CEMA) is employed jointly with a kernel tracking algorithm. It is found that during the initial formation of the flame, many ignition kernels form in mixtures with low scalar dissipation rate and large contribution from AI mode. As they quickly grow, they merge into a single flame front that becomes increasingly more diffusion-assisted over time, balancing the AI mode. Turbulence is shown to have a significant enhancing effect in lean premixed flames, but further analysis is required to fully characterize it. These findings are relevant for the industrial premixer studied, and also for novel micromix concepts that may be used in the next generation of GT combustion systems.

ADVANCED PROPULSION SYSTEMS↗

Post-irradiation Heating Tests of As-Irradiated AGR-3/4 TRISO Fuel Compacts

Four post-irradiation heating tests of fuel compacts from the U.S. Advanced Gas Reactor (AGR)-3/4 irradiation experiment were completed. In addition to tristructural isotropic (TRISO)-coated driver fuel, each compact contained designed-to-fail (DTF) particles with fuel kernels coated only in pyrocarbon so as to simulate exposed kernels. Tests at 1600/1700°C, 1400°C, and 1200°C were performed to measure fission product release as a function of time and temperature. Silver release was highest in the 1200°C test, supporting the observation that silver release rates are highest in the 1100–1300°C range. Compared to tests of AGR-1 compacts with no exposed kernels, the Cs-134 and Kr-85 releases were noticeably higher in AGR-3/4. The exposed kernels’ contributions to Eu and Sr release are inconclusive, due to the difficulty in distinguishing among the combined effects of higher irradiation temperatures in these particular AGR-3/4 compacts, the presence of the DTF particles, and the Fuel Accident Condition Simulator (FACS) test temperatures. These data can be used to make inferences about fission product retention in exposed kernels as a function of time and temperature.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Multiphysics Degradation Modeling of Energy Storage Materials via RKPM with a Neural Network-Enhancement

In energy storage materials, strong electrochemical-mechanical coupling and highly anisotropic material properties contribute to the formation and propagation of micro-cracking during charge/discharge cycling, resulting in reduced performance and service life. A coupled electro-chemo-mechanical reproducing kernel particle method (RKPM) formulation is developed, and a patch-test is formulated to certify optimal convergence of the proposed RKPM method for the coupled physics system. With microstructural images supplied by the National Renewable Energy Laboratory (NREL), pixel-based model construction by RKPM is then used to represent the complex material microstructures for modeling the coupled physics of these systems. Further, a neural network-enhanced reproducing kernel particle method (NN-RKPM) [1, 2] is introduced to effectively model damage and crack propagation in the material microstructures; the location, orientation, and solution transition near a localization are automatically captured by superimposed block-level NN optimizations. This NN enrichment approach allows for effective modeling of localizations via a fixed background discretization, relieving tedious efforts for adaptive refinement in traditional mesh-based methods. Applications to the heterogeneous microstructures of Li-ion battery cathodes will be presented to demonstrate the effectiveness of the proposed methods. Reference: [1] Baek, J., Chen, J. S., Susuki, K., "Neural Network enhanced Reproducing Kernel Particle Method for Modeling Localizations," International Journal for Numerical Methods in Engineering, Vol. 123, pp 4422-4454, https://doi.org/10.1002/nme.7040, 2022. [2] Baek, J., Chen, J. S., "A Neural Network-Based Enrichment of Reproducing Kernel Approximation for Modeling Brittle Fracture", Computer Methods in Applied Mechanics and Engineering Vol. 410, 116590, 2024.

electro-chemo-mechanical coupling↗

Gaussian Process Regression under Computational and Epistemic Misspecification

Gaussian process regression is a classical kernel method for function estimation and data interpolation. In large data applications, computational costs can be reduced using low-rank or sparse approximations of the kernel. This paper investigates the effect of such kernel approximations on the interpolation error. We introduce a unified framework to analyze Gaussian process regression under important classes of computational misspecification: Karhunen-Loève expansions that result in low-rank kernel approximations, multiscale wavelet expansions that induce sparsity in the covariance matrix, and finite element representations that induce sparsity in the precision matrix. Furthermore, our theory also accounts for epistemic misspecification in the choice of kernel parameters.

Gaussian process regression↗

Forward variable selection enables fast and accurate dynamic system identification with Karhunen-Loève decomposed Gaussian processes

A promising approach for scalable Gaussian processes (GPs) is the Karhunen-Loève (KL) decomposition, in which the GP kernel is represented by a set of basis functions which are the eigenfunctions of the kernel operator. Such decomposed kernels have the potential to be very fast, and do not depend on the selection of a reduced set of inducing points. However KL decompositions lead to high dimensionality, and variable selection thus becomes paramount. This paper reports a new method of forward variable selection, enabled by the ordered nature of the basis functions in the KL expansion of the Bayesian Smoothing Spline ANOVA kernel (BSS-ANOVA), coupled with fast Gibbs sampling in a fully Bayesian approach. It quickly and effectively limits the number of terms, yielding a method with competitive accuracies, training and inference times for tabular datasets of low feature set dimensionality. Theoretical computational complexities are O ( N P 2 ) in training and O ( P ) per point in inference, where N is the number of instances and P the number of expansion terms. The inference speed and accuracy makes the method especially useful for dynamic systems identification, by modeling the dynamics in the tangent space as a static problem, then integrating the learned dynamics using a high-order scheme. The methods are demonstrated on two dynamic datasets: a ‘Susceptible, Infected, Recovered’ (SIR) toy problem, along with the experimental ‘Cascaded Tanks’ benchmark dataset. Comparisons on the static prediction of time derivatives are made with a random forest (RF), a residual neural network (ResNet), and the Orthogonal Additive Kernel (OAK) inducing points scalable GP, while for the timeseries prediction comparisons are made with LSTM and GRU recurrent neural networks (RNNs) along with the SINDy package.

Hayes, Kyle↗

Formulation of full state feedback for infinite order structural systems

The estimation of exact displacement and displacement rate feedback kernels from finite dimensional control solutions based on finite element structural models is discussed. These kernels are then transformed to equivalent curvature and curvature rate feedback kernels. These curvature kernels are augmented with single point displacement and rotation feedback to account for rigid body motions. A growing class of sensors known as area-averaging sensors is used to measure the curvature and curvature rate state functions. The output of area-averaging sensors equals the convolution of all structural curvature states with the spatial sensitivity function of the sensors. Transforming the discrete feedback gains into continuous feedback kernels and employing area-averaging sensors make it possible to implement full state feedback for infinite order structural systems.

Miller, David W.↗

Design and Analysis of Architectures for Structural Health Monitoring Systems

During the two-year project period, we have worked on several aspects of Health Usage and Monitoring Systems for structural health monitoring. In particular, we have made contributions in the following areas. 1. Reference HUMS architecture: We developed a high-level architecture for health monitoring and usage systems (HUMS). The proposed reference architecture is shown. It is compatible with the Generic Open Architecture (GOA) proposed as a standard for avionics systems. 2. HUMS kernel: One of the critical layers of HUMS reference architecture is the HUMS kernel. We developed a detailed design of a kernel to implement the high level architecture.3. Prototype implementation of HUMS kernel: We have implemented a preliminary version of the HUMS kernel on a Unix platform.We have implemented both a centralized system version and a distributed version. 4. SCRAMNet and HUMS: SCRAMNet (Shared Common Random Access Memory Network) is a system that is found to be suitable to implement HUMS. For this reason, we have conducted a simulation study to determine its stability in handling the input data rates in HUMS. 5. Architectural specification.

Mukkamala, Ravi↗

3DRT-MPASS

Data from all current JPL missions are stored in files called SPICE kernels. At present, animators who want to use data from these kernels have to either read through the kernels looking for the desired data, or write programs themselves to retrieve information about all the needed objects for their animations. In this project, methods of automating the process of importing the data from the SPICE kernels were researched. In particular, tools were developed for creating basic scenes in Maya, a 3D computer graphics software package, from SPICE kernels.

Lickly, Ben↗

An Approach to Retrieve BRDF from Satellite and Airborne Measurements of Surface-Reflected Radiance Based on Decoupling of Atmospheric Radiative Transfer and Surface Reflection

Bi-directional Reflection Distribution Function (BRDF) defines anisotropy of the surface reflection. It is required to specify the boundary condition for radiative transfer (RT) modeling. Measurements of reflected radiance by satellite- and air-borne sensors provide information about anisotropy of surface reflection. Atmospheric correction needs to be performed to derive BRDF from the reflected radiance. Common approach for BRDF retrievals consists of the use of kernel-based BRDF and RT modeling that needs to be done anew at every step of the iterative process. The kernels’ weights are obtained by minimization of the difference between measured and modeled radiance. This study develops a new method of retrieving kernel-based BRDF that requires RT calculations to be done only once. The method employs the exact analytical expression of radiance at any atmospheric level through the solutions of two auxiliary atmosphere-only RT problems and the surface-reflected radiance at the surface level. The latter is related to BRDF and solutions of the auxiliary RT problems by a Fredholm integral equation of the second kind. The approach requires to perform RT calculations one time before the iterations. It can use observations taken at different atmospheric conditions assuming that surface conditions remain unchanged during the time span of observations. The algorithm accurately catches zero weights of the kernels that may be a concern if the number of kernels is greater than 3 in current mainstream approaches. The study presents numerical tests of the BRDF retrieval algorithm for various surface and atmospheric conditions.

Radkevich, Alexander↗

ExtremeMETA: High-speed Lightweight Image Segmentation Model by Remodeling Multi-channel Metamaterial Imagers

Deep neural networks (DNNs) have heavily relied on traditional computational units, such as CPUs and GPUs. However, this conventional approach brings significant computational burden, latency issues, and high power consumption, limiting their effectiveness. This has sparked the need for lightweight networks such as ExtremeC3Net. Meanwhile, there have been notable advancements in optical computational units, particularly with metamaterials, offering the exciting prospect of energy-efficient neural networks operating at the speed of light. Yet, the digital design of metamaterial neural networks (MNNs) faces precision, noise, and bandwidth challenges, limiting their application to intuitive tasks and low-resolution images. In this study, we proposed a large kernel lightweight segmentation model, ExtremeMETA. Based on ExtremeC3Net, our proposed model, ExtremeMETA maximized the ability of the first convolution layer by exploring a larger convolution kernel and multiple processing paths. With the large kernel convolution model, we extended the optic neural network application boundary to the segmentation task. To further lighten the computation burden of the digital processing part, a set of model compression methods was applied to improve model efficiency in the inference stage. The experimental results on three publicly available datasets demonstrated that the optimized efficient design improved segmentation performance from 92.45 to 95.97 on mIoU while reducing computational FLOPs from 461.07 MMacs to 166.03 MMacs. The large kernel lightweight model ExtremeMETA showcased the hybrid design’s ability on complex tasks.

large convolution kernel↗

Neutron and x-ray computed tomography of a natural uranium tristructural isotropic (TRISO) fuel compact

A natural uranium-based, unirradiated tristructural isotropic (TRISO) fuel compact was nondestructively imaged using both X-ray (XCT) and neutron computed tomography (nCT). While XCT of compacts can provide information on fuel kernels, imaging artifacts preclude examination of the graphite matrix. In this work, nCT was used for the first time on a TRISO compact to examine the graphite matrix. A crack was clearly resolved within the graphite matrix, proving that nCT is a viable tool for nondestructive volumetric examination of the matrix material in TRISO fuel compacts. The XCT and nCT data were then fused together to create a more comprehensive dataset containing both matrix and fuel kernels.

36 - MATERIALS SCIENCE↗

Post-irradiation examination of AGR-3/4 TRISO fuel compacts using three-dimensional X-ray computed tomography

The AGR-3/4 irradiation tests combined the third and fourth planned irradiation experiments in the US Department of Energy’s Advanced Gas Reactor (AGR) testing campaign of tri-structural isotropic (TRISO) fuel compacts. In this article, we present post-irradiation examination (PIE) using X-ray computed tomography (XCT) of two unirradiated and two irradiated compacts from the AGR-3/4 irradiation tests. The irradiated compacts studied (compact 7–1 and compact 12–4) represent the upper and lower limit of burnup within the AGR-3/4 irradiation experiment. This article presents a detailed quantitative analysis on the post-irradiation structure of TRISO fuel compacts. Various quantitative parameters including shape, size, and packing of kernels, and their spatial distribution, were utilized to gain insights into the structural changes caused by irradiation. The equivalent diameter and sphericity were found to increase and decrease, respectively, in irradiated compact 7–1 due to its higher burnup. Nearest neighbor distance between fuel kernels decreased after irradiation, suggesting irradiation-induced shrinkage of graphitic matrix. Furthermore, each compact in AGR-3/4 irradiation tests contained 20 designed-to-fail (DTF) fuel particles that were meant to act as a source of fission product release to the experiment test train. Furthermore, in the present work, all DTF fuel particles in the four compacts studied were identified, and it was found that they exhibited larger kernel swelling in compact 12–4 and smaller kernel swelling in compact 7–1, compared to the driver particles.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Unsteady aerodynamic loads on pitching aerofoils represented by Gaussian body force distributions

The actuator line model (ALM) is an approach commonly used to represent lifting and dragging devices like wings and blades in large-eddy simulations (LES). The crux of the ALM is the projection of the actuator point forces onto the LES grid by means of a Gaussian regularisation kernel. The minimum width of the kernel is constrained by the grid size; however, for most practical applications like LES of wind turbines, this value is an order of magnitude larger than the optimal value that maximises accuracy. This discrepancy motivated the development of corrections for the actuator line, which, however, neglect the effect of unsteady spanwise shed vorticity. In this work we develop a model for the impact of spanwise shed vorticity on the unsteady loading of an aerofoil modelled as a Gaussian body force distribution, where the model is applicable within the regime of unsteady attached flow. The model solution is derived both in the time and frequency domain and features an explicit dependence on the Gaussian kernel width. We verify the model with ALM-LES for both pitch steps and periodic pitching. The model solution is compared with Theodorsen theory and validated with both computational fluid dynamics using body fitted grids and experiment. It is concluded that the optimal kernel width for unsteady aerodynamics is approximately 40 % of the chord. The ALM is able to predict the magnitude of the unsteady loading up to a reduced frequency of 𝑘 ≈ 0.2.

17 WIND ENERGY↗

ICED: An Integrated CGRA Framework Enabling DFVS-Aware Acceleration

oarse-grained reconfigurable arrays (CGRAs) are a promising solution to enable energy-efficient acceleration of applications from different domains. By leveraging reconfiguration at the functional level, they can adapt to significantly different computational patterns. Existing CGRA mapping approaches extract instruction-level parallelism, exploit loop-pipelining opportunities, guarantee the data dependency, and target high throughput of a given loop. However, the recurrence data-dependency in the DFG and the mismatch between required and available computing/communication resources complicate the mapping, and might lead to significant unbalances in the utilization of the CGRA's tiles. This results in wasted power for tiles with low utilization. Applying dynamic voltage and frequency scaling (DVFS) can potentially solve this challenge and improve energy efficiency by adjusting voltage and frequency of different tiles independently. CGRAs have also been successful in accelerating data-dependent streaming applications. However, in these applications, the execution time of each kernel in the pipeline might dynamically vary depending on the characteristics of the input. This also leads to under-utilization of resources for the dynamically changing kernels that do not limit the application throughput. DVFS can also improve energy efficiency for these applications by dynamically changing the voltage and frequency levels of tiles that host non performance-constraining kernels. This paper proposes ICEDTEA -- an integrated DVFS-aware framework to map applications on CGRAs that support power islands. ICEDTEA proposes a CGRA architecture supporting DVFS islands at varying granularity (from a single tile to a group of tiles) and the related DVFS-aware compilation and mapping toolchain. ICEDTEA is the first work that introduces DVFS support for spatio-temporal CGRAs at power-island levels. The experimental evaluation shows that ICEDTEA improves average utilization by 2.3$\times$ and energy-efficiency by 1.32$\times$ over a conventional CGRA. With streaming applications, ICEDTEA improves energy efficiency by 1.12$\times$ over a state-of-the-art CGRA that introduces partial dynamic reconfiguration to adapt to variations in kernels' throughput.

Tan, Cheng↗

A Performance-Portable MultiGPU Implementation of 3D Euler Equations using ProtoX and IRIS

Computational scientists often face challenges when developing and optimizing code for high-performance computing (HPC), especially when trying to leverage GPUs. Given the heterogeneity of the nodes that comprise many modern HPC facilities, considerable demand exists for performance portable solutions for the core computational kernels used in many scientific computing libraries. In this work, we demonstrate a fourth-order finite volume method–based implementation of the Euler equations, which are an integral part of computational fluid dynamics. Our performance-portable multiGPU implementation for Euler equations uses ProtoX to generate kernels and IRIS for portability. ProtoX is a domain-specific language that uses a structured-grid partial differential equation library called Proto as its front end and the SPIRAL code generation system as its back end to generate optimized kernels for different architectures. Optimized kernels generated by ProtoX are orchestrated through the IRIS intelligent runtime system to provide portability. Two levels of optimizations within the IRIS runtime— directed acyclic graph fusion and task fusion—are explored to efficiently utilize computing resources in a multiGPU environment. Performance improvement through these optimizations is showcased by comparing the base ProtoX-IRIS implementation on AMD GPUs (Frontier node) and on NVIDIA GPUs (NVIDIA DGX-1).

Mankad, Het↗

JACC: Leveraging HPC Meta-Programming and Performance Portability with the Just-in-Time and LLVM-based Julia Language

We present JACC (Julia for Accelerators), the first high-level, and performance-portable model for the just-in-time and LLVM-based Julia language. JACC provides a unified and lightweight front end across different back ends available in Julia, enabling the same Julia code to run efficiently on many HPC CPU and GPU targets. We evaluated the performance of JACC for common HPC kernels as well as for the most computationally demanding kernels used in applications, HPCCG, a supercomputing benchmark test for sparse domains, and HARVEY, a blood flow simulator to assist in the diagnosis and treatment of patients suffering from vascular diseases. We carried out the performance analysis on the most advanced US DOE supercomputers: Aurora, Frontier, and Perlmutter. Overall, we show that JACC has a negligible overhead versus vendor-specific solutions, reporting GPU speedups with no extra cost to programmability.

Valero-Lara, Pedro↗