Search NASA⌕ Search

SEARCH · Search NASA

Results for “Computer implementation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 397 records · Page 22

Coulomb Interaction-Driven Entanglement of Electrons on Helium

The generation and evolution of entanglement in many-body systems is an active area of research that spans multiple fields, from quantum information science to the simulation of quantum many-body systems encountered in condensed matter, subatomic physics, and quantum chemistry. Motivated by recent experiments exploring quantum information processing systems with electrons trapped above the surface of cryogenic noble gas substrates, we theoretically investigate the generation of entanglement between two electrons via their unscreened Coulomb interaction. The model system consists of two electrons confined in separate electrostatic traps that establish microwave-frequency quantized states of their motion. We compute the motional energy spectra of the electrons, as well as their entanglement, by diagonalizing the model Hamiltonian with respect to a single-particle Hartree product basis. We also compare our results with the predictions of an effective Hamiltonian. The computational procedure outlined here can be employed for device design and guidance of experimental implementations. In particular, the theoretical tools developed here can be used for fine-tuning and optimization of control parameters in future experiments with electrons trapped above the surface of superfluid helium or solid neon. Published by the American Physical Society 2024

Physics↗

Tardigrade-examples V0.1.0

Tardigrade-examples is a repository of computational workflows that exercise the Tardigrade software package. The Tardigrade software package is an implementation of Eringen’s micromorphic continuum theory with capabilities to support multiscale material modeling. These capabilities include homogenization through the Micromorphic Filter, calibration of micromorphic material models, and macroscale simulation in Tardigrade-MOOSE. This repository investigates continuum upscaling of various direct numerical simulations (DNS) conducted in Abaqus finite element (FE), Ratel FE, and GEOS material point method (MPM) software. Verification of the upscaling workflow is first investigated by considering DNS of trivial stress states for homogeneous materials, results of which indicate that classical continuum behavior is recovered as expected. DNS of heterogeneous materials are then considered.

Allard, Thomas↗

Genesis: A Compiler Framework for Hamiltonian Simulation on Hybrid CV-DV Quantum Computers

We introduce Genesis, the first compiler designed to support Hamiltonian Simulation on hybrid continuous-variable (CV) and discrete-variable (DV) quantum computing systems. Genesis is a two-level compilation system. At the first level, it decomposes an input Hamiltonian into basis gates using the native instruction set of the target hybrid CV-DV quantum computer. At the second level, it tackles the mapping and routing of qumodes/qubits to implement long-range interactions for the gates decomposed from the first level. Rather than a typical implementation that relies on SWAP primitives similar to qubit-based (or DV-only) systems, we propose an integrated design of connectivity-aware gate synthesis and beamsplitter SWAP insertion tailored for hybrid CV-DV systems. We also introduce an OpenQASM-like domain-specific language (DSL) named CVDV-QASM to represent Hamiltonian in terms of Pauli-exponentials and basic gate sequences from the hybrid CVDV gate set. Genesis has successfully compiled several important Hamiltonians, including the Bose-Hubbard model, Z2−Higgs model, Hubbard-Holstein model, Heisenberg model and Electron-vibration coupling Hamiltonians, which are critical in domains like quantum field theory, condensed matter physics, and quantum chemistry. Our implementation is available at Genesis-CVDV-Compiler https://github.com/ruadapt/Genesis-CVDV-Compiler

Chen, Henry↗

Genesis Data Card Schema, Template and Supporting Tools

Genesis Data Cards provide a standardized template and schema for documenting scientific datasets in support of discovery, access, interoperability, reusability, governed use, and AI usability. This release of the Genesis Data Card repository includes a versioned Markdown template, a LinkML schema with generated Pydantic and JSON artifacts, schema documentation, and example completed data cards. Validation tooling is provided to ensure that completed data cards conform to the schema prior to submission. Accompanying documentation for the structured metadata is provided as a Field Reference Guide. The schema and accompanying template provided in this repository address the call for actionable context that enables humans and AI systems to find, access, interpret, cite, and reuse data, and, when appropriate, integrate it into AI and machine learning workflows. The data card is intended to serve as a common metadata artifact intended to support standardized, cross-program dataset documentation across Department of Energy (DOE)-aligned efforts, including but not limited to Genesis Mission-related implementations, the Office of Science, National Nuclear Security Administration (NNSA), and Advanced Simulation and Computing (ASC) data governance and stewardship initiatives.

data card↗

Improving the Parameterization of Cloud and Rain Microphysics in E3SM using Novel Observationally-Constrained Bayesian Approach (Final Technical Report)

In this project, we sought to develop new cloud and rain microphysics frameworks within the Energy Exascale Earth System Model (E3SM). This work encompassed two primary avenues of research: 1) Further development of a Bayesian-based scheme called BOSS (Bayesian Observationally-constrained Statistical-physical Scheme) to represent cloud and rain microphysics, testing it in realistic high-resolution cloud models, and implementing it in E3SM; 2) Development of a methodology utilizing machine learning to enable computationally tractable use of tractable use of Markov chain Monte Carlo sampling for Bayesian parameter estimation in Earth system and cloud models. In this project, we adapted the BOSS microphysics scheme, originally formulated for rain-only, to include all liquid-phase microphysical processes for cloud and rain, in particular the processes that mediate between these two categories, for example the conversion from cloud to rain through collision and coalescence of drops. We constrained the scheme via comparison and testing against a detailed model that explicitly represents the evolution of cloud and rain particles, called a bin microphysics scheme.

54 ENVIRONMENTAL SCIENCES↗

Benchmarking of three DWM-based wake models at below-rated wind speeds

Wind turbine wake models are essential tools for predicting power losses and structural loads in wind farms. Among these, the dynamic wake meandering (DWM) model, included as a recommended approach in the International Electrotechnical Commission design standard, is a widely used engineering-fidelity method that balances accuracy and computational cost. This study compares the performance of three DWM-based wake model implementations (from the Technical University of Denmark, the National Renewable Energy Laboratory, and the Institute for Energy Technology) under below-rated wind speed conditions. Model predictions of wake flow, power output, and structural loads for a four-turbine row are evaluated across different ambient turbulence levels and wind-direction misalignments and compared against high-fidelity large-eddy simulation results. All three models captured the overall wake evolution and mean turbine performance with reasonable accuracy; their predicted time-averaged thrust and power were typically within 5 %–10 % of the large-eddy simulation benchmark. However, notable differences emerged in wake structure and unsteady load predictions, with discrepancies increasing for turbines further downstream. These differences highlight the importance of modelling choices such as wake summation and turbulence treatment, which strongly influence power-deficit and fatigue-load predictions. Comparison with large-eddy simulations reveals each approach's strengths and weaknesses, indicating where improvements are needed. Overall, the findings point to specific refinements for DWM models to improve their fidelity, ultimately enabling more robust wake predictions for wind farm design and operation.

17 WIND ENERGY↗

Computer program product for classifying materials

Systems and methods for classifying materials utilizing one or more sensor systems, which may implement a machine learning system in order to identify or classify each of the materials, which may then be sorted into separate groups based on such an identification or classification. The machine learning system may utilize a neural network, and be previously trained to recognize and classify certain types of materials.

Kumar, Nalin↗

Geometry-aware training of factorized layers in tensor Tucker format

Reducing parameter redundancies in neural network architectures is crucial for achieving feasible computational and memory requirements during train and inference of large networks. Given its easy implementation and flexibility, one promising approach is layer factorization, which reshapes weight tensors into a matrix format and parameterizes it as the product of two rank-r matrices. However, this family of approaches often requires an initial full-model warm-up phase, prior knowledge of a feasible rank, and it is sensitive to parameter initialization.In this work, we introduce a novel approach to train the factors of a Tucker decomposition of the weight tensors. Our training proposal proves to be optimal in locally approximating the original unfactorized dynamics and stable for the initialization. Furthermore, the rank of each mode is dynamically updated during training.We provide a theoretical analysis of the algorithm, showing convergence, approximation and local descent guarantees. The method's performance is further illustrated through a variety of experiments, showing remarkable training compression rates and comparable or even better performance than the full baseline and alternative layer factorization strategies.

Zangrando, Emanuele [Gran Sasso Science Institute ↗

QRCODE: Massively parallelized real-time time-dependent density functional theory for periodic systems

We present a new software module, QRCODE (Quantum Research for Calculating Optically Driven Excitations), for massively parallelized real-time time-dependent density functional theory (RT-TDDFT) calculations of periodic systems in the open-source Qbox software package. Our approach utilizes a custom implementation of a fast Fourier transformation scheme that significantly reduces inter-node message passing interface (MPI) communication of the major computational kernel and shows impressive scaling up to 16,344 CPU cores. In addition to improving computational performance, QRCODE contains a suite of various time propagators for accurate RT-TDDFT calculations. As benchmark applications of QRCODE, we calculate the current density and optical absorption spectra of hexagonal boron nitride (h-BN) and photo-driven reaction dynamics of the ozone-oxygen reaction. We also calculate the second and higher harmonic generation of monolayer and multi-layer boron nitride structures as examples of large material systems. Our optimized implementation of RT-TDDFT in QRCODE enables large-scale calculations of real-time electron dynamics of chemical and material systems with enhanced computational performance and impressive scaling across several thousand CPU cores.

97 MATHEMATICS AND COMPUTING↗

Cyber-Informed Engineering for Strategic Planning

The CIE Guide for Engaging Organizational Leadership: Board and C-Suite describes how to apply CIE at the senior-level of organizational management. This guidance integrates CIE concepts into theory and practice of guiding coalition leadership to improve permeation of CIE concepts throughout organizational culture. This guide incorporates feedback from CIE community of practice volunteers in a study of how business administrative and strategic planning and management guidance and course materials can be applied to CIE implementation throughout an organization's principles and processes. This guide defines the interests guiding senior leadership, managers and supervisors, and technicians and workers involved in creating and operating and maintaining systems, and cultural/ historical/ political assumptions or influences shaped the landscape that could facilitate or impede development of CIE. The included guidance for organizational strategic management and leadership can be imbedded into contract guidance based on the CIE implementation guide and lessons learned from industry application.

97 MATHEMATICS AND COMPUTING↗

Quantum Annealing for Real-World Machine Learning Applications

Optimizing the training of a machine learning pipeline is important for reducing training costs and improving model performance. One such optimizing strategy is quantum annealing, which is an emerging computing paradigm that has shown potential in optimizing the training of a machine learning model. The implementation of a physical quantum annealer has been realized by D-Wave systems and is available to the research community for experiments. Recent experimental results on a variety of machine learning applications have shown interesting results especially under the conditions where the performance of classical machine learning techniques are limited such as limited training data and high dimensional features. This chapter explores the application of D-Wave’s quantum annealer for optimizing machine learning pipelines for real-world classification problems. We review the application domains on which a physical quantum annealer has been used to train machine learning classifiers. We discuss and analyze the experiments performed on the D-Wave quantum annealer for applications such as image recognition, remote sensing imagery, security, computational biology, biomedical sciences, and physics. We discuss the possible advantages and the problems for which quantum annealing is likely to be advantageous over classical computation.

Kumar nath, Rajdeep↗

Performance Improvements of the Griffin Solvers in FY24

The Griffin code is a MOOSE-based reactor physics application jointly developed by Idaho National Laboratory and Argonne National Laboratory under the Department of Energy Office of Nuclear Energy Nuclear Energy Advanced Modeling and Simulation Program. This fiscal year, we have made significant efforts to improve the performance of transport solver options and cross-section generation for the efficient use of Griffin in advanced reactor applications. For the HFEM-PN solver, the residual evaluations of HFEM kernels were optimized by utilizing the pre- computed averaged cross sections for individual elements. Numerical integration involving the evaluation of basis functions at quadrature points was bypassed by facilitating precomputed element mass matrices for response matrices. Red-black iterations were improved by introducing a new generalized minimum residual based solver. The memory usage of response matrix storage was significantly reduced by applying basis function rotations on interfaces and calculating volumetric odd-parity moments on the fly. Additionally, the adjoint flux and transient calculation capabilities of the HFEM-PN solver were successfully implemented and verified using the TWIGL benchmark problem. For the DFEM-SN solver, memory footprint and computation time were significantly reduced by not treating angular flux vectors as the MOOSE nonlinear system vectors. Specifically for IQS, scalar adjoint weighting was introduced to further eliminate angular adjoint flux storage in the MOOSE auxiliary system. It was demonstrated through the three-dimensional Advanced Burner Test Reactor core problem that the memory usage for transient calculations with the IQS method was reduced by over 7.5× compared to before the optimizations. For the self-shielding application programming interface, a new double-heterogeneity treatment method, named the Bell Function-Based Analytic Two-Region Slowing Down Method, was developed to efficiently flux-volume homogenize TRISO particles with the matrix. Additionally, optimizations were made to hyper- fine group (HFG) slowing down calculations by pretabulating collision probability coefficients and grouping isotopes, significantly reducing the computational time for calculating scattering sources per HFG. Lastly, the pin power reconstruction module was extended to account for temporal behavior in a microreactor analysis problem, specifically for a control drum transient. Verification tests for each of these improvements demonstrated significant performance enhancements and memory reduction.

22 - GENERAL STUDIES OF NUCLEAR REACTORS↗

CI/CD Efforts for Validation, Verification and Benchmarking OpenMP Implementations

Software developers must adapt to keep up with the changing capabilities of platforms so that they can utilize the power of High-Performance Computers (HPC), including exascale systems. OpenMP, a directive-based parallel programming model, allows developers to include directives to existing C, C++, or Fortran code to allow node level parallelism without compromising performance. This paper describes our CI/CD efforts to provide easy evaluation of the support of OpenMP across different compilers using existing testsuites and benchmark suites on HPC platforms. Our main contributions include (1) the set of a Continuous Integration (CI) and Continuous Development (CD) workflow that captures bugs and provides faster feedback to compiler developers, (2) an evaluation of OpenMP (offloading) implementations supported by AMD, HPE, GNU, LLVM, and Intel, and (3) evaluation of the quality of compilers across different heterogeneous HPC platforms. With the comprehensive testing through the CI/CD workflow, we aim to provide a comprehensive understanding of the current state of OpenMP (offloading) support in different compilers and heterogeneous platforms consisting of CPUs and GPUs from NVIDIA, AMD, and Intel.

Jarmusch, Aaron↗

IRI Technology Landscape – A survey of re-usable components and methodologies

This document describes technical implementation details on network access schemes connecting API-driven workflows to supercomputer centers. API-driven workflows are a central theme in connected computing, since they bring the terminal-mainframe' access pattern present since the 1970s up to the task of interfacing with modern web browser technologies. Both security (HTTPS/TLS/IPSec/VPNs/public key cryptography/digital signatures) and network protocol stacks (HTTP-REST APIs, tokens, gRPC, SRTP) have evolved to the point where implementing API-driven workflows is possible using stable, secure off-the-shelf software.

97 MATHEMATICS AND COMPUTING↗

Ensemble Kalman filter for data assimilation coupled with low-resolution computations techniques applied in fluid dynamics

This paper presents an innovative Reduced-order model (ROM) for merging experimental and simulation data using data assimilation (DA) to estimate the "True" state of a fluid dynamics system, leading to more accurate predictions. Our methodology introduces a novel approach by implementing the ensemble Kalman filter (EnKF) within a reduced-dimensional framework, grounded in a robust theoretical foundation and applied to fluid dynamics. To address the substantial computational demands of DA, the proposed ROM employs low-resolution (LR) techniques to drastically reduce computational costs. This innovative approach involves downsampling datasets for DA computations, followed by an advanced reconstruction technique based on low-cost singular value decomposition (lcSVD). The lcSVD method, a key innovation in this paper, has never been applied to DA before and offers a highly efficient way to enhance resolution with minimal computational resources. Our results demonstrate significant reductions in both computation time and RAM usage through these LR techniques without compromising the accuracy of the estimations. For instance, in a turbulent test case, for a data compression rate of 15.9, the LR approach can achieve a speed-up of 13.7 and a RAM compression of 90.9% while maintaining a low relative root mean square error (RRMSE) of 2.6%, compared to 0.8% in the high-resolution (HR) reference. Furthermore, we highlight the effectiveness of the EnKF in estimating and predicting the state of fluid flow systems based on limited observations and given low-fidelity numerical data. This paper highlights the potential of the proposed DA method in fluid dynamics applications, particularly for improving computational efficiency in CFD and related fields. Its ability to balance accuracy with low computational and memory costs makes it especially suitable for large-scale and real-time applications, such as environmental monitoring or engineering design. This method will be incorporated into ModelFLOWs-app.

Data Assimilation↗

A cohesive zone treatment for the material point method involving problems of large deformation and damage

A new algorithm is described that permits the use of cohesive zones in the material point method for problems involving large deformation and fracture. In contrast to previous cohesive zone implementations, this method does not utilize massless surface-element particles. Instead, cohesive tractions are computed using the shape function mappings from a reference grid configuration in combination with explicitly defined particle surface normals and surface positions. These normals and relative surface positions are updated each time step according to particle deformation. The tractions are converted to cohesive forces using the nodal areas and mapped back to particles using the same reference shape function mappings. These forces are then remapped by conventional particle-to-grid interpolation as external forces using the current-configuration shape-function mappings. This allows highly compliant cohesive zones to function over jump displacements larger than a grid cell. Upon damage, these interfaces can revert to conventional multi-field contact surfaces. This approach is general and readily applies to two and three dimensions as well as being compatible with damage-field gradient partitioning offering exceptional computational flexibility. The framework for this method enables other capabilities, such as improved contact precision using explicitly defined surface normals and positions, and a method to mitigate spurious material damage at weak discontinuities between stiff brittle materials and soft or compliant materials.

Cohesive zone↗

Minimization of Cathode|Solid-Electrolyte Interfacial Delamination through the Application of Interphase Layers

Next-generation lithium-ion batteries are expected to use solid electrolytes (SEs) to enable higher energy density and extreme fast-charge capabilities. One major mode of degradation at the cathode|SE interface is delamination between the cathode active materials and SEs, which leads to performance decay. Experimental observations indicate that implementation of interphase layers can minimize the cathode|SE delamination induced capacity fade. A multiscale computational methodology is developed here to investigate the applicability of boron substituted lithium carbonate (Li 2+x B x C 1–x O 3 , x = 0.5, or LBCO) to minimize the delamination at the cathode|SE interface. Atomistic simulations indicate that the fracture energies at both the cathode|LBCO and LBCO|SE interfaces are higher than those at the cathode|SE interface, which reduces the extent of delamination. Mesoscale simulations indicate that, apart from increasing the fracture energy, decreasing the evolution of strain energy by lowering the elastic modulus of the interphase layer can also minimize the extent of delamination at the cathode|SE interface. However, the adoption of an interphase layer with high ionic conductivity is necessary to minimize the ohmic losses during operation at higher current densities. This study provides guidance on selecting interphase layers with specific properties and thicknesses to minimize both interfacial delamination and impedance growth.

LBCO↗

Efficient multimode Wigner tomography

Abstract Advancements in quantum system lifetimes and control have enabled the creation of increasingly complex quantum states, such as those on multiple bosonic cavity modes. When characterizing these states, traditional tomography scales exponentially with the number of modes in both computational and experimental measurement requirement, which becomes prohibitive as the system size increases. Here, we implement a state reconstruction method whose sampling requirement instead scales polynomially with system size, and thus mode number, for states that can be represented within such a polynomial subspace. We demonstrate this improved scaling with Wigner tomography of multimode entangled W states of up to 4 modes on a 3D circuit quantum electrodynamics (cQED) system. This approach performs similarly in efficiency to existing matrix inversion methods for 2 modes, and demonstrates a noticeable improvement for 3 and 4 modes, with even greater theoretical gains at higher mode numbers.

Science & Technology - Other Topics↗