Search NASA⌕ Search

SEARCH · Search NASA

Results for “coded computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 271 records · Page 15

ORCHA: A performance portability system for extreme heterogeneity

Heterogeneity is the prevalent trend in the rapidly evolving high-performance computing (HPC) landscape in both hardware and application software. The diversity in hardware platforms, currently comprising various accelerators and a future possibility of specializable chiplets, poses a significant challenge for scientific software developers aiming to harness optimal performance across different computing platforms while maintaining the quality of solutions when their applications are simultaneously growing more complex. Code synthesis and code generation can provide mechanisms to mitigate this challenge. We have developed a divide and conquer approach where different aspects of performance are handled by different stand-alone tools that are interfaced with the application through generated code. This portability system, ORCHA, enables users to configure and orchestrate their computations among available resources on a platform by specifying a high-level recipe, thereby permitting a many-to-many paradigm where each recipe results in a different variant of the application. The core design goal is to let users decide the application’s hardware mapping and orchestration by editing only the high-level recipe—without modifying the maintained source code or binding the application to a particular runtime system. Tools in ORCHA distribution are: CG-Kit for translating the recipe into an execution graph; Milhoja to execute the graph by orchestrating data and task movement among hardware resources; and Macroprocessor that enables users to define their own code-shorthand for higher composability and easier management of code variants. Additionally, the design of ORCHA permits tools to work in a plug-and-play mode where the application can build and run without CG-Kit and Milhoja, and either tool can be swapped out for other tools with similar capabilities by modifying the code generation portion of ORCHA. In this paper, we describe the design of ORCHA and the role that code-generation plays in isolating applications from tools. We demonstrate the breadth of configurations ORCHA enables with a case study in which an application configuration is realized on three distinct hardware mappings—a GPU-centric, a CPU/GPU balanced, and a CPU/GPU concurrent layouts by using different recipes.

Lee, Youngjun↗

Binder-benchmarking

SAND2025-07593O Binder-benchmarking evaluates the speed and memory impacts of C++, Python, and Matlab code binders. As a repository, it provides a way to locally run computation-based and memory-based benchmark suites on pybind11 and nanobind-based code in a Docker image. The software runs simple-speed and memory benchmarks on primitive navigation and integration exemplar algorithms. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Walker II, Michael [Sandia National Lab. (SNL-CA),↗

Subject-specific modeling framework for particle deposition using computational fluid dynamics

Quantifying particle deposition and dose in the respiratory tract requires a physiologically realistic representation and reproducible computational workflows. However, existing modeling frameworks, such as the International Commission on Radiological Protection (ICRP) compartmental models and the Multiple Path Particle Dosimetry (MPPD) tool, lack detailed deposition profiles and subject-specific capabilities. The combination of advances in computer vision algorithms applied to the respiratory tract and Computational Fluid and Particle Dynamics (CFPD) allows high-fidelity simulations of particle behavior in anatomically accurate geometries derived from individual CT scans. The segmentation, preprocessing, and file preparation task for a CFPD simulation was often time-consuming, and no prior studies to-date have yet presented a fully automated framework. This work presents a fully automated workflow to obtain individualized particle deposition profiles in the human respiratory tract. The pipeline starts with segmenting upper and lower airway geometries using morphological and deep learning-based methods, generating three-dimensional (3D) models from CT imaging data. Next, a series of algorithms are presented to quality check and prepare the 3D geometry for a CFD or CFPD simulation. The preprocessing step includes correcting geometric artifacts, enforcing a physically consistent mesh, and automatically identifying and capping multiple outlets, which is required for CFD/CFPD simulations. These processed models are then input into open-source (OpenFOAM) or commercial (StarCCM+) CFD solvers, where flow and transient particle transport equations — including turbulence and particle–wall interactions are solved under realistic breathing conditions. Finally, the resulting particle deposition profiles can be integrated with Monte Carlo radiation transport codes and state-of-the-art computational phantoms to assess organ-specific absorbed doses in scenarios of radioactive aerosol inhalation. The presented work streamlines respiratory tract segmentation, preprocessing for CFD/CFPD simulations, and integration with dose assessment workflows, reducing manual intervention and improving access to high-fidelity, subject-specific modeling. The high precision in predicted particle deposition and dose distributions can improve personalized treatment strategies in respiratory medicine and refine dose estimates for radiation protection.

AI↗

Characterization and Optimization of the Fitting of Quantum Correlation Functions

This case study presents a characterization and optimization of an application code for extracting parton distribution functions from high energy electron-proton scattering data. Profiling this application code reveals that the phase-space density computation accounts for 93% of the overall execution time for a single iteration on a single core. When executing multiple iterations in parallel on a multicore system, the application spends 78% of its overall execution time idling due to load imbalance. We address these issues by first transforming the application code from Python to C++ and then tackling the application load imbalance via a hybrid scheduling strategy that combines dynamic and static scheduling. These techniques result in a 62% reduction in CPU idle time and a 2.46x speedup in overall execution time per node. In addition, the typically enabled power-management mechanisms in supercomputers (e.g., AMD Turbo Core, Intel Turbo Boost, and RAPL) can significantly impact intra-node scalability when more than 50% of the CPU cores are used. This finding underscores the importance of understanding system interactions with power management, as they can adversely impact application performance, and highlights the necessity of intra-node scaling tests to identify performance degradation that inter-node scaling tests might otherwise overlook.

Chuang, Pi-Yueh [Virginia Tech,Dept. of Computer S↗

Flag Gadgets Based on Classical Codes

Fault-tolerant syndrome extraction is a key ingredient in implementing fault-tolerant quantum computation. While conventional methods use a number of extra qubits that are linear in the weight of the syndrome, several improvements have been introduced using flag gadgets. In this work, we develop a framework to design flag gadgets using classical codes. Using this framework, we show how to perform fault-tolerant syndrome extraction for any stabilizer code with arbitrary distance using exponentially fewer qubits than conventional methods when qubit measurement and reset are relatively slow compared to a round of error correction. In particular, our method requires only ( 2 t + 1 ) t ⌈ log 2 ( w ) ⌉ flag qubits to fault-tolerantly measure a weight- w stabilizer. We further take advantage of the saving provided by our construction to fault-tolerantly measure multiple stabilizers using a single gadget and show that it maintains the same exponential advantage when it is used to fault-tolerantly extract the syndromes of quantum low-density parity-check codes. Using the developed framework, we perform computer-assisted search to find several small examples where our constructions reduce the number of qubits required. These small examples may be relevant to near-term experiments on small-scale quantum computers. Published by the American Physical Society 2024

Anker, Benjamin↗

Strong Scalability Analysis of the Albany Land Ice code on HPC Architectures

Scalability is a critical factor in High-Performance Computing (HPC), where optimizing resource usage has a direct impact on cost-effectiveness and time-efficiency. This report presents a strong scaling performance study of the Albany Land Ice (ALI) code across different HPC architectures, towards determining the best configuration to use when running large-scale simulation ensembles.

97 MATHEMATICS AND COMPUTING↗

Impact of anisotropy on TRISO fuel performance

Manufacturing of tristructural isotropic (TRISO) particles involves the deposition of pyrolytic carbon (PyC) and silicon carbide (SiC) layers using the fluidized bed chemical vapor deposition (CVD) process. The CVD process is known to generate polycrystalline layers with crystallographic textures, which imparts anisotropic thermophysical properties to the layers. Past studies have shown the risk for particle failure increases with an increase in anisotropy. The limit beyond which the anisotropy of PyC layers becomes unacceptable due to failure risk has been identified as a high-priority knowledge gap. This work presents a first systematic study on the effects of anisotropic thermal and mechanical properties on TRISO fuel performance. This computational study, performed using the fuel performance code BISON, investigates how the anisotropy in elasticity and thermal properties affect the stresses, temperature, and failure of a TRISO particle. The influence of other factors, such as operating temperature and particle geometry on the anisotropy effects, also has been analyzed. The studies utilize the recently published anisotropic elasticity and thermal behavior models for TRISO PyC and SiC layers implemented using tensors with full anisotropic capability. The spherical TRISO particles with anisotropic properties were found to have greater maximum tensile stress and significantly higher failure probability than the spherical particles with isotropic properties. In conclusion, the fuel performance predicted using these recently developed models was found to be comparable with the performance obtained using the historical models.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Shock tube simulations for the three-layer Richtmyer–Meshkov instability with single-mode and multimode perturbations

While the canonical two-component, single-mode Richtmyer–Meshkov instability (RMI) has been extensively studied, relatively less work has focused on the effects of an additional intermediate-density middle layer. This work investigates such three-material RMI configurations at two Atwood number scenarios using the ARES hydrodynamics code. After validation against previous experimental and computational studies, setups corresponding to recent three-layer shock tube experiments are simulated. Cases with both single-mode and multimode perturbations are studied to quantify mixing across the interface between the materials with highest and intermediate density. In particular, this work is able to comprehensibly examine differences between two- and three-dimensional setups for the single-mode and multimode problems. Observations from previous two-layer investigations still apply in the three-layer setup, but over the time horizons considered, there appears to be insufficient nonlinear mode coupling to create significant differences between two- and three-dimensional simulations following the first passage of a shock. Finally, additional reshock simulations have additional nonlinear growth that does result in expected differences between two- and three-dimensional cases in this three-layer setup, but significant differences do not manifest during the time horizon studied.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Calculating the space-charge-limited current density for nonplanar geometries by simulating the charge-free electric field

Calculating the space-charge-limited-current density (SCLCD) for a complicated diode geometry often requires computationally expensive particle-in-cell (PIC) codes. Here, this paper addresses this issue by using the charge-free electric field $E_0$ calculated using COMSOL Multiphysics to determine local and global SCLCD. The SCLCD obtained by using the surface average of $|E_0|^2$ on the cathode recovers theoretical results for one-dimensional (1D) planar, cylindrical, and tip-to-tip geometries in appropriate limits. We next compared tip-to-tip calculations with the SCLCD obtained using the PIC code Empire. The SCLCD calculated using COMSOL agreed well with Empire for flatter 1D tip-to-tip geometries and diverged with increasing sharpness. Physically, Empire predicts lower SCLCD than COMSOL because the electrons spread due to concentrated space-charge at the tip, whereas theory assumes that the electrons follow the charge-free electric field lines. We further assess the behavior of the SCLCD for tips protruding from the centers of flat, circular plates of various areas. Larger plate areas with constant tip size recover the 1D planar SCLCD globally and 1D tip-to-tip SCLCD locally, while reducing the difference between Empire and COMSOL calculations since larger plates capture more of the emitted electrons, reducing SCLCD suppression due to beam spreading. These results show that charge-free electric field simulations can be used to determine the SCLCD without needing to simulate particle dynamics in PIC.

Wright, Jack K. [Purdue Univ., West Lafayette, IN ↗

Linearised Fokker–Planck collision model for gyrokinetic simulations

We introduce a gyrokinetic, linearised Fokker–Planck collision model that satisfies conservation laws and is accurate at arbitrary collisionalities. The differential test-particle component of the operator is exact; the integral field-particle component is approximated using a spherical harmonic and a modified Laguerre polynomial expansion developed by Hirshman and Sigmar (1976 Phys. Fluids 19 1532). The numerical methods of the implementation in the δf-gyrokinetic code stella (Barnes et al 2019 J. Comput. Phys. 391 365–80) are discussed, and conservation properties of the operator are demonstrated. The collision model is then benchmarked against the collision model of the gyrokinetic solver GS2 in the limiting cases of a reduced test-particle collision operator and energy- and momentum-conserving operator. The accuracy of the full collision model is investigated by solving the parallel Spitzer-Härm problem for the transport coefficients. It is shown that retaining collisional energy flux and higher-order terms in the field-particle operator reduces errors in the transport coefficients from 10%–25% for a simple momentum- and energy-conserving model to under 1%.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

UOWDetection

The code provides machine learning, data collection, and computer science tools used to detect undocumented orphaned wells in United States from aerial imagery.

Kim, Anastasiia↗

msdlive-cli-distro

MSD-LIVE, the MultiSector Dynamics – Living, Intuitive, Value-adding, Environment, is a flexible and scalable data and code management system combined with a distributed computational platform that will enable MSD researchers to document and archive their data, run their models and analysis tools, and share their data, software, and multi-model workflows within a robust Community of Practice. MSD-LIVE will facilitate a new open, collaborative, resource-rich, technology-facilitated, community-driven way of doing MSD research.

Lansing, Carina↗

CFDverify

Estimating the discretization error of computational fluid dynamics (CFD) or other scientific codes as part of solution verification is often a non-trivial part of the analysis process. Methods can be complicated and may include assumptions/qualifications that need to be checked during analysis. CFD analysts, therefore, are likely to make errors when trying to conduct this necessary analysis on their own in not knowing about the best method for their problem, not correctly implementing a method, or in not have diagnostic tools to determine if the method was correctly applied.

Weinmeister, Justin [Oak Ridge National Laboratory↗

Photoneutron Production Using an Electron Linear Accelerator for Applications in Neutron Imaging

Photoneutron production is possible using an electron linear accelerator and a target capable of generating photonuclear reactions. A short pulse neutron source can be useful for neutron imaging dynamic experiments. This study is aimed at the feasibility of photoneutron production using a 20 MeV electron linear accelerator and tungsten and depleted uranium targets of various thicknesses. MCNP6 (Monte Carlo N-Particle) code will be used to develop a computational model to estimate total neutron yield, and this will be verified at the Idaho State University’s Accelerator Center. After verification of the neutron yield and energy spectra, an additional MCNP6 model will be developed to analyze the neutron imaging processes. This study will potentially prove it is possible to conduct multi-mode imaging experiments on the anticipated Scorpius electron linear accelerator at the Nevada National Security Sites.

43 PARTICLE ACCELERATORS↗

Reproducibility of fixed-node diffusion Monte Carlo across diverse community codes: The case of water–methane dimer

Fixed-node diffusion quantum Monte Carlo (FN-DMC) is a widely trusted many-body method for solving the Schrödinger equation, known for its reliable predictions of material and molecular properties. Furthermore, its excellent scalability with system complexity and near-perfect utilization of computational power make FN-DMC ideally positioned to leverage new advances in computing to address increasingly complex scientific problems. Even though the method is widely used as a computational gold standard, reproducibility across the numerous FN-DMC code implementations has yet to be demonstrated. This difficulty stems from the diverse array of DMC algorithms and trial wave functions, compounded by the method’s inherent stochastic nature. Here, this study represents a community-wide effort to assess the reproducibility of the method, affirming that yes, FN-DMC is reproducible (when handled with care). Using the water–methane dimer as the canonical test case, we compare results from eleven different FN-DMC codes and show that the approximations to treat the non-locality of pseudopotentials are the primary source of the discrepancies between them. In particular, we demonstrate that, for the same choice of determinantal component in the trial wave function, reliable and reproducible predictions can be achieved by employing the T-move, the determinant locality approximation, or the determinant T-move schemes, while the older locality approximation leads to considerable variability in results. These findings demonstrate that, with appropriate choices of algorithmic details, fixed-node DMC is reproducible across diverse community codes—highlighting the maturity and robustness of the method as a tool for open and reliable computational science.

Della Pia, Flaviano [Univ. of Cambridge (United Ki↗

SCALE HTR-PROTEUS Benchmark Model

This dataset contains input and result files of computational simulations of HTR-PROTEUS benchmark with the latest version of SCALE code system. The simulations cover criticality control rod worth calculations as well as sensitivity analysis and uncertainty quantification. Users wanting to reproduce results from this dataset are required to obtain a license to the SCALE code system for which details on the distribution can be found here: https://www.ornl.gov/scale/releases

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

QECC-Synth: A Layout Synthesizer for Quantum Error Correction Codes on Sparse Architectures

Quantum Error Correction (QEC) codes are essential for achieving fault-tolerant quantum computing (FTQC). However, their implementation faces significant challenges due to disparity between required dense qubit connectivity and sparse hardware architectures. Current approaches often either underutilize QEC circuit features or focus on manual designs tailored to specific codes and architectures, limiting their capability and generality. In response, we introduce QECC-Synth, an automated compiler for QEC code implementation that addresses these challenges. We leverage the ancilla bridge technique tailored to the requirements of QEC circuits and introduces a systematic classification of its design space flexibilities. We then formalize this problem using the MaxSAT framework to optimize these flexibilities. Evaluation shows that our method significantly outperforms existing methods while demonstrating broader applicability across diverse QEC codes and hardware architectures.

Yin, Keyi [University of California, San Diego]↗

An introduction to Spent Nuclear Fuel decay heat for Light Water Reactors: a review from the NEA WPNCS

This paper summarized the efforts performed to understand decay heat estimation from existing spent nuclear fuel (SNF), under the auspices of the Working Party on Nuclear Criticality Safety (WPNCS) of the OECD Nuclear Energy Agency. Needs for precise estimations are related to safety, cost, and optimization of SNF handling, storage, and repository. The physical origins of decay heat (a more correct denomination would be decay power) are then introduced, to identify its main contributors (fission products and actinides) and time-dependent evolution. Due to limited absolute prediction capabilities, experimental information is crucial; measurement facilities and methods are then presented, highlighting both their relevance and our need for maintaining the unique current full-scale facility and developing new ones. The third part of this report is dedicated to the computational aspect of the decay heat estimation: calculation methods, codes, and validation. Different approaches and implementations currently exist for these three aspects, directly impacting our capabilities to predict decay heat and to inform decision-makers. Finally, recommendations from the expert community are proposed, potentially guiding future experimental and computational developments. One of the most important outcomes of this work is the consensus among participants on the need to reduce biases and uncertainties for the estimated SNF decay heat. If it is agreed that uncertainties (being one standard deviation) are on average small (less than a few percent), they still substantially impact various applications when one needs to consider up to three standard deviations, thus covering more than 95% of cases. The second main finding is the need of new decay heat measurements and validation for cases corresponding to more modern fuel characteristics: higher initial enrichment, higher average burnup, as well as shorter and longer cooling time. Similar needs exist for fuel types without public experimental data, such as MOX, VVER, or CANDU fuels. A third outcome is related to SNF assemblies for which no direct validation can be performed, representing the vast majority of cases (due to the large number of SNF assemblies currently stored, or too short or too long cooling periods of interest). A few solutions are possible, depending on the application. For the final repository, systematic measurements of quantities related to decay heat can be performed, such as neutron or gamma emission. This would provide indications of the SNF decay heat at the time of encapsulation. For other applications (short- or long-term cooling), the community would benefit from applying consistent and accepted recommendations on calculation methods, for both decay heat and uncertainties. This would improve the understanding of the results and make comparisons easier.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗