Search NASA⌕ Search

SEARCH · Search NASA

Results for “Implementation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 523 records · Page 29

Parallel Runtime Interface for Fortran (PRIF): A Multi-Image Solution for LLVM Flang

Fortran compilers that provide support for Fortran’s native parallel features often do so with a runtime library that depends on details of both the compiler implementation and the communication library, while others provide limited or no support at all. This paper introduces a new generalized interface that is both compiler- and runtime-library-agnostic, providing flexibility while fully supporting all of Fortran’s parallel features. The Parallel Runtime Interface for Fortran (PRIF) was developed to be portable across shared- and distributed-memory systems, with varying operating systems, toolchains and architectures. It achieves this by defining a set of Fortran procedures corresponding to each of the parallel features defined in the Fortran standard that may be invoked by a Fortran compiler and implemented by a runtime library. PRIF aims to be used as the solution for LLVM Flang to provide parallel Fortran support. This paper also briefly describes our PRIF prototype implementation: Caffeine.

Bonachea, Dan↗

Exploring code portability solutions for HEP with a particle tracking test code

Traditionally, high energy physics (HEP) experiments have relied on x86 CPUs for the majority of their significant computing needs. As the field looks ahead to the next generation of experiments such as DUNE and the High-Luminosity LHC, the computing demands are expected to increase dramatically. To cope with this increase, it will be necessary to take advantage of all available computing resources, including GPUs from different vendors. A broad landscape of code portability tools—including compiler pragma-based approaches, abstraction libraries, and other tools—allow the same source code to run efficiently on multiple architectures. In this paper, we use a test code taken from a HEP tracking algorithm to compare the performance and experience of implementing different portability solutions. While in several cases portable implementations perform close to the reference code version, we find that the performance varies significantly depending on the details of the implementation. Achieving optimal performance is not easy, even for relatively simple applications such as the test codes considered in this work. Several factors can affect the performance, such as the choice of the memory layout, the memory pinning strategy, and the compiler used. The compilers and tools are being actively developed, so future developments may be critical for their deployment in HEP experiments.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Optimizing resource allocation in Miscanthus breeding via sparse testing designs for genomic prediction

Phenotyping high-biomass perennial crops is laborious and the rate of genetic gain in conventional perennial crop breeding programs is typically low. So, it is especially important to identify methods that produce efficiency gains in the breeding process. Miscanthus is a C4 perennial grass with favorable characteristics for producing biomass as a feedstock for biofuels and diverse bio-based products. Increasing biomass yield will increase profitability and environmental benefits, so it is a key target for Miscanthus breeding. In addition, the identification of well-adapted genotypes across a wide range of environmental conditions requires the establishment of multi-environment trials (METs). Sparse testing is a genomic prediction-based strategy that reduces the phenotyping costs in METs by selecting a subset of genotypes to evaluate in a subset of environments and then predicts the performance of the unobserved genotype-environment combinations. A Miscanthus sacchariflorus (MSA) population comprising 336 genotypes observed across three environments was analyzed implementing sparse testing designs. Three prediction models considering main effects (environments, genotypes, genomic) and interaction effects (genotype-by-environment; G×E interaction) were implemented for forecasting dry biomass yield (YDY), total culm (TCM), average internode length (AIL), and culm node number (CNN). Multiple calibration sets based on different compositions and sizes were considered to evaluate performance in terms of the predictive ability (PA) and the mean square error (MSE) for a fixed testing set size. The training set size ranged from 52 to 112 to predict a fixed set of 224 unobserved genotypes across all three environments. The results showed that the model accounting for G×E interaction consistently presented the highest PA and the lowest MSE: for CNN (PA: ~0.77, MSE: ~0.5) and YDY (PA: ~0.70, MSE: ~1.3) while for TCM and AIL these ranged from ~0.28 to 0.41 and ~1.3 to 4.3, respectively. Overall, varying training sets and allocation strategies did not affect PA and MSE, with 52 non-overlapping and 0 overlapping genotypes per environment as the optimal cost-effective allocation framework. This suggests that implementing sparse testing designs could significantly reduce phenotyping costs by fivefold, without compromising PA in breeding programs for perennial crops such as Miscanthus.

Miscanthus sacchariflorus (MSA)↗

SineKAN: Kolmogorov-Arnold Networks using sinusoidal activation functions

Recent work has established an alternative to traditional multi-layer perceptron neural networks in the form of Kolmogorov-Arnold Networks (KAN). The general KAN framework uses learnable activation functions on the edges of the computational graph followed by summation on nodes. The learnable edge activation functions in the original implementation are basis spline functions (B-Spline). Here, we present a model in which learnable grids of B-Spline activation functions are replaced by grids of re-weighted sine functions (SineKAN). We evaluate numerical performance of our model on a benchmark vision task. We show that our model can perform better than or comparable to B-Spline KAN models and an alternative KAN implementation based on periodic cosine and sine functions representing a Fourier Series. Further, we show that SineKAN has numerical accuracy that could scale comparably to dense neural networks (DNNs). Compared to the two baseline KAN models, SineKAN achieves a substantial speed increase at all hidden layer sizes, batch sizes, and depths. Current advantage of DNNs due to hardware and software optimizations are discussed along with theoretical scaling. Additionally, properties of SineKAN compared to other KAN implementations and current limitations are also discussed.

Reinhardt, Eric↗

Energy Leaders: The Catalyst for Strategic Energy Management

This study investigates the crucial role energy leaders play in driving strategic energy management (SEM) and accelerating cost savings within a manufacturing organization and consequently, the industrial sector. Whereas energy efficiency can be seen as an innovative business practice with irrefutable cost benefits, its effective implementation requires strategic leadership and a structured approach. This research analyzes data collected from 120 participants representing 71 companies attending the Energy Bootcamp events organized by the U.S. Department of Energy’s (DOE) Better Plants program. The collected data focused on the state of SEM implementation, the presence and responsibilities of energy leaders, and the formation and function of energy teams. The findings reveal a significant gap between the perceived importance of SEM and its actual adoption, highlighting the need for strong leadership to drive behavioral changes by championing energy efficiency initiatives. Results indicate that effective energy leaders possess a diverse skill set, including the ability to secure top management buy-in, foster a culture of energy consciousness, and collaborate across departments. This study emphasizes the importance of empowering energy leaders with clearly defined roles and responsibilities as well as the authority to build and lead cross-functional energy teams. Furthermore, integrating energy management into existing organizational structures and leveraging readily available resources are identified as key factors for successful implementation. This research underscores how dedicated leadership and effective SEM practices help achieve industrial energy efficiency goals, providing practical insights for organizations seeking to improve performance and contribute to a resilient future.

energy leader↗

Synthesis of DOTA-Based 43 Sc Radiopharmaceuticals Using Cyclotron-Produced 43 Sc as Exemplified by [ 43 Sc]Sc-PSMA-617 for PSMA PET Imaging

The implementation of theranostics in oncologic nuclear medicine has exhibited immense potential in improving patient outcomes in prostate cancer with the implementation of [ 68 Ga]Ga-PSMA-11 PET and [ 177 Lu]Lu-PSMA-617 into clinical practice. However, the correlation between radiopharmaceutical biodistributions seen with [ 68 Ga]Ga-PSMA-11 PET imaging and downstream [ 177 Lu]Lu-PSMA-617 therapy remains imperfect. This suggests that prostate cancer theranostics could potentially be further refined through the implementation of true theranostics, tandem pairs of diagnostic and therapeutic radiopharmaceuticals that utilize the same ligand and element, thus yielding identical pharmacokinetics. The radioscandiums are one such group of true theranostic radiopharmaceuticals. The radioscandiums consist of two β+ emitting scandium isotopes ( 43 Sc/ 44 Sc), as well as a β − emitting therapeutic isotope ( 47 Sc), which can all conjugate with PSMA-targeting PSMA-617. This potential has led to extensive investigations into the production of the radioscandiums as well as pre-clinical assessments with several ligands; however, there is a lack of literature extensively describing the complete synthesis of scandium radiopharmaceuticals. which therefore limits the accessibility of radioscandium research in theranostics. As such, this work aims to present an easily translatable protocol for the synthesis of [ 43 Sc]Sc-PSMA-617 from a [ 42 Ca]CaCO 3 starting material, including target formation, nuclear production via 42 Ca(d,n) 43 Sc reaction, chemical separation, radiolabeling, solvent reformulation, and target recycling.

62 RADIOLOGY AND NUCLEAR MEDICINE↗

AthenaK: A Performance-portable Version of the Athena++ Adaptive Mesh Refinement Framework

We describe AthenaK: a new implementation of the Athena++ block-based adaptive mesh refinement framework using the Kokkos programming model. Finite volume methods for Newtonian, special relativistic, and general relativistic (GR) hydrodynamics and magnetohydrodynamics (MHD), and GR-radiation hydrodynamics and MHD, as well as a module for evolving Lagrangian tracer or charged test particles (e.g., cosmic rays) are implemented using the framework. In two companion papers, we describe (1) a new solver for the Einstein equations based on the Z4c formalism, and (2) a GRMHD solver in dynamical spacetimes also implemented using the framework, enabling new applications in numerical relativity. By adopting Kokkos, the code can be run on virtually any hardware, including CPUs, GPUs from multiple vendors, and emerging Advanced RISC Machine processors. AthenaK shows excellent performance and weak scaling, achieving over 1 billion cell updates per second for hydrodynamics in three dimensions on a single NVIDIA Grace Hopper processor. It does this with a typical parallel efficiency of 80% on 65,536 AMD GPUs on the OLCF Frontier system. Such performance portability enables AthenaK to leverage modern exascale computing systems for challenging applications in astrophysical fluid dynamics, numerical relativity, and multimessenger astrophysics.

79 ASTRONOMY AND ASTROPHYSICS↗

thornado+FLASH-X: A Hybrid Discontinuous Galerkin–Implicit-explicit and Finite-volume Framework for Neutrino-radiation Hydrodynamics in Core-collapse Supernovae

We present neutrino-transport algorithms implemented in the toolkit for high-order neutrino-radiation hydrodynamics (thornado) and their coupling to self-gravitating hydrodynamics within the adaptive mesh refinement–based multiphysics simulation framework FLASH-X. thornado, developed primarily for simulations of core-collapse supernovae (CCSNe), employs a spectral, six-species two-moment formulation with algebraic closure and special-relativistic observer corrections accurate to $\mathcal{O}(v/c)$, and uses discontinuous Galerkin (DG) methods for phase-space discretization combined with implicit-explicit time stepping. A key development is a nonlinear neutrino–matter coupling algorithm based on nested fixed-point iteration with Anderson acceleration, enabling fully implicit treatment of collisional processes, including energy-coupling interactions such as neutrino–electron scattering and pair production. Coupling to finite-volume (FV) hydrodynamics is achieved through a hybrid DG-FV representation of the fluid variables and operator-split evolution within FLASH-X. The implementation is verified using basic transport tests with idealized opacities and relaxation and deleptonization problems with tabulated microphysics. Spherically symmetric CCSN simulations demonstrate accuracy and robustness of the coupled scheme, including close agreement with the CCSN simulation code Chimera. An axisymmetric CCSN simulation further demonstrates the viability of DG-based neutrino transport for multidimensional supernova modeling within FLASH-X. thornado’s neutrino-transport solver is GPU-enabled using OpenMP offloading or OpenACC, and all CCSN applications included in this work use the GPU implementation. Together, these results establish a foundation for future enhancements in physics fidelity, numerical algorithms, and computational performance, for increasingly realistic large-scale CCSN simulations.

Endeve, Eirik [Oak Ridge National Laboratory (ORNL↗

Refactoring the elastic–viscous–plastic solver from the sea ice model CICE v6.5.1 for improved performance

This study focuses on the performance of the elastic–viscous–plastic (EVP) dynamical solver within the sea ice model, CICE v6.5.1. The study has been conducted in two steps. First, the standard EVP solver was extracted from CICE for experiments with refactored versions, which are used for performance testing. Second, one refactored version was integrated and tested in the full CICE model to demonstrate that the new algorithms do not significantly impact the physical results. The study reveals two dominant bottlenecks, namely (1) the number of Message Parsing Interface (MPI) and Open Multi-Processing (OpenMP) synchronization points required for halo exchanges during each time step combined with the irregular domain of active sea ice points and (2) the lack of single-instruction, multiple-data (SIMD) code generation. The standard EVP solver has been refactored based on two generic patterns. The first pattern exposes how general finite differences on masked multi-dimensional arrays can be expressed in order to produce significantly better code generation by changing the memory access pattern from random access to direct access. The second pattern takes an alternative approach to handle static grid properties. The measured single-core performance improvement is more than a factor of 5 compared to the standard implementation. The refactored implementation of strong scales on the Intel® Xeon® Scalable Processors series node until the available bandwidth of the node is used. For the Intel® Xeon® CPU Max series, there is sufficient bandwidth to allow the strong scaling to continue for all the cores on the node, resulting in a single-node improvement factor of 35 over the standard implementation. This study also demonstrates improved performance on GPU processors.

58 GEOSCIENCES↗

Replication Package for "Union-Find and Usability: A Case Study and Analysis of Rust Formal Verifier

This replication package is a case study on automated deductive verification for Rust for practical programs. It is a companion artifact to a corresponding usability study on verification titled "Union-Find and Usability: A Case Study and Analysis of Rust Formal Verifiers". It seeks to answer the question "Can Rust developers today use Rust verifiers to verify their code?". To answer this question, the study contrasts the verification experience of two mature Rust verifiers, Creusot and Prusti, by using the tools to develop a verified implementation of union-find in Rust. The union-find implementation is based on real-world code as used in the popular egg E-graph library. The artifact consists of two different verified libraries, one using Creusot and one using Prusti. The libraries have similar Rust interfaces and high-level proofs but differ in their details: Creusot and Prusti have different annotation languages and support different proof styles. Each implementation can be verified with its respective tool and compiles as a traditional Rust development.

Sarracino, John↗

Automation of Nanoparticle Synthesis Processes in a Plasma Environment Using LabVIEW

This work presents an automated control system for the synthesis of nanomaterials by plasma-enhanced chemical vapor deposition (PECVD), implemented using the LabVIEW software environment. The main objective of the study is to develop an integrated hardware-software platform that enables sequential control of the key stages of the PECVD process, including vacuum chamber preparation, pressure monitoring, working gas supply, plasma ignition, power matching, cyclic nanomaterial growth, and optical monitoring of nanoparticles in the plasma environment. The use of LabVIEW made it possible to integrate actuator control, experimental parameter acquisition, and realtime process visualization within a single automated system. The automated cycle begins with evacuation of the reaction chamber to a predefined base pressure. Transition to the next stage is permitted only after the specified pressure threshold has been reached, ensuring reproducible initial conditions for each experiment. The program then controls the supply of the working gas through mass flow controllers (MFCs). In this work, two gas-flow control modes were considered: analog control using a 0-5 V voltage signal and digital communication via RS-232 interface. It was shown that the analog approach requires accurate scaling of the control voltage, since applying 5 V corresponds to full-scale opening of the controller and results in the maximum gas flow. In contrast, the RS232 interface enables the gas flow rate to be specified directly in sccm, improving the accuracy, flexibility, and convenience of gas-environment control. After pressure stabilization, LabVIEW initiates RF plasma ignition and executes the RF matching algorithm aimed at minimizing reflected power and improving the stability of the plasma process. A separate software module implements the cyclic nanomaterial growth mode, in which the plasma-on time, plasma duration, and total number of synthesis cycles are predefined. This approach makes it possible to control material accumulation on the substrate and to correlate the process parameters with the morphological characteristics of the resulting nanostructures. The final module of the system is designed for optical monitoring of the nanoparticle cloud density in dusty plasma. For this purpose, the change in the intensity of laser radiation passing through the plasma region is recorded using a photodetector and a Keithley 2401 measuring unit connected to LabVIEW via RS-232 interface. The difference between the initial and modified optical signal intensity is used as a diagnostic parameter characterizing the formation and temporal evolution of nanoparticles. The developed system demonstrates that LabVIEW can be effectively applied not only for the automation of individual instruments, but also for the implementation of a complete digital control cycle for PECVD-based nanomaterial synthesis.

PECVD↗

Adapter Signaling Evaluations on EVs [Slides]

In case J3400 and J3400/1 define a basic analog signaling approach for DC charging adapters to communicate over-temperature events to both the EV and EVSE. They also require the EV to implement mitigation and corrective actions in cases where the EVSE does not respond to adapter thermal signals. In this study, multiple production vehicles were evaluated for responsiveness to assess field readiness and ensure consistent performance in accordance with the SAE J1772, J3400 and J3400/1 standards. A series of test cases were developed and executed across different vehicles, using an adapter thermal breakout fixture designed by NLR for evaluations. The results revealed notable variations in behavior: some vehicles expected to comply with J1772 were found to be non-responsive under certain conditions. Additionally, while some EV OEMs appear to align with the J3400 standard, discrepancies exist due to its evolving nature, resulting in inconsistent implementation of the latest requirements. These findings highlight the need for alignment within standards organizations to ensure consistent interpretation, encourage compliance, and reduce potential confusion across implementations.

33 ADVANCED PROPULSION SYSTEMS↗

Innovating the next generation of commercial smart building software

Nearly 30% of commercial building energy use is wasted due to equipment faults and HVAC controls problems. The result is increased emissions, compromised comfort and productivity, and less reliable coordination of building power needs with a clean grid. The energy impact alone represents $17 billion in potential savings. Today’s smart building software provides a robust solution to address these operational deficiencies. Energy management and information systems (EMIS) are saving up to 9% on average, with two-year paybacks. They are being incorporated into energy management processes, commissioning services, and utility programs. As effective as they are, two barriers prevent even deeper benefits; limited personnel to fix problems once they are identified, and the expense and time to manually implement changes in control systems. In partnership with the research community, the EMIS industry is developing new capabilities to overcome these barriers. Moving beyond siloed products for either fault detection and diagnostics, or optimal control, these new capabilities empower users to not only automatically identify faults, but also to push corrective action, and control improvements to their buildings. In this paper, several areas for enhancements are documented: ‘one-time’ correction of faults such as setpoints, schedules, and economizer lockouts; short-term active testing for automated proportional integral derivative (PID) loop tuning and functional testing; and continuous supervisory control for demand flexibility and year-round efficiency. Results are presented from a pair of partner implementations out of a dozen providers integrating these enhancements into their products, including field tests from across the country, and insights into operator acceptance and integration into operations and maintenance practices.

Casillas, Armando↗

Distributed Multi-GPU Community Detection on Exascale Computing Platforms

Community detection is a fundamental operation in graph mining, and by uncovering hidden structures and patterns within complex systems it helps solve fundamental problems pertaining to social networks, such as information diffusion, epidemics, and recommender systems. Scaling graph algorithms for massive networks becomes challenging on modern distributed-memory multi-GPU (Graphics Processing Unit) systems due to limitations such as irregular memory access patterns, load imbalances, higher communication-computation ratios, and cross-platform support. We present a novel algorithm HiPDPL-GPU (Distributed Parallel Louvain) to address these challenges. We conduct experiments involving different partitioning techniques to achieve an optimized performance of HiPDPL-GPU on the two largest supercomputers: Frontier and Summit. Remarkably, HiPDPL-GPU processes a graph with 4.2 billion edges in less than 3 minutes using 1024 GPUs. Qualitatively, the performance of HiPDPL-GPU is similar or better compared to other state-of-the-art CPU- and GPU-based implementations. While prior GPU implementations have predominantly employed CUDA, our first-of-its-kind implementation for community detection is cross-platform, accommodating both AMD and NVIDIA GPUs.

Sattar, Naw Safrin↗

Optimum Operation of Microgrid Systems that Employ PV Solar and Battery Systems

On a global scale, over 1.3 billion lack access to electricity (85% in rural areas), and approximately 2.8 billion people rely on traditional biomass for cooking. The World Health Organization estimates that household air pollution from inefficient stoves causes more premature deaths than malaria, tuberculosis, and HIV/AIDS. Increasing demand for energy has led to dramatic increases in emissions. The need for reliable electricity and limiting emissions drives research on Resilient Hybrid Energy Systems (RHESs), which provide cleaner energy by combining wind, solar, and biomass energy with traditional fossil energy, increasing production efficiency and reliability and reducing generating costs and emissions. Microgrids have been shown as an efficient means of implementing RHESs, with some focused mainly on reducing the environmental impact of electric power generation. The technical challenges of designing, implementing, and applying microgrids involve conducting a cradle-to-grave Life Cycle Analysis (LCA) to evaluate these systems’ environmental and economic performance under diverse operating conditions to evaluate resiliency. A sample RHES was developed and used to demonstrate the implementation in rural applications, where the system can provide reliable electricity for heating, cooling, lighting, and pumping clean water. The model and findings can be utilized by other regions around the globe facing similar challenges.

Nagapurkar, Prashant↗

Demystifying Solar in WAP

Over the last two years, the National Renewable Energy Laboratory has identified the pathways for implementing solar in the Weatherization Assistance Program and Low Income Home Energy Assistance Program and practices for successful implementation. Learn about successful implementation models from around the network and the new resources that can give you a head start in planning solar into your program, including a demonstration of tools to help evaluate solar potential and cost-effectiveness at the household level, as well as a new tool from the U.S. Department of Energy National Community Solar Partnership to help connect households receiving LIHEAP assistance to community solar subscriptions that guarantee energy bill savings.

ENERGY PLANNING, POLICY, AND ECONOMY,SOLAR ENERGY↗

Applicability of the Milestones Approach to Deployments of Transportable Nuclear Power Plants (TNPPs)

Transportable nuclear power plants (TNPPs) can provide potential benefits to countries embarking on nuclear programs, offering reduced infrastructure requirements, shorter timeframes for implementation, cost savings and greater deployment flexibility than larger conventional reactors. However, the deployment of a TNPP in a Host State comes with the obligation to establish sufficient regulatory, institutional, and technical infrastructure, which, among others, includes a legal and regulatory framework and a competent regulatory body to implement a State’s safeguards obligations. This paper considers how the unique technical and deployment features of TNPPs may affect the process of preparing for and implementing safeguards in nuclear newcomer countries. Evaluating this issue through the lens of the IAEA’s Milestones Approach, this paper discusses some potential implications arising from the shortening of some milestones phases due to reduced construction or licensing time for TNPPs, and the need for increased cooperation between Host States and Supplier States in preparing for and meeting certain safeguards obligations. These considerations are potentially relevant to various stakeholders: newcomer States considering TNPP deployment; the States and companies that supply such reactors; as well as organizations that support international safeguards capacity building.

Siserman-Gray, Ioana-Cristina↗

A Proposed Evaluation Framework for New and Emerging Low Embodied-Carbon Concrete Technologies

New opportunities for carbon reductions in buildings create a strong need for a common framework and method for those who design, build and influence construction to evaluate lifecycle carbon reductions from design decisions and technology choices. These opportunities include a wide range of low-embodied-carbon concrete materials being rapidly developed and introduced to the market. How to evaluate these newer materials and technologies has become critical for both public- and private-sector actors seeking to decarbonize building constructions by leveraging the Infrastructure Investment and Jobs Act (IIJA) and Inflation Reduction Act (IRA) funds. We propose an evaluation framework to assess the lifecycle carbon reductions from adoption of these technologies, including a subset of key “must have” (1) technical criteria (embodied carbon level, technology development stage); (2) market criteria (market size, scalability); and (3) financial criteria (cost of technology implementation compared to businessas-usual) from a range of options. We discuss how to use the framework and illustrate it using a “heatmap,” rating score and short case study of a promising technology. We also propose a plan to implement this framework that includes (1) standardized measurement and validation methods for verifying emission reductions from these technologies, and (2) avenues to implement real world demonstrations. We conclude with recommendations for next steps on framework refinement and commercialization strategy development.

Singh, Reshma↗