Search NASA⌕ Search

SEARCH · Search NASA

Results for “Portable application”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Scintillation light detection in polycrystalline diamond using single photon detectors

Here, this study investigates the scintillation properties of polycrystalline diamond for particle detection applications, particularly in neutron and alpha radiation environments. Polycrystalline diamonds provide a cost-effective alternative to monocrystalline diamonds while retaining essential detection properties. Photoluminescence measurements were performed to analyze emission spectra, revealing distinct characteristics based on impurity content and crystallinity. Scintillation responses were assessed using Silicon Photomultipliers (SiPMs), demonstrating the capability of polycrystalline diamond powders to respond to alpha irradiation, albeit with reduced resolution compared to traditional scintillators. A prototype neutron detector was developed by combining diamond powder with neutron-sensitive 6 LiF, and its performance was evaluated through experimental testing and Geant4 simulations. The findings indicate that polycrystalline diamond-based detectors can achieve significant detection efficiency while remaining insensitive to gamma radiation, offering potential for portable neutron detection applications.

47 OTHER INSTRUMENTATION↗

Compact fiber-coupled narrowband two-mode squeezed light source

Quantum correlated states of light, such as squeezed states, are a fundamental resource for the development of quantum technologies, as they are needed for applications in quantum metrology, quantum computation, and quantum communications. It is thus critical to develop compact, efficient, and robust sources to generate such states. Here, we report on a compact, narrowband, fiber-coupled source of two-mode squeezed states of light at 795 nm based on four-wave mixing (FWM) in an 85 Rb atomic vapor. The source is designed in a small modular form factor, with two input fiber-coupled beams, the seed and pump beams required for the FWM, and two output fibers, one for each of the modes of the squeezed state. The system is optimized for low pump power (135 mW) to achieve a maximum intensity-difference squeezing of 4.4 dB after the output of fibers at an analysis frequency of 1 MHz. Furthermore, the narrowband nature of the source makes it ideal for atomic-based quantum sensing and quantum networking configurations that rely on atomic quantum memories. Such a source paves the way for a versatile and portable platform for applications in quantum information science.

Jain, Umang [University of Oklahoma, Norman, OK (U↗

Back to functional hydrides: Effects of neutron-irradiated microstructure on hydrogen retention in yttrium hydride

Functional hydrides are promising candidates for advanced nuclear reactors, particularly in portable or transportable applications, due to their high hydrogen-retention capabilities, enabling efficient neutron moderation, and compact reactor design. However, hydrogen mobility in hydrides at elevated irradiation temperatures poses significant technological challenges, necessitating a comprehensive understanding of their irradiation behavior. Furthermore, this study investigated the microstructural and chemical stability of neutron-irradiated yttrium hydrides to assess their hydrogen-retention capacity. A targeted literature review was also conducted to contextualize neutron-irradiation effects on functional hydrides with regards to structural stability and hydrogen retention. Experimental characterizations revealed that, at high temperatures, irradiated hydrides retained their phase stability, which was likely enhanced by irradiation-induced microstructure evolution. Notably, an amorphous yttrium and oxygen -rich surface layer was present at the free surface of the hydride. Its thickness decreased while a continuous crystalline Y-O-rich layer was formed with increasing neutron damage. Additionally, the number density of dislocation loops and cavities generally increased as a function of neutron dose. First-principles calculations of hydrogen behavior within yttrium vacancy clusters in yttrium hydrides demonstrated vacancy-size-dependent hydrogen stability and configuration, highlighting the role of vacancy geometry in regulating hydrogen retention. Thus, the presence of irradiation-induced dislocation loops and cavities were hypothesized to improve hydrogen retention. Collectively, these findings advance the understanding of hydride behavior under neutron irradiation as well as their technological readiness for portable or transportable nuclear reactors.

Defect clusters↗

A bending test protocol for characterizing the mechanical performance of flexible photovoltaics

Flexible photovoltaics (PV) represents a promising research field with great potential for wearable, portable, and indoor applications. Significant progress has been made in recent years, with flexible emerging PV reporting power conversion efficiency (PCE) over 24%. Yet, there is a need for a unifying protocol to assess PV performance, compare research results, and evaluate the state- of-the-art achievements in flexible PV. In this paper, a consensus protocol is presented for measuring PCE over 1,000 bending cycles under 1% strain. Moreover, several good practice guidelines are proposed including bending procedures, flexibility testing with and without encapsulation, and ambient conditions during testing (e.g., temperature, humidity, illumination). Notably, the importance of uniform application of bending radius and the testing of parallel and perpendicular orientations of bending axis with respect to the direction of the electric current are emphasized. These recommendations aim to promote consistency in device comparison and allow for better reproducibility.

Fukuda, Kenjiro↗

Time-temperature history and input files for ExaCA v2.0 scaling, performance, and demonstration simulations

The files in this data repository are used in various sections of the manuscript "ExaCA v2.0: A versatile, scalable, and performance portable cellular automata application for additive manufacturing solidification" by Rolchigo et al. (DOI: 10.1016/j.commatsci.2025.113734). The README file references dataset numbers as given in the manuscript's Table 2, as well as the manuscript's relevant subsections.

36 MATERIALS SCIENCE↗

Evaluating integration and performance of containerized climate applications on a Hewlett Packard Enterprise Cray system

Containers have taken over large swaths of cloud computing as the most convenient way of packaging and deploying applications. The features that containers offer for packaging and deploying applications translate to high performance computing (HPC) as well. At The National Oceanic and Atmospheric Administration, containers provide an easy way to build and distribute complex HPC applications, allowing faster collaboration, portability, and experiment computer environment reproducibility amongst the scientific community. The challenge arises when applications rely on message passing interface (MPI). This necessitates investigation into how to properly run these applications with their own unique requirements and produce performance on par with native runs. We investigate the MPI performance for benchmarks and containerized climate models for various containers covering selection of compiler and MPI library combinations from the Cray provided programming environments on the Cray XC supercomputer GAEA. Performance from the benchmarks and the climate models shows that for the most part containerized applications perform on par with the natively built applications when the system optimized Cray MPICH libraries are bound into the container, and the hybrid model containers have poor performance in comparison. We also describe several challenges and our solutions in running these containers, particularly challenges with heterogeneous jobs for the containerized model runs.

Abraham, Subil↗

GEOS: A performance portable multi-physics simulation framework for subsurface applications

GEOS is a simulation framework focused on solving tightly coupled multi-physics problems with an initial emphasis on subsurface reservoir applications. Currently, GEOS supports capabilities for studying carbon sequestration, geothermal energy, hydrogen storage, and related subsurface applications. The unique aspect of GEOS that differentiates it from existing reservoir simulators is the ability to simulate tightly coupled compositional flow, poromechanics, fault slip, fracture propagation, and thermal effects, etc. Extensive documentation is available on the GEOS documentation pages (GEOS Documentation, 2024). Note that GEOS, as presented here, is a complete rewrite of the previous incarnation of the GEOS referred to in (Settgast et al., 2017).

58 GEOSCIENCES↗

CHARM-SYCL & IRIS: A Tool Chain for Performance Portability on Extremely Heterogeneous Systems

Performance portability is becoming crucial as high-performance computing systems become increasingly heterogeneous. We have many options for CPUs and accelerators (e.g., GPUs) but also for non-Von Neumann architectures such as field-programmable gate arrays. This paper presents the CHARM-SYCL unified programming environment for multiple accelerator types as a performance-portable programming environment. It uses the IRIS library developed at Oak Ridge National Laboratory as the back end accelerator runtime. IRIS has a high-performance scheduler to distribute tasks across accelerators. This design allows us to run an application from the same source on multiple systems with multiple configurations. We provide three types of portability with CHARM-SYCL: Portable Workflow, Compiler and Runtime Portability, and Application and Performance Portability. We implement a Monte Carlo simulation benchmark code on the CHARM-SYCL execution environment and demonstrate that our programming environment can accommodate extremely heterogeneous systems.

Fujita, Norihisa↗

Creating Apptainer Workflows with Docker-Compose-like Utilities

Creating Apptainer Workflows with Docker-Compose-like Utilities In this presentation, I will explore the utilization of a tool called process-compose, inspired by docker-compose, to create Apptainer-based services. This approach allows for easy deployment and management of fully containerized applications on High Performance Computing (HPC) systems without requiring elevated privileges. Benefits to the Ecosystem: By incorporating process-compose and Apptainer, I aim to address several key challenges in the HPC ecosystem: Simplified Workflow Management: Process-compose provides a user-friendly interface for defining and managing complex containerized application services, reducing the setup time and lowering the barrier to entry for new users. Enhanced Portability: Apptainer ensures that containerized applications can run consistently across different HPC environments, promoting greater portability and reducing compatibility issues. Process-compose is also a single binary that does not need to be installed by admin level users. Community Driven Solutions: This approach aligns with the goals of the High Performance Software Foundation (HPSF) to advance community-driven solutions. By sharing our experiences and insights, I hope to foster collaboration and innovation within the HPC community. Increased Productivity: The combination of process-compose and Apptainer streamlines the serve deployment process, allowing researchers and developers to focus more on their scientific work rather than the intricacies of system or service administration. Through this presentation, attendees will gain valuable insights into the practical implementation of containerized workflows on HPC systems, learn about the benefits of using process-compose and Apptainer, and understand how these tools can contribute to a more efficient HPC ecosystem.

97 - MATHEMATICS AND COMPUTING↗

Advances in ArborX to support exascale applications

ArborX is a performance portable geometric search library developed as part of the Exascale Computing Project (ECP). In this paper, we explore a collaboration between ArborX and a cosmological simulation code HACC. Large cosmological simulations on exascale platforms encounter a bottleneck due to the in-situ analysis requirements of halo finding, a problem of identifying dense clusters of dark matter (halos). This problem is solved by using a density-based DBSCAN clustering algorithm. With each MPI rank handling hundreds of millions of particles, it is imperative for the DBSCAN implementation to be efficient. In addition, the requirement to support exascale supercomputers from different vendors necessitates performance portability of the algorithm. We describe how this challenge problem guided ArborX development, and enhanced the performance and the scope of the library. We explore the improvements in the basic algorithms for the underlying search index to improve the performance, and describe several implementations of DBSCAN in ArborX. Further, we report the history of the changes in ArborX and their effect on the time to solve a representative benchmark problem, as well as demonstrate the real world impact on production end-to-end cosmology simulations.

97 MATHEMATICS AND COMPUTING↗

Portable and Cost-Effective Device for Reliable Detection of Counterfeit and Non-compliant Refrigerants in Diverse Applications

Counterfeit refrigerants pose significant challenges to safety, system reliability, and operational effectiveness due to their harmful contaminants or incompatible chemical compositions. Utilizing these noncompliant products can lead to reduced efficiency, equipment failures, and expensive repairs. Additionally, heightened demand for alternative refrigerants during the industry's transition has created supply gaps, enabling counterfeit products to proliferate. Accurate detection and analysis tools are therefore essential to verify refrigerant authenticity and ensure system integrity in diverse applications. This paper presents the development of a portable device designed for reliable identification and detailed analysis of refrigerant composition. By integrating precision gas sampling, controlled pressure regulation, and automated sensor technology, the device not only detects deviations from standard refrigerant properties but also provides a comprehensive composition breakdown. Pre-calibrated sensors measure the refrigerant gas to identify specific concentrations and contaminants, with an intuitive LED-based indicator system ensuring quick interpretation of results. The user-friendly interface enables operators to select refrigerant types for targeted testing, further enhancing accuracy and usability for field technicians. Comprehensive testing was conducted on mildly flammable A2L refrigerants, showcasing the device’s robustness and adaptability in analyzing composition and detecting discrepancies. The device demonstrated consistent accuracy across a range of refrigerant samples, affirming its reliability in diverse operational environments. Its design minimizes contamination risks during sampling and provides detailed composition results within 90 seconds, ensuring efficient and precise analysis. With a projected price point under $150, the proposed solution delivers affordability alongside its lightweight portability and straightforward operation. Unlike complex and costly alternatives, such as gas chromatography systems, this device provides an accessible option for technicians, customs personnel, and industry operators in need of quick and effective refrigerant verification. Compatible with both current formulations and emerging refrigerant technologies, the device addresses critical counterfeit detection needs across a range of applications. By delivering accurate composition analysis and counterfeit identification, this innovation enhances system performance, safety, and operational reliability in crucial industries.

Cheekatamarla, Praveen [ORNL] (ORCID:0000000248827↗

ORCHA: A performance portability system for extreme heterogeneity

Heterogeneity is the prevalent trend in the rapidly evolving high-performance computing (HPC) landscape in both hardware and application software. The diversity in hardware platforms, currently comprising various accelerators and a future possibility of specializable chiplets, poses a significant challenge for scientific software developers aiming to harness optimal performance across different computing platforms while maintaining the quality of solutions when their applications are simultaneously growing more complex. Code synthesis and code generation can provide mechanisms to mitigate this challenge. We have developed a divide and conquer approach where different aspects of performance are handled by different stand-alone tools that are interfaced with the application through generated code. This portability system, ORCHA, enables users to configure and orchestrate their computations among available resources on a platform by specifying a high-level recipe, thereby permitting a many-to-many paradigm where each recipe results in a different variant of the application. The core design goal is to let users decide the application’s hardware mapping and orchestration by editing only the high-level recipe—without modifying the maintained source code or binding the application to a particular runtime system. Tools in ORCHA distribution are: CG-Kit for translating the recipe into an execution graph; Milhoja to execute the graph by orchestrating data and task movement among hardware resources; and Macroprocessor that enables users to define their own code-shorthand for higher composability and easier management of code variants. Additionally, the design of ORCHA permits tools to work in a plug-and-play mode where the application can build and run without CG-Kit and Milhoja, and either tool can be swapped out for other tools with similar capabilities by modifying the code generation portion of ORCHA. In this paper, we describe the design of ORCHA and the role that code-generation plays in isolating applications from tools. We demonstrate the breadth of configurations ORCHA enables with a case study in which an application configuration is realized on three distinct hardware mappings—a GPU-centric, a CPU/GPU balanced, and a CPU/GPU concurrent layouts by using different recipes.

Lee, Youngjun↗

Exascale workflow applications and middleware: An ExaWorks retrospective

Exascale computers offer transformative capabilities to combine data-driven and learning-based approaches with traditional simulation applications to accelerate scientific discovery and insight. However, these software combinations and integrations are difficult to achieve due to the challenges of coordinating and deploying heterogeneous software components on diverse and massive platforms. Here, we present the ExaWorks project, which addresses many of these challenges. We developed a workflow Software Development Toolkit (SDK), a curated collection of workflow technologies that can be composed and interoperated through a common interface, engineered following current best practices, and specifically designed to work on HPC platforms. ExaWorks also developed PSI/J, a job management abstraction API, to simplify the construction of portable software components and applications that can be used over various HPC schedulers. The PSI/J API is a minimal interface for submitting and monitoring jobs and their execution state across multiple and commonly used HPC schedulers. We also describe several leading and innovative workflow examples of ExaWorks tools used on DOE leadership platforms. Furthermore, we discuss how our project is working with the workflow community, large computing facilities, and HPC platform vendors to address the requirements of workflows sustainably at the exascale.

97 MATHEMATICS AND COMPUTING↗

Dual Channel Dual Staging: Hierarchical and Portable Staging for GPU-Based In-Situ Workflow

In-situ workflows have emerged as an attractive approach for addressing data movement challenges at very large scales. Since GPU-based architectures dominate the HPC landscapes, porting these in-situ workflows, and, specifically, the inter-application data exchange, to GPU-based systems can be challenging. Technologies such as GPUDirect RDMA (GDR), which is typically used for I/O in GPU applications as an optimization that circumvents the CPU overhead, can be leveraged to support bulk data exchanges between GPU applications. However, current GDR design often lacks performance portability across HPC clusters built with different hardware configurations. Furthermore, the local CPU may also be effectively used as an auxiliary communication mechanism to offload data exchanges. In this paper, we present a dual channel dual staging approach for efficient, scalable, and performance-portable inter-application data exchange for in-situ workflows. This approach exploits the data access pattern within in-situ workflows along with the inherent execution asynchrony to accelerate data exchanges and, at the same time, improve performance portability. Specifically, the dual channel dual staging method leverages both the local CPU and the remote data staging server to build a hierarchical joint staging area and uses this staging area to transform blocking inter-application bulk data exchanges into best-effort local data movements between GPU and CPU. The dual channel dual staging is implemented as a portability extension of the Dataspaces-GPU staging framework. We present an experimental evaluation of its performance, portability, and scalability using this implementation on three leadership GPU clusters. The evaluation results demonstrate that the dual channel dual staging method saves up to 75% in data-exchange time compared to host-based, GDR, and alternate portable designs, while maintaining scalability (up to 512 GPUs) and performance portability across the three platforms.

Zhang, Bo [University of Utah]↗

Coherent diffraction imaging in the undergraduate laboratory

We present an undergraduate optics instructional laboratory designed to teach skills relevant to a broad range of modern scientific and technical careers. In this laboratory project, students image a custom aperture using coherent diffraction imaging, while learning principles and skills related to digital image processing and computational imaging, including multidimensional Fourier analysis, iterative phase retrieval, noise reduction, finite dynamic range, and sampling considerations. After briefly reviewing these imaging principles, we describe the required experimental materials and setup for this project. Our experimental apparatus is both inexpensive and portable, and a software application we developed for interactive data analysis is freely available.

Porter, J. Nicholas↗

CG-Kit: Code Generation Toolkit for performant and maintainable variants of source code applied to Flash-X hydrodynamics simulations

CG-Kit is a new Code Generation tool-Kit that we have developed as a part of the solution for portability and maintainability for multiphysics computing applications. The development of CG-Kit is rooted in the urgent need created by the shifting landscape of high-performance computing platforms and the algorithmic complexities of a particular large-scale multiphysics application: Flash-X. To efficiently use computing resources on a heterogeneous node, an application must have a map of computation to resources and a mechanism to move the data and computation to the resources according to the map. Most existing performance portability solutions are focussed on abstracting the expression of computations so that a unified source code can be specialized to run on different resources. However, such an approach is insufficient for a code like Flash-X, which has a multitude of code components that can be assembled in various permutations and combinations to form different instances of applications. Similar challenges apply to any code that has composability, where a single specified way of apportioning work among devices may not be optimal. Additionally, use cases arise where the optimal control flow of computation may differ for different devices while the underlying numerics remain identical. This combination leads to unique challenges including handling an existing large code base in Fortran and/or C/C++, subdivision of code into a great variety of units supporting a wide range of physics and numerical methods, different parallelization techniques for distributed and shared memory systems and accelerator devices, and heterogeneity of computing platforms requiring coexisting variants of parallel algorithms. All of these challenges demand that scientific software developers apply existing knowledge about domain applications, algorithms, and computing platforms to determine custom abstractions and granularity for code generation. There is a critical lack of tools to tackle those problems. CG-Kit is designed to fill this gap by providing a user with the ability to express their desired control flow and computation-to-resource map in the form a pseudocode-like recipe. It consists of standalone tools that can be combined into highly specific and, we argue, highly effective portability and maintainability toolchains. Here we present the design of our new tools: parametrized source trees, control flow graphs, and recipes. The tools are implemented in Python. They are agnostic to the programming language of the source code targeted for code generation. In conclusion, we demonstrate the capabilities of the toolkit with two examples, first, multithreaded variants of the basic AXPY operation, and second, variants of parallel algorithms within a hydrodynamics solver, called Spark, from Flash-X that operates on block-structured adaptive meshes.

Algorithmic portability↗

Machine Learning-Enabled Wearable Piezoelectric Acoustic Sensor for Real-Time Breast Abnormality Detection

In contemporary society, breast health has become a significant public health concern, particularly among women. According to statistics from the World Health Organization, both the incidence and mortality rates of breast tumors have steadily increased in recent years. Therefore, effective early-stage screening and postoperative monitoring are essential for maintaining breast health. However, conventional clinical diagnostic modalities are typically bulky, operationally complex, and unsuitable for continuous real-time monitoring, which limits their use in portable and everyday health management applications. To address these limitations, this study proposes a machine learning-integrated wearable piezoelectric sensing platform as an auxiliary tool for breast health assessment. The device consists of a PDMS matching layer embedded with flexible silver nanowires, a P(VDF-TrFE) piezoelectric layer, and a multi-channel low-noise signal acquisition circuit. It is capable of acquiring acoustic echo signals from tissue-mimicking environments and automatically evaluating signal validity using a convolutional neural network (CNN). By integrating piezoelectric sensing with deep learning-based signal analysis, the proposed system achieves a signal-to-noise ratio exceeding 70 dB and a real-time classification accuracy above 96% under controlled conditions. These results demonstrate that the platform provides a compact, portable, and intelligent approach for wearable sensing of mechanical heterogeneity and highlight its potential for future development in continuous biomedical monitoring technologies.

He, Shuaitong↗

Injection Locking of Gigahertz‐Frequency Surface Acoustic Wave Phononic Crystal Oscillator

Low-noise gigahertz (GHz) frequency sources are essential for applications in signal processing, sensing, and telecommunications. Surface acoustic wave (SAW) resonator-based oscillators offer compact form factors and low-phase noise due to their short mechanical wavelengths and high-quality (Q) factors. However, their small footprint makes them vulnerable to environmental variation, resulting in their poor long-term frequency stability. Injection locking is widely used to suppress frequency drift of lasers and oscillators by synchronizing to an ultra-stable reference. Here, injection locking of a 1-GHz SAW phononic-crystal oscillator is demonstrated, achieving 40-dB phase noise reduction at low offset frequencies and unperturbed low noise at large offset frequencies. Compared to a free-running SAW oscillator, which typically exhibits frequency drifts of several hundred hertz over minutes, the injection-locked oscillator reduces the frequency deviation to below 0.35 Hz. The locking range and oscillator dynamics is also investigated in the injection pulling region. The demonstrated injection-locked SAW oscillator could find applications in high-performance portable telecommunications and sensing systems.

injection locking↗