Search NASA⌕ Search

SEARCH · Search NASA

Results for “performance analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Insights from Optimizing HPL Performance on Exascale Systems: A Comparative Analysis of Panel Factorization

High performance LINPACK (HPL) remains the primary benchmark for evaluating supercomputing performance. It includes many parts with substantial internal complexity, and its performance is affected by a large number of parameters that interact in ways that are difficult to predict on large-scale heterogeneous supercomputer systems. We present a comprehensive performance analysis of HPL on Frontier, the world’s first exascale supercomputer, which achieved HPL performance of 1.35 exaflops. Through empirical parameter tuning, detailed modeling, and comparative evaluation, we uncover critical performance insights, share lessons learned, and outline best practices for effective parameter tuning on exascale systems. We introduce and evaluate two novel PDFACT strategies: a dedicated-thread (DT) variant and a GPU-based variant (GPUPDFACT) implementation using HIP cooperative groups, demonstrating that GPU-based factorization outperforms conventional CPU-based PDFACT on Frontier’s architecture. Our findings establish key performance factors for HPL on exascale systems and offer valuable guidance for future high-performance computing and benchmarking efforts.

Lu, Hao [ORNL] (ORCID:000000018941870X)↗

DESI DR1 Ly$α$ forest: 3D full-shape analysis and cosmological constraints

We perform an analysis of the full shapes of Lyman-$α$ (Ly$α$) forest correlation functions measured from the first data release (DR1) of the Dark Energy Spectroscopic Instrument (DESI). Our analysis focuses on measuring the Alcock-Paczynski (AP) effect and the cosmic growth rate times the amplitude of matter fluctuations in spheres of $8$$h^{-1}\text{Mpc}$, $fσ_8$. We validate our measurements using two different sets of mocks, a series of data splits, and a large set of analysis variations, which were first performed blinded. Our analysis constrains the ratio $D_M/D_H(z_\mathrm{eff})=4.525\pm0.071$, where $D_H=c/H(z)$ is the Hubble distance, $D_M$ is the transverse comoving distance, and the effective redshift is $z_\mathrm{eff}=2.33$. This is a factor of $2.4$ tighter than the Baryon Acoustic Oscillation (BAO) constraint from the same data. When combining with Ly$α$ BAO constraints from DESI DR2, we obtain the ratios $D_H(z_\mathrm{eff})/r_d=8.646\pm0.077$ and $D_M(z_\mathrm{eff})/r_d=38.90\pm0.38$, where $r_d$ is the sound horizon at the drag epoch. We also measure $fσ_8(z_\mathrm{eff}) = 0.37\; ^{+0.055}_{-0.065} \,(\mathrm{stat})\, \pm 0.033 \,(\mathrm{sys})$, but we do not use it for cosmological inference due to difficulties in its validation with mocks. In $Λ$CDM, our measurements are consistent with both cosmic microwave background (CMB) and galaxy clustering constraints. Using a nucleosynthesis prior but no CMB anisotropy information, we measure the Hubble constant to be $H_0 = 68.3\pm 1.6\;\,{\rm km\,s^{-1}\,Mpc^{-1}}$ within $Λ$CDM. Finally, we show that Ly$α$ forest AP measurements can help improve constraints on the dark energy equation of state, and are expected to play an important role in upcoming DESI analyses.

Cuceu, Andrei [LBL, Berkeley; Chicago U., KICP] (O↗

Understanding parton evolution in matter from renormalization group analysis

We perform a renormalization group (RG) analysis of collinear hadron production in deep inelastic scattering on nuclei. We consider the limit where the parent parton energy E is large, while the medium opacity remains small. We identify the fixed order and leading enhanced medium contributions to the semi-inclusive cross sections and derive RG equations that resum multiple emissions near the endpoints of the splitting functions at first order in opacity. These evolution equations treat the same type of radiation enhancement in matter as the modified Dokshitzer-Gribov-Lipatov-Altarelli-Parisi approach, but differ in the way one regulates the collinear divergences. They provide a unique analytic insight into the problem of resummation and a faster and more efficient path to phenomenology. The new RG evolution framework is applied to study fragmentation in eA reactions.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Constraints on the U ( 1 ) B − L model from global QCD analysis

We perform the first global QCD analysis of electron-nucleon deep-inelastic scattering and related high-energy data including the beyond the Standard Model U ( 1 ) B − L gauge boson, Z ′ . Contrary to the dark photon case, we find no improvement in the χ 2 relative to the baseline result. The finding allows us to place exclusion limits on the coupling constant of the Z ′ with mass in the range M Z ′ = 2 to 160 GeV. Published by the American Physical Society 2025

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Isospin dependence of the nuclear EMC effect from a global QCD analysis

We perform a new global QCD analysis of unpolarized parton distribution functions (PDFs) in the nucleon from proton, deuteron, and A = 3 data, including recent measurements of He 3 / D and H 3 / D cross section ratios from the MARATHON experiment at Jefferson Lab. Simultaneously inferring the PDFs and nucleon off-shell corrections allows both to be determined consistently, without theoretical assumptions about the isospin dependence of nuclear effects. The analysis provides strong evidence for the need of nucleon off-shell corrections to describe the A = 3 data, with large isoscalar and a suggestion of nonzero isovector contributions in A ≤ 3 nuclei. We find that the extracted EMC ratios of nuclear to nucleon structure functions for A = 2 and 3 differ from those naively extrapolated from heavy nuclei down to low A .

Cocuzza, C. [William & Mary] (ORCID:00000003492292↗

Independent Analyses of Antares R1 Core Design

This report summarizes the reactor analysis work performed by Oak Ridge National Laboratory (ORNL) for Antares Nuclear Incorporated's R1 Mark-1 heat pipe reactor design. The work was performed under a collaboration through the Department of Energy GAIN Nuclear Energy Voucher program. The purpose of this work is to perform an independent reactor analysis of the Antares reactor design and, where possible, compare ORNL's results with the results obtained by Antares. The overarching goal is to provide Antares with an independent review and calculation of their design, thereby contributing to Antares' mission to further their reactor design concept. A report with all ORNL technical results and proprietary details was provided to Antares Nuclear Incorporated separately; the present report provides a summary of the work completed, and all proprietary information is omitted. The independent reactor analysis was performed using both the SCALE code system for neutronics analysis and Flownex for thermal hydraulics analysis.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Open Source Software Prevalence Ingest Tool

The OSSP Ingest Tool accepts user-input organizational information, ingests IT/OT asset lists in Excel format, and ingests the associated CycloneDX SBOM's. It then performs analytics demonstrating the ability to answer the follow research questions: o RQ1. Ability to identify all OSS services running on, and all OSS components present within, an OT device o RQ1a: Ability to differentiate multiple versions of the same OSS component within each OT device. o RQ1b: Ability to differentiate running from not-running OSS components. o RQ1c: Ability to differentiate based on the originator of the component, because a supplier may have modified it after retrieval from the upstream software source. o RQ2. Ability to correlate the identity of a single OSS component across multiple OT devices, mitigating common name variations such as differences in capitalization, '-' vs '_', and so on. o RQ3. Ability to perform subset analysis of OSS components across multiple OT devices o RQ3a: Ability to perform subset analysis across OSS libraries, generating density & distribution graphs to identify commonly-used libraries and outliers. o RQ3b: Ability to perform subset analysis of a single OSS library, generating density & distribution by CI sector, by device type, by device make/model, and/or by firmware version. o RQ3c: Ability to perform subset analysis by grouping OSS libraries according to programming language, then overlay with RQ4b. o RQ3d: Ability to perform subset analysis by OSS upstream source, providing insight into degree of modifications performed by suppliers. o RQ4. Ability to identify dependencies (transitive and direct) of each differentiated OSS library within each OT device, and enable RQ1,2,3 iteratively for dependencies. o RQ1. Ability to identify all OSS services running on, and all OSS components present within, an OT device o RQ1a: Ability to differentiate multiple versions of the same OSS component within each OT device. o RQ1b: Ability Page

Kapadia, Shayna [Lawrence Livermore National Labor↗

TDCOSMO - XVI. Measurement of the Hubble constant from the lensed quasar WGD 2038–4008

Time-delay cosmography is a powerful technique to constrain cosmological parameters, particularly the Hubble constant (H0). The TDCOSMO Collaboration is performing an ongoing analysis of lensed quasars to constrain cosmology using this method. In this work, we obtain constraints from the lensed quasar WGD 2038−4008 using new time-delay measurements and previous mass models by TDCOSMO. This is the first TDCOSMO lens to incorporate multiple lens modeling codes and the full time-delay covariance matrix into the cosmological inference. The models are fixed before the time delay is measured, and the analysis is performed blinded with respect to the cosmological parameters to prevent unconscious experimenter bias. We obtain DΔ t = 1.68−0.38+0.40 Gpc using two families of mass models, a power-law describing the total mass distribution, and a composite model of baryons and dark matter, although the composite model is disfavored due to kinematics constraints. In a flat ΛCDM cosmology, we constrain the Hubble constant to be H0 = 65−14+23 km s−1 Mpc−1. The dominant source of uncertainty comes from the time delays, due to the low variability of the quasar. Future long-term monitoring, especially in the era of the Vera C. Rubin Observatory’s Legacy Survey of Space and Time, could catch stronger quasar variability and further reduce the uncertainties. This system will be incorporated into an upcoming hierarchical analysis of the entire TDCOSMO sample, and improved time delays and spatially-resolved stellar kinematics could strengthen the constraints from this system in the future.Key words: gravitational lensing: strong / cosmological parameters / distance scale⋆ Corresponding author; kcwong19@gmail.com.⋆⋆ NHFP Einstein fellow.

79 ASTRONOMY AND ASTROPHYSICS↗

A database and meta-analysis on the performance of exploding pusher implosions conducted at OMEGA

A database of 222 exploding pusher implosions conducted at the OMEGA Laser Facility is presented. The dataset consists of glass-shell capsules filled with varying pressures of D 2 , T 2 , and 3 He, which were imploded using square laser pulses with intensities ranging from 1 to 1 × 10 15 W/cm 2 . The database includes measurements of bang times, ion temperatures, and yields from the DD, D 3 He, and DT fusion reactions. A semi-analytic exploding pusher model is introduced, which effectively captures the observed trends in the data. This model predicts that the measurements scale according to a power-law relation based on the initial capsule and laser conditions. A generalized power-law scaling relation is directly fit to each dataset, providing a useful interpolation of the entire database. Overall, the database provides a valuable resource to estimating bang times, temperatures, and yields for the design of future experiments. Additionally, it provides a diverse set of data for validating more advanced implosion physics models.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Transport and losses of energetic particles in tokamaks in the presence of Alfvén activity using the new full orbit TAPaS code coupled to FAR3d

Recent developments and tools integrated into the TAPaS code are presented, enabling realistic scenario simulations of particle dynamics within experimental tokamak magnetic equilibria. In particular, the enhanced capabilities of TAPaS enable seamless coupling with external simulations, provided the metric and equilibrium magnetic field of the external code are known. Coupling TAPaS with the gyro-fluid code FAR3d, the transport and losses of energetic particles in the presence Alfvén eigenmodes (AEs) in DIII-D plasma discharge #159243 were investigated. Detailed analyses of prompt losses with and without collisions were performed. Then, further analysis was performed in the presence of electromagnetic perturbations resulting from AEs activity. The results indicate that, for the energies and the initial conditions considered here, the presence of AEs enhances the particle losses.

Alfven eigenmodes↗

PWR Core Analysis for Cycle Extension and Uprates with LEU+ Accident Tolerant Fuel and 80 GWd/Tonne Burnup Limit

The U.S. Nuclear Regulatory Commission has recently drafted a rule enabling fuel burnup increase in light water reactors up to 80 GWd/t. In conjunction with use of fuel enrichment up to 10%, and accident tolerant fuel (ATF), this is anticipated to facilitate 24-month cycles in PWRs, along with further power uprates. In this paper, PWR core analysis is performed for 20% increased PWR power output along with cycle extension up to 24 months, in combination with use of chromia-doped fuel and chromium-coated clad, considered to be the most near-term ATF concepts. In combination, these lead to challenging conditions with a core average discharge burnup of up to ~74 GWd/t, challenging even the 80 GWd/t burnup limit. Analysis is performed using the 2-step method with POLARIS (within SCALE) used for the lattice calculations and PARCS for the core calculations. Core designs are first baselined for current operating conditions (LEU, 62 GWd/t discharge burnup limit) and then derived that meet cycle constraints on power distribution and the updated lead pin discharge burnup limit while maintaining at least two batches of fuel in the core. Gadolina loadings in fuel pins of up to 8% are used, with enrichment zoning both within the core and, to a limited extent, within assemblies. Here, doped fuel with coated cladding can utilize the same core designs as the reference UOX cores, exhibiting slightly lower burnup due to higher fuel density, which also offsets the slight reactivity penalty from the doping and coating. For the analysis performed here, doped fuel enabled a core with 24-month cycle and 20% uprate to stay within the 80 GWd/t lead pin discharge burnup limit.

LEU+↗

Studying CPU and memory utilization of applications on Fujitsu A64FX and Nvidia Grace Superchip

ARM-based manycore CPU architectures are well-positioned to provide the rising memory throughput requirements of modern data intensive scientific applications in High Performance Computing (HPC). The Fujitsu A64FX CPU platform is based on the ARM v8.2A architecture, and is the processor of the flagship Japanese supercomputer - "Fugaku", which was previously ranked as the #1 supercomputer in the world according to the Top500 list. The Nvidia Grace superchip features 144 Neoverse V2 cores based on the ARMv9 architecture with 4x128b SVE2, providing exceptional computational power. The chip supports up to 480GB of memory, making it ideal for AI, machine learning, and scientific computing workloads. In this paper, we conduct a thorough performance exploration of a variety of parallel bandwidth-sensitive benchmarks and applications compiled with the native Fujitsu compiler on a Fugaku A64FX compute node and ARM (LLVM) Compiler on an NVIDIA Grace superchip compute node, engaging all the computational cores per cluster using OpenMP multithreading (assuming the cores can drive the available bandwidth). Our ultimate goals are to study the resource utilization of scientific applications and benchmarks on A64FX and Grace superchip, considering graph application scenarios ( GAP Benchmark suite) and eleven appli- cation proxies from the Rodinia heterogeneous benchmark suite (considering domains such as Data Mining, Bioinformatics, Fluid Dynamics, Pattern Recognition, etc.). Through exhaustive performance monitoring, we quantify the resource utilization of diverse OpenMP-based HPC applications on both the Fujitsu A64FX and the Nvidia Grace Superchip platforms.

benchmarking, Performance Analysis, High performan↗

Temperature‐Dependent Crystallization in Two‐Step Perovskite Deposition Revealed by In Situ GIWAXS and Machine Learning‐Guided Analysis

The performance and stability of perovskite solar cells are strongly governed by the crystallization behavior of their active layer. In two-step sequential deposition, early-stage film formation plays a decisive role in determining final phase purity and device quality. Guided by a data-driven analysis of nearly 39 000 devices in the FAIR perovskite database, we identified solvent-mediated quenching and thermal processing as key variables affecting power conversion efficiency (PCE), particularly in two-step fabrication. Here, to investigate these effects in real time, we designed and implemented a custom-built, temperature-controlled spin-coating system, enabling precise thermal modulation during precursor deposition. Using this platform, we performed in situ GIWAXS measurements to study the crystallization dynamics of FA 0.5 MA 0.5 PbI 3 films over a temperature range of 30°C–90°C. Our results reveal a non-monotonic relationship between spin-coating temperature and α-phase formation, governed by the interplay between precursor interdiffusion, PbI 2 crystallinity, and δ-phase suppression. The custom thermal control enabled us to isolate and quantify these competing effects during the earliest stages of film formation, providing mechanistic insight into how spin-coating temperature governs both phase purity and kinetic pathways in two-step perovskite systems. Temperature-dependent SEM and photovoltaic device measurements further demonstrate that early-stage crystallization pathways directly translate into differences in morphology, charge-transport continuity, and device performance. These findings inform targeted strategies for optimizing deposition protocols to balance rapid nucleation, phase stability, and device performance.

Saadawy, Ahmed [King Fahd University of Petroleum ↗

Geant4-based Analysis of Faraday Cup Performance for PIP-II Laser Wire Scanner System

The Proton Improvement Plan-II (PIP-II) accelerator upgrade at Fermilab marks a significant advancement in high-energy physics research. This initiative aims to enhance Fermilab's accelerator complex by replacing the existing linear accelerator (linac) with a warm front end (WFE) capable of accelerating H- beams up to 2.1 MeV. Subsequently, a superconducting linac (SCL), that further accelerates these beams up to 800 MeV. To accurately measure the transverse beam profile, traditional wire scanners will be utilized in the WFE section, while Laser wire scanners will be implemented along the SCL. The Faraday cup for the Laser wire scanners has been designed using the GEANT4 simulation toolkit. This poster presents a detailed analysis of its performance along the SCL, focusing on electron absorption, secondary electron emission, backscattering, etc.

Wijethunga, S. A.K.↗

Phosphoproteomics Modifications in Women with Rheumatoid Arthritis─Application of Web-Based Software to Enhance Data Visualization

Individuals with rheumatoid arthritis (RA) are at increased risk of functional disability, cardiovascular disease, and obesity, all of which are influenced by dysregulated skeletal muscle. Here, this pilot study aims to identify phosphoproteomics changes in RA skeletal muscle and visualize modifications through development of a web-based app designed to promote user-friendly data interpretation and visualization. NanoLC–MS/MS analysis was performed on vastus lateralis biopsies from three women with RA and matched healthy controls. Differential analysis was performed using the Limma R package. Kinase substrate enrichment analysis (KSEA) predicted changes in kinase activity. RA muscle displayed 35 upregulated and 60 downregulated phosphosites, including the cytoskeletal proteins TTN (Ser33201, Ser33013, Ser20925), NEB (Ser2219, Thr254, Ser33013, Ser20925), FLNA (Ser1459), and LASP1 (Ser146). Compared to healthy controls, KSEA predicted decreased activity of several kinases in RA muscle, including PRKACA and CDKs. All such changes were visualized by use of our web-based app. Overall, phosphoproteome analysis reveals signaling alterations in RA skeletal muscle linked to cytoskeletal proteins, representing candidate disease biomarkers; these modifications can be explored through use of our web-based software.

phosphoproteomics↗

Spotlight: efficient automated global optimization in rietveld analysis of diffraction data

Performing reliable Rietveld analysis on tens or hundreds of powder diffraction datasets from parametric or time-resolved experiments often poses a bottleneck in extracting meaningful results from the data. While automated analysis of data has recently been demonstrated, high temperature annealing studies, during which phase transformations occur and lattice parameters may change due to repartitioning of elements, are prime examples where automation by a simple phase identification from a database of room temperature structures or automation by sequential refinements is likely to fail. To enable reliable, efficient, automated Rietveld analysis, we present a Python package named Spotlight , building on established Rietveld packages such as MAUD, GSAS , or GSAS-II , which extends the refinement of best fit parameters to a global optimization using an ensemble of optimizers leveraging hierarchical parallel execution on high-performance computing clusters. Spotlight further enables the efficient design of refinement plans through the iterative automated machine-learning of a surrogate for the refinement on which the global optimizations are performed until results from the surrogate converge to the response surface data. We demonstrate Spotlight with the analysis of uranium molybdenum and Ti–6Al–4V datasets, as well as in two open-source tutorials analyzing aluminium oxide and lead sulphate.

36 MATERIALS SCIENCE↗