Search NASA⌕ Search

SEARCH · Search NASA

Results for “ART”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

OptiBench: An Optimization Benchmark Tool for Renewable Energy Problems

We propose a benchmark framework and visualization tool, OptiBench, for analyzing the performance of state-of-the-art optimization solvers across a variety of optimization problems in renewable energy research. Our framework is designed from the ground up in the Julia programming language and enables analysis at scale on high performance computing (HPC) systems. Our visualization tool allows effortless evaluation of optimization solver performance, robustness, and accuracy through intuitive plots, e.g., performance profiles, heat maps, and distribution plots. We have tested three benchmark suites relevant to the modeling of renewable energy systems, viz., CUTEst, PGLib-OPF, and WaterTAP water treatment optimization problems. We illustrate benchmarking of CUTEst using OptiBench on the National Renewable Energy Laboratory's (NREL) HPC Kestrel. Our findings indicate that MA57 HSL linear solver demonstrated the best overall performance for an experimental IPOPT implementation. Our work is ongoing and we intend to add support for more optimization solvers and benchmark test suites in the future.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Leveraging AI for Productive and Trustworthy HPC Software: Challenges and Research Directions

We discuss the challenges and propose research directions for using AI to revolutionize the development of high-performance computing (HPC) software. AI technologies, in particular large language models, have transformed every aspect of software development. For its part, HPC software is recognized as a highly specialized scientific field of its own. We discuss the challenges associated with leveraging state-of-the-art AI technologies to develop such a unique and niche class of software and outline our research directions in the two US Department of Energy–funded projects for advancing HPC Software via AI: Ellora and Durban.

Teranishi, Keita [ORNL] (ORCID:0000000166472690)↗

Measurement of $t\overline{t}$ production in association with additional b -jets in the eμ final state in proton–proton collisions at $\sqrt{s}$ = 13 TeV with the ATLAS detector

This paper presents measurements of top-antitop quark pair ($t\overline{t}$) production in association with additional b-jets. The analysis utilises 140 fb -1 of proton–proton collision data collected with the ATLAS detector at the Large Hadron Collider at a centre-of-mass energy of 13 TeV. Fiducial cross-sections are extracted in a final state featuring one electron and one muon, with at least three or four b-jets. Results are presented at the particle level for both integrated cross-sections and normalised differential cross-sections, as functions of global event properties, jet kinematics, and b-jet pair properties. Observable quantities characterising b-jets originating from the top quark decay and additional b-jets are also measured at the particle level, after correcting for detector effects. The measured integrated fiducial cross sections are consistent with $t\overline{t}$$b\overline{b}$ predictions from various next-to-leading-order matrix element calculations matched to a parton shower within the uncertainties of the predictions. State of-the-art theoretical predictions are compared with the differential measurements; none of them simultaneously describes all observables. Differences between any two predictions are smaller than the measurement uncertainties for most observables.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Theory uncertainties of the irreducible background to VBF Higgs production

Higgs boson production through gluon fusion in association with two jets is an irreducible background to Higgs boson production through vector boson fusion, one of the most important channels for analyzing and understanding the Higgs boson properties at the Large Hadron Collider. Despite a range of available simulation tools, precise predictions for the corresponding final states are notoriously hard to achieve. Using state-of-the-art fixed-order calculations as the baseline for a comparison, we perform a detailed study of similarities and differences in existing event generators. We provide consistent setups for the simulations that can be used to obtain identical parametric precision in various programs used by experiments. We find that NLO calculations for the two-jet final state are essential to achieve reliable predictions.

Chen, Xuan [Shandong U.] (ORCID:0000000167722196)↗

Strangeness enhancement at its extremes: multiple (multi-)strange hadron production in pp collisions at \(\sqrt{s}=5.02\) TeV

The probability to observe a specific number of strange and multi-strange hadrons (nS), denoted as P(nS), is measured by ALICE at midrapidity (|y| < 0.5) in $$\sqrt{s}=5.02$$ TeV proton-proton (pp) collisions, dividing events into several multiplicity-density classes. Exploiting, for the first time, a technique based on counting the number of strange-particle candidates event-by-event, this measurement allows one to extend the study of strangeness production beyond the mean of the distribution. This constitutes a new test bench for production mechanisms, probing events with a large imbalance between strange and non-strange content. The analysis of a large-statistics data sample makes it possible to extract P(nS) up to a maximum nS of 7 for $${\text{K}}_{\text{S}}^{0}$$, 5 for Λ and $$\overline{\Lambda }$$, 4 for Ξ− and $${\overline{\Xi } }^{+}$$, and 2 for Ω− and $${\overline{\Omega } }^{+}$$. From this, the probability of producing strange hadron multiplets per event is calculated, thereby enabling the extension of the study of strangeness enhancement to extreme situations where several strange quarks hadronize in a single event at midrapidity. Moreover, comparing hadron combinations with different u and d quark compositions and equal overall s quark content, the contribution to the enhancement pattern coming from non-strangeness related mechanisms is isolated. The results are compared with state-of-the-art phenomenological models implemented in commonly used Monte Carlo event generators, including PYTHIA 8 Monash 2013, PYTHIA 8 with QCD-based Color Reconnection and Rope Hadronization (QCD-CR + Ropes), and EPOS LHC, which incorporates both partonic interactions and hydrodynamic evolution. These comparisons show that the new approach dramatically enhances the sensitivity to the different underlying physics mechanisms modeled by each generator.

Abualrob, I J↗

Higher-order tails and RG flows due to scattering of gravitational radiation from binary inspirals

Abstract We establish and develop a novel methodology to treat higher-order non-linear effects of gravitational radiation that is scattered from binary inspirals, which employs modern scattering-amplitudes methods on the effective picture of the binary as a composite particle. We spell out our procedure to study such effects: assembling tree amplitudes via generalized-unitarity methods and employing the closed-time-path formalism to derive the causal effective actions, which encompass the full conservative and dissipative dynamics. We push through to a new state of the art for these higher-order effects, up to the third subleading tail effect, at order$$ {G}_N^5 $$ G N 5 and the 5-loop level, which corresponds to the 8.5PN order. We formulate the consequent dissipated energy for these higher-order corrections, and carry out a renormalization analysis, where we uncover new subleading RG flow of the quadrupole coupling. For all higher-order tail effects we find perfect agreement with partial observable results in PN and self-force theories, where available.

Physics↗

Measurements of W + W − production cross-sections in pp collisions at $\sqrt{s}=13$ TeV with the ATLAS detector

Measurements of W + W − → e ± νμ ∓ ν production cross-sections are presented, providing a test of the predictions of perturbative quantum chromodynamics and the electroweak theory. The measurements are based on data from pp collisions at $\sqrt{s}$ = 13 TeV recorded by the ATLAS detector at the Large Hadron Collider in 2015–2018, corresponding to an integrated luminosity of 140 fb −1 . The number of events due to top-quark pair production, the largest background, is reduced by rejecting events containing jets with b-hadron decays. An improved methodology for estimating the remaining top-quark background enables a precise measurement of W + W − cross-sections with no additional requirements on jets. The fiducial W + W − cross-section is determined in a maximum-likelihood fit with an uncertainty of 3.1%. The measurement is extrapolated to the full phase space, resulting in a total W + W − cross-section of 127 ± 4 pb. Differential cross-sections are measured as a function of twelve observables that comprehensively describe the kinematics of W + W − events. The measurements are compared with state-of-the-art theory calculations and excellent agreement with predictions is observed. A charge asymmetry in the lepton rapidity is observed as a function of the dilepton invariant mass, in agreement with the Standard Model expectation. A CP-odd observable is measured to be consistent with no CP violation. Limits on Standard Model effective field theory Wilson coefficients in the Warsaw basis are obtained from the differential cross-sections.

Accelerator Physics↗

Measurements and interpretations of W ± Z production cross-sections in pp collisions at $\sqrt{s}=13$ TeV with the ATLAS detector

Measurements of integrated and differential cross-sections for W ± Z production in proton-proton collisions are presented. The data collected by the ATLAS detector at the Large Hadron Collider from 2015 to 2018 at a centre-of-mass energy of $\sqrt{s}=13$ TeV are used, corresponding to an integrated luminosity of 140 fb −1 . The W ± Z candidate events are reconstructed using leptonic decay modes of the gauge bosons into electrons or muons. The integrated cross-section per lepton flavour for the production of W ± Z is measured in the detector fiducial region with a relative precision of 4%. The measured value is compared with the Standard Model prediction at a precision of up to next-to-next-to-leading-order in QCD and next-to-leading-order in electroweak. Cross-sections for W + Z and W − Z production and their ratio are presented. The W ± Z production is also measured differentially as functions of various kinematic variables, including new observables sensitive to CP-violation effects. All measurements are compared with state-of-the-art Standard Model predictions from fixed-order calculations or Monte Carlo generators based on next-to-leading-order matrix elements interfaced with parton showers. An effective field theory interpretation of the measurements is performed, considering both CP-conserving and CP-violating dimension-6 operators modifying the W ± Z production. In the absence of observed deviations from the Standard Model, limits on CP-conserving Wilson coefficients are extracted using the transverse mass of the W ± Z system. For CP-violating coefficients a machine learning approach is used to construct an observable with enhanced sensitivity to CP-violation effects.

hadron-hadron scattering↗

The Ghent Hybrid model in NuWro: a new neutrino single-pion production model in the GeV regime

Neutrino-induced single-pion production constitutes an essential interaction channel in modern neutrino oscillation experiments, with its products building up a significant fraction of the observable hadronic final states. Frameworks of oscillation analyses strongly rely on Monte Carlo neutrino event generators, which provide theoretical predictions of neutrino interactions on nuclear targets. Thus, it is crucial to integrate state-of-the-art single-pion production models with Monte Carlo simulations to prepare for the upcoming systematics-dominated landscape of neutrino measurements. In this work, we present the implementation of the Ghent Hybrid model for neutrino-induced single-pion production in the NuWro Monte Carlo event generator. The interaction dynamics includes coherently-added contributions from nucleon resonances and a non-resonant background, merged into the pythia branching predictions in the deep-inelastic regime, as instrumented by NuWro. This neutrino-nucleon interaction model is fully incorporated into the nuclear framework of the generator, allowing it to account for the influence of both initial- and final-state nuclear medium effects. We compare the predictions of this integrated implementation with recent pion production data from accelerator-based neutrino experiments. The results of the novel model show improved agreement of the generator predictions with the data and point to the significance of the refined treatment of the description of pion-production processes beyond the ∆ region.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

On the Trotter Error in Many-body Quantum Dynamics with Coulomb Potentials

Efficient simulation of many-body quantum systems is central to advances in physics, chemistry, and quantum computing, with a key question being whether the simulation cost scales polynomially with the system size. Here, in this work, we analyze many-body quantum systems with Coulomb interactions, which are fundamental to electronic and molecular systems. We prove that Trotterization for such unbounded Hamiltonians achieves a 1/4-order convergence rate, with explicit polynomial dependence on the number of particles. The result holds for all initial wavefunctions in the domain of the Hamiltonian, and the 1/4-order convergence rate is optimal, as previous work has numerically demonstrated that it can be saturated by a specific initial ground state. The main challenges arise from the many-body structure and the singular nature of the Coulomb potential. Our proof strategy differs from prior state-of-the-art Trotter analyses, addressing both difficulties in a unified framework. Our analysis treats the Coulomb potential as an unbounded operator without modification or regularization, and does not rely on spatial discretization, making it compatible with both first- and second-quantized circuit constructions.

Fang, Di [Duke Univ., Durham, NC (United States)]↗

Empirical Characterization and Modeling of Cohesive – to – Adhesive Shear Fracture Mode Transition due to Increased Adhesive Layer Thicknesses of Fiber Reinforced Composite Single – Lap Joints

Here, to ensure a strong adhesive bond, most standards and adhesive manufacturers specify a maximum adhesive gap of 1 mm when bonding fiber reinforced composite structures. In manufacturing large components, such as joining two halves of wind turbine blades, meeting this gap tolerance specification is impractical; gaps larger than 10 mm are common in large adhesively bonded composite structures using state-of-the-art manufacturing techniques. Currently, there is a lack of fundamental understanding of the failure mechanics of adhesive gaps larger than 3 mm. To create such understanding, glass fiber - acrylic thermoplastic composite panels bonded using different epoxy adhesives within single-lap joint samples with adhesive thicknesses of 0.1 mm, 0.3 mm, 1 mm, 3 mm, 5 mm, and 10 mm were sheared to failure. A transition from cohesive to adhesive failure was observed to occur about 1 mm to 3 mm joint thicknesses. Plotting the shear stress normalized by the ratio of the joint width to thickness as a function of the joint thickness normalized by the joint length is shown to result in the ability to fit simple empirically derived models of the cohesive-to-adhesive failure transition, regardless of the adhesive. Furthermore, using these normalized variables, all the observed cohesively failed specimens collapse to a single master curve, as do the adhesively failed specimens.

36 MATERIALS SCIENCE↗

Improving coastal water level estimation by merging nadir-only satellite altimetry data into a hydrodynamic model

Providing robust real time flood warnings is of paramount importance to coastal communities. Although state-of-the-art hydrodynamic models are capable of robustly predicting Coastal Water Levels (CWL), unresolved drivers affecting level fluctuations are often not represented by the model governing equations. This work evaluates a novel method to improve the performance of the ADvanced CIRCulation (ADCIRC) hydrodynamic model by assimilating observations from four nadir-only satellite altimetry missions against a set of National Oceanic and Atmospheric Administration (NOAA) gauge stations located across the entire U.S. East Coast. Two different types of simulations were performed – Open Loop (OL) and Data Assimilation (DA). Five different simulations were performed where four different satellite altimetry observations were assimilated individually and combined with two different scenarios – with and without considering the data quality flags. Results indicate that, despite their limited spatial coverage, merging nadir-only observations into ADCIRC from the newly launched Surface Water and Ocean Topography (SWOT)’s nadir altimeter can improve the model performance at 76% of the gauge locations, whereas Sentinel-6 improves 73% of the total locations, Jason-3 74%, and SARAL 21%. Furthermore, combining observations from SWOT-nadir, Jason-3, and Sentinel-6 can improve the ADCIRC performance at more than 80% of the gauge locations for 107-day simulation. Nadir-only satellite altimetry observations can be useful for improving the model performance even if flagged as “poor quality” near the coast. When the flagged data are disregarded, SWOT can improve ADCIRC at 78%, Sentinel-6 at 73%, Jason-3 at 53%, and SARAL at 21% of the gauge locations. The ability to improve the model simulations largely depends on the availability of a satellite overpass nearby. Therefore, model performance can be further enhanced if satellite observations are available during a storm surge event, stressing the importance of frequent satellite overpasses.

Aafnan Bhuiyan, Soelem↗

GraphTango: A Hybrid Representation Format for Efficient Streaming Graph Updates and Analysis

Abstract Streaming graph processing performs batched updates and analytics on a time-evolving graph. The underlying representation format of the graph largely determines the throughputs of these updates and analytics phases. Existing representation formats usually employ variations of hash tables or adjacency lists. However, a recent study showed that the adjacency-list-based approaches perform poorly on heavy-tailed graphs, and the hash table-based approaches suffer on short-tailed graphs. We propose GraphTango, a hybrid representation format that provides excellent update and analytics throughput regardless of the graph’s degree distribution. GraphTango dynamically switches among three different formats based on a vertex’s degree: (i) Low-degree vertices store the edges directly with the neighborhood metadata, confining accesses to a single cache line, (2) Medium-degree vertices use adjacency lists, and (3) High-degree vertices use hash tables as well as adjacency lists. In this case, the adjacency list provides fast traversal during the analytics phase, while the hash table provides constant-time lookups during the update phase. We further optimized the performance by designing an open-addressing-based hash table that fully utilizes every fetched cache line. In addition, we developed a thread-local lock-free memory pool that allows fast growing/shrinking of the adjacency lists and hash tables in a multi-threaded environment. We evaluated GraphTango with the help of the SAGA-Bench framework and compared it with four other representation formats: Stinger, Degree-aware Robin Hood Hashing, and two adjacency list-based formats with different workload balancing scheme. On average, GraphTango provides 4.5x higher insertion throughput, 3.2x higher deletion throughput, and 1.1x higher analytics throughput over the next best format. Furthermore, we integrated GraphTango with the state-of-the-art graph processing frameworks DZiG and RisGraph. Compared to the vanilla DZiG and vanilla RisGraph , [ GraphTango + DZiG ] and [ GraphTango + RisGraph ] reduces the average batch processing time by 2.3x and 1.5x, respectively.

Ahmed, Alif↗

Field–potential finite-difference time-domain (FiPo FDTD) technique for computational electromagnetics

Modeling light–matter interactions at the nanoscale requires accurate handling of coupled quantum and electromagnetic systems. This coupling requires information about the electric scalar potential Φ and the magnetic vector potential A, which are not typically calculated in standard computational electromagnetics implementations. To that end, we have developed a field–potential finite-difference time-domain (FiPo FDTD) algorithm, which solves a set of first-order equations for Φ and A alongside equations for the electric and magnetic fields E and H. The FiPo Basic code is essentially conventional FDTD, but with an added module that calculates the potentials. The FiPo Hybrid code self-consistently calculates both fields and potentials and is particularly suitable for coupling with quantum electronic transport solvers because it can be sourced by the potentials themselves. To terminate the domain and mimic infinite space, we have derived and implemented a convolutional perfectly matched layer (CPML) absorbing boundary condition for FiPo FDTD whose performance is on par with state-of-the-art CPMLs for standard FDTD. We present FiPo simulation results on several example systems.

Avazpour, L. [University of Wisconsin-Madison, WI ↗

230Th/234U radiochronometry for uranium materials by alpha spectrometry for nuclear forensics analysis

We have developed an alpha spectrometry method for 230Th/234U radiochronometry to determine the model separation age for uranium materials. This method offers an alternative to the more commonly used, but significantly more resource intensive multi-collector inductively coupled plasma mass spectrometry (MC-ICP-MS) while providing similar accuracy and precision. The development included radiochemical separation of uranium and thorium, calibration of the 232U tracer and the alpha spectrometry measurement. The method was validated by analyzing a certified reference material, CRM 125-A, which is a uranium oxide pellet assay, isotopic, and radiochronometric standard. The results were compared to those obtained by our routine MC-ICP-MS 230Th/234U radiochronometry analysis method to assess the performance parameters for alpha spectrometry compared to the state-of-the-art technique.

Macsik, Zsuzsanna [Los Alamos National Laboratory ↗

Simultaneous optimal system and controller design for multibody systems with joint friction using direct sensitivities

Abstract Real-world multibody systems are often subject to phenomena like friction, joint clearances, and external events. These phenomena can significantly impact the optimal design of the system and its controller. This work addresses the gradient-based optimization methodology for multibody dynamic systems with joint friction using a direct sensitivity approach. The Brown–McPhee model has been used to characterize the joint friction in the system. This model is suitable for the study due to its accuracy for dynamic simulation and its compatibility with sensitivity analysis. This novel methodology supports codesign of the multibody system and its controller, which is especially relevant for applications like robotics and servo-mechanical systems, where the actuation and design are highly dependent on each other. Numerical results are obtained using a software package written in Julia with state-of-the-art libraries for automatic differentiation and differential equations. Three case studies are provided to demonstrate the attractive properties of simultaneous optimal design and control approach for certain applications.

Verulkar, Adwait↗

Scalable training of trustworthy and energy-efficient predictive graph foundation models for atomistic materials modeling: a case study with HydraGNN

We present our work on developing and training scalable, trustworthy, and energy-efficient predictive graph foundation models (GFMs) using HydraGNN, a multi-headed graph convolutional neural network architecture. HydraGNN expands the boundaries of graph neural network (GNN) computations in both training scale and data diversity. It abstracts over message passing algorithms, allowing both reproduction of and comparison across algorithmic innovations that define nearest-neighbor convolution in GNNs. This work discusses a series of optimizations that have allowed scaling up the GFMs training to tens of thousands of GPUs on datasets consisting of hundreds of millions of graphs. Our GFMs use multitask learning (MTL) to simultaneously learn graph-level and node-level properties of atomistic structures, such as energy and atomic forces. Using over 154 million atomistic structures for training, we illustrate the performance of our approach along with the lessons learned on two state-of-the-art US Department of Energy (US-DOE) supercomputers, namely the Perlmutter petascale system at the National Energy Research Scientific Computing Center and the Frontier exascale system at Oak Ridge Leadership Computing Facility. The HydraGNN architecture enables the GFM to achieve near-linear strong scaling performance using more than 2000 GPUs on Perlmutter and 16,000 GPUs on Frontier.

97 MATHEMATICS AND COMPUTING↗

Mixed-precision numerics in scientific applications: survey and perspectives

The explosive demand for artificial intelligence (AI) workloads has led to a significant increase in silicon area dedicated to lower-precision computations on recent high-performance computing hardware designs. However, mixed-precision capabilities, which can achieve performance improvements of up to 8x compared to double-precision in extreme compute-intensive workloads, remain largely untapped in most scientific applications. A growing number of efforts have shown that mixed-precision algorithmic innovations can deliver superior performance without sacrificing accuracy. These developments should prompt computational scientists to seriously consider whether their scientific modeling and simulation applications could benefit from the acceleration offered by new hardware and mixed-precision algorithms. In this survey, we (1) review progress across diverse scientific domains—fluid dynamics, weather and climate, quantum chemistry, and computational genomics—that have begun adopting mixed-precision strategies; (2) examine state-of-the-art algorithmic techniques such as iterative refinement, splitting and emulation schemes, and adaptive precision solvers; (3) assess their implications for accuracy, performance, and resource utilization; and (4) survey the emerging software ecosystem that enables mixed-precision methods at scale. We conclude with perspectives and recommendations on cross-cutting opportunities, domain-specific challenges, and the role of co-design between application scientists, numerical analysts, and computer scientists. Collectively, this survey underscores that mixed-precision numerics can reshape computational science by aligning algorithms with the evolving landscape of hardware capabilities.

Graphics processing units↗