Search NASA⌕ Search

SEARCH · Search NASA

Results for “Generative AI”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

619 records · Page 35

Measurements and interpretations of W ± Z production cross-sections in pp collisions at $\sqrt{s}=13$ TeV with the ATLAS detector

Measurements of integrated and differential cross-sections for W ± Z production in proton-proton collisions are presented. The data collected by the ATLAS detector at the Large Hadron Collider from 2015 to 2018 at a centre-of-mass energy of $\sqrt{s}=13$ TeV are used, corresponding to an integrated luminosity of 140 fb −1 . The W ± Z candidate events are reconstructed using leptonic decay modes of the gauge bosons into electrons or muons. The integrated cross-section per lepton flavour for the production of W ± Z is measured in the detector fiducial region with a relative precision of 4%. The measured value is compared with the Standard Model prediction at a precision of up to next-to-next-to-leading-order in QCD and next-to-leading-order in electroweak. Cross-sections for W + Z and W − Z production and their ratio are presented. The W ± Z production is also measured differentially as functions of various kinematic variables, including new observables sensitive to CP-violation effects. All measurements are compared with state-of-the-art Standard Model predictions from fixed-order calculations or Monte Carlo generators based on next-to-leading-order matrix elements interfaced with parton showers. An effective field theory interpretation of the measurements is performed, considering both CP-conserving and CP-violating dimension-6 operators modifying the W ± Z production. In the absence of observed deviations from the Standard Model, limits on CP-conserving Wilson coefficients are extracted using the transverse mass of the W ± Z system. For CP-violating coefficients a machine learning approach is used to construct an observable with enhanced sensitivity to CP-violation effects.

hadron-hadron scattering↗

AI Model Benchmarking for Nonproliferation Applications: Steel Thread Benchmarking Task Force Technical Report (Rev. 2)

Steel Thread is a NA-22 venture that seeks to build trustworthy, reliable AI models that can be used in a wide variety of nonproliferation tasks. A key aspect of building these models is developing appropriate benchmarks and evaluation methods, which will enable the venture to identify and adapt models to provide the most value in the nonproliferation domain. Benchmarks must be relevant to key tasks in this domain, such as question answering, information retrieval, document summarization and classification, consensus analysis, and image and data analysis. This report 1) provides an overview of benchmark design, evaluation, and challenges; 2) reviews a variety of open benchmarks, with a focus on language models and tasks; and 3) identifies benchmarks that are most relevant to Steel Thread. This report is intended to serve as a basis for further efforts to classify and evaluate benchmarks and their correlation with success on nonproliferation-specific tasks. The Steel Thread venture has defined benchmarks to be a particular combination of a dataset (or datasets) and a metric (or metrics) conceptualized as representing one or more specific tasks or sets of abilities for a specific modality. It is adopted by a research community as a shared framework for comparing methods.1 It includes 1) Data: Labeled (a designated subset not used for training, which could be all the data), 2) Metric: A way to quantify performance, 3) Task/Ability: What the benchmark is testing, 4) Protocol: A structured and repeatable evaluation process, 5) Baseline/Reference Model: For comparison; could be statistical, rule-based, SME-derived, or another model, and 6) Maintenance Plan: to update with new information over time; important for long-term utility. For further clarity, the definition includes what a benchmark, in this context, is not. It is not a corpus of training data, specific to a model (it is intended to apply to a range of models), a universal evaluation of performance, a guarantee that the ‘top’ model on the leaderboard will be the best fit for every specific use case, an all-encompassing proof of a model’s universal quality, nor is it a one-size-fits-all measure of success. It does not cover every real-world constraint (like operational, ethical, or cost considerations), a systems integration test, or a unit test. This definition was inspired by and resulted from discussions within the Steel Thread Benchmarking Task Force. This group was formed to define what we would mean as a benchmark within Steel Thread but persisted as the need to develop a thorough understanding of the large and expanding existing benchmarking space. This technical report is a result of the group’s divide and conquer approach to exploring this space. The release of benchmarks might not be progressing as quickly as model development, but it is moving very fast, as many benchmarks quickly become saturated, when state-of-the-art models score so close to the benchmark’s ceiling that their results are virtually indistinguishable. At that point, the test no longer differentiates between new systems, so researchers usually stop reporting scores as the benchmark no longer informs about improvements from the next generation of models. In the OpenAI announcement of GPT-5, they reported results on six flagship public benchmarks (AIME 2025, SWE-bench Verified, Aider Polyglot, MMMU, HealthBench Hard, GPQA) but the full system-card covers roughly thirty-five separate evaluations, comprising hundreds of test task items in total. There have been some efforts to summarize benchmarks in specific fields, like for text-to-image generation, but these surveys have had a narrow methodology scope. Therefore, a comprehensive survey of all benchmarks or even all benchmarks that could be relevant to Steel Thread is outside of the scope of this report. We chose some specific benchmarks to investigate in detail.

97 MATHEMATICS AND COMPUTING↗

A deep generative model for deciphering cellular dynamics and in silico drug discovery in complex diseases

Human diseases are characterized by intricate cellular dynamics. Single-cell transcriptomics provides critical insights, yet a persistent gap remains in computational tools for detailed disease progression analysis and targeted in silico drug interventions. Here we introduce UNAGI, a deep generative neural network tailored to analyse time-series single-cell transcriptomic data. This tool captures the complex cellular dynamics underlying disease progression, enhancing drug perturbation modelling and screening. When applied to a dataset from patients with idiopathic pulmonary fibrosis, UNAGI learns disease-informed cell embeddings that sharpen our understanding of disease progression, leading to the identification of potential therapeutic drug candidates. Validation using proteomics reveals the accuracy of UNAGI’s cellular dynamics analysis, and the use of the fibrotic cocktail-treated human precision-cut lung slices confirms UNAGI’s predictions that nifedipine, an antihypertensive drug, may have anti-fibrotic effects on human tissues. UNAGI’s versatility extends to other diseases, including COVID, demonstrating adaptability and confirming its broader applicability in decoding complex cellular dynamics beyond idiopathic pulmonary fibrosis, amplifying its use in the quest for therapeutic solutions across diverse pathological landscapes.

Neural Network↗

Simultaneous Unbinned Differential Cross-Section Measurement of Twenty-Four Z+jets Kinematic Observables with the ATLAS Detector

Z boson events at the Large Hadron Collider can be selected with high purity and are sensitive to a diverse range of QCD phenomena. As a result, these events are often used to probe the nature of the strong force, improve Monte Carlo event generators, and search for deviations from standard model predictions. All previous measurements of Z boson production characterize the event properties using a small number of observables and present the results as differential cross sections in predetermined bins. In this analysis, a machine learning method called omnifold is used to produce a simultaneous measurement of twenty-four Z +jets observables using 139 fb -1 of proton-proton collisions at √s =13 TeV collected with the ATLAS detector. Unlike any previous fiducial differential cross-section measurement, this result is presented unbinned as a dataset of particle-level events, allowing for flexible reuse in a variety of contexts and for new observables to be constructed from the twenty-four measured observables.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Atomate2: modular workflows for materials science

High-throughput density functional theory (DFT) calculations have become a vital element of computational materials science, enabling materials screening, property database generation, and training of “universal” machine learning models. While several software frameworks have emerged to support these computational efforts, new developments such as machine learned force fields have increased demands for more flexible and programmable workflow solutions. This manuscript introduces atomate2, a comprehensive evolution of our original atomate framework, designed to address existing limitations in computational materials research infrastructure. Key features include the support for multiple electronic structure packages and interoperability between them, along with generalizable workflows that can be written in an abstract form irrespective of the DFT package or machine learning force field used within them. Our hope is that atomate2's improved usability and extensibility can reduce technical barriers for high-throughput research workflows and facilitate the rapid adoption of emerging methods in computational material science.

97 MATHEMATICS AND COMPUTING↗

Discovery of a nearby radio relic in the low-mass, merging cluster Abell 4067

Shock waves generated during cluster mergers offer a powerful probe of how large-scale structure grows and evolves in the Universe. As part of the MeerKAT-South Pole Telescope (SPT) survey, we report the discovery of a single arc-like radio relic in the galaxy cluster Abell 4067 ($z=0.099$), one of the lowest-mass clusters known to host such a structure. MeerKAT UHF-band (0.58–1.09 GHz) observations reveal a relic with a largest linear size of $\sim 1.48\pm 0.02$ Mpc, located at a projected distance of 0.95 Mpc from the cluster centre. XMM–Newton X-ray data show that the relic’s position and orientation relative to the intracluster medium (ICM) elongation are consistent with a merger-driven shock-wave scenario. The relic has an estimated radio power of $3.10\pm 0.03\times 10^{24}$ W Hz$^{-1}$ at 150 MHz. When placed in the $P_{150\, \mathrm{MHz}}$–$M_{500}$ scaling relation, the Abell 4067 relic appears less luminous compared to relics in more massive clusters, suggesting an association with weak merger shocks. This finding supports the idea that relics in low-mass clusters may form through less energetic merger events, leading to weak merger shocks. The latter is supported by the absence of a detectable central radio halo in Abell 4067, which reinforces the idea that luminous radio haloes are not a universal outcome of cluster mergers and highlights the role of cluster mass, merger energetics, and evolutionary stage in shaping diffuse radio emission in the ICM.

79 ASTRONOMY AND ASTROPHYSICS↗

MeVPrtl: An Event Generator for Dark Sector Particles in the Short-Baseline Neutrino Program

MeVPrtl is a modular event generator of beyond the Standard Model (BSM) physics particles developed for use in the Short-Baseline Neutrino (SBN) Program. A large class of BSM physics models predict that new particles could be produced in the intense Booster Neutrino Beam (BNB) and Neutrinos at the Main Injector (NuMI) beams at Fermilab, travel to the SBN Program detectors, and decay into Standard Model (SM) particles. These new physics models are motivated by dark matter, the neutrino mass scale, and a solution to the strong CP problem. MeVPrtl provides an interface to implement the overlapping phenomenology of these models, and to connect them with meson flux inputs and object outputs used by the SBN Program's LArSoft-based detector simulation. Implementations for the Higgs portal, heavy neutral lepton, and heavy QCD axion models exist within MeVPrtl. In this paper these implementations and their validation, as well as details of the MeVPrtl interface, are specified.

Abratenko Ao, P.↗