Search NASA⌕ Search

SEARCH · Search NASA

Results for “common framework”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Theory of ab initio downfolding with arbitrary-range electron-phonon coupling

Ab initio downfolding describes the electronic structure of materials within a low-energy subspace, often around the Fermi level. Typically starting from mean-field calculations, this framework allows for the calculation of one- and two-electron interactions, and the parametrization of a many-body Hamiltonian representing the active space of interest. The subsequent solution of such Hamiltonians can provide insights into the physics of strongly correlated materials. While phonons can substantially screen electron-electron interactions, electron-phonon coupling has been commonly ignored within ab initio downfolding, and when considered, this is done only for short-range coupling. Here we propose a theory of ab initio downfolding that accounts for short- and long-range electron-phonon coupling on equal footing. Our practical computational implementation is readily compatible with current downfolding approaches. We apply our approach to polar materials MgO and GeTe, and we reveal the importance of both short-range and long-range electron-phonon coupling in determining the magnitude of electron-electron interactions. Our results show that in the static limit, phonons reduce the on-site repulsion between electrons by 40% for MgO and by 79% for GeTe. Our framework also predicts that overall attractive nearest-neighbor interactions arise between electrons in GeTe, consistent with superconductivity in this material.

Tubman, Norm M↗

Unsupervised multimodal fusion of in-process sensor data for advanced manufacturing process monitoring

Effective monitoring of manufacturing processes is crucial for maintaining product quality and operational efficiency. Modern manufacturing environments often generate vast amounts of complementary multimodal data, including visual imagery from various perspectives and resolutions, hyperspectral data, and machine health monitoring information such as actuator positions, accelerometer readings, and temperature measurements. However, fusing and interpreting this complex, high-dimensional data presents significant challenges, particularly when labeled datasets are unavailable or impractical to obtain. This paper presents a novel approach to multimodal sensor data fusion in manufacturing processes, inspired by the Contrastive Language-Image Pre-training (CLIP) model. We leverage contrastive learning techniques to correlate different data modalities without the need for labeled data, overcoming limitations of traditional supervised machine learning methods in manufacturing contexts. Our proposed method demonstrates the ability to handle and learn encoders for five distinct modalities: visual imagery, audio signals, laser position (x and y coordinates), and laser power measurements. By compressing these high-dimensional datasets into low-dimensional representational spaces, our approach facilitates downstream tasks such as process control, anomaly detection, and quality assurance. The unsupervised nature of our method makes it broadly applicable across various manufacturing domains, where large volumes of unlabeled sensor data are common. We evaluate the effectiveness of our approach through a series of experiments, demonstrating its potential to enhance process monitoring capabilities in advanced manufacturing systems. This research contributes to the field of smart manufacturing by providing a flexible, scalable framework for multimodal data fusion that can adapt to diverse manufacturing environments and sensor configurations. The proposed method paves the way for more robust, data-driven decision-making in complex manufacturing processes.

Contrastive Learning↗

Data and scripts associated with a manuscript modeling microbial regulation of priming effects

This data package is associated with the publication “Modeling Microbial Regulatory Feedback in Organic Matter Decomposition Identifies Copiotrophic Traits as Key Drivers of Positive Priming” published as a preprint on BioRXiv by Ahamed et al. (2026); https://doi.org/10.1101/2024.08.11.607483. The package contains MATLAB scripts and saved simulation outputs used to implement a cybernetic model of microbial regulation during complex organic matter (OM) decomposition governing priming effects. It includes models of (i) single microbial functional groups (copiotrophic or oligotrophic degraders) and (ii) binary consortia composed of degraders and non-degraders with contrasting or common growth traits. Simulation results were generated using Monte Carlo analyses, with randomized key model parameters across a range of environmental mixing fractions of complex and labile OM. The dataset was created to provide a transparent and reusable computational framework for systematically exploring how microbial growth traits, metabolic regulation, and community composition influence OM decomposition dynamics and priming effects. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to a readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes the variable definitions. This package includes: (1) annotated MATLAB code implementing the system of ordinary differential equations and cybernetic control laws; (2) saved output files containing data (e.g., biomass, substrates, enzyme levels, priming metrics); and (3) scripts for processing saved outputs and regenerating figures. Specifically, the data package contains three main MATLAB scripts: runPrimingModel.m, runPlotData.m, and runPlotSuppFigS1.m, along with this readme and supporting documentation. Users should begin with runPrimingModel.m, which contains the annotated code implementing the system of ordinary differential equations and cybernetic control laws. This script runs the Monte Carlo simulations of microbial OM decomposition and allows users to modify microbial trait definitions, adjust parameter distributions, or define new community configurations. Simulation outputs are automatically saved as .mat files in the folder named SavedData, which stores all pre-generated results included in this package. The second script, runPlotData.m, reads files from the SavedData folder and processes them to regenerate the figures presented in the manuscript. The third script, runPlotSuppFigS1.m, specifically generates Figure S1 in the Supplementary Material of the manuscript. The package also includes the aforementioned files in non-proprietary .txt format. If users intend to use them, they should first save the files in their respective .m or .mat formats prior to execution in MATLAB.

Biomass concentration↗

Universal framework for simultaneous tomography of quantum states and SPAM noise

We present a general denoising algorithm for performing simultaneous tomography of quantum states and measurement noise. This algorithm allows us to fully characterize state preparation and measurement (SPAM) errors present in any quantum system. Our method is based on the analysis of the properties of the linear operator space induced by unitary operations. Given any quantum system with a noisy measurement apparatus, our method can output the quantum state and the noise matrix of the detector up to a single gauge degree of freedom. We show that this gauge freedom is unavoidable in the general case, but this degeneracy can be generally broken using prior knowledge on the state or noise properties, thus fixing the gauge for several types of state-noise combinations with no assumptions about noise strength. Such combinations include pure quantum states with arbitrarily correlated errors, and arbitrary states with block independent errors. This framework can further use available prior information about the setting to systematically reduce the number of observations and measurements required for state and noise detection. Our method effectively generalizes existing approaches to the problem, and includes as special cases common settings considered in the literature requiring an uncorrelated or invertible noise matrix, or specific probe states.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Mechanism of O 2 /NO-Promoted Oxidative C–C Bond Cleavage in Linear Alkanes

Selective oxidation of alkanes to oxygenated products remains a fundamental challenge, particularly if the goal is to promote C−C bond cleavage while minimizing formation of CO 2 . Oxidation conditions that use O 2 and nitrogen oxides (NO x ) have been shown to be very effective in promoting radical-mediated functionalization and oxidative carbon−carbon cleavage in saturated hydrocarbon polymers, such as polyethylene. Here, we investigate the mechanism of O 2 /NO-mediated oxidation of ndecane as a prototypical linear alkane substrate. These reactions enable identification and quantification of reactive intermediates, including nitrites, nitrates, alcohols, and ketones. Under the reaction conditions, these species convert into common ketone intermediates that evolve into α-diketones and other α-functionalized ketones, which undergo further conversion into carboxylic acids. Infrared (IR) spectroscopy indicates that HDPE oxidation proceeds through similar key intermediates. Together, these findings establish a mechanistic framework for NO x -mediated alkane oxidation and provide a foundation for the development of broadly applicable oxidative transformations.

Anions↗

Adaptive optical correction for in vivo two-photon fluorescence microscopy with neural fields

Adaptive optics restore ideal imaging performance in complex samples by measuring and correcting optical aberrations but often require custom-built microscopes with carefully aligned wavefront sensing/shaping devices and can be susceptible to sample motion. Here we describe NeAT, a computational framework using neural fields for adaptive optics two-photon fluorescence microscopy. NeAT estimates wavefront aberration and recovers sample structure from a 3D image stack without requiring external datasets for training. Incorporating motion correction in learning and correcting conjugation errors commonly found in commercial microscopes, NeAT is designed for deployment in biological laboratories for in vivo imaging. We validate NeAT’s performance using a custom-built microscope with a wavefront sensor under varying signal-to-noise ratios, aberration and motion conditions. With a commercial microscope, we demonstrate real-time aberration correction for in vivo morphological and functional imaging in the living mouse brain, with NeAT improving the signal and accuracy of glutamate and calcium imaging of synapses and neurons.

Kang, Iksung↗

Neutrino flavor instabilities in neutron star mergers with moment transport: Slow, fast, and collisional modes

Determining where, when, and how neutrino flavor oscillations must be included in large-scale simulations of hot and dense astrophysical environments is an enduring challenge that must be tackled to obtain accurate predictions. Here, using an angular moment-based linear stability analysis framework, we examine the different kinds of flavor instabilities that can take place in the context of the postprocessing of a neutron star merger simulation, with a particular focus on the collisional flavor instability and a careful assessment of several commonly used approximations. First, neglecting anisotropies of the neutrino field, we investigate the extent to which commonly used monoenergetic growth rates reproduce the results obtained from a full multienergy treatment. Contrary to the large discrepancies found in core-collapse supernova environments, we propose a simple combination of energy-averaged estimates that reproduces the multienergy growth rates in our representative simulation snapshot. We then quantify the impact of additional physical effects, including nuclear many-body corrections, scattering opacities, and the inclusion of the vacuum term in the neutrino Hamiltonian. Finally, we include the neutrino distribution anisotropies, which allows us to explore, for the first time in a multienergy setting, the interplay between collisional, fast, and slow modes in a moment-based neutron star merger simulation. We find that, despite a dominance of the fast instability in most of the simulation volume, certain regions exhibit only a collisional instability, while others, especially at large distances, exhibit a slow instability that is largely underestimated if anisotropic effects are neglected.

neutrino oscillations↗

DOE FAIR Surrogate Benchmarks Supporting AI and Simulation Research (SBI Surrogate Benchmark Initiative) (Final Report)

Computational Science is being revolutionized by integrating AI and simulation and, in particular, by deep learning surrogate models that can replace all or part of traditional large‐scale HPC computations. Such surrogates can achieve remarkable performance improvements, as much as several orders of magnitude, and save both compute time and energy. The Surrogate Benchmark Initiative (SBI) project creates a community repository and FAIR (Findable, Accessible, Interoperable, and Reusable) data ecosystem for HPC application surrogate benchmarks. The SBI team comes from Argonne National Laboratory (ANL), Indiana University (IU), Rutgers University, the University of Tennessee, Knoxville (UTK), and the University of Virginia(UVA). SBI repositories include data, code, and all relevant collateral artifacts, that the science and engineering community needs to use and reuse these data sets and surrogates. SBI repositories generate active research from both participants in SBI and the broader AI and domain science communities. This project develops surrogates that use several different neural nets to learn and quickly infer the results of simulations and data systems and capture them as surrogate benchmarks with a rich set of metadata, covering. Data; Model; Metrics specification; Machine specification; Science, Speed, Power Results, We research FAIR metadata for these benchmarks. We develop application surrogate examples as benchmarks across many fields (ANL, UTK, IU, UVA). We also study non Surrogate benchmarks that have many common features and similar issues regarding FAIRness. We work with MLCommons (UVA, UTK), which is a major machine learning benchmarking activity where we get metadata ontologies, software, and benchmarks, benchmarks have datasets, models, and metadata, and they need a technical framework developed by UTK and Rutgers and deployed by UVA. We study features of Surrogates, including performance, training set size, and uncertainty quantification (Rutgers, UVA and IU).

97 MATHEMATICS AND COMPUTING↗

Automation of Vulnerability and Patch Management: Information Extraction, Association, and Optimization

Vulnerability and patch management is an integral part of a robust cybersecurity program, yet it grows increasingly complex due to the sheer amount of data that must be analyzed. Particularly in Operational Technology (OT) environments, analysis must be done manually because of the lack of automated solutions. Additionally, there are many steps in this process, from the initial discovery of the vulnerability to the implementation of its remediation, and each step in the process requires different data in order to be performed effectively. In this work, we provide approaches and strategies to assist operators in industrial or OT environments throughout the vulnerability management cycle. Security advisories provide key information about mitigation strategies, or actions that can be taken when a patch is unavailable or cannot be installed. Details of these strategies are not shared in public vulnerability databases and must be found manually. We approach this problem by designing a solution to automatically identify that information within vendor security advisories and retrieve it for operator use. We start with an approach that requires domain-specific knowledge of certain frequently-seen reference websites. Next, an approach that can work on an arbitrary website but relies on certain keywords. Finally, an approach that uses Natural Language Processing (NLP) methods and does not require specific knowledge or keywords. Each of these approaches is more general than its predecessor; we demonstrate high accuracy for all approaches Advisories also often contain details of affected products in non-standard or natural language formats. While this information can be easily understood when read by an operator, the non-standard format acts as a barrier to effective automation. We provide an approach for the first step in this process: identifying vendors in security advisories and mapping them to a standard framework for representing digital assets and software products. We evaluate five established string similarity algorithms, plus one of our own design that combines string similarity and information theory, on the task of mapping vendors to their corresponding entries in the Common Platform Enumeration (CPE) repository. Our results show that our proposed metric outperforms all others. Due to the constraints on time, finances, and personnel for organizations, Large Language Models (LLMs) may seem like attractive opportunities for security operators to speed up information gathering; however, it is still not clear whether LLMs can handle vulnerability management tasks well. To answer this question, we perform an empirical study of LLMs’ ability to provide consistent, accurate information about vulnerabilities in order to guide organizations in their adoption of LLMs. We observe poor performance for all models tested, suggesting that these models are not well-suited to the consistent retrieval of accurate vulnerability information. Finally, once vulnerabilities have been identified and any additional information has been obtained, operators must decide which remediation actions to implement based on their available resources. This already-complex problem becomes even more so when we consider that a vulnerability may have multiple avenues for remediation. We formulate this scenario as two knapsack problems and provide solutions, which we then compare against several existing strategies for vulnerability prioritization seen in real operational environments.

McClanahan, Kylie↗

CONFLUX: A standardized framework to calculate reactor antineutrino flux

Nuclear fission reactors are abundant sources of antineutrinos for neutrino physics experiments. The flux and spectrum of antineutrinos emitted by a reactor can indicate its activity and composition, suggesting potential applications of neutrino measurements beyond fundamental scientific studies that may be valuable to society. The utility of reactor antineutrinos for applications and fundamental science is dependent on the availability of precise predictions of these emissions. For example, in the last decade, disagreements between reactor antineutrino measurements and models have inspired revision of reactor antineutrino calculations and standard nuclear databases as well as searches for new fundamental particles not predicted by the Standard Model of particle physics. Past predictions and descriptions of the methods used to generate them are documented to varying degrees in the literature, with different modeling teams incorporating a range of methods, input data, and assumptions. The resulting difficulty in accessing or reproducing past models and reconciling results from differing approaches complicates the future study and application of reactor antineutrinos. The CONFLUX (Calculation Of Neutrino FLUX) software framework is a neutrino prediction tool built with the goal of simplifying, standardizing, and democratizing the process of reactor antineutrino flux calculations. CONFLUX includes three primary methods for calculating the antineutrino emissions of nuclear reactors or individual beta decays that incorporate common nuclear data and beta decay theory. The software is prepackaged with the current nuclear databases, including ENDF.B/VIII, JEFF-3.3, and ENSDF, and it includes the capability to predict time-dependent reactor emissions, adjust nuclear database or beta decay inputs/assumptions, and propagate related sources of uncertainty. Here, this paper describes the CONFLUX software structure, details the methods used for flux and spectrum calculations, and provides examples of potential use cases.

Zhang, Xianyi [Lawrence Livermore National Laborat↗

Mg vacancy and impurity-limited MgO single crystal thermal conductivity

Magnesium oxide (MgO) exhibits one of the highest thermal conductivities among oxides and is widely used as a dielectric material and substrate in semiconductor devices, in refractory applications, and as a promising filler in thermal interface materials for electronics. Its high thermal conductivity may be sensitive to impurity and defects, yet this influence is still uncertain. Here, in this study, the impact of the common impurities, i.e., Al, Ca, Ti, V, Fe, Si, B, Nb, Zr, Na, and K, as well as Mg and O vacancies on phonon scattering and thermal conductivity of MgO is studied using a fully first-principles T-matrix framework. It is found that B, Nb, and Zr impurities, along with Mg vacancies, lead to exceptionally strong reductions in thermal conductivity. By contrast, O vacancies and other impurities have modest to minimal impacts. Leveraging the T-matrix results, we reassess the perturbative, mass-only formalism whose use is pervasive in the literature and show that neglecting bond disorder does not necessarily lead to underestimation: for all transition-metal impurities studied, bond perturbations partially cancel mass disorder, causing the traditional perturbative model to overestimate scattering. We propose a simple modified perturbative expression that incorporates both mass and bond disorder and closely reproduces the T-matrix trends. Our predicted low-temperature trends by including phonon-impurity and phonon-boundary scattering match reasonably well with experiments. This work provides an in-depth study of impurity- and vacancy-limited thermal conductivity of MgO and suggests that reported “high-purity” MgO values have likely not yet reached the intrinsic upper limit, which may be substantially higher.

36 - MATERIALS SCIENCE↗

FAIR Surrogate Benchmarks Supporting AI and Simulation Research (Final Report)

Computational Science is being revolutionized by integrating AI and simulation and, in particular, by deep learning surrogate models that can replace all or part of traditional large‐scale HPC computations. Such surrogates can achieve remarkable performance improvements, as much as several orders of magnitude, and save both compute time and energy. The Surrogate Benchmark Initiative (SBI) project creates a community repository and FAIR (Findable, Accessible, Interoperable, and Reusable) data ecosystem for HPC application surrogate benchmarks. The SBI team comes from Argonne National Laboratory (ANL), Indiana University (IU), Rutgers University, the University of Tennessee, Knoxville (UTK), and the University of Virginia (UVA). SBI repositories include data, code, and all relevant collateral artifacts that the science and engineering community need to use and reuse these data sets and surrogates. SBI repositories generate active research from both the participants in SBI and the broad community of AI and domain scientists. This project develops surrogates that use several different neural nets to learn and quickly infer the results of simulations and data systems and captures them as surrogate benchmarks with a rich set of metadata covering: Data; Model; Metrics specification; Machine specification; and Science, Speed, and Power Results. We research FAIR metadata for these benchmarks. We develop application surrogate examples as benchmarks across many fields (ANL, UTK, IU, UVA). We also study non-Surrogate benchmarks that have many common features and similar issues as regards FAIRness. We work with MLCommons (UVA, UTK), which is a major machine learning benchmarking activity where we get metadata ontologies, software, and benchmarks, Benchmarks have datasets, models, and metadata and they need a technical framework developed by UTK and Rutgers and deployed by UVA. We study features of Surrogates including performance, training set size, and uncertainty quantification (Rutgers, UVA and IU).

97 MATHEMATICS AND COMPUTING↗

Mining for Metal–Organic Systems: Chemistry Frontiers of Th-, U-, and Zr-Materials

The conceptual framework presented in this Perspective overviews the design principles of innovative thorium-based materials that could address urgent needs of the medicinal, nuclear energy, and waste remediation sectors from the lens of zirconium and uranium analogs. We survey the intersections of Zr, Th, and U chemistry with a focus on how the intrinsic behavior of each metal translates to broader material properties, including, but not limited to, structural and topological diversity, preferential metal–ligand binding, and reactivity. On the example of several classes of materials, including organometallic complexes, polyoxometalates, and the primary focus of this Perspective, metal–organic frameworks (MOFs), the design principles that govern the preparation of Zr-, Th-, and U-compounds, including oxophilicity, variation in oxidation states, and stable coordination environments have been considered. Further, we highlight how the impact of the mentioned variables may shift throughout the progression from discrete molecular systems to extended structures. We discuss the common assumption that zirconium-organic materials are typically considered a close analog of thorium-based congeners in areas such as material design and preparation. Through consideration of fundamental chemistry principles, we shed light on the relationships between Zr-, Th-, and U-based materials and highlight how a critical analysis of their distinct properties can be used to target a desired material performance. Finally, we provide a detailed understanding of Th-based materials chemistry by anchoring their fundamental properties between two well-studied reference points, zirconium- and uranium-containing analogs.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

ClassNMSW- a real-time classification approach for non-recycled municipal solid waste using hyperspectral imaging

Real-time classification of non-recycled municipal solid waste (NMSW) is essential for efficient valorization. This study introduces ClassNMSW, a comprehensive framework for classifying 22 NMSW subclasses under industrial constraints by using hyperspectral imaging (HSI). A primary innovation of this work is the development of a variance-controlled spectral extraction algorithm. Unlike traditional methods that rely on simple averaging, this approach systematically investigates the extent of pixel extraction to minimize the loss of critical chemical information while maximizing data reduction thus ensuring high spectral fidelity with low computational cost. The approach developed in this work integrates automated, computer-vision-based background removal, eliminating the need for the manual thresholding common in current literature. To resolve ambiguities among chemically similar subclasses, a tiered classification and multi-camera fusion strategy (NIR17 and NIR22) is implemented. Results demonstrate that ClassNMSW achieves an object-wise weighted accuracy of 98.70% for single-sensor configurations and 100% under sensor fusion. A novel rolling-window strategy satisfies desired end-to-end latency of <2 s, satisfying the strict deterministic requirements of high-speed industrial sorting environments. The ClassNMSW framework provides a scalable foundation for advancing circularity and resource recovery in large-scale waste valorization operations.

99 - GENERAL AND MISCELLANEOUS↗

Nodeman: A Node Management Tool For Hpc Clusters

NodeMan is a command line tool to manage nodes in an HPC cluster. At it's core, it is an extensible framework composed of bash scripting and GNU parallel. HPC System Administrator will find it useful in that it encapsulates desired functions and allows them to be assembled in a way familiar to administrators - through pipes. In fact, NodeMan functions can work with common command line tools as long as they use stdin/stdout. System Administrators can construct moderately complex logic and filtering on a compact command line that would normally require a substantial shell script. In the spirit of clush and pdsh, it is able to run commands remotely on nodes. Additionally, NodeMan is more flexible. For example, it can interact with IPMI and naturally processes node lists for orchestrating different tools. The library of useful pre-built functions is growing. System administrators can easily create new functions and make it their own.

Serr, ScottM↗

Women in Nuclear Security: The United Arab Emirates Story

Historically in the United Arab Emirates, security personnel stationed at nuclear sites were typically men with military or policing backgrounds. Although these groups bring important, transferrable skills and experience to the nuclear security mission, some Emirati stakeholders recognize the increased benefit of a greater diversity of voices and backgrounds within the nuclear security team hierarchy. By recruiting from a wider range of disciplines and encouraging gender diversity, the security team could develop broader insight into security challenges, and a greater variety of perspectives could bring forth more innovative solutions. When aligning a diverse, multinational workforce with a common mission, all members of a security organization must fundamentally change their individual approaches to the work; the entire staff must be on the same page. This can be particularly challenging for some people who find it difficult to transition from a framework of beliefs because of past experiences based on prescriptive, exclusionary, hierarchical thought. Also, national cultural values are learned early, held deeply, and change slowly over the course of generations, and inherent biases rooted in these values can create barriers to the collective progress of a team. Creating a more open, diverse, and transparent culture for the team by accepting a unique, egalitarian, professional nuclear security culture that all members adhere to can be a challenging task for some. In this article, I share my experiences related to professional roadblocks, personal challenges, and lessons learned, particularly those regarding the empowerment of women in science and engineering, while developing and establishing the United Arab Emirates nuclear security program. I also discuss the support (and underrepresentation) of women, including those currently in critical leadership positions, and their careers in nuclear security.

Barakah Nuclear Power Plant↗

SO(3)-invariant PCA with application to molecular data

Principal component analysis (PCA) is a fundamental technique for dimensionality reduction and denoising; however, its application to three-dimensional data with arbitrary orientations -- common in structural biology -- presents significant challenges. A naive approach requires augmenting the dataset with many rotated copies of each sample, incurring prohibitive computational costs. In this paper, we extend PCA to 3D volumetric datasets with unknown orientations by developing an efficient and principled framework for SO(3)-invariant PCA that implicitly accounts for all rotations without explicit data augmentation. By exploiting underlying algebraic structure, we demonstrate that the computation involves only the square root of the total number of covariance entries, resulting in a substantial reduction in complexity. We validate the method on real-world molecular datasets, demonstrating its effectiveness and opening up new possibilities for large-scale, high-dimensional reconstruction problems.

Fraiman, Michael [Tel Aviv Univ., Tel Aviv (Israel↗

An Evaluation of Representation Learning Methods in Particle Physics Foundation Models

We present a systematic evaluation of representation learning objectives for particle physics within a unified framework. Our study employs a shared transformer-based particle-cloud encoder with standardized preprocessing, matched sampling, and a consistent evaluation protocol on a jet classification dataset. We compare contrastive (supervised and self-supervised), masked particle modeling, and generative reconstruction objectives under a common training regimen. In addition, we introduce targeted supervised architectural modifications that achieve state-of-the-art performance on benchmark evaluations. This controlled comparison isolates the contributions of the learning objective, highlights their respective strengths and limitations, and provides reproducible baselines. We position this work as a reference point for the future development of foundation models in particle physics, enabling more transparent and robust progress across the community.

Chen, Michael [Caltech]↗