Search NASA⌕ Search

SEARCH · Search NASA

Results for “federated machine learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

OLCF’s Advanced Computing Ecosystem (ACE): FY25 Update for Ongoing Efforts

The advent of widespread use of artificial intelligence (AI) and machine learning (ML) models in science, coupled with fast data production rates of scientific instruments strain the traditional batch-oriented high-performance computing (HPC) environment. As scientific exploration continues to require more data and faster processing and analysis, new emerging technologies and capabilities to enable cross-facility and time-sensitive workflows are required for seamless integration of HPC and experimental facilities. The Advanced Computing Ecosystem (ACE) is a strategic initiative within the Oak Ridge Leadership Computing Facility (OLCF) established in 2024 to support the development of cutting-edge technologies to advance computational research and infrastructure at OLCF and across the Department of Energy (DOE). Several DOE initiatives are spearheading the evolution of the scientific landscape by blurring facility boundaries and connecting the user facilities to advance scientific capabilities and ensure energy dominance. The DOE Integrated Research Infrastructure (IRI) program is one example that is laying a foundation to support complex cross-facility workflows. The IRI program aims to integrate diverse computational resources, data infrastructures, and scientific instruments to facilitate collaboration and accelerate scientific discovery. The Interconnected Science Ecosystem (INTERSECT) initiative at Oak Ridge National Laboratory (ORNL) is another example that aims to revolutionize scientific research through AI-driven, interconnected autonomous laboratories and research facilities. Finally, the American Science Cloud (AmSC), recently announced in the “One Big Beautiful Bill”, aims to leverage prior infrastructure efforts of the IRI and automation and AI efforts of INTERSECT (and others) to build a federated, AI-augmented AmSC platform to unify the DOE’s computing, experimental, and data resources to catalyze scientific innovation.

97 MATHEMATICS AND COMPUTING↗

Apomixis in Farmers’ Fields: Overview, Case Studies from Forage Grasses and Considerations for Future Apomictic Crops

Apomixis occurs naturally in several commercially important species from diverse plant families. While in some of these species apomixis is yet to be exploited in breeding schemes aimed at fixing heterosis, genetic progress and cultivar development, in other species apomixis has been integrated at different stages of breeding. Some of the most relevant examples come from the subfamily Panicoideae, the second largest subfamily of the Poaceae, and are the main focus of this review. The subfamily encompasses many tropical and sub-tropical grasses and grains of worldwide economic importance. Apomictic tropical forages are prime examples of how apomixis can be used and exploited in the development of marketable cultivars, which are essential to the meat and milk production industries globally. The main commercial forages used as grass pastures covering millions of hectares in tropical and sub-tropical regions are polyploids exhibiting gametophytic apomixis that belong to the genus Urochloa spp. (brachiariagrasses) and to the species Megathyrsus maximus (guineagrass). Buffel grass (Cenchrus ciliaris) and Paspalum spp. are other important apomictic forages bred and used in these regions. Breeding involves large germplasm collections from the centers of origin of the species, and for most of them, sexually reproducing diploid plants have been found. Chromosomically duplicated plants that maintain sexual reproduction are used in crosses with apomictic genotypes for the development and selection of cultivars to be marketed or used as progenitors in subsequent breeding cycles. The peculiarities of each genus/species breeding programs, the cultivars obtained from these programs, and the impact of use of marker assisted selection in cultivar development are presented. In addition, the test or implementation of new technologies such as high throughput phenotyping, and the use of machine learning methods for trait prediction and genomic selection are positively impacting the selection and speed of development of new polyploid apomictic cultivars. Furthermore, genetic transformation techniques, including genome editing, provide an additional layer for design of tailor-made, customer-oriented cultivars.

Cenchrus↗

Utilizing AI and Spatial Data to Identify & Rapidly Disseminate Energy Infrastructure Insights

GeoGov Summit Final Presentation entitled "Utilizing AI and Spatial Data to Identify & Rapidly Disseminate Energy Infrastructure Insights". Maintaining the integrity of energy infrastructure plays a critical role in ensuring energy security. Robust foundational AI models using data from federal, state, industry, and other sources can help address integrity risk management & mitigation issues as well as evaluate extended use strategies. Trusted foundational models can help with industry adoption and accelerate innovation by enhancing integrity predictions, reduce costs, and informing infrastructure build-out. Coordination, collaboration & data sharing to develop robust models to aid in: Optimizing operations; Minimizing costs; Ensuring energy security.

Advanced Infrastructure Integrity Model (AIIM)↗

Earth Independent Medical Operations (EIMO) Datascope: Challenges and Potential Solutions

Data flows and storage/retrieval capacity are severely constrained during missions in space and challenges will become even greater during exploration class missions. There is a need for an artificial intelligence (AI)-based clinical decision support system (CDSS) to monitor and analyze data to provide real-time consultative support for crew medical officer (CMO) decision-making. EIMO is defined as the gradual transition of medical care and decision making from terrestrial to space-based assets, enabling support of astronaut health and performance and reducing overall mission risk. While a hallmark of this paradigm shift from low-earth orbit is that on-board care will increasingly become the responsibility of the astronauts for primary management and decision making, terrestrial assets will continue to be paramount in pre-mission screening and planning, as well as prevention, health maintenance and long-term care contingencies. New capabilities and systems that enable progressively more robust and resilient systems and crews will be necessary to reduce risk and increase probability of deep space exploration mission success. An aspiration for EIMO is to develop AI-enhanced solutions for analysis of crew health & performance data and to facilitate clinical decision support for autonomous medical operations. A “system of systems” approach is envisioned whereby EIMO will deploy AI-supported natural language processing and machine learning (ML) techniques to utilize embedded reference databases and real-time data streams [input vectors] from multiple data sources. Constituent input vectors may include environmental controls, countermeasures data, behavioral data, physiologic wearables, point-of-care laboratory tests, personalized medical records, inventory trade space risk assessments, COTS medical databases, and ground support inputs. An ideal AI capability would possess trained fusion algorithms to cross reference input vectors with medical ‘knowledge’ [cultivated database] to stratify relevant data streams for predictive and actionable capabilities. In addition, EIMO will feature mobility, in that it can be accessed and can push/pull data within and between multiple vehicles/habitats. Large amounts and variable sources of data can be leveraged to diagnose, inform treatment strategies, and potentially predict medical events and performance decrements. Inclusion of advanced training tools using extended reality will enable increasingly autonomous medical care to aid a CMO when ground support is unavailable or time-delayed beyond required action window, e.g., emergent medical situations. EIMO CDSS would require very large datasets to train pre-flight and significant amounts of data are needed to support ML via in-flight CDSS operations. An additional challenge will be to find sufficient data to train a model relevant to astronaut demographics. The rapid, accelerating evolution of this field creates a propitious solution space to leverage multi-modal AI through public-private partnership(s). The status of multi-modal AI systems today would preclude their use for long duration missions as they remain unreliable and are subject to “digital hallucinations” and other errors that could pose operational risk. A federated labs structure is being considered to test and optimize data flow from the multiple input vectors leading to field testing in suitable ground/flight analogs. Critical to the success of an EIMO CDSS will be integration and interoperability and success will be defined by a system that can serve as an in-flight medical consult for the CMO providing critical support during medical contingencies. Benefits to terrestrial medicine may be significant as an outflow of the EIMO medical system, particularly for remote areas and communities lacking significant infrastructure, personnel and resources.

J Lemery↗

Earth Independent Medical Operations (EIMO) Datascope: Challenges and Potential Solutions

Data flows and storage/retrieval capacity are severely constrained during missions in space and challenges will become even greater during exploration class missions. There is a need for an artificial intelligence (AI)-based clinical decision support system (CDSS) to monitor and analyze data to provide real-time consultative support for crew medical officer (CMO) decision-making. EIMO is defined as the gradual transition of medical care and decision making from terrestrial to space-based assets, enabling support of astronaut health and performance and reducing overall mission risk. While a hallmark of this paradigm shift from low-earth orbit is that on-board care will increasingly become the responsibility of the astronauts for primary management and decision making, terrestrial assets will continue to be paramount in pre-mission screening and planning, as well as prevention, health maintenance and long-term care contingencies. New capabilities and systems that enable progressively more robust and resilient systems and crews will be necessary to reduce risk and increase probability of deep space exploration mission success. An aspiration for EIMO is to develop AI-enhanced solutions for analysis of crew health & performance data and to facilitate clinical decision support for autonomous medical operations. A “system of systems” approach is envisioned whereby EIMO will deploy AI-supported natural language processing and machine learning (ML) techniques to utilize embedded reference databases and real-time data streams [input vectors] from multiple data sources. Constituent input vectors may include environmental controls, countermeasures data, behavioral data, physiologic wearables, point-of-care laboratory tests, personalized medical records, inventory trade space risk assessments, COTS medical databases, and ground support inputs. An ideal AI capability would possess trained fusion algorithms to cross reference input vectors with medical ‘knowledge’ [cultivated database] to stratify relevant data streams for predictive and actionable capabilities. In addition, EIMO will feature mobility, in that it can be accessed and can push/pull data within and between multiple vehicles/habitats. Large amounts and variable sources of data can be leveraged to diagnose, inform treatment strategies, and potentially predict medical events and performance decrements. Inclusion of advanced training tools using extended reality will enable increasingly autonomous medical care to aid a CMO when ground support is unavailable or time-delayed beyond required action window, e.g., emergent medical situations. EIMO CDSS would require very large datasets to train pre-flight and significant amounts of data are needed to support ML via in-flight CDSS operations. An additional challenge will be to find sufficient data to train a model relevant to astronaut demographics. The rapid, accelerating evolution of this field creates a propitious solution space to leverage multi-modal AI through public-private partnership(s). The status of multi-modal AI systems today would preclude their use for long duration missions as they remain unreliable and are subject to “digital hallucinations” and other errors that could pose operational risk. A federated labs structure is being considered to test and optimize data flow from the multiple input vectors leading to field testing in suitable ground/flight analogs. Critical to the success of an EIMO CDSS will be integration and interoperability and success will be defined by a system that can serve as an in-flight medical consult for the CMO providing critical support during medical contingencies. Benefits to terrestrial medicine may be significant as an outflow of the EIMO medical system, particularly for remote areas and communities lacking significant infrastructure, personnel and resources.

Medical Operations↗

Towards Sustainable Aviation With Efficient Airspace Operations

In November 2021, the Federal Aviation Administration published the United States Aviation Climate Action Plan to accelerate innovation across the U.S. aviation ecosystem. In response, NASA has established the Sustainable Flight National Partnership to engage with industry, academia, and other agencies to accomplish net-zero carbon emissions by 2050. As part of the SFNP Mission, NASA is conducting a series of real world operational demonstrations in the current National Airspace System with a focus on delivering real world sustainability benefits. This paper presents the current work and future plan for the SFNP Ops Demo Series and describes the cloud based infrastructure used for virtual deployment of decision support tools to flight operators and Air Traffic Controllers to help improve operational efficiency of the National Airspace System. Validation results are shared along with the key sustainability benefits such as jet fuel savings and reduction in CO2 emissions. The operational efficiency metrics such as delay savings are also reported.

Sustainable Flight National Partnership↗

Reconstruction of atmospheric neutrinos in DUNE’s horizontal-drift far-detector module

This paper reports on the capabilities in reconstructing and identifying atmospheric neutrino interactions in one of the Deep Underground Neutrino Experiment’s (DUNE) far detector modules, a liquid argon time projection chamber (LArTPC) with horizontal drift (FD-HD) of ionization electrons. The reconstruction is based upon the workflow developed for DUNE’s long-baseline oscillation analysis, with some necessary machine-learning models’ retraining and the addition of features relevant only to atmospheric neutrinos such as the neutrino direction reconstruction. Where relevant, the impact of the detection of the charged particles of the hadronic system is emphasized, and comparisons are carried out between the case when lepton-only information is considered in the reconstruction (as is the case for many neutrino oscillation experiments), versus when all particles identified in the LArTPC were included. Three neutrino direction reconstruction methods have been developed and studied for the atmospheric analyses: using lepton-only information, using all reconstructed particles, and using only correlations from reconstructed hits. The results indicate that incorporating more than just lepton information significantly improves the resolution of both neutrino direction and energy reconstruction. The angle reconstruction algorithms developed in this work result in no strong dependence on particle direction for reconstruction efficiencies or neutrino flavor identification. This comprehensive review of the reconstruction of atmospheric neutrinos in DUNE’s FD-HD LArTPC is the first step towards developing a first neutrino oscillation sensitivity analysis, which will ready DUNE for its first measurements.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Integrated System Planning: Emerging Software Requirements in the Power Industry

Power system planning software remains fragmented across organizational boundaries, with specialized tools for capacity expansion, production cost modeling, power flow, and dynamic analysis operating on incompatible data models and assumptions. This article argues that the fragmentation is not merely a technical problem but a predictable consequence of Conway's law: software architectures mirror the departmental structures within which they are developed. Regulatory milestones like Federal Energy Regulatory Commission (FERC) Order 888 formalized these divisions, but the roots trace back to the distinct engineering disciplines-mechanical, chemical, and electrical-that staffed generation and transmission planning departments in vertically integrated utilities. As the industry moves toward integrated system planning (ISP) that coordinates generation, transmission, and distribution investment decisions, the software ecosystem must evolve accordingly. We identify five categories of software requirements to enable this transition: coherent data inputs decoupled from individual applications, unified and extensible data schemas, modular component representations that support multiple abstraction levels, lifecycle management of planning datasets, and well-defined application programming interface (API) contracts that separate data exchange from algorithmic control. We examine how these requirements interact with three common workflow patterns-serial gate clearing, sequential multiapplication, and convergence oriented-and discuss the interface design principles each demands. We then outline a vision for platform-based planning architectures where specialized analytical services compose through standardized interfaces and where artificial intelligence (AI)/machine learning (ML) tools augment decision support within a disciplined software infrastructure. The practices proposed here offer a path from today's siloed tool collections toward collaborative planning ecosystems capable of handling the complexity of modern power system transformation.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Real World Applications of AI/ML in Optimizing Airspace Operations

As National Airspace System (NAS) is going through the Digital Transformation journey, data science and analytics methods can significantly contribute to improving the traditional physics-based decision-making tools. The adoption of AI/ML methods will not only help accelerate Federal Aviation Administration (FAA)’s vision of Info-centric NAS but also contribute to the overall objective of sustainable aviation. AI/ML can improve the ground and airspace operations by enhancing the accuracy of current decision-making tools used by the airlines and the FAA to manage traffic on the ground and in the air. Huge amount of data that gets collected during a flight. AI/ML methods can extract information from this data and provide valuable insights to make better operational decisions. NASA has partnered with the FAA and commercial airlines such as American and Southwest Airlines on this effort and has successfully demonstrated the benefits of using ML in real world environment by reducing delays and optimizing ground operations at the US airports. In 2022 itself, NASA demonstrated real-world benefits (over 24K lbs. of fuel savings, over 76.6K lbs. CO2 emission savings, and several hours of delay savings) by deploying ML based prediction models to optimize ground operations at Dallas/Fort Worth International and Dallas Love Field Airports in Texas. These tools are being deployed on the cloud for broader deployment, adaptability, and scalability. NASA is developed a reference implementation of the cloud-based platform to significantly lower the bar to development and distribution of these digital services for aviation. In this talk, I will share information about the Digital Information Platform project, the novel AI/ML based approaches used for optimizing ground operations and the opportunities to partner with NASA on these demonstrations.

air traffic management↗

2022 Spring Internship Exit Presentation

As efforts of the National Aeronautics and Space Administration (NASA) and the Federal Aviation Administration (FAA) continue to digitize the air traffic management (ATM) domain, there is countless times of need for downstream natural language processing (NLP) tasks such as named entity recognition, text summarization, classification, and more. Although there are a plethora of open-sourced pre-trained transformer models in the NLP field such as BERT, RoBERTa, XLNet, and GPT-3, these models are trained on general corpora and perform poorly on domain-specific terminology and phraseology seen in ATM documents such as Notice to Airmen (NOTAMs) and Letters of Agreement (LoA). Our proposed research objective will be to first gather a large corpus of air traffic management related documents, orders, notices, books, technical papers, conference papers, articles, and other miscellaneous sources of text data from the FAA, NASA, and accredited conference and publication societies. After gathering this data, many steps will have to be taken to collate and preprocess the data into a format understandable by our test transformer models. Thirdly, we will set up training pipelines to train the RoBERTa model on its unsupervised training task masked language modelling (MLM) using resources provided by the NASA Advanced Supercomputing (NAS) facilities. Finally, these fine-tuned transformer models will be evaluated on their performance on down-stream NLP tasks as mentioned above, to show whether they will be effective when working with ATM related data or not. Once complete, this model could be made open-sourced on the HuggingFace website, where the rest of the ATM community can access and utilize this tool.

NLP↗

Forced Component Estimation Statistical Method Intercomparison Project (ForceSMIP)

Anthropogenic climate change is unfolding rapidly, yet its regional manifestation can be obscured by internal variability. A primary goal of climate science is to identify the externally forced climate response from among the noise of internal variability. Separating the forced response from internal variability can be addressed in climate models by using a large ensemble to average over different possible realizations of internal variability. However, with only one realization of the real world, it is a major challenge to isolate the forced response directly in observations. In the Forced Component Estimation Statistical Method Intercomparison Project (ForceSMIP), contributors used existing and newly developed statistical and machine learning methods to estimate the forced response over 1950–2022 within individual realizations of the climate system. Participants used neural networks, linear inverse models, fingerprinting methods, and low-frequency component analysis, among other approaches. These methods were trained using large ensembles from multiple climate models and then applied to observations. Here, we evaluate method performance within large ensembles and investigate the estimates of the forced response in observations. Our results show that many different types of methods are skillful for estimating the forced response in climate models, though the relative skill of individual methods varies depending on the variable and evaluation metric. Methods with comparable skill in models can give a wide range of estimates of the forced response pattern in observations, illustrating the epistemic uncertainty in forced response estimates. ForceSMIP gives new insights into the forced response in observations, its uncertainty, and methods for its estimation.

Climate attribution↗

Precision calibration of calorimeter signals in the ATLAS experiment using an uncertainty-aware neural network

The ATLAS experiment at the Large Hadron Collider explores the use of modern neural networks for a multi-dimensional calibration of its calorimeter signal defined by clusters of topologically connected cells (topo-clusters). The Bayesian neural network (BNN) approach not only yields a continuous and smooth calibration function that improves performance relative to the standard calibration but also provides uncertainties on the calibrated energies for each topo-cluster. The results obtained by using a trained BNN are compared to the standard local hadronic calibration and to a calibration provided by training a deep neural network. The uncertainties predicted by the BNN are interpreted in the context of a fractional contribution to the systematic uncertainties of the trained calibration. They are also compared to uncertainty predictions obtained from an alternative estimator employing repulsive ensembles.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

NEPATEC2.0: NEPA Text Corpus v2.0

The National Environmental Policy Act of 1969, as amended (NEPA), is a major environmental law in the United States, requiring Federal agencies to consider and document potential environmental impacts before deciding on a proposed action. Modernization of NEPA and permitting processes faces significant challenges due to the lack of standardized formats and interoperable systems for organizing and sharing NEPA-related information across agencies. Much of the information gathered during NEPA reviews is written into documents such as categorical exclusions, environmental assessments, and environmental impact statements, then filed in predominately independent agency file stores that may or may not be publicly accessible. The application of metadata and data standards, such as those recommended by the Council on Environmental Quality (CEQ), to NEPA documents offers a shared vocabulary and structure for key entities like projects, processes, and documents that can streamline information exchange and enhance collaboration across systems. In this work, we publicly release NEPATEC2.0, an expanded corpus of NEPA documents with associated metadata. NEPATEC2.0 encompasses approximately 120,000 documents from 60,000 projects prepared by more than 60 different agencies. Modeled to align with CEQ metadata standards, NEPATEC2.0 promotes consistency in environmental reviews and supports the ongoing effort to modernize permitting technologies by facilitating more transparent, efficient, and data-driven decision-making. Importantly, NEPATEC2.0 demonstrates the possibilities and limitations of large language model-based prompting to extract information from NEPA documents at scale.

environmental review↗

NEPATEC v2.0: Standardized Metadata and Text Corpus of National Environmental Policy Act Documents

The National Environmental Policy Act of 1969, as amended (NEPA), is a major environmental law in the United States, requiring Federal agencies to consider and document potential environmental impacts before deciding on a proposed action. Modernization of NEPA and permitting processes faces significant challenges due to the lack of standardized formats and interoperable systems for organizing and sharing NEPA-related information across agencies. Much of the information gathered during NEPA reviews is written into documents such as categorical exclusions, environmental assessments, and environmental impact statements, then filed in predominately independent agency file stores that may or may not be publicly accessible. The application of metadata and data standards, such as those recommended by the Council on Environmental Quality (CEQ), to NEPA documents offers a shared vocabulary and structure for key entities like projects, processes, and documents that can streamline information exchange and enhance collaboration across systems. In this work, we publicly release NEPATEC2.0, an expanded corpus of NEPA documents with associated metadata. NEPATEC2.0 encompasses approximately 120,000 documents from 60,000 projects prepared by more than 60 different agencies. Modeled to align with CEQ metadata standards, NEPATEC2.0 promotes consistency in environmental reviews and supports the ongoing effort to modernize permitting technologies by facilitating more transparent, efficient, and data-driven decision-making. Importantly, NEPATEC2.0 demonstrates the possibilities and limitations of large language model-based prompting to extract information from NEPA documents at scale.

54 ENVIRONMENTAL SCIENCES↗

Measurements and interpretations of W ± Z production cross-sections in pp collisions at $\sqrt{s}=13$ TeV with the ATLAS detector

Measurements of integrated and differential cross-sections for W ± Z production in proton-proton collisions are presented. The data collected by the ATLAS detector at the Large Hadron Collider from 2015 to 2018 at a centre-of-mass energy of $\sqrt{s}=13$ TeV are used, corresponding to an integrated luminosity of 140 fb −1 . The W ± Z candidate events are reconstructed using leptonic decay modes of the gauge bosons into electrons or muons. The integrated cross-section per lepton flavour for the production of W ± Z is measured in the detector fiducial region with a relative precision of 4%. The measured value is compared with the Standard Model prediction at a precision of up to next-to-next-to-leading-order in QCD and next-to-leading-order in electroweak. Cross-sections for W + Z and W − Z production and their ratio are presented. The W ± Z production is also measured differentially as functions of various kinematic variables, including new observables sensitive to CP-violation effects. All measurements are compared with state-of-the-art Standard Model predictions from fixed-order calculations or Monte Carlo generators based on next-to-leading-order matrix elements interfaced with parton showers. An effective field theory interpretation of the measurements is performed, considering both CP-conserving and CP-violating dimension-6 operators modifying the W ± Z production. In the absence of observed deviations from the Standard Model, limits on CP-conserving Wilson coefficients are extracted using the transverse mass of the W ± Z system. For CP-violating coefficients a machine learning approach is used to construct an observable with enhanced sensitivity to CP-violation effects.

hadron-hadron scattering↗

Search for squarks and gluinos in pp collisions at $\sqrt{s} = 13$ TeV and 13.6 TeV in events with $\tau$-leptons, jets and missing transverse momentum using the ATLAS detector

A search for R-parity-conserving supersymmetry in events with large missing transverse momentum, jets and at least one hadronically decaying $\tau$-lepton is presented. Both gluino and squark pair production are considered, with the cascade decay of each gluino or squark producing either a $\tau$-slepton or a $\tau$-sneutrino. Three channels are examined, requiring either exactly one hadronically decaying $\tau$-lepton and no other leptons, exactly one hadronically decaying $\tau$-lepton and at least one other lepton, or two or more hadronically decaying $\tau$-leptons. Analyses in the three channels are optimised independently and combined statistically. Two separate analysis strategies, either a cut-and-count or machine-learning approach, are used. The search uses 140 and 51.8 of pp collision data recorded by the ATLAS detector at the Large Hadron Collider during 2015–2018 at TeV and 2022–2023 at TeV, respectively. Gluino masses below 2.25 TeV and squark masses up to 1.7 TeV are excluded

Aad, G. [CNRS/IN2P3] (ORCID:0000000266654934)↗

Search for new physics in final states with semivisible jets or anomalous signatures using the ATLAS detector

A search is presented for hadronic signatures of beyond the Standard Model (BSM) physics, with an emphasis on signatures of a strongly coupled hidden dark sector accessed via resonant production of a 𝑍′ mediator. The ATLAS experiment dataset collected at the Large Hadron Collider from 2015 to 2018 is used, consisting of proton-proton collisions at $\sqrt{𝑠}$ = 13 TeV and corresponding to an integrated luminosity of 140 fb −1 . The 𝑍′ mediator is considered to decay to two dark quarks, which each hadronize and decay to showers containing both dark and Standard Model particles, producing a topology of interacting and noninteracting particles within a jet known as “semivisible.” Machine learning methods are used to select these dark showers and reject the dominant background of mismeasured multijet events, including an anomaly detection approach to preserve broad sensitivity to a variety of BSM topologies. A resonance search is performed by fitting the transverse mass spectrum based on a functional form background estimation. No significant excess over the expected background is observed. Results are presented as limits on the production cross section of semivisible jet signals, parametrized by the fraction of invisible particles in the decay and the 𝑍′ mass, and by quantifying the significance of any generic Gaussian-shaped mass peak in the anomaly region.

particle dark matter↗

Weakly supervised anomaly detection for resonant new physics in the dijet final state using proton-proton collisions at $\sqrt{s}$ = 13 TeV with the ATLAS detector

An anomaly detection search for narrow-width resonances beyond the Standard Model that decay into a pair of jets is presented. The search is based on 139 fb −1 of proton-proton collisions at $\sqrt{s}$ = 13 TeV recorded during 2015–2018 with the ATLAS detector at the Large Hadron Collider. The analysis is optimized without a particular signal model and aims to be sensitive to a broad range of new physics. It uses two different machine learning strategies to estimate the background in different signal regions. In each region, a weakly supervised classifier is trained to distinguish this background model from data. The analysis focuses on events with high transverse momentum jets reconstructed as large-radius jets. The mass and substructure of these jets are used as inputs to the classifiers. After a classifier-based selection, the distribution of the invariant mass of the two jets is used to search for potential local excesses. The model-independent results of both the anomaly detection methods show no signs of significant local excesses. In addition to model-independent results, a representative set of signal models is injected into the data, and the sensitivity of the methods to these scenarios is reported.

Aad, G. [Aix-Marseille Université] (ORCID:00000002↗