Search NASA⌕ Search

SEARCH · Search NASA

Results for “Louvain method”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Towards scaling community detection on distributed-memory heterogeneous systems

Distributed multi-GPU systems pose significant challenges and opportunities for efficient execution of parallel applications. Graph algorithms are generally characterized by irregular memory accesses, low computation to communication ratios, and load balancing problems that are especially hard to address on multi-GPU systems. Graph community detection is an important problem in the emerging domain of graph analytics with numerous applications. In this paper, we present our ongoing work on distributed-memory multi-GPU implementation for graph community detection. Our work parallelizes the widely used (albeit serial) Louvain method on distributed multi-GPU platforms. Supported by an extensive set of experiments on a multi-GPU enabled supercomputer (OLCF Summit) and a single compute node (Nvidia DGX-2®), we demonstrate competitive performance to existing distributed-memory CPU-based implementation, and up to 6.5 better results than Nvidia RAPIDS® CUGRAPH. To the best of our knowledge, this work represents the first effort for community detection on distributed multi-GPU systems. Our approach and related findings can be extended to numerous other iterative graph algorithms on multi-GPU systems.

97 MATHEMATICS AND COMPUTING↗

Direction-optimizing Label Propagation Framework for Structure Detection in Graphs: Design, Implementation, and Experimental Analysis

Label Propagation is not only a well-known machine learning algorithm for classification but also an effective method for discovering communities and connected components in networks. We propose a new Direction-optimizing Label Propagation Algorithm (DOLPA) framework that enhances the performance of the standard Label Propagation Algorithm (LPA), increases its scalability, and extends its versatility and application scope. As a central feature, the DOLPA framework relies on the use of frontiers and alternates between label push and label pull operations to attain high performance. It is formulated in such a way that the same basic algorithm can be used for finding communities or connected components in graphs by only changing the objective function used. Additionally, DOLPA has parameters for tuning the processing order of vertices in a graph to reduce the number of edges visited and improve the quality of solution obtained. We present the design and implementation of the enhanced algorithm as well as our shared-memory parallelization of it using OpenMP. We also present an extensive experimental evaluation of our implementations using the LFR benchmark and real-world networks drawn from various domains. Compared with an implementation of LPA for community detection available in a widely used network analysis software, we achieve at most five times the F-Score while maintaining similar runtime for graphs with overlapping communities. We also compare DOLPA against an implementation of the Louvain method for community detection using the same LFR-graphs and show that DOLPA achieves about three times the F-Score at just 10% of the runtime. For connected component decomposition, our algorithm achieves orders of magnitude speedups over the basic LP-based algorithm on large-diameter graphs, up to 13.2× speedup over the Shiloach-Vishkin algorithm, and up to 1.6× speedup over Afforest on an Intel Xeon processor using 40 threads.

97 MATHEMATICS AND COMPUTING↗

Use of Graph Theory and Neural Networks for Microstructural Classification

Recent advances in materials data analytics have provided new avenues for determining process-structure-property (PSP) linkages in a variety of materials. Machine learning techniques including few-shot learning have increased the efficiency of classifying microscopy images for the purposes of material characterization. Modifications in segmentation also show potential in improving the accuracy of our current pyCHIP classifier. Replacing previous encoders trained on ImageNet with those trained on microscopy images like MicroNet has initially shown better performance at classifying images of irradiated samples. Additionally, different normalization approaches were tested to show no discernable effect on classification. The Louvain method for community detection is analyzed on a set of irradiated samples with different parameters to determine which proved beneficial under what circumstances. We suggest that microscopy experiments be automated in the future using a combination of these techniques to enable high-throughput analyses.

36 MATERIALS SCIENCE↗

Exploring the Landscape of Distributed Graph Clustering on Leadership Supercomputers

The rapid growth of large-scale datasets in fields like biology and social networks has driven the need for advanced graph analytics techniques. Community detection, a fundamental task in graph analytics, identifies closely connected groups of nodes within a network, providing valuable insights across various disciplines. This study focuses on two classic community detection methods, the Louvain algorithm and Markov Clustering (MCL), and evaluates the performance of two prominent distributed community detection algorithms: HiPDPL-GPU, our prior implementation, and HipMCL. We conduct experiments on GPU-accelerated heterogeneous HPC systems, Summit and Frontier, to assess their performance under varying conditions. Our objective is to identify the strengths and weaknesses of these algorithms in terms of scalability, and quality of solutions. We evaluate these algorithms on a diverse set of 70+ networks spanning 13 domains, with sizes ranging up to 4.2 billion edges. Our results demonstrate that HiPDPL-GPU consistently outperforms HipMCL, especially for large-scale networks. HiPDPL-GPU achieves significantly faster runtimes (47x to 1439x), higher modularity scores, and improved scalability. These findings highlight HiPDPL-GPU as a promising solution for efficient and effective large-scale graph analytics in diverse application domains, and provide insights into the feasibility of using MCL-based approaches for certain application domains.

Community detection, graph algorithms↗

High-throughput computation of electric polarization in solids via Berry flux diagonalization

Electric polarization in the absence of an externally applied electric field is a key property of polar materials, but the standard interpolation-based ab initio approach to compute polarization differences within the modern theory of polarization presents challenges for automated high-throughput calculations. Berry flux diagonalization [J. Bonini et al., Phys. Rev. B 102, 045141 (2020)] has been proposed as an efficient and reliable alternative, though it has yet to be widely deployed. Here, we assess Berry flux diagonalization using ab initio calculations of a large set of materials, introducing and validating heuristics that ensure branch alignment with a minimal number of intermediate interpolated structures. Our automated implementation of Berry flux diagonalization succeeds in cases where prior interpolation-based workflows fail due to band-gap closures or branch ambiguities. Benchmarking with ab initio calculations of 176 candidate ferroelectrics, we demonstrate the efficacy of the approach on a broad range of insulating materials and obtain accurate effective polarization values with fewer interpolated structures than prior automated interpolation-based workflows. Our real-space heuristics that can predict gauge stability a priori from ionic displacements enable a general automated framework for reliable polarization calculations and efficient high-throughput screening of chemically and structurally diverse polar insulators. These results establish Berry flux diagonalization as a robust and efficient method to compute the effective polarization of solids and to accelerate the data-driven discovery of functional polar materials.

Poteshman, Abigail N. [University of Chicago, IL (↗

Atomate2: modular workflows for materials science

High-throughput density functional theory (DFT) calculations have become a vital element of computational materials science, enabling materials screening, property database generation, and training of “universal” machine learning models. While several software frameworks have emerged to support these computational efforts, new developments such as machine learned force fields have increased demands for more flexible and programmable workflow solutions. This manuscript introduces atomate2, a comprehensive evolution of our original atomate framework, designed to address existing limitations in computational materials research infrastructure. Key features include the support for multiple electronic structure packages and interoperability between them, along with generalizable workflows that can be written in an abstract form irrespective of the DFT package or machine learning force field used within them. Our hope is that atomate2's improved usability and extensibility can reduce technical barriers for high-throughput research workflows and facilitate the rapid adoption of emerging methods in computational material science.

97 MATHEMATICS AND COMPUTING↗

Updates on MURAVES Project at Mt. Vesuvius

The MUon RAdiography of VESuvius (MURAVES) project aims to employ muography imaging techniques to investigate the internal structure of the summit of Mount Vesuvius, an active volcano located near Naples, Italy. This paper reports recent advancements in data analysis and simulation tools that significantly improve the quality and reliability of the experiment’s results. A new track selection method, referred to as the Golden Selection, has been developed to identify high-quality muon tracks by applying an improved χ 2 -based criterion. This method enhances the signal-to-background ratio and improves the resolution of the resulting muographic images. Moreover, the simulation framework has been upgraded through the integration of the MULDER (MUon simuLation for DEnsity Reconstruction) library, which consolidates the functionalities of previously used libraries into a single, unified platform. MULDER enables efficient and accurate modeling of muon flux variations induced by topographical features. A good agreement is observed between the simulated and measured muon flux maps, validating the effectiveness of the new analysis and simulation approaches.

Cosmic rays↗

Refining Jets for CMS Run 3 using Fast Simulation

As the LHC moves into its high-luminosity phase, the CMS experiment must handle more complex data collected at much higher rates. While the Geant4-based simulation application (FullSim) provides highly accurate simulation to complement real data, FullSim’s intensive consumption of computing resources becomes an increasing liability as the rates increase, while faster tools offer an advantage. The fast MC production application (FastSim) delivers a complete simulation with a factor of 10 speedup over FullSim, but introduces inaccuracies in some observables. A specialized refinement method, Fast Perfekt, employs machine learning to improve the accuracy of FastSim. An initial report of this work focused on the refinement of jet flavor tagging observables. This article presents an update on the refinement, focusing on PUPPI jets with Run 3 data-taking conditions. Refinement is extended to include jet transverse momentum as well as its propagation to missing transverse momentum. A gridbased framework and real-time monitoring system have been developed to facilitate optimization and scaling of the refinement to a large number of target variables.

Güngördü, Açelya Deniz [Istanbul Tech. U.]↗

Cs X Si 15 P 21 ( X = Sn or Pb): Polar Noncentrosymmetric Si–P Frameworks Stabilized by Covalent X –P Bonding

Metal silicon phosphides composed of earth-abundant Si and P tend to exhibit semiconducting properties and adopt diverse crystal structures with relatively small additions of structure-directing elements. The potential of silicon phosphide materials in nonlinear optical applications has been hindered by the inability to systematically produce noncentrosymmetric structures with such a flexible framework. Here, in this work, two isostructural compounds with a novel noncentrosymmetric structure were made possible by the inclusion of elements with stereochemically active lone pairs (Sn 2+ and Pb 2+ ). The structures were determined through single-crystal and synchrotron powder X-ray diffraction. Analysis of chemical bonding in real space through the electron localization function revealed stereochemically active Pb 2+ and Sn 2+ species in a trigonal pyramidal coordination with {Pb/Sn}–P bonds. Such covalent bonding between Pb and P is quite uncommon in extended solids and has been reported in a few rare instances. Band structure calculations and linear optical measurements confirm the semiconducting nature of Cs X Si 15 P 21 ( X = Sn or Pb). The synthesis was optimized to yield high-purity polycrystalline samples. The nonlinear optical properties show promising second-harmonic generation (SHG) coefficients from the Kurtz–Perry method. First-principles calculations of the nonlinear optical properties support the experimentally determined SHG values and provide moderate values of birefringence, suggesting Cs X Si 15 P 21 could be phase-matchable and practical nonlinear optical materials in the mid-IR region.

crystal structure↗

Model-agnostic search for dijet resonances with anomalous jet substructure in proton–proton collisions at $\sqrt{s}$ = 13 TeV

This paper presents a model-agnostic search for narrow resonances in the dijet final state in the mass range 1.8-6 TeV. The signal is assumed to produce jets with substructure atypical of jets initiated by light quarks or gluons, with minimal additional assumptions. Search regions are obtained by utilizing multivariate machine-learning methods to select jets with anomalous substructure. A collection of complementary anomaly detection methods - based on unsupervised, weakly supervised, and semisupervised algorithms - are used in order to maximize the sensitivity to unknown new physics signatures. These algorithms are applied to data corresponding to an integrated luminosity of 138 fb -1 , recorded by the CMS experiment at the LHC, at a center-of-mass energy of 13 TeV. No significant excesses above background expectations are seen. Exclusion limits are derived on the production cross section of benchmark signal models varying in resonance mass, jet mass, and jet substructure. Many of these signatures have not been previously sought, making several of the limits reported on the corresponding benchmark models the first ever. When compared to benchmark inclusive and substructure-based search strategies, the anomaly detection methods are found to significantly enhance the sensitivity to a variety of models.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Machine learning method for enforcing variable independence in background estimation with LHC data: ABCDisCoTEC

A novel solution is presented for the problem of estimating the backgrounds of a signal search using observed data while simultaneously maximizing the sensitivity of the search to the signal. The 'ABCD method' provides a reliable framework for background estimation by partitioning events into one signal-enhanced region (A) and three background-enhanced control regions (B, C, and D) via two smoothly varying, statistically independent variables. In practice, even slight correlations between the two variables can significantly undermine the method's performance. Thus, choosing appropriate variables by hand can present a formidable challenge, especially when background and signal differ only subtly. To address this issue, the ABCD with distance correlation (ABCDisCo) method was developed to construct two learned variables via a neural network trained to provide strong signal-background discrimination with small values of the distance correlation (DisCo) measure between the two learned variables. However, relying solely on minimizing the DisCo can result in learned variables that may not have distributions of background events that are smoothly varying and localized at extreme values, as necessary for the validity of the background estimation. The ABCDisCo training enhanced with closure (ABCDisCoTEC) method is introduced to solve this issue by directly minimizing the nonclosure, expressed as a dedicated differentiable loss term. This extended method is applied to a data set of proton-proton collisions at a center-of-mass energy of 13 TeV recorded by the CMS detector at the CERN Large Hadron Collider. Additionally, given the complexity of the minimization problem with constraints on multiple loss terms, the modified differential method of multipliers is applied and shown to greatly improve the stability and robustness of the ABCDisCoTEC method, compared to grid search hyperparameter optimization procedures.

Hayrapetyan, Aram [Yerevan Phys. Inst.]↗

Machine-learning techniques for model-independent searches in dijet final states

Anomaly detection methods used in a recent search for new phenomena by CMS at the CERN LHC are presented. The methods use machine learning to detect anomalous jets produced in the decay of new massive particles without depending on a specific theory model. The effectiveness of these approaches in enhancing sensitivity to various simulated signal samples is studied and compared using data collected in proton–proton collisions at a center-of-mass energy of 13 TeV. In an example analysis, the capabilities of anomaly detection methods are further demonstrated by identifying large-radius jets consistent with Lorentz-boosted hadronically decaying top quarks in a model-agnostic framework.

CMS↗

Measurement of the dineutrino system kinematic variables in dileptonic top quark pair production in proton-proton collisions at $\sqrt{s}=13$ TeV

Differential top quark pair production cross sections are measured in the dilepton final states e + e − , μ + μ − , and e ± μ ∓ , as a function of kinematic variables of the two-neutrino system: the transverse momentum $p^{vv}_{\textrm{T}}$ of the dineutrino system, the minimum distance in azimuthal angle between $\vec{p}^{vv}_{\textrm{T}}$ and leptons, and in two dimensions in bins of both observables. The measurements are performed using CERN LHC proton-proton collisions at $\sqrt{s}=13$ TeV, recorded by the CMS detector between 2016 and 2018, corresponding to an integrated luminosity of 138 fb −1 . The measured cross sections are unfolded to the particle level using an unregularized least squares method. Results are compared with predictions by the standard model of particle physics, and found to be in agreement with theoretical calculations as well as Monte Carlo simulations.

Hadron-Hadron Scattering↗

A method for correcting the substructure of multiprong jets using the Lund jet plane

Many analyses at the CERN LHC exploit the substructure of jets to identify heavy resonances produced with high momenta that decay into multiple quarks and/or gluons. This paper presents a new technique for correcting the substructure of simulated large-radius jets from multiprong decays. The technique is based on reclustering the jet constituents into several subjets such that each subjet represents a single prong, and separately correcting the radiation pattern in the Lund jet plane of each subjet using a correction derived from data. The data presented here correspond to an integrated luminosity of 138 fb −1 collected by the CMS experiment between 2016–2018 at a center-of-mass energy of 13 TeV. The correction procedure improves the agreement between data and simulation for several different substructure observables of multiprong jets. This technique establishes, for the first time, a robust calibration for the substructure of jets with four or more prongs, enabling future measurements and searches for new phenomena containing these signatures.

Hadron-Hadron Scattering↗

Observation of 𝑡⁢𝑊⁢𝑍 Production at the CMS Experiment

The first observation of single top quark production in association with a 𝑊 and a 𝑍 boson in proton-proton collisions is reported. The analysis uses data at center-of-mass energies of 13 and 13.6 TeV recorded with the CMS detector at the CERN LHC, corresponding to a total integrated luminosity of 200 fb −1 . Events with three or four charged leptons, which can be electrons or muons, are selected. Advanced machine-learning algorithms and improved reconstruction methods, compared to an earlier analysis, result in an unprecedented sensitivity to 𝑡⁢𝑊⁢𝑍 production. The measured cross sections for 𝑡⁢𝑊⁢𝑍 production are 248 ± 52 fb and 242 ± 77 fb for $\sqrt{s}$ =13 and 13.6 TeV, respectively. The signal is established with a statistical significance of 5.8 standard deviations, with 3.5 expected, compared to the background-only hypothesis.

Hayrapetyan, Aram [Yerevan Physics Institute]↗

Reweighting simulated events using machine-learning techniques in the CMS experiment

Data analyses in particle physics rely on an accurate simulation of particle collisions and a detailed simulation of detector effects to extract physics knowledge from the recorded data. Event generators together with a GEANT -based simulation of the detectors are used to produce large samples of simulated events for analysis by the LHC experiments. These simulations come at a high computational cost, where the detector simulation and reconstruction algorithms have the largest CPU demands. This article describes how machine-learning (ML) techniques are used to reweight simulated samples obtained with a given set of parameters to samples with different parameters or samples obtained from entirely different simulation programs. The ML reweighting method avoids the need for simulating the detector response multiple times by incorporating the relevant information in a single sample through event weights. Results are presented for reweighting to model variations and higher-order calculations in simulated top quark pair production at the LHC. This ML-based reweighting is an important element of the future computing model of the CMS experiment and will facilitate precision measurements at the High-Luminosity LHC.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Improved measurements of the TeV-PeV extragalactic neutrino spectrum from joint analyses of IceCube tracks and cascades

The IceCube South Pole Neutrino Observatory has discovered the presence of a diffuse astrophysical neutrino flux at energies of TeV and beyond using neutrino induced muon tracks and cascade events from neutrino interactions. Here, we present two analyses sensitive to neutrino events in the energy range 1 TeV to 10 PeV, using more than 10 years of IceCube data. Both analyses consistently reject a neutrino spectrum following a single power-law with a significance of > 4⁢𝜎 in favor of a broken power law. We describe the methods implemented in the two analyses, the spectral constraints obtained, and the validation of the robustness of the results. Additionally, we report the detection of a muon neutrino in the medium energy starting events sample, or MESE, with an energy of 11.4$^{+2.46}_{−2.53}$ PeV, the highest energy neutrino observed by IceCube to date. The results presented here show insights into the spectral shape of astrophysical neutrinos, which has important implications for inferring their production processes in a multimessenger picture.

astroparticles↗

Measurements of $\textrm{t}\overline{\textrm{t}}\textrm{W}$ differential cross sections and the leptonic charge asymmetry at $\sqrt{s}=13$ TeV

Measurements of properties of top quark-antiquark pair production in association with a W boson in proton-proton collisions at a center-of-mass energy of 13 TeV are presented, using a data sample corresponding to an integrated luminosity of 138 fb −1 , recorded by the CMS experiment at the CERN LHC. Events are selected based on the presence of either two leptons with the same electric charge or three leptons, and multiple jets and b-tagged jets. We present measurements of differential production cross sections as a function of kinematic variables sensitive to different aspects of the process modeling, using a multivariate discriminator in the two-lepton selection region and a simple selection-based method in the three-lepton region. The normalized cross section measurements are generally consistent with the standard model expectations, while we observe larger values compared to the expectations in the absolute cross section measurements, consistent with previous inclusive cross section measurements. In addition, we measure the leptonic charge asymmetry of this process, obtaining an observed value of ${A}_c^{\ell }=-{0.19}_{-0.18}^{+0.16}$, consistent with the expectation of −0.085 ± 0.006 predicted by next-to-leading order simulations.

Hadron-Hadron Scattering↗