Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallelism”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 811 records · Page 45

Cryogenic light detectors with thermal signal amplification for 0 νββ search experiments

As a step towards the realization of cryogenic-detector experiments to search for neutrinoless double-beta decay (such as CROSS, BINGO, and CUPID), we investigated a batch of 10 Ge light detectors (LDs) assisted by Neganov-Trofimov-Luke (NTL) signal amplification. Each LD was assembled with a large cubic light-emitting crystal (45 mm side) using the recently developed CROSS mechanical structure. The detector array was operated at milli-Kelvin temperatures in a pulse-tube cryostat at the Canfranc underground laboratory in Spain. We achieved good performance with scintillating bolometers from CROSS, made of Li 2 100 MoO 4 crystals and used as reference detectors of the setup, and with all LDs tested (except for a single device that encountered an electronics issue). No leakage current was observed for 8 LDs with an electrode bias up to 100 V. Operating the LDs at an 80 V electrode bias applied in parallel, we obtained a gain of around 9 in the signal-to-noise ratio of these devices, allowing us to achieve a baseline noise RMS of O(10 eV). Thanks to the strong current polarization of the temperature sensors, the time response of the devices was reduced to around half a millisecond in rise time. The achieved performance of the LDs was extrapolated via simulations of pile-up rejection capability for several configurations of the CUPID detector structure. Despite the sub-optimal noise conditions of the LDs (particularly at high frequencies), we demonstrated that the NTL technology provides a viable solution for background reduction in CUPID.

47 OTHER INSTRUMENTATION↗

Low power on-chip data transmission for wafer-scale monolithic active pixel sensors

Here, this paper details the implementation of the digital pulse shaping subsystem within the Backbone Transmission Line Encoding (BTLE) driver, a low-power, long-distance on-chip data transmission solution designed in a 65 nm CMOS process. Digital pulse shaping is critical for minimizing inter-symbol interference (ISI) caused by bandwidth limitations of on-chip interconnects, especially in wafer-scale monolithic active pixel sensors (MAPS). A duobinary encoder coupled with a parallelized polyphase finite impulse response (FIR) filter is used for efficient shaping of the transmitted signal spectrum. This reconfigurable architecture achieves reliable 160 Mb/s data transfer over a 10 cm on-chip link, as validated by simulations demonstrating low power consumption (FoM 37.3 fJ/bit/mm of transmission line length) and effective ISI mitigation.

47 OTHER INSTRUMENTATION↗

High-rate heavy-ion tracking using MWPCs and PPACs with an Anti-Discharge Unit

This work reports on the development of two robust, heavy-ion beam tracking concepts operating at low pressure (< 15 Torr) for high rate applications (> 200 kHz). The first concept consists of a Multi-Wire Proportional Counter (MWPC) with a central anode consisting of 12 μm Au-plated Tungsten wires spaced 1 mm from each other. The anode grid is sandwiched between two segmented cathodes aligned orthogonally in the two dimensions for (x,y) particle localization. The second detector is a Parallel-Plate Avalanche Counter (PPAC) that uses the same readout geometry as the MWPC, but replaces the wire anode with a 150 nm silver layer deposited on both sides of a thin (< 1 mg/cm 2 ) polypropylene foil. Additionally, the bias circuitry for the PPAC anode central foil is equipped with an Anti-Discharge Unit (ADU) to prevent transitions from proportional operation to streamer formation, thereby avoiding damaging discharges. The localization capability of both detectors was tested with a low-rate alpha-particle source (241-Am). A position resolution of < 1 mm (FWHM) was achieved under stable, high-gas-gain (> 1000) operating conditions. Their performance in terms of detection efficiency as a function of the isotope charge (Z) was determined by irradiating the detectors with a cocktail beam (Z ≤ 15) with energy of ∼ 100 MeV/u. Full detection efficiency is maintained for all available fragments under optimal operational conditions (i.e., voltage bias). Full detection efficiency was achieved at rates above 200 kHz by irradiating the detectors with a 1 cm diameter 238 U beam at an energy of 143 MeV/u.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Reconstruction of neutrino events in the Accelerator Neutrino Neutron Interaction Experiment. Part I

The Accelerator Neutrino Neutron Interaction Experiment (ANNIE) was designed to reconstruct neutrino events from the Fermilab Booster Neutrino Beam (BNB) with the parallel goals of measuring neutron production in interactions with oxygen and serving as a testbed for new technology. The ANNIE detector consists of a 26-ton water Cherenkov target tank instrumented with conventional photomultiplier tubes (PMTs), a downstream tracking muon spectrometer, and an upstream double wall of plastic scintillator to serve to veto charged particles incoming from neutrino events that occur upstream of the experimental setup. ANNIE has also deployed multiple Large-Area Picosecond PhotoDetectors (LAPPDs) and a test vessel of water-based liquid scintillator (WbLS). This paper describes the event reconstruction performance of the detector before implementation of these novel technologies, which will serve as a baseline against which their impact can be measured. That said, even the techniques used for event reconstruction using only the conventional PMT array and muon spectrometer are significantly different than those used in other water Cherenkov detectors due to the small size of ANNIE (which makes nanosecond-scale timing not as useful as in a large detector) and the availability of reconstruction information from the tracking muon spectrometer. We demonstrate that combining the information from these two elements into a single fit using only pattern recognition yields a muon vertex uncertainty of 60 cm, a directional uncertainty of 13.2 degrees, and energy reconstruction uncertainty of about 10% for BNB muon neutrino Charged Current Zero Pion (CC0π) events.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Large-scale real-time signal processing in physics experiments: the ALICE TPC FPGA pipeline

For LHC Run 3, the ALICE Time Projection Chamber was upgraded to operate in continuous readout mode. Interaction rates of up to 50 kHz in Pb-Pb collisions require real-time processing of more than 3 TB s -1 of raw detector data. This requirement is met by a custom FPGA-based processing pipeline that performs the complete front-end data treatment fully in-stream, including common-mode correction, pedestal subtraction, ion-tail filtering, zero suppression, and dense data packing. A central element of the design is a highly parallel common-mode correction algorithm operating directly on the streaming data. It robustly identifies signal-free readout channels on a time-bin basis and applies pad-dependent scaling to compensate for local variations in capacitive coupling in the GEM readout. In combination with pedestal subtraction and ion-tail filtering, this enables accurate baseline restoration under extreme high-occupancy conditions, preventing signal loss while efficiently suppressing noise prior to zero suppression. The pipeline operates continuously at the full detector bandwidth and reduces the raw input rate of approximately 3 TB s -1 to about 900 GBps for Pb-Pb collisions at the target interaction rate. Overall, it represents a large-scale FPGA-based real-time signal-processing implementation for high-energy physics detector readout.

Digital signal processing (DSP)↗

The sweeper spectrometer for neutron invariant-mass spectroscopy at FRIB

Neutron invariant-mass spectroscopy (NIMS) is a key technique for studying unbound and weakly bound nuclei at the limits of stability. At the Facility for Rare Isotope Beams (FRIB), such measurements are performed using the Sweeper spectrometer, a large-gap, high-rigidity dipole system coupled to the MoNA-LISA neutron detector arrays. To meet the demands imposed by higher beam energies (>130 MeV/u) and the broad cocktail-beam selection available at FRIB, the spectrometer has recently been upgraded to improve particle-identification and detection performance. Upstream of the reaction target, a plastic scintillator with Silicon photomultiplier (SiPM) readout provides the global trigger and time reference, two parallel plate avalanche counters (PPACs) track the trajectories of incoming beam particles, and a silicon PIN detector measures the energy loss, ΔE, for charge (Z) identification. After the Sweeper magnet, the trajectories of the reaction products are tracked by two micro-pattern drift chambers (MPDCs), their charge (Z) is identified by a Frisch-grid ionization chamber (FG-IC), and their mass-to-charge ratio (A/Q) is deduced by time-of-flight measurement using a fast plastic scintillator read out by an array of photomultiplier tubes (PMTs). The detection system also incorporates the Modular Neutron Array (MoNA) for neutron detection and the CAESium-iodide scintillator ARray (CAESAR) for high-efficiency γ-ray measurements to enable full kinematic reconstruction. Performance was evaluated using a cocktail beam around 37 Al accelerated at E ≈ 130 MeV/u during the first FRIB campaign, demonstrating the readiness of the upgraded system for future studies of nuclei at and beyond the neutron drip line.

Particle identification methods↗

The birth of a field through a marriage of scales: urban meteorological modeling meets regional climate modeling

Urban meteorological modeling and regional climate modeling developed largely independently, with each discipline addressing different aspects of atmospheric processes across space and time. In this review, I show that urban climate modeling did not arise as a simple scaling extension of urban meteorological modeling or an add-on to regional climate modeling, but instead emerged through the selective inheritance of complementary strengths from its parents after both disciplines reached sufficient methodological maturity. By tracing the parallel evolution of these disciplines, I demonstrate that urban climate modeling inherited the physically explicit treatment of the built environment developed within urban meteorological modeling, and the hierarchical scale translation and climatological framing that matured within regional climate modeling, allowing urban effects to influence climate-relevant outcomes. Building on this synthesis, I propose an objective, and methodologically-grounded definition of urban climate modeling that distinguishes it from urban meteorological modeling. This distinction is increasingly important as urban climate data informs decisions with long-lived societal consequences, and as ambiguity in terminology risks conflating fundamentally different modeling frameworks with distinct physical meaning and decision relevance. In addition to a clarifying definition, this review outlines a research framework for advancing urban climate modeling through scale-aware coupling strategies that preserve physically consistent urban-atmosphere interactions.

regional climate↗

Enhancing quantum annealing accuracy through replication-based error mitigation *

Abstract Quantum annealers like those manufactured by D-Wave Systems are designed to find high quality solutions to optimization problems that are typically hard for classical computers. They utilize quantum effects like tunneling to evolve toward low-energy states representing solutions to optimization problems. However, their analog nature and limited control functionalities present challenges to correcting or mitigating hardware errors. As quantum computing advances towards applications, effective error suppression is an important research goal. We propose a new approach called replication based mitigation (RBM) based on parallel quantum annealing (QA). In RBM, physical qubits representing the same logical qubit are dispersed across different copies of the problem embedded in the hardware. This mitigates hardware biases, is compatible with limited qubit connectivity in current annealers, and is well-suited for currently available noisy intermediate-scale quantum annealers. Our experimental analysis shows that RBM provides solution quality on par with previous methods while being more flexible and compatible with a wider range of hardware connectivity patterns. In comparisons against standard QA without error mitigation on larger problem instances that could not be handled by previous methods, RBM consistently gets better energies and ground state probabilities across parameterized problem sets.

Djidjev, Hristo N. (ORCID:0000000192868824)↗

Model-based, in-situ, non-destructive qualification and certification of parts made by autonomous additive manufacturing

To address the significant productivity challenges associated with the qualification and certification (Q&C) tasks of additively manufactured (AM) parts, which have traditionally relied on rigorous post‐build inspection and testing, we propose an integrated framework that combines model‐based qualification and certification (MBQ&C) with autonomous additive manufacturing (AAM). MBQ&C employs high‐fidelity predictive models, developed within the Integrated Computational Materials Engineering (ICME) paradigm, to simulate process–structure–property–performance relationships for assessing a part’s fitness for use. Since predictive models are commonly machine learning (ML)-based or reduced-order surrogates of validated physics models, they run efficiently, enabling timely inference. In parallel, the self-driving AAM utilises ML-based adaptive, closed‐loop control strategies to avoid, mitigate, or repair defects and anomalies during fabrication, thereby increasing the likelihood of producing acceptable parts. A key feature of the combined AAM-MBQ&C framework is that predictive models explicitly incorporate defects or anomalies that persist after the build, using instance-specific data captured via in-situ sensing. This customisation enables a build‐specific assessment of fitness for use, rather than relying on nominal or generic parameters. Such individualised evaluation provides a robust basis for Q&C-related acceptance decisions relating to each build. Additionally, the rapid solution capabilities of ML or reduced-order models enable the determination of a part’s suitability for service shortly after build completion. As the framework matures, it has the potential to substantially reduce reliance on conventional point‐design approaches—such as time‐consuming post‐build computed tomography scanning and costly destructive testing. Thus, the AAM-MBQ&C framework represents a transformative, scalable strategy for quality assurance of AM components, as parts produced within a stable, validated, and certified envelope can be certified with reduced testing. Key benefits include: (1) significant gains in Q&C productivity through efficient, model-centric assessment; (2) performance-based classification of defects into critical and non-critical categories; (3) the ability to predict potential deviations in the performance of parts affected by real-time, adaptive process control interventions relative to those produced under a certified process, and (4) the enabling of virtual Q&C for service environments that are difficult, hazardous, or impractical to access or reproduce experimentally. Collectively, these capabilities strengthen the business case for AM, particularly for high‐consequence and mission‐critical applications. Finally, although this work focuses on powder-based AM, the proposed techniques could be extended to AM processes employing alternative feedstock forms.

Gunasegaram, Dayalan↗

Review of solar-enabled desalination and implications for zero-liquid-discharge applications

Abstract The production of freshwater from desalinating abundant saline water on the planet is increasingly considered a climate change adaptation measure. Yet, there are challenges associated with the high cost, intensive energy demand, and environmental implications of desalination. Effective integration of solar energy generation and freshwater production can address both issues. This review article highlights recent key advances in such integration achieved in a joint-research university-national laboratory partnership under the auspices of the United States Department of Energy and parallel efforts worldwide. First, an overview of current and emerging desalination technologies and associated pretreatment, brine treatment, and valorization technologies that together can result in zero-liquid-discharge systems is presented, and their technological readiness levels are evaluated. Then, advanced modeling techniques and new software platforms that enable optimization of solar-desalination applications with the dual objective of cost and environmental impact minimization are discussed.

14 SOLAR ENERGY↗

Predicting nonequilibrium Green’s function dynamics and photoemission spectra via nonlinear integral operator learning

Understanding the dynamics of nonequilibrium quantum many-body systems is an important research topic in a wide range of fields across condensed matter physics, quantum optics, and high-energy physics. However, numerical studies of large-scale nonequilibrium phenomena in realistic materials face serious challenges due to intrinsic high-dimensionality of quantum many-body problems and the absence of time-invariance. The nonequilibrium properties of many-body systems can be described by the dynamics of the correlator, or the Green's function of the system, whose time evolution is given by a high-dimensional system of integro-differential equations, known as the Kadanoff–Baym equations (KBEs). The time-convolution term in KBEs, which needs to be recalculated at each time step, makes it difficult to perform long-time numerical simulation. In this paper, we develop an operator-learning framework based on recurrent neural networks (RNNs) to address this challenge. We utilize RNNs to learn the nonlinear mapping between Green's functions and convolution integrals in KBEs. By using the learned operators as a surrogate model in the KBE solver, we obtain a general machine-learning scheme for predicting the dynamics of nonequilibrium Green's functions. Besides significant savings per each time step, the new methodology reduces the temporal computational complexity from $O(N_t^3)$ to $O(N_t)$ where N t is the number of steps taken in a simulation, thereby making it possible to study large many-body problems which are currently infeasible with conventional KBE solvers. Through various numerical examples, we demonstrate the effectiveness of the operator-learning based approach in providing accurate predictions of physical observables such as the reduced density matrix and time-resolved photoemission spectra. Moreover, our framework exhibits clear numerical convergence and can be easily parallelized, thereby facilitating many possible further developments and applications.

97 MATHEMATICS AND COMPUTING↗

SAGIPS: a physics-inspired scalable asynchronous generative inverse-problem solver

Abstract Solving large-scale inverse problems using deep-learning algorithms have become an essential part of modern research and industrial applications. The complexity of the underlying inverse problem may require the utilization of high performance computing systems which poses a challenge on the algorithmic design of the inverse problem solver. Most deep learning algorithms require, due to their design, custom parallelization techniques in order to be resource efficient while showing a reasonable convergence. In this paper we introduce a S calable A synchronous G enerative I nverse P roblem S olver (SAGIPS) on high-performance computing systems. We present a workflow that utilizes an asynchronous ring-allreduce algorithm to transfer the gradients of the generator network across multiple GPUs. Experiments with a scientific proxy application demonstrate that SAGIPS shows near linear weak scaling, together with a convergence quality that is comparable to traditional methods. The approach presented here allows leveraging Generative Adverserial Network across multiple GPUs, promising advancements in solving complex inverse problems at scale.

97 MATHEMATICS AND COMPUTING↗

Geometric GNNs for charged particle tracking at GlueX

Nuclear physics experiments are aimed at uncovering the fundamental building blocks of matter. The experiments involve high-energy collisions that produce complex events with many particle trajectories. Tracking charged particles resulting from collisions in the presence of a strong magnetic field is critical to enable the reconstruction of particle trajectories and precise determination of interactions. It is traditionally achieved through combinatorial approaches that scale worse than linearly as the number of hits grows. Since particle hit data naturally form a point cloud and can be structured as graphs, graph neural networks (GNNs) emerge as an intuitive and effective choice for this task. In this study, we evaluate the GNN model for track finding on the data from the GlueX experiment at Jefferson Lab. We use simulation data to train the model and test on both simulation and real GlueX measurements. We demonstrate that GNN-based track finding outperforms the currently used traditional method at GlueX in terms of segment-based efficiency at a fixed purity while providing faster inferences. We show that the GNN model can achieve significant speedup by processing multiple events in batches, which exploits the parallel computation capability of graphical processing units (GPUs). Finally, we compare the GNN implementation on GPU and field-programmable gate array and describe the trade-off.

batched GNN pipeline↗

Latent Twins

Over the past decade, scientific machine learning has transformed the development of mathematical and computational frameworks for analyzing, modeling, and predicting complex systems. From inverse problems to numerical partial differential equations (PDEs), dynamical systems, and model reduction, these advances have pushed the boundaries of what can be simulated. Yet they have often progressed in parallel, with representation learning and algorithmic solution methods evolving largely as separate pipelines. With Latent Twins, we propose a unifying mathematical framework that creates a hidden surrogate in latent space for the underlying equations. Whereas digital twins mirror physical systems in the digital world, Latent Twins mirror mathematical systems in a learned latent space governed by operators. Through this lens, classical modeling, inversion, model reduction, and operator approximation all emerge as special cases of a single principle. We establish the fundamental approximation properties of Latent Twins for both ordinary differential equations (ODEs) and PDEs and demonstrate the framework across three representative settings: (i) canonical ODEs, capturing diverse dynamical regimes; (ii) a PDE benchmark using the shallow-water equations, contrasting Latent Twin simulations with deep operator network and forecasts with a four-dimensional variational method baseline; and (iii) a challenging real-data geopotential reanalysis dataset, reconstructing and forecasting from sparse, noisy observations. Latent Twins provide a compact, interpretable surrogate for solution operators that evaluate across arbitrary time gaps in a single-shot, while remaining compatible with scientific pipelines such as assimilation, control, and uncertainty quantification. Looking forward, this framework offers scalable, theory-grounded surrogates that bridge data-driven representation learning and classical scientific modeling across disciplines.

Latent Twins↗

Supercharging simulation-based inference for Bayesian optimal experimental design

Abstract Bayesian optimal experimental design (BOED) seeks to maximize the expected information gain (EIG) of experiments. This requires a likelihood estimate, which in many settings is intractable. Simulation-based inference (SBI) provides powerful tools for this regime. However, existing work explicitly connecting SBI and BOED is restricted to a single contrastive EIG bound. We show that the EIG admits multiple formulations which can directly leverage modern SBI density estimators, encompassing neural posterior, likelihood, and ratio estimation. Building on this perspective, we define a novel EIG estimator using neural likelihood estimation. Further, we identify optimization as a key bottleneck of gradient based EIG maximization and show that a simple multi-start parallel gradient ascent procedure can substantially improve reliability and performance. With these innovations, our SBI-based BOED methods are able to match or outperform by up to 22% existing state-of-the-art approaches across standard BOED benchmarks.

97 MATHEMATICS AND COMPUTING↗

Poplar: a phylogenomics pipeline

Motivation Generating phylogenomic trees from the genomic data is essential in understanding biological systems. Each step of this complex process has received extensive attention and has been significantly streamlined over the years. Given the public availability of data, obtaining genomes for a wide selection of species is straightforward. However, analyzing that data to generate a phylogenomic tree is a multistep process with legitimate scientific and technical challenges, often requiring a significant input from a domain-area scientist. Results We present Poplar, a new, streamlined computational pipeline, to address the computational logistical issues that arise when constructing the phylogenomic trees. It provides a framework that runs state-of-the-art software for essential steps in the phylogenomic pipeline, beginning from a genome with or without an annotation, and resulting in a species tree. Running Poplar requires no external databases. In the execution, it enables parallelism for execution for clusters and cloud computing. The trees generated by Poplar match closely with state-of-the-art published trees. The usage and performance of Poplar is far simpler and quicker than manually running a phylogenomic pipeline. Availability and implementation Freely available on GitHub at https://github.com/sandialabs/poplar. Implemented using Python and supported on Linux.

Koning, Elizabeth [Sandia National Laboratories (S↗

SAIGE-GPU: accelerating genome- and phenome-wide association studies using GPUs

Genome-wide association studies (GWAS) at biobank scale are computationally intensive, especially for admixed populations requiring robust statistical models. SAIGE is a widely used method for generalized linear mixed-model GWAS but is limited by its CPU-based implementation, making phenome-wide association studies impractical for many research groups. We developed SAIGE-GPU, a GPU-accelerated version of SAIGE that replaces CPU-intensive matrix operations with GPU-optimized kernels. The core innovation is distributing genetic relationship matrix calculations across GPUs and communication layers. Applied to 2068 phenotypes from 635 969 participants in the Million Veteran Program, including diverse and admixed populations, SAIGE-GPU achieved a 5-fold speedup in mixed model fitting on supercomputing infrastructure and cloud platforms. We further optimized the variant association testing step through multi-core and multi-trait parallelization. Deployed on Google Cloud Platform and Azure, the method provided substantial cost and time savings. Source code and binaries are available for download at https://github.com/saigegit/SAIGE/tree/SAIGE-GPU-1.3.3. A code snapshot is archived at Zenodo for reproducibility (DOI: [10.5281/zenodo.17642591]). SAIGE-GPU is available in a containerized format for use across HPC and cloud environments and is implemented in R/C++ and runs on Linux systems.

Rodriguez, Alex [Argonne National Laboratory (ANL)↗

Genome evolution and transcriptome plasticity is associated with adaptation to monocot and dicot plants in Colletotrichum fungi

Colletotrichum fungi infect a wide diversity of monocot and dicot hosts, causing diseases on almost all economically important plants worldwide. Colletotrichum is also a suitable model for studying gene family evolution on a fine scale to uncover events in the genome associated with biological changes. Here we present the genome sequences of 30 Colletotrichum species covering the diversity within the genus. Evolutionary analyses revealed that the Colletotrichum ancestor diverged in the late Cretaceous in parallel with the diversification of flowering plants. We provide evidence of independent host jumps from dicots to monocots during the evolution of Colletotrichum, coinciding with a progressive shrinking of the plant cell wall degradative arsenal and expansions in lineage-specific gene families. Comparative transcriptomics of 4 species adapted to different hosts revealed similarity in gene content but high diversity in the modulation of their transcription profiles on different plant substrates. Combining genomics and transcriptomics, we identified a set of core genes such as specific transcription factors, putatively involved in plant cell wall degradation. These results indicate that the ancestral Colletotrichum were associated with dicot plants and certain branches progressively adapted to different monocot hosts, reshaping the gene content and its regulation.

59 BASIC BIOLOGICAL SCIENCES↗