Search NASASearch

SEARCH · Search NASA

Results for “Model Counting”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Morphotype-resolved characterization of microalgal communities in a nutrient recovery process with ARTiMiS flow imaging microscopy

Microalgae-driven nutrient recovery represents a promising technology for phosphorus removal from wastewater while simultaneously generating biomass that can be valorized to offset treatment costs. As full-scale processes come online, system parameters including biomass composition must be carefully monitored to optimize performance and prevent culture crashes. In this study, flow imaging microscopy (FIM) was leveraged to characterize microalgal community composition in near real-time at a full-scale municipal wastewater treatment plant (WWTP) in Wisconsin, USA, and population and morphotype dynamics were examined to identify relationships between water chemistry, biomass composition, and system performance. Two FIM technologies, FlowCam and ARTiMiS, were evaluated as monitoring tools. ARTiMiS provided a more accurate estimate of total system biomass, and estimates derived from particle area as a proxy for biovolume yielded better approximations than particle counts. Deep learning classification models trained on annotated image libraries demonstrated equivalent performance between FlowCam and ARTiMiS, and convolutional neural network (CNN) classifiers proved significantly more accurate when compared to feature table-based dense neural network (DNN) models. Across a two-year study period, Scenedesmus spp. appeared most important for phosphorus removal, and were negatively impacted by elevated temperatures and increase in nitrite/nitrate concentrations. Chlorella and Monoraphidium also played an important role in phosphorus removal. For both Scenedesmus and Chlorella, smaller morphological types were more often associated with better system performance, whereas larger morphotypes likely associated with stress response(s) correlated with poor phosphorus recovery rates. Furthermore, these results demonstrate the potential of FIM as a critical technology for high-resolution characterization of industrial microalgal processes.

59 BASIC BIOLOGICAL SCIENCES

Two-Tower Quantum Matrix Chain Multiplication: Trading Qubits for Depth

Matrix chain multiplication -- computing $\mathcal{W} = M^{(0)}\cdots M^{(K-1)}$ where $M^{(k)} \in \mathbb{R}^{P_k \times P_{k+1}}$-- arises in scientific computing, machine learning, and graph analysis. Despite the importance of this problem, for chains of distinct matrices, the classical number of operations grows linearly with the chain length $K$ and polynomially in the matrix dimensions. We present \emph{Two-Tower Matrix Multiplication}, a quantum subroutine that encodes the product $\mathcal{W}$ of the $K$ matrices into a quantum state in circuit depth $\mathcal{O}(\max_{k} \mathrm{polylog} (P_k P_{k+1}))$, which is independent of~$K$ within the QRAM-based state-preparation model, whereas the qubit count is $\mathcal{O}\bigl(\sum_{k} \log P_k \bigr)$; the total gate count remains linear in $K$, so the gain is in the circuit depth. The construction interleaves state-preparation operators across two layers; within each layer, all operators act on disjoint registers and execute in parallel. This subroutine can be specialized for the chain-vector case, which computes the product of $K-1$ matrices applied to a vector. We prove the correctness of the subroutine for all $K$ and provide two implementations using the Qiskit and QCLAB frameworks. The subroutine is applicable to any downstream quantum algorithm that operates on a matrix encoded in the statevector, including norm estimation, graph-matrix powers, linear system solving, and quantum machine learning kernels.

Antonioli, Giacomo [Pisa U.] (ORCID:00090000668703

Allometric and Mobile Terrestrial LiDAR Modeling of Aboveground Woody Biomass of Populus in Coppice Production

Poplars ( Populus spp.) and their hybrids are increasingly being grown in coppice production to generate bioenergy feedstocks at frequent intervals. Allometric equations are re-quired to predict aboveground biomass (AGB) of coppiced individuals with minimal field measurements. Likewise, remote sensing tools like LiDAR (light detection and ranging) can be used if models are available to predict AGB from point cloud data. Therefore, this study sought to develop equations to predict dry woody AGB from field measurements and LiDAR data from coppiced poplar field trials containing eastern cottonwood ( P. del-toides ) and hybrid poplar taxa. We found that taxa-specific allometric models containing the summed basal area of the three largest stems in the coppice provided the best predictive model, with stem height and stem count failing to provide additional explanatory power. The best predictive LiDAR-based model was independent of taxa but had slightly lower adjusted R 2 and higher RMSE than the allometric model. It contained four parameters including crown volume, leaf area index, variance of height returns, and the top point density (i.e., density metric 9 or the proportion of points in the highest point interval when the point cloud is evenly divided into ten vertical intervals). In total, these models can be used to quickly and efficiently estimate dry woody AGB of Populus coppice systems for bioenergy feedstock production.

AGB

Tensor decompositions for count data that leverage stochastic and deterministic optimization

There is growing interest to extend low-rank matrix decompositions to multi-way arrays, or tensors. One fundamental low-rank tensor decomposition is the canonical polyadic decomposition (CPD). The challenge of fitting a low-rank, nonnegative CPD model to Poisson-distributed count data is of particular interest. Several popular algorithms use local search methods to approximate the maximum likelihood estimator (MLE) of the Poisson CPD model. Here, this work presents two new algorithms that extend state-of-the-art local methods for Poisson CPD. Hybrid GCP-CPAPR combines Generalized Canonical Decomposition (GCP) with stochastic optimization and CP Alternating Poisson Regression (CPAPR), a deterministic algorithm, to increase the probability of converging to the MLE over either method used alone. Restarted CPAPR with SVDrop uses a heuristic based on the singular values of the CPD model unfoldings to identify convergence toward optimizers that are not the MLE and restarts within the feasible domain of the optimization problem, thus reducing overall computational cost when using a multi-start strategy. We provide empirical evidence that indicates our approaches outperform existing methods with respect to converging to the Poisson CPD MLE.

CPAPR

Soil biogeochemical properties and metrics of tree-mycorrhizal dominance for a 25-Ha forest in South Central Indiana, USA.

This data package contains a dataset used in the papers “Seeing the forest for all the trees: Mycorrhizal-associated nutrient economies are modulated by stem density and the synchrony between overstory and understory communities” and “Mycorrhizal associations of tree species influence soil nitrogen dynamics via effects on soil acid–base chemistry”. Four csv files are included along with a dataset. The dataset features chemical soil properties for a single sampling campaign within the 25 Ha Lilly-Dickey Woods Smithsonian Forest Global Earth Observatory (ForestGEO) plot in South Central Indiana, USA (ldw_dat_raw.csv). Also included are separate files focused on pH (pH_data.csv), carbon and nitrogen (CN_data.csv), and nitrification rates (Nitrification_data.csv). These variables are commonly associated with the tree-mycorrhizal dominance of forest stands. In these data subsets, each soil variable was matched to a 10 meter radius neighborhood wherein metrics of tree-mycorrhizal dominance (basal area, stem count, importance value, etc.) were calculated. Models between these soil variables and dominance metrics were used to investigate how different assessments of mycorrhizal associated nutrient economies (MANE) capture these relationships. This research was performed as a part of the Smithsonian ForestGEO project. This data package can be used to explore spatial variability in soil chemistry within a mature hardwood forest, or it can be combined with the included tree data, other fine-scale spatial information, or other tree inventory data for the site to evaluate how soil chemistry varies with tree community composition or edaphic or topographic properties.

Craig, Matthew [ORNL] (ORCID:0000000288907920)

StreamGen: Connecting Populations of Streams and Shells to Their Host Galaxies

In this work, we study how the abundance and dynamics of populations of disrupting satellite galaxies change systematically as a function of host galaxy properties. We apply a theoretical model of the phase-mixing process to classify intact satellite galaxies and stellar streamlike and shell-like debris in ∼1500 Milky Way–mass systems generated by a semi-analytic galaxy formation code, SatGen. In particular, we test the effect of host galaxy halo mass, disk mass, ratio of disk scale height to length, and stellar feedback model on disrupting satellite populations. We find that the counts of tidal debris are consistent across all host galaxy models, within a given host mass range, and that all models can have streamlike debris on low-energy orbits, consistent with that observed around the Milky Way. However, we find a preference for streamlike debris on lower-energy orbits in models with a thicker (lower-density) host disk or on higher-energy orbits in models with a more massive host disk. Importantly, we observe significant halo-to-halo variance across all models. These results highlight the importance of simulating and observing large samples of Milky Way–mass galaxies and accounting for variations in host properties when using disrupting satellites in studies of near-field cosmology.

dark matter

Informing solar blind radioluminescence imaging through a calibrated spectrum

While direct radiation detection methods offer great insight into the origin of ionizing particles and photons, their use to locate contaminated areas or concealed radioactive sources can lead to undue exposure of personnel and equipment to ionizing radiation or the potential for contamination. These same sources induce ultraviolet (UV) optical photon fluorescence in air – a process referred to as radioluminescence – that may be imaged from low dose regions over larger attenuation lengths than ionizing radiation. However, most optical detection methods are limited to low lighting conditions to image the more abundant ultraviolet-A (UV-A) photons. To extend this capability to room light or daytime conditions, the solar blind region (Ultraviolet-C (UV-C), <280 nm) can be tapped. Though the emission yield of UV-C photons is roughly two orders of magnitude lower than that in the UV-A regime, the UV-C offers dramatic improvements in signal-to-noise ratios under bright lighting conditions due to decreased background interferences. The yield of specific UV-C lines, if present in the literature at all, varies widely, which has a large impact in modeling and analyzing standoff UV-C measurement scenarios. Thus, we have captured improved radioluminescence spectra over 250–400 nm and identified observed emission peaks in ambient air. Many of the UV-C photons produced by ionizing radiation excitation result from high-energy, molecular nitrogen Gaydon-Herman transitions which have had limited study to date for this application. Relating these findings to published UV-A yields, we estimate emissions between 0.15–0.19 photons/MeV over 250–280 nm. Additionally, we also use a commercial corona-discharge imaging camera to demonstrate outdoor UV-C radiation mapping of alpha and gamma emitters from 50 and 75 m standoffs, respectively. The imaged “counts” are compared to optically modelled values and show the same trend over distance.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND

Detecting outbreaks using a spatial latent field

In this paper, we present a method for estimating the infection-rate of a disease as a spatial-temporal field. Our data comprises time-series case-counts of symptomatic patients in various areal units of a region. We extend an epidemiological model, originally designed for a single areal unit, to accommodate multiple units. The field estimation is framed within a Bayesian context, utilizing a parameterized Gaussian random field as a spatial prior. We apply an adaptive Markov chain Monte Carlo method to sample the posterior distribution of the model parameters condition on COVID-19 case-count data from three adjacent counties in New Mexico, USA. Our results suggest that the correlation between epidemiological dynamics in neighboring regions helps regularize estimations in areas with high variance (i.e., poor quality) data. Using the calibrated epidemic model, we forecast the infection-rate over each areal unit and develop a simple anomaly detector to signal new epidemic waves. Our findings show that anomaly detector based on estimated infection-rates outperforms a conventional algorithm that relies solely on case-counts.

Safta, Cosmin [Sandia National Laboratories (SNL-C

An Integral-based Technique to Accelerate the Monte Carlo Radiative Transfer Computation for Supernovae

We present an integral-based technique (IBT) algorithm to accelerate supernova (SN) radiative transfer calculations. The algorithm utilizes “integral packets,” which are calculated by the path integral of the Monte Carlo (MC) energy packets, to synthesize the observed spectropolarimetric signal at a given viewing direction in a 3D time-dependent radiative transfer program. Compared to the event-based technique (EBT) proposed by M. Bulla et al., our algorithm significantly reduces the computation time and increases the MC signal-to-noise ratio (S/N). Using a 1D spherical symmetric Type Ia SN ejecta model DDC10 and its derived 3D model, the IBT algorithm has successfully passed the verification of spherical symmetry and cross comparison on a 3D SN model with the direct-counting technique and EBT. Notably, with our algorithm implemented in the 3D MC radiative transfer code SEDONA, the computation time is faster than EBT by a factor of 10−30, and the S/N is better by a factor of 1.5−3, with the same number of MC quanta.

79 ASTRONOMY AND ASTROPHYSICS

Uranium particle age dating, aggregation, and model age best estimators

We present important aspects of uranium particle age dating by Large-Geometry Secondary Ion Mass Spectrometry (LG-SIMS) that can introduce bias and increase model age uncertainties, especially for small, young, and/or low-enriched particles. This metrology is important for applications related to International Nuclear Safeguards. We explore influential factors related to model age estimation, including the effects of evolving surface chemistry on inter-element measurements of particles (e.g., Th and U), detector background, and aggregation methods using simulated and actual particle samples. We introduce a new model age estimator, called “mid68”, that supplements 95% confidence intervals, providing a “best estimate” and uncertainty about the most likely age. The mid68 estimator can be calculated using the Feldman and Cousins method or Bayesian methods and provides a value with a symmetric uncertainty that can be used for calculations and approximate aggregation of processed model age values when the raw data and correction factors are not available. For particles yielding low 230 Th counts amidst nonzero detector background, their underlying model age probability distributions are asymmetric, so the mid68 estimator provides additional robust information regarding the underlying model age likelihood. This study provides a comprehensive and timely examination of critical aspects of uranium particle age dating as more laboratories establish particle chronometry capabilities.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Bounds on galaxy stochasticity from halo occupation distribution modeling

The joint probability distribution of matter overdensity and galaxy counts in cells is a powerful probe of cosmology, and the extent to which variance in galaxy counts at fixed matter density deviates from Poisson shot noise is not fully understood. The lack of informed bounds on this stochasticity is currently the limiting factor in constraining cosmology with the galaxy–matter probability distribution function (PDF). We investigate stochasticity in the conditional distribution of galaxy counts along lines of sight with fixed matter density, and we present a halo occupation distribution (HOD)-based approach for obtaining plausible ranges for stochasticity parameters. To probe the high-dimensional space of possible galaxy–matter connections, we derive a set of HODs that conserve the galaxies’ linear bias and number density to produce RED M A G I C-like galaxy catalogs within the A BACUS S UMMIT suite of N -body simulations. We study the impact of individual HOD parameters and cosmology on stochasticity and perform a Monte Carlo search in HOD parameter space subject to the constraints on bias and density. In mock catalogs generated by the selected HODs, shot noise in galaxy counts spans both sub-Poisson and super-Poisson values, ranging from 80% to 133% of Poisson variance for cells with mean matter density. Nearly all of the derived HODs show a positive relationship between local matter density and stochasticity. For galaxy catalogs with higher stochasticity, modeling galaxy bias to second order is required for an accurate description of the conditional PDF of galaxy counts at fixed matter density. The presence of galaxy assembly bias also substantially extends the range of stochasticity in the super-Poisson direction. This HOD-based approach leverages degrees of freedom in the galaxy–halo connection to obtain informed bounds on nuisance model parameters and can be adapted to study other parametrizations of shot noise in galaxy counts, in particular to motivate prior ranges on stochasticity for cosmological analyses.

Britt, Dylan (ORCID:000000019905601X)

The Poisson tensor completion non-parametric differential entropy estimator

We introduce the Poisson tensor completion (PTC) estimator, a non-parametric differential entropy estimator. The PTC estimator leverages inter-sample relationships to compute a low-rank Poisson tensor decomposition of the frequency histogram. Our crucial observation is that the histogram bins are an instance of a space partitioning of counts and thus can be identified with a spatial Poisson process. The Poisson tensor decomposition leads to a completion of the intensity measure over all bins—including those containing few to no samples—and leads to our proposed PTC differential entropy estimator. A Poisson tensor decomposition models the underlying distribution of the count data and guarantees non-negative estimated values and so can be safely used directly in entropy estimation. Our estimator is the first tensor-based estimator that exploits the underlying spatial Poisson process related to the histogram explicitly when estimating the probability density with low-rank tensor decompositions for the purpose of tensor completion. Furthermore, we demonstrate that our PTC estimator is a substantial improvement over standard histogram-based estimators for sub-Gaussian probability distributions because of the concentration of norm phenomenon.

42 ENGINEERING

The Poisson tensor completion parametric estimator

We introduce the Poisson tensor completion (PTC) estimator that exploits inter-sample relationships to compute a low-rank Poisson tensor decomposition of the frequency histogram for samples of a multivariate distribution. Our crucial observation is that the histogram bins are an instance of a space partitioning of counts and thus can be identified with a spatial non-homogeneous Poisson process. The Poisson tensor decomposition leads to a completion of the mean measure over all bins—including those containing few to no samples—and leads to our proposed estimator. A Poisson tensor decomposition models the underlying distribution of the count data and guarantees non-negative estimated values obviating the need for additional constraints to ensure non-negativity. Furthermore, we demonstrate that our PTC estimator is a substantial improvement over standard histogram-based estimators for sub-Gaussian probability distributions because of the concentration of norm phenomenon.

97 MATHEMATICS AND COMPUTING

LandScan mosaic enables high-resolution gridded population estimates with explicit uncertainty

Gridded population datasets represent high-resolution distributions of human occupancy, enabling informed decision-making across a broad range of fields. These data products are valuable for assessing environmental risk, urban development, disaster preparedness and resource allocation—areas where accurate population estimates directly enhance policy effectiveness and optimize resource distribution. Despite the importance of gridded population datasets, traditional population modeling approaches often overlook inherent uncertainties in the estimation process. This limitation can create a false sense of certainty in population estimates, potentially leading to flawed decisions by those who rely on the data. To address this methodological gap, we introduce a probabilistic machine learning modeling framework, LandScan Mosaic, that explicitly incorporates uncertainty into the population modeling process. Our approach systematically quantifies uncertainty in three key modeling parameters of the LandScan HD gridded population dataset: building use types, floor counts, and occupancy rates. By employing Monte Carlo simulations, we propagate these uncertainties through the modeling process, yielding probability distributions of population counts in place of deterministic point estimates. We demonstrate the practical application of this framework in Iloilo City, Philippines, using structured decision-making techniques and our probabilistic estimates to identify and prioritize areas most affected by projected flooding, supporting targeted interventions that address both economic and social risks. In doing so, we propose a population-specific approach for incorporating confidence into structured decision making processes. Through a comparative analysis with conventional deterministic approaches and point estimate approaches, including LandScan HD and WorldPop, we evaluate how the incorporation of machine learning and uncertainty influences decision rankings. This research advances population distribution modeling by offering a robust, quantitative approach that explicitly accounts for uncertainty in the underlying data, along with guidance for how users can apply uncertainty in their decision-making.

Environmental sciences

Optical galaxy cluster mock catalogs with realistic projection effects: Validations with the SDSS clusters

Galaxy clusters identified in optical imaging surveys suffer from projection effects: Physically unassociated galaxies along a cluster’s line of sight can be counted as its members and boost the observed richness (the number of cluster members). To model the impact of projection on cluster cosmology analyses, we apply a halo occupation distribution model to 𝑁-body simulations to simulate the red galaxies contributing to cluster members, and we use the number of galaxies in a cylinder along the line of sight (counts in cylinders) to model the impact of projection on cluster richness. We compare three projection models: uniform, quadratic, and Gaussian, and we convert between them by matching their effective cylinder volumes. We validate our mock catalogs using SDSS redMaPPer clusters’ data vectors, including counts vs richness, stacked lensing signal, spectroscopic redshift distribution of member galaxies, and richness remeasured on a redshift grid. We find the former two are insensitive to the projection model, while the latter two favor a quadratic projection model with a width of ≈180 ℎ −1 Mpc (equivalent to the volume of a uniform model with a width of 100 ℎ −1 Mpc and a Gaussian model with a width of 110 ℎ −1 Mpc, or a Gaussian redshift error of 0.04). Furthermore, our framework provides an efficient and flexible way to model optical cluster data vectors, paving the way for a simulation-based joint analysis for clusters, galaxies, and shear.

79 ASTRONOMY AND ASTROPHYSICS

Arctic Impact Identification with Less Data Using Variable Relationships: An Exploratory Express LDRD project.

Regional impacts from sea ice loss can be challenging to separate from internal climate variability, potentially requiring thousands of ensemble members. East Asian wintertime cooling has been linked to sea ice loss from present day conditions in the Polar Amplification Model Intercomparison Project with these large ensemble counts. This cooling is theorized to arise from a strengthened Siberian High and East Asian Jet response. The strengthened Siberian High can be detected with one fifth the ensemble members needed for the East Asian wintertime cooling in a single model. We thus hypothesize that leveraging relationships between multiple variables in a conditional pathways-based approach would reduce the number of required ensemble members to conclusively attribute East Asian wintertime cooling to future sea ice concentrations. In all analyzed cases, confidence was increased when evaluating sea ice loss’s responsibility for the joint effects of East Asian cooling, East Asian Jet strengthening, and Siberian High strengthening over just East Asian cooling. However, we were not able to confidently attribute future East Asian wintertime cooling to sea ice loss in a single model. We found that significant intra-ensemble variability within single Earth System Models (ESMs) produced highly uncertain forcing response models upon which attribution results were undermined. We were able to show that ensemble mean seasonally averaged metrics from multiple ESMs greatly improved the accuracy of the forcing response linear models and exposed the necessity of all three steps in the pathway (sea ice area, Siberian High pressure, and East Asian Jet speed) for accurate prediction of East Asian wintertime cooling. Although all three steps were necessary, East Asian wintertime cooling possesses a large dependence on the Siberian High pressure, which weakens the confidence associated with overall strong joint-attribution comparing present day and future scenarios. We believe transitioning the pathway nodes to relative changes between the Siberian High and Aleutian Low as well as between the midlatitude westerlies and subtropical jet in the East Asianj Jet region may be able to produce significant attribution more fully dependent upon all three steps. Ultimately, this research demonstrates the simple extensibility of conditional pathways-based attribution to sea ice loss forcing on the Earth system.

54 ENVIRONMENTAL SCIENCES

Computer vision-based rock bolt detection in orthomosaic imagery obtained in the Waste Isolation Pilot Plant underground facility

Assessing structural integrity of large underground tunnel facilities is often a time consuming and human-labor intensive task. Thus, research using various modes of sensing and automated detection of key structural components in mines is posed to aid in safety assessments and establishing overall structural health. We propose an approach utilizing off-the-shelf camera and lidar technology fixed to a custom sensing platform, image stitching techniques, fine-tuned object detection models, and specialized model-inference methods to automatically detect, count, and map roof bolts for assessment of structural safety in man-made underground tunnels. Results show a novel workflow for effective object counting in orthomosaic tunnel ceiling images generated from collections in GPS-denied mining environments. Additionally, we demonstrate effective fine-tuning of EfficientDet object detectors utilizing state-of-the-art image augmentation techniques known as the mosaic and mixup transformations. Our work is demonstrated on sensed data and imagery collected from the Department of Energy (DOE) Waste Isolation Pilot Plant (WIPP) where miles of tunnel ceiling must be assessed for structural integrity.

42 ENGINEERING

MoE-Inference-Bench: Performance Evaluation of Mixture of Expert Large Language and Vision Models

Mixture of Experts (MoE) models have enabled the scaling of Large Language Models (LLMs) and Vision Language Models (VLMs) by achieving massive parameter counts while maintaining computational efficiency. However, MoEs introduce several inference-time challenges, including load imbalance across experts and the additional routing computational overhead. To address these challenges and fully harness the benefits of MoE, a systematic evaluation of hardware acceleration techniques is essential. We present MoE-Inference-Bench, a comprehensive study to evaluate MoE performance across diverse scenarios. We analyze the impact of batch size, sequence length, and critical MoE hyperparameters such as FFN dimensions and number of experts on throughput. We evaluate several optimization techniques on Nvidia H100 GPUs, including pruning, Fused MoE operations, speculative decoding, quantization, and various parallelization strategies. Our evaluation includes MoEs from the Mixtral, DeepSeek, OLMoE and Qwen families. The results reveal performance differences across configurations and provide insights for the efficient deployment of MoEs.

Chitty-Venkata, Krishna Teja