Search NASA⌕ Search

SEARCH · Search NASA

Results for “random testing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Multivariate Testing of Sampling Techniques to Address Class Imbalance in Building Use Type Classification

This study addresses the challenges inherent in building use type classification, particularly focusing on the issue of class imbalance in the training datasets for machine learning classifiers. We comprehensively analyze the efficacy of various class-balancing sampling techniques. Employing Monte Carlo simulations and Bayesian optimization, we evaluated the performance of multiple sampling methods, including Random Oversampling, Random Undersampling, SMOTE, Borderline-SMOTE, and ADASYN, across a dataset encompassing nine southeastern coastal states of the United States. Our findings reveal that simple random over- and undersampling techniques outperform more sophisticated methods. Additionally, we show inherent value in creating an imbalance in training data to effectively train a machine learning classifier for distinguishing between residential and nonresidential buildings. This study provides valuable guidance for future research on building use type classification research and lays essential groundwork for developing attribute-rich building stock datasets.

Adams, Daniel↗

Architecture and performance of Perlmutter's 35 PB ClusterStor E1000 all-flash file system

NERSC's newest system, Perlmutter, features a 35 PB all-flash Lustre file system built on HPE Cray ClusterStor E1000. Here, we present its architecture, early performance figures, and performance considerations unique to this architecture. We demonstrate the performance of E1000 OSSes through low-level Lustre tests that achieve over 90% of the theoretical bandwidth of the SSDs at the OST and LNet levels. We also show end-to-end performance for both traditional dimensions of I/O performance (peak bulk-synchronous bandwidth) and nonoptimal workloads endemic to production computing (small, incoherent I/Os at random offsets) and compare them to NERSC's previous system, Cori, to illustrate that Perlmutter achieves the performance of a burst buffer and the resilience of a scratch file system. Finally, we discuss performance considerations unique to all-flash Lustre and present ways in which users and HPC facilities can adjust their I/O patterns and operations to make optimal use of such architectures.

97 MATHEMATICS AND COMPUTING↗

Electric dipole polarizability of 58 Ni

The electric dipole strength distribution in 58 Ni between 6 and 20 MeV has been determined from proton inelastic scattering experiments at very forward angles at RCNP, Osaka. The experimental data are rather well reproduced by quasiparticle random-phase approximation calculations including vibration coupling, despite a mild dependence on the adopted Skyrme interaction. They allow an estimate of the experimentally inaccessible high-energy contribution above 20 MeV, leading to an electric dipole polarizability α D ⁡ ( 58 Ni) = 3.48 (31)⁢ fm 3 . This serves as a test case for recent extensions of coupled-cluster calculations with chiral effective field theory interactions to nuclei with two nucleons on top of a closed-shell system.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Mock data sets for the Eboss and DESI Lyman-α forest surveys

We present a publicly-available code to generate sets of mock Lyman-α (Lyα) forest data that have realistic large-scale correlations including those due to the Baryonic Acoustic Oscillations (BAO). The primary purpose of these mocks is to test the analysis procedures of the Extended Baryon Oscillation Survey (eBOSS) and the Dark Energy Spectroscopy Instrument (DESI) surveys. The transmitted flux fraction, F(λ), of background quasars due to Lyα absorption in the intergalactic medium (IGM) is simulated using the Fluctuating Gunn-Petterson Approximation (FGPA) applied to Gaussian random fields produced through the use of fast Fourier transforms (FFT). The output includes the IGM-Lyα transmitted flux fraction along quasar lines of sight and a catalog of high-column-density systems appropriately placed at high-density regions of the IGM. This output serves as input to additional code that superimposes the IGM tranmission on realistic quasar spectra, adds absorption by high-column-density systems and metals, and simulates instrumental transmission and noise. Redshift space distortions (RSD) of the flux correlations are implemented by including the large-scale velocity-gradient field in the FGPA resulting in a correlation function of F(λ) that can be accurately predicted. One hundred realizations have been produced over the 14,000 deg 2 DESI survey footprint with 100 quasars per deg 2 . The analysis of these realizations shows that the correlations of F(λ) follows the prediction within the accuracy of eBOSS survey. Here, the most time-consuming part of the mock production occurs before application of the FGPA, and the existing pre-FGPA forests can be used to easily produce new mock sets with modified redshift-dependent bias parameters or observational conditions.

79 ASTRONOMY AND ASTROPHYSICS↗

Intrinsic nonlocality of spin- and polarization-resolved probabilities in strong-field quantum electrodynamics

Spin and polarization are central to precision tests of fundamental physics and for interpreting radiation from astrophysical sources and ultraintense laser-matter experiments. Here, focusing on the fundamental process of nonlinear Compton scattering, we demonstrate that a key assumption underlying current strong-field quantum electrodynamics models, i.e., that emission can be treated as an instantaneous random event sampled from a local differential rate, is inconsistent once emission angles, electron spin, and/or photon polarization are resolved. Namely, even in strictly constant and uniform fields , the resulting fully differential distribution is sign indefinite, yielding negative inferred probabilities. The physical reason is that the photon emission probability builds up over a finite length of the electron trajectory, the formation region, during which the electron direction changes by roughly the same small angle that defines the radiation cone. Therefore, we put forward a new method where we integrate over this formation region analytically to obtain a physically consistent electron spin and photon polarization model. We show that the implementation of our model is compatible with existing Monte Carlo and particle-in-cell workflows. Simulations of a GeV-class electron-laser collision accessible at current petawatt facilities and of emission in a pulsarlike magnetic field are shown to reveal spin and polarization patterns that differ even qualitatively from state-of-the-art local models. In particular, our new model predicts substantial angle-dependent circular photon polarization where the well-known collinear-emission approach yields none, and a pronounced helicity bias in the recoiling electrons absent from current predictions. These findings have direct implications for upcoming strong-field QED experiments and for interpreting polarized radiation from extreme astrophysical environments.

astrophysical electromagnetic fields↗

Probabilistic Analysis of Long-Term Degradation of Microwave Cavity Flow Sensor

We are investigating a microwave resonant cavity transducer for flow sensing in the vessel of a high temperature fluid advanced reactor (AR), such as a molten salt cooled reactor (MSCR) or a sodium fast reactor (SFR). This transducer is a hollow metallic cylindrical cavity, with the flat wall of the cylinder flexible enough to undergo microscopic deflection due to dynamic fluid pressure. Membrane deflection leads to a shift in the resonant frequency, which can be detected with a spectrum analyzer. We have performed a proof-of-concept experiment of flow sensing with the transducer in liquid sodium at 340°C in impinging liquid jet geometry. The transducer remained in liquid sodium for 70 days. After removal, no structural damage was observed, and the expected transducer response was verified in a water test. Because long-term (multi-year) experimental tests of transducer resilience to harsh environment are not practical, we have developed a probabilistic model of creep to estimate transducer resilience to the harsh environment. The probabilistic model considers diffusion creep under the condition of high temperature and low stress, where the stress and temperature are allowed to be random variables with Gaussian distributions. Using the probabilistic model, we estimate inelastic membrane deflections due to creep for several temperature ranges. We conclude that for temperatures less than 650°C, creep has negligible long-term effect on the transducer performance. Since a yellowish residue was observed on the transducer surface after 70 days of immersion in liquid sodium, we have investigated possible evidence of corrosion. Chromium depletion is a typical indicator of the corrosion process in stainless steel. Scraping off a residue from the transducer and performing scanning electron microscopy (SEM) with energy dispersive analysis (EDS) did not find any chromium in the residue. Approximately 60% of the residue consisted of copper, which can be attributed to contamination of sodium due to powder residue from machining of copper and brass components of the transducer.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Nanoionics Drastically Accelerating Mass Transfer at Elevated Temperatures over 750 °C

Nanoionics were previously considered thermally unstable and infeasible for devices operating above 500 °C. Here, we elucidate the design principle for establishing stable nanoionics from various oxides. We utilized reversible solid oxide cells (SOCs) as the test bed and implemented nanoionics using atomic layer deposition (ALD). We demonstrate a straightforward, interface-controlled, practical approach to render a conformal, ∼15 nm thick ALD film, which initially thermodynamically favors the formation of a solid solution with the substrate into surface nanoionics with single or double layers of nanograins with random crystal orientations. The nanoionics exhibited conductivity estimated to be 7 orders of magnitude higher than that of their bulk-scale counterpart. They demonstrated conformability with uniform grain sizes of ∼15 nm, even after electrochemical operation for ∼500 h at 750 °C and 1000 h at 850 °C. The thermal stability and conductivity of such nanoionics represent a conceptual and technological framework in nanoionics.

atomic layer deposition↗

RatXcan: A framework for cross-species integration of genome-wide association and gene expression data

Genome-wide association studies (GWAS) have implicated specific alleles and genes as risk factors for numerous complex traits. However, translating GWAS results into biologically and therapeutically meaningful discoveries remains extremely challenging. Most GWAS results identify noncoding regions of the genome, suggesting that differences in gene regulation are the major driver of trait variability. To better integrate GWAS results with gene regulatory polymorphisms, we previously developed PrediXcan (also known as “transcriptome-wide association studies” orTWAS), which maps SNPs to predicted gene expression using GWAS data. In this study, we developed RatXcan, a framework that extends this methodology to outbred heterogeneous stock (HS) rats. RatXcan accounts for the close familial relationships among HS rats by modeling the relatedness with a random effect that encodes the genetic relatedness. RatXcan also corrects for polygenic-driven inflation because of the equivalence between a relatedness random effect and the infinitesimal polygenic model. To develop RatXcan, we trained transcript predictors for 8,934 genes using reference genotype and expression data from five rat brain regions. We found that the cis genetic architecture of gene expression in both rats and humans was sparse and similar across brain tissues. We tested the association between predicted expression in rats and two example traits (body length and BMI) using phenotype and genotype data from 5,401 densely genotyped HS rats and identified a significant enrichment between the genes associated with rat and human body length and BMI. Thus, RatXcan represents a valuable tool for identifying the relationship between gene expression and phenotypes across species and paves the way to explore shared biological mechanisms of complex traits.

Genetics & Heredity↗

Data-Driven Closures and Assimilation for Stiff Multiscale Random Dynamics

Here, we introduce a data-driven and physics-informed framework for propagating uncertainty in stiff, multiscale random ordinary differential equations (RODEs) driven by correlated (colored) noise. Unlike systems subjected to Gaussian white noise, a deterministic equation for the joint probability density function (PDF) of RODE state variables does not exist in closed form. Moreover, such an equation would require as many phase-space variables as there are states in the RODE system. To alleviate this curse of dimensionality, we instead derive exact, albeit unclosed, reduced-order PDF (RoPDF) equations for low-dimensional observables/quantities of interest. The unclosed terms take the form of state-dependent conditional expectations, which are directly estimated from data at sparse observation times. However, for systems exhibiting stiff, multiscale dynamics, data sparsity introduces regression discrepancies that compound during RoPDF evolution. This is overcome by introducing a kinetic-like defect term to the RoPDF equation, which is learned by assimilating in sparse, low-fidelity RoPDF estimates. Two assimilation methods are considered, namely nudging and deep neural networks, which are successfully tested against Monte Carlo simulations.

97 MATHEMATICS AND COMPUTING↗

Estimating the CO 2 Fertilization Effect on Extratropical Forest Productivity From Flux‐Tower Observations

Abstract The land sink of anthropogenic carbon emissions, a crucial component of mitigating climate change, is primarily attributed to the CO 2 fertilization effect on global gross primary productivity (GPP). However, direct observational evidence of this effect remains scarce, hampered by challenges in disentangling the CO 2 fertilization effect from other long‐term confounding drivers, particularly climatic changes. Here, we introduce a novel statistical approach to separate the CO 2 fertilization effect on photosynthetic carbon uptake using eddy covariance (EC) records across 38 extratropical forest sites. We find the median stimulation rate of GPP to be 3.2 ± 0.9 gC m −2 yr −1 ppm −1 (or 16.4 ± 4.2% per 100 ppm) under increasing atmospheric CO 2 across these sites, respectively. To validate the robustness of our findings, we test our statistical method using factorial simulations of an ensemble of process‐based land surface models. We address additional factors, including nitrogen deposition and land management, that may impact plant productivity, potentially confounding the attribution to the CO 2 fertilization effect. Assuming these site‐specific effects offset to some extent across sites as random factors, the estimated median value still reflects the strength of the CO 2 fertilization effect. However, disentanglement of these long‐term effects, often inseparable by timescale, requires further causal research. Our study provides direct evidence that the photosynthetic stimulation is maintained under long‐term CO 2 fertilization across multiple EC sites. Such observation‐based quantification is key to constraining the long‐standing uncertainties in the land carbon cycle under rising CO 2 concentrations.

Environmental Sciences & Ecology↗

Evaluation of the Turbine Integrated Mortality Reduction (TIMR SM ) Technology as a Smart Curtailment Approach (Final Summary Report)

Wind energy is a crucial technology for achieving net-zero emissions by 2050. However, the growth and deployment of wind energy in North America have led to the deaths of many bat species due to operating wind turbines. Hundreds of thousands of bats are estimated to die at wind turbines annually in North America. Operational minimization, which includes feathering turbine blades and curtailment, has been documented to reduce bat fatality effectively. Curtailment refers to altering turbine operation based on wind speed, time of year, temperature, sensors, and activity models. However, when turbines are curtailed, they do not generate power, resulting in energy loss and revenue for wind energy facilities. The Electric Power Research Institute (EPRI) funded the development of Turbine Integrated Mortality Reduction (TIM SM ) Technology, which curtails turbine operation when bats are detected. The initial TIMR system research showed promising results, with an 85% reduction in overall bat fatalities and a 91% reduction for the little brown bat. However, these results were based on a single site during one fall season, and it was unclear if similar results could be replicated at other wind energy facilities. This research aimed to validate the TIMR system results from the prior field study at a second site in the U.S., estimate the power production and reduction in bat mortality at turbines with installed TIMR systems relative to blanket curtailment and fully operational turbines, test the TIMR system in two calendar years and during the summer and fall periods, and evaluate the operational and commercial characteristics of the TIMR system for potential wind industry adoption. The study was conducted at a 500.9-MW wind energy facility in southeast Adair County, Iowa. Three experimental treatments were involved in this randomized block design study: TIMR, Curtailment at 5.0 m/s, and Normal Operation. In 2021, three treatments were used at 18 turbines, expanding to four treatments across 36 turbines in 2022. The TIMR system worked as designed throughout the entire study; however, because of unexpected wind turbine operational challenges in 2021, there was not sufficient sample size to evaluate the treatment differences. In 2022, there were significant differences in fatality levels between treatment types and normal operating turbines. Curtailment at 5.0 m/s reduced fatalities by 30.8% compared to normal operations, and TIMR decreased fatalities by 48.6% compared to normal operations. Two different methods were used to evaluate the differences in energy loss for each treatment. The TIMR system resulted in 1.3% to 1.6 % annual energy loss in 2021 and 1.0% to 1.2 % in 2022. The Curtailment at 5.0 m/s resulted in 0.6% to 0.8 % annual energy loss in 2021 and 0.5% to 0.6 % in 2022. The project achieved all the stated objectives and demonstrated that TIMR is an effective technology that balances bat fatality reduction with energy generation. The results will support the deployment of TIMR and other acoustic sensor-based technologies. The research provides valuable insights into the impact of different treatments on fatality rates and energy outputs, contributing to the ongoing efforts to mitigate the environmental impact of wind energy.

17 WIND ENERGY↗

Quantifying mean, variability, and uncertainty in indoor radon exposure in Pennsylvania using random forest and quantile regression forest models

Radon is a naturally occurring radioactive gas that poses a serious health risk as the primary cause of lung cancer in non-smokers. Despite the well-known adverse association with health outcomes, current radon exposure assessments are limited to county-level or average-level estimates, which fail to capture regional variability. This study uses Machine Learning models, including Random Forest (RF) and Quantile Regression Forest (QRF), to estimate the indoor radon concentrations at the ZCTA (Zip code tabulation area)-level and characterize uncertainties in model estimates. Incorporating geological, meteorological, and building-specific data, the models aim to improve radon risk assessment by capturing mean exposure, variability, and extreme concentration levels. Processed radon test data (n = 718,111) were analyzed using average, variability, and quantile prediction methods. Models that estimate the average radon exposure at the ZCTA-level can yield promising model-fit results, but they do not capture the underlying variability of indoor radon exposure within a ZCTA. We utilize volatility analyses to identify characteristics indicative of high variability of indoor radon exposure. We also show that a QRF model can be used to estimate upper quantiles of residential radon exposure, thereby uncovering localized areas of elevated exposure that were not apparent in mean estimates. The results highlighted the need for a deep characterization of exposure risk and show that regions with moderate average exposure levels could still harbor extreme outliers with implications for evaluating health risks. Utilizing multiple radon exposure models allows for a deeper characterization of radon risk within a geographic area and can better identify high-risk areas. The results from this study provide a foundation for developing mitigation strategies and examining associations between radon exposure and health outcomes at fine scales. Future research should extend the geographic scope and incorporate additional environmental risk factors to establish a comprehensive framework for risk assessment.

Lee, Heechan [ORNL]↗

CMB low multipole alignments across WMAP and Planck data releases

ABSTRACT The first observations of the cosmic microwave background (CMB) from NASA's Wilkinson Microwave Anisotropy Probe (WMAP) led to finding ‘alignment’ anomalies not expected from fluctuations in the isotropic cosmological model. We study the data of all 8 full-sky public releases since then to test for anomalous alignments and shapes of the first 60 multipoles, i.e. over the range $2\le l \le 61$. We use rotationally invariant and covariant statistics to test isotropy of all subsequent WMAP data releases, along with those from the ESA’s Planck mission. Anomalous alignments among the multipoles $l=1, 2, 3$ are very consistent and robust. More alignments are detected, some of them new, while significance is diluted by the large range of the search. Power entropy, a measure of the randomness of the multipoles, is consistently anomalous at about $2\sigma$ level or better across all data releases. It appears that the CMB is not as random as the cosmological principle predicts on large angular scales.

Patel, Sanjeet Kumar↗

Adaptive Interface-PINNs (AdaI-PINNs) for inverse problems: Determining material properties for heterogeneous systems

Here, we determine spatially varying discontinuous material properties using a domain-decomposition based physics-informed neural networks (PINNs) framework named the Adaptive Interface-PINNs or AdaI-PINNs (Roy et al., 2024). We propose the use of distinct neural networks for the field variables and material properties within each material, utilizing adaptive activation functions. While the neural networks across different materials share the same weights and biases, their activation functions are uniquely tailored using a hyperparameter that influences the slope of the activation function. The proposed framework is tested on several one-dimensional and two-dimensional benchmark examples, and its performance is compared with conventional PINNs and existing domain-decomposition PINNs frameworks, namely, the Multi-domain physics-informed neural network (M-PINN), and the eXtended physics-informed neural networks (XPINNs). The results demonstrate that the proposed approach can determine randomly distributed discontinuous material properties with an L 2 error of $\mathscr{O}$ (10 -3 ) for the material property and the root-mean-square error of $\mathscr{O}$ (10 -3 ) for the primary variable while the other approaches yield errors that are approximately two orders of magnitude larger (that is, $\mathscr{O}$ (10 -1 )). Moreover, the spatial distribution of material properties obtained using the proposed framework is in close agreement with the true distribution, whereas the other approaches fare much worse. Additionally, the proposed approach is approximately 40% faster than its competitors, indicating its potential as a robust alternative for solving inverse problems in heterogeneous materials.

36 MATERIALS SCIENCE↗

Efficient online quantum circuit learning with no upfront training

Optimization is a promising candidate for studying the utility of variational quantum algorithms (VQAs). However, evaluating cost functions using quantum hardware introduces runtime overheads that limit exploration. Surrogate-based methods can reduce calls to a quantum computer, yet existing approaches require hyperparameter pre-training and have been tested only on small problems. Here, we show that surrogate-based methods can enable successful optimization at scale, without pre-training, by using radial basis function interpolation (RBF) to construct an adaptive, hyperparameter-free surrogate. Using the surrogate as an acquisition function drives hardware queries to the vicinity of the true optima. For 16-qubit random 3-regular Max-Cut instances with the Quantum Approximate Optimization Algorithm (QAOA), our method outperforms state-of-the-art approaches, without considering their upfront training costs. Furthermore, we successfully optimize QAOA circuits for 127-qubit random Ising models on an IBM processor using 10 4 −10 5 measurements. Strong empirical performance demonstrates the promise of automated surrogate-based learning for large-scale VQA applications.

97 MATHEMATICS AND COMPUTING↗

Efficient Decision Trees for Tensor Regressions

Here, we proposed the tensor-input tree (TT) method for scalar-on-tensor and tensor-on-tensor regression problems. We first address scalar-on-tensor problem by proposing scalar-output regression tree models whose input variables are tensors (i.e., multi-way arrays). We devised and implemented fast randomized and deterministic algorithms for efficient fitting of scalar-on-tensor trees, making TT competitive against tensor-input GP models (Yu, Li, and Liu; Sun et al.). Based on scalar-on-tensor tree models, we extend our method to tensor-on-tensor problems using additive tree ensemble approaches. Theoretical justification and extensive experiments, including testing robustness to entrywise input tensor noise, are provided on real and synthetic datasets to illustrate the performance of TT. Our implementation is provided at https://github.com/hrluo/TensorDecisionTreeRegressor. Supplementary materials for this article are available online.

Decision tree regressions↗

CaloTrilogy: Toward a Breakthrough in One-Step, End-to-End, Physics-Guided Shower Generation for Modern Calorimeters

High-precision calorimeter simulation at current and future colliders imposes rapidly growing computational demands, motivating the development of machine-learning surrogates for traditional Monte Carlo tools such as Geant4. Flow matching and diffusion-based generative models have become leading approaches for high-dimensional fast simulation because of their sample quality, but typically require ${\cal O}(100)$ function evaluations at inference and often rely on auxiliary networks to constrain global observables, compromising streamlined end-to-end generation. We introduce a unified framework that improves the balance between speed, shower quality, and physics fidelity. The method combines: (i) an average velocity field integrator that enables sampling in one or a few evaluations; (ii) a learned generative prior in shower space, constructed from data rather than random noise; and (iii) physics-guided loss terms that impose inductive biases on key observables during training. These elements are training time regularizers, preserving end-to-end inference with no additional cost. With only one or a few evaluation steps, the model achieves shower quality competitive with state-of-the-art flow and diffusion approaches, tested on several public high granularity calorimeter datasets. The results demonstrate inter-layer shower structure consistent with the underlying physics, providing a strong candidate for future fast simulation workflows.

Jiang, Cheng [Edinburgh U.]↗

ATAT: Astronomical Transformer for time series and Tabular data

Context. The advent of next-generation survey instruments, such as theVera C. RubinObservatory and its Legacy Survey of Space and Time (LSST), is opening a window for new research in time-domain astronomy. The Extended LSST Astronomical Time-Series Classification Challenge (ELAsTiCC) was created to test the capacity of brokers to deal with a simulated LSST stream. Aims. Our aim is to develop a next-generation model for the classification of variable astronomical objects. We describe ATAT, the Astronomical Transformer for time series And Tabular data, a classification model conceived by the ALeRCE alert broker to classify light curves from next-generation alert streams. ATAT was tested in production during the first round of the ELAsTiCC campaigns. Methods. ATAT consists of two transformer models that encode light curves and features using novel time modulation and quantile feature tokenizer mechanisms, respectively. ATAT was trained on different combinations of light curves, metadata, and features calculated over the light curves. We compare ATAT against the current ALeRCE classifier, a balanced hierarchical random forest (BHRF) trained on human-engineered features derived from light curves and metadata. Results. When trained on light curves and metadata, ATAT achieves a macro F1 score of 82.9 ± 0.4 in 20 classes, outperforming the BHRF model trained on 429 features, which achieves a macro F1 score of 79.4 ± 0.1. Conclusions. The use of transformer multimodal architectures, combining light curves and tabular data, opens new possibilities for classifying alerts from a new generation of large etendue telescopes, such as theVera C. RubinObservatory, in real-world brokering scenarios.

Astronomy & Astrophysics↗