Search NASA⌕ Search

SEARCH · Search NASA

Results for “normalization method”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Using DAPPER to extract the photon strength function of 58 Fe using the inverse Oslo and shape methods

The photon strength function of 58 Fe has been extracted using both the Oslo and Shape methods from particle–γ coincidence data measured using the Detector Array for Photons, Protons, and Exotic Residues, which probes nuclei utilizing (d,p) reactions in inverse kinematics. Four particle–γ coincidence matrices, each constructed with different treatments of the γ–ray energies, are explored in order to observe the impact on the resulting nuclear level density and photon strength. The final photon strength function reported is found to agree well with previous Oslo measurements of other iron isotopes. Systematic uncertainties are included, using different model parameters and their reported errors to perform the Oslo method normalization. The model-independent Shape method is explored and the functional form of the photon strength function obtained is in agreement with the Oslo method results. A low-energy enhancement is not reported for 58 Fe in this work given possible subtraction issues originating from strongly populated states.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Using DAPPER to extract the photon strength function of 58Fe using the inverse Oslo and shape methods

The photon strength function of 58 Fe has been extracted using both the Oslo and Shape methods from particle-γ coincidence data measured using the Detector Array for Photons, Protons, and Exotic Residues, which probes nuclei utilizing (d,p) reactions in inverse kinematics. Four particle-γ coincidence matrices, each constructed with different treatments of the γ-ray energies, are explored in order to observe the impact on the resulting nuclear level density and photon strength. The final photon strength function reported is found to agree well with previous Oslo measurements of other iron isotopes. Systematic uncertainties are included, using different model parameters and their reported errors to perform the Oslo method normalization. The model-independent Shape method is explored and the functional form of the photon strength function obtained is in agreement with the Oslo method results. A low-energy enhancement is not reported for 58 Fe in this work given possible subtraction issues originating from strongly populated states.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

OmicsMLMentor: A Web Application for Guided Machine Learning Analysis of Omics Data

Expression-based omics technologies (e.g. proteomics, metabolomics, transcriptomics, etc.) increasingly rely on supervised and unsupervised machine learning (ML) models to find key biomolecules distinguishing conditions, identify natural groupings in biological data, or generate predictions for outcomes of interest. Fitting ML models to omics data presents several challenges, including handling missing data, selecting a normalization method, choosing a valid model, and optimizing hyperparameters, all requiring statistical programming skills to address these challenges. Thus, the open-source web application SLOPE was designed to lower the barrier to ML modeling for omics data. SLOPE supports the fitting of 15 ML models (10 supervised and 5 unsupervised) tailored to omics datasets, such as proteomics, metabolomics, lipidomics, and transcriptomics. SLOPE offers several omics-specific features, including methods for handling missingness (imputation, conversion, removal), normalization tests, ranking of models based on the structure of a user’s data and user input, and optimal hyperparameter selections using cross-validation splits. By streamlining ML workflows for omics analysis, SLOPE address critical gaps in existing online web tools, facilitating a broader adoption of these models for omics research. Here, SLOPE is applied to data from a lignin exposure study to highlight the workflow for fitting both supervised and unsupervised models to data.

lipidomics↗

Evaluation of normalization strategies for mass spectrometry-based multi-omics datasets

Introduction Data normalization is crucial for multi-omics integration, reducing systematic errors and maximizing the likelihood of discovering true biological variation. Most studies assess normalization for a single omics type or use datasets from separate experiments. Few address time-course data, where normalization might bias temporal differentiation. In this study, we compared common normalization methods and a machine learning approach, Systematical Error Removal using Random Forest (SERRF), using multi-omics datasets generated from the same experiment—even from the same cell lysate. Objectives To develop a straightforward process to assess normalization effects and identify the most robust methods across multi-omics datasets. Methods We analyzed metabolomics, lipidomics, and proteomics datasets from primary human cardiomyocytes and motor neurons exposed to acetylcholine-active compounds over time. Normalization effectiveness was evaluated based on improvement in QC features consistency and observing the change in treatment and time-related variance. Results Probabilistic Quotient Normalization (PQN) and Locally Estimated Scatterplot Smoothing (LOESS) QC were identified as optimal for metabolomics and lipidomics, while PQN, Median, and LOESS normalization excelled for proteomics. These methods consistently enhanced QC feature consistency in metabolomics and lipidomics, and preserved time-related variance or treatment-related variance in proteomics, demonstrating their effectiveness and robustness. SERRF normalization, applied only to metabolomics in this study, outperformed other methods in some datasets but inadvertently masked treatment-related variance in others. Conclusion Our evaluation identified PQN and LoessQC as the top methods for metabolomics and lipidomics, and PQN, Median, and Loess normalization for proteomics, in multi-omics integration in a temporal study.

60 APPLIED LIFE SCIENCES↗

Causes and consequences of experimental variation in Nicotiana benthamiana transient expression

Infiltration of Agrobacterium tumefaciens into Nicotiana benthamiana has become a foundational technique in plant biology, enabling efficient delivery of transgenes in planta with technical ease, robust signal, and relatively high throughput. Despite transient expression’s prevalence in disciplines such as synthetic biology, little work has been done to describe and address the variability inherent in this system, a concern for experiments that rely on highly quantitative readouts. In a comprehensive analysis of N. benthamiana agroinfiltration experiments, we model sources of variability that affect transient expression. Our findings emphasize the need to validate normalization methods under the specific conditions of each study, as distinct normalization schemes do not always reduce variation either within or between experiments. Using a dataset of 1915 plants collected over three years, we develop a model of variation in N. benthamiana transient expression, using power analysis to determine the number of individual plants required for a given effect size. Drawing on our longitudinal data, these findings inform practical guidelines for minimizing variability through strategic experimental design and power analysis, providing a foundation for more robust and reproducible use of N. benthamiana in quantitative plant biology and synthetic biology applications.

Tang, Sophia N. [Joint BioEnergy Institute (JBEI),↗

X-ray Absorption Spectroscopy of Dilute Metalloenzymes at X-ray Free-Electron Lasers in a Shot-by-Shot Mode

X-ray absorption spectroscopy (XAS) of 3d transition metals provides important electronic structure information for many fields. However, X-ray-induced radiation damage under physiological temperature has prevented using this method to study dilute aqueous systems, such as metalloenzymes, as the catalytic reaction proceeds. Here we present a new approach to enable operando XAS of dilute biological samples and demonstrate its feasibility with K-edge XAS spectra from the Mn cluster in photosystem II and the Fe–S centers in photosystem I. This approach combines highly efficient sample delivery strategies and a robust signal normalization method with high-transmission Bragg diffraction-based spectrometers at X-ray free-electron lasers (XFELs) in a damage-free, shot-by-shot mode. These photon-out spectrometers have been optimized for discriminating the metal Mn/Fe Kα fluorescence signals from the overwhelming scattering background present on currently available detectors for XFELs that lack suitable energy discrimination. We quantify the enhanced performance metrics of the spectrometer and discuss its potential applications for acquiring time-resolved XAS spectra of biological samples during their reactions at XFELs.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Improved 140 Nd Production for the 140 Nd/ 140 Pr In Vivo Generator through Target Recycling and Radiochemical Optimization

Theranostic strategies that utilize f-block therapeutic radionuclides, including 161 Tb, 177 Lu, 225 Ac, and 227 Th, suffer from a shortage of positron emission tomography (PET) imaging counterparts in the same chemical space and often rely on 68 Ga as a surrogate. The 140 Nd/ 140 Pr in vivo PET generator, which belongs to the f-block, may address this issue and can be produced via the 141 Pr(p,2n) 140 Nd production route by using medium-energy cyclotrons. However, impurities in the target material, including stable Nd, and the inherent difficulty of adjacent lanthanide separations limit the achievable radionuclidic and chemical purity of 140 Nd. In this work, we address these challenges through the purification and recycling of praseodymium target material and optimization of Nd/Pr separation. The resulting purified 140Nd was evaluated using DOTA and Macropa chelators via radiolabeling and in vitro stability studies. A target material purification and recycling method was developed for the monoisotopic 141 Pr starting material to remove stable Nd impurities, yielding 90.3 ± 4.7% (n = 3) recovery. The purified 141 Pr was isolated as Pr 6 O 11 and irradiated with 24 MeV protons (20.07 MeV at the target surface) at 20 μA for 4 h, which produced 1417.0 ± 83.4 MBq (38.3 ± 2.2 mCi) of 140 Nd at the end of bombardment (EOB). The produced 140 Nd was purified through an optimized DGA normal method to recover 71.6 ± 6.3% pure 140 Nd. The amount of stable Nd reduced progressively in each target purification cycle from >340 ppm without purification to <250 ppb after three cycles, while other measured metallic impurities were below 30 ppb. This improvement in target purity was reflected in the direct increase of apparent molar activity (AMA), when purified 140 Nd was evaluated with DOTA and Macropa chelators. AMA of [ 140 Nd]Nd-DOTA and [ 140 Nd]Nd-Macropa increased from 70.3 MBq/μmol (1.9 mCi/μmol) and 74 MBq/μmol (2.0 mCi/μmol) to 8025.3 MBq/μmol (216.9 mCi/μmol) and 8473.0 MBq/μmol (229.0 mCi/μmol), respectively, after the third target purification cycle. Further evaluation of chelator-labeled 140 Nd showed that [ 140 Nd]Nd-DOTA was stable in phosphate-buffered saline (PBS), saline, human serum, and mouse serum, whereas [140Nd]Nd-Macropa was stable in all except human serum. This work established a practical methodological advance for the production of 140 Nd/ 140 Pr in vivo PET generators, combining optimized target recycling and radiochemical separation to enable scaled-up and high-molar activity 140 Nd suitable for preclinical imaging. These advances support broader development of 140 Nd/ 140 Pr as a robust PET analogue, especially for f-block therapeutics.

Irradiation↗

Real-Time Anomaly Detection for Beyond Standard Model Searches in ProtoDUNE Horizontal Drift

This paper summarizes work conducted throughout a SULI internship at Fermi National Accelerator Laboratory focused on building an unsupervised machine learning model for real-time anomaly detection in ProtoDUNE Horizontal Drift. Using simulated data, we trained an autoencoder model on a pure cosmic dataset, and evaluated it on both cosmic and neutrino events---making the model an anomaly detector. The goal was to make a model which matches or exceeds the current ADC Simple Window trigger algorithm so that our model can perform at the same rate but provide sensitivity to potential beyond-the-Standard-Model (BSM) signatures. In the end, we were able to construct a model which slightly exceeds the capabilities of the ADC Simple Window while remaining completely unsupervised, achieving $31.9 \pm 0.2$\% ($26.6 \pm 0.2$\%) $\nu$ efficiency at 5 Hz (2 Hz), a 3.6 (3.2) percentage point increase. Additionally, $17.5 \pm 0.3$\% ($18.3 \pm 0.3$\%) of the events that passed the autoencoder at 5 Hz (2 Hz) were missed by the current trigger algorithm. Future work will investigate alternative normalization methods, including quantile transformation, and evaluate the model on ProtoDUNE-HD detector-glitch data if that data becomes available.

Wilson, Cameron C. [Cincinnati U., RWC]↗

Real-Time Anomaly Detection for Beyond Standard Model Searches in ProtoDUNE Horizontal Drift

This paper summarizes work conducted throughout a SULI internship at Fermi National Accelerator Laboratory focused on building an unsupervised machine learning model for real-time anomaly detection in ProtoDUNE Horizontal Drift. Using simulated data, we trained an autoencoder model on a pure cosmic dataset, and evaluated it on both cosmic and neutrino events---making the model an anomaly detector. The goal was to make a model which matches or exceeds the current ADC Simple Window trigger algorithm so that our model can perform at the same rate but provide sensitivity to potential beyond-the-Standard-Model (BSM) signatures. In the end, we were able to construct a model which slightly exceeds the capabilities of the ADC Simple Window while remaining completely unsupervised, achieving $31.9 \pm 0.2$\% ($26.6 \pm 0.2$\%) $\nu$ efficiency at 5 Hz (2 Hz), a 3.6 (3.2) percentage point increase. Additionally, $17.5 \pm 0.3$\% ($18.3 \pm 0.3$\%) of the events that passed the autoencoder at 5 Hz (2 Hz) were missed by the current trigger algorithm. Future work will investigate alternative normalization methods, including quantile transformation, and evaluate the model on ProtoDUNE-HD detector-glitch data if that data becomes available.

Wilson, Cameron C. [Cincinnati U., RWC]↗

Real-Time Anomaly Detection for Searches Beyond the Standard Model in the ProtoDUNE Horizontal Drift Detector

This paper summarizes work conducted throughout a SULI internship at Fermi National Accelerator Laboratory focused on building an unsupervised machine learning model for real-time anomaly detection in ProtoDUNE Horizontal Drift. Using simulated data, we trained an autoencoder model on a pure cosmic dataset, and evaluated it on both cosmic and neutrino events—making the model an anomaly detector. The goal was to make a model which matches or exceeds the current ADC Simple Window trigger algorithm so that our model can perform at the same rate but provide sensitivity to potential beyond-the-Standard-Model (BSM) signatures. In the end, we were able to construct a model which slightly exceeds the capabilities of the ADC Simple Window while remaining completely unsupervised, achieving 31.9 ± 0.2% (26.6 ± 0.2%) ν efficiency at 5 Hz (2 Hz), a 3.6 (3.2) percentage point increase. Additionally, 17.5 ± 0.3% (18.3 ± 0.3%) of the events that passed the autoencoder at 5 Hz (2 Hz) were missed by the current trigger algorithm. Future work will investigate alternative normalization methods, including quantile transformation, and evaluate the model on ProtoDUNE-HD detector-glitch data if that data becomes available.

Wilson, C. [Cincinnati U., RWC]↗

Large-Volume Injection and Assessment of Reference Standards for n -Alkane δD and δ 13 C Analysis via Gas Chromatography Isotope Ratio Mass Spectrometry

Compound-specific stable isotope analysis of hydrogen (δD) and carbon (δ 13 C) in organic compounds is a valuable tool in biogeochemical research. A key limitation of this method is the relatively large amount of sample required to achieve desirable precision. We developed a large-volume (20 μL) injection method that allows for high throughput analysis of less concentrated samples and tested it for δ 13 C and δD measurements of n-alkanes. We also conducted a comparison of reference standards and assessed several methods to normalize and correct n-alkane δD and δ13C measurements. The mean precision of the δD method based on 233 environmental n-alkane samples (two to three replications per sample) is 4.0‰ (1σ, estimated from the weighted mean of the pooled unbiased standard deviations) and 0.46‰ (1σ) for δ 13 C from 37 environmental samples (two to three replications per sample). The evaluation of reference standards shows that the use of n-alkane standards with large offsets in δD values in adjacent n-alkane chains can lead to biases in measurement correction. The large-volume injection method shows good reproducibility of δ 13 C and δD measurements of n-alkanes and reduces the required sample concentration by about 80%. We propose that for δD measurements, a reference standard set should be used in which each reference standard has a limited range of δD values and no adjacent n-alkane chains, to minimize memory effects.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Towards a data-driven model of hadronization using normalizing flows

We introduce a model of hadronization based on invertible neural networks that faithfully reproduces a simplified version of the Lund string model for meson hadronization. Additionally, we introduce a new training method for normalizing flows, termed MAGIC, that improves the agreement between simulated and experimental distributions of high-level (macroscopic) observables by adjusting single-emission (microscopic) dynamics. Our results constitute an important step toward realizing a machine-learning based model of hadronization that utilizes experimental data during training. Finally, we demonstrate how a Bayesian extension to this normalizing-flow architecture can be used to provide analysis of statistical and modeling uncertainties on the generated observable distributions.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Risk Ratio and Risk Difference Estimation in Case-cohort Studies

Background: In case-cohort studies with binary outcomes, ordinary logistic regression analyses have been widely used because of their computational simplicity. However, the resultant odds ratio estimates cannot be interpreted as relative risk measures unless the event rate is low. The risk ratio and risk difference are more favorable outcome measures that are directly interpreted as effect measures without the rare disease assumption. Methods: We provide pseudo-Poisson and pseudo-normal linear regression methods for estimating risk ratios and risk differences in analyses of case-cohort studies. These multivariate regression models are fitted by weighting the inverses of sampling probabilities. Also, the precisions of the risk ratio and risk difference estimators can be improved using auxiliary variable information, specifically by adapting the calibrated or estimated weights, which are readily measured on all samples from the whole cohort. Finally, we provide computational code in R (R Foundation for Statistical Computing, Vienna, Austria) that can easily perform these methods. Results: Through numerical analyses of artificially simulated data and the National Wilms Tumor Study data, accurate risk ratio and risk difference estimates were obtained using the pseudo-Poisson and pseudo-normal linear regression methods. Also, using the auxiliary variable information from the whole cohort, precisions of these estimators were markedly improved. Conclusion: The ordinary logistic regression analyses may provide uninterpretable effect measure estimates, and the risk ratio and risk difference estimation methods are effective alternative approaches for case-cohort studies. These methods are especially recommended under situations in which the event rate is not low.

60 APPLIED LIFE SCIENCES↗

A Pseudoreversible Normalizing Flow for Stochastic Dynamical Systems with Various Initial Distributions

Here, we present a pseudoreversible normalizing flow method for efficiently generating samples of the state of a stochastic differential equation (SDE) with various initial distributions. The primary objective is to construct an accurate and efficient sampler that can be used as a surrogate model for computationally expensive numerical integration of SDEs, such as those employed in particle simulation. After training, the normalizing flow model can directly generate samples of the SDE’s final state without simulating trajectories. The existing normalizing flow model for SDEs depends on the initial distribution, meaning the model needs to be retrained when the initial distribution changes. The main novelty of our normalizing flow model is that it can learn the conditional distribution of the state, i.e., the distribution of the final state conditional on any initial state, such that the model only needs to be trained once and the trained model can be used to handle various initial distributions. This feature can provide a significant computational saving in studies of how the final state varies with the initial distribution. Additionally, we propose to use a pseudoreversible network architecture to define the normalizing flow model, which has sufficient expressive power and training efficiency for a variety of SDEs in science and engineering, e.g., in particle physics. We provide a rigorous convergence analysis of the pseudoreversible normalizing flow model to the target probability density function in the Kullback–Leibler divergence metric. Numerical experiments are provided to demonstrate the effectiveness of the proposed normalizing flow model.

97 MATHEMATICS AND COMPUTING↗

Efficient Monte Carlo event generation for neutrino-nucleus exclusive cross sections

Modern neutrino-nucleus cross section computations need to incorporate sophisticated nuclear models to achieve greater predictive precision. However, the computational complexity of these advanced models often limits their practicality for experimental analyses. To address this challenge, we introduce a new Monte Carlo method utilizing normalizing flows to generate surrogate cross sections that closely approximate those of the original model while significantly reducing computational overhead. As a case study, we built a Monte Carlo event generator for the neutrino-nucleus cross section model developed by the Ghent group. This model employs a Hartree-Fock procedure to establish a quantum mechanical framework in which both the bound and scattering nucleon states are solutions to the mean-field nuclear potential. The surrogate cross sections generated by our method demonstrate excellent accuracy with a relative effective sample size of more than 98.4%, providing a computationally efficient alternative to traditional Monte Carlo sampling methods for differential cross sections.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Welch Method and Bootstrapping Applied to Subcritical Gamma Noise

We measured the prompt neutron decay constant 𝛼 of the CROCUS zero-power reactor at the Swiss Federal Institute of Technology Lausanne using cross-power spectral density (CPSD) analysis of gamma-gamma correlations from two trans-stilbene organic scintillators positioned near the reactor core. We measured critical and subcritical states, with water levels ranging from 960 mm (critical) to 800 mm (𝜌=−1.4 $ subcritical). Our analysis used the Welch method, dividing signal segments for fast Fourier transform (FFT) frequency analysis and applying bootstrapping uncertainty quantification that uses Welch-defined segments. Results demonstrated a clear increase in the measured 𝛼 as reactor reactivity decreased, distinguishing critical from subcritical conditions. At the 960-mm critical level, 𝛼 was estimated at 155.9 ± 0.7 s −1 , and for the 800-mm subcritical level, 𝛼 increased significantly to 367.3 ± 6.9 s –1 . A linear regression of subcritical states yielded a critical estimate of 154.0 ± 3.1 s –1 , aligning with the static 𝛼 estimate at critical. The bootstrapping method produced normally distributed 𝛼 estimates, confirming data consistency. The gamma CPSD 𝛼 estimates clearly distinguish reactor states and improve monitoring of zero-power reactors. The future deployment of modular and microreactors as potential candidates for noise analysis is demonstrated in CROCUS, particularly zero-power mock-ups of new designs. The improvement of noise analysis in the subcritical domain from this work will support experimental data for reactor deployment and procedure.

CROCUS↗

Comparison of multi-stage air treatment process divided by the same temperature and enthalpy difference

The multi-stage air treatment system has been proposed recently, and lower grade chilled/hot water could be used and energy efficiency could be improved. However, it has not been studied which division method of air treatment processes has higher energy efficiency. In this study, the model to calculate the energy consumption of multi-stage air treatment process is introduced, and the effects of two division methods, i.e. multi-stage air treatment process divided by the same temperature difference (ST method) or same enthalpy difference (SE method) between inlet and outlet at each stage, under 9 different air inlet parameters in the 2-stage and 3-stage air treatment processes are analysed and compared. The results show that (1) the system energy consumption of the SE method is generally lower than that of the ST method; (2) there is generally a larger energy consumption reduction rate of SE method when the air relative humidity is 70% compared to relative humidity of 50% and 90%; (3) the difference between ST method and SE method is not great, so both methods can be used for the design of multi-stage treatment system although SE method is normally recommended.

Wang, Wentao↗

Distributed Coordination of Networked Microgrids for Voltage Support in Bulk Power Grids

The increasing deployment of distributed energy resources (DERs) and microgrids (MGs) in power distribution systems has enabled the adjustment of reactive power consumption as seen at the substation, which can be used to provide voltage support for the bulk power system (BPS). Leveraging this new capability will provide greater resiliency to the power system as a whole. Here, the goal of this paper is to develop and compare three different algorithms, namely distributed optimal power flow, distributed consensus algorithm, and fully decentralized collaborative autonomy for unbalanced distribution systems for microgrid coordination. These algorithms use networked MGs to support the BPS voltage when a contingency at the bulk grid results in abnormally low voltages, which may be a precursor to voltage collapse. Our comparative analysis includes both qualitative and quantitative assessments of the three algorithms and a discussion of the trade-offs between the decentralized and distributed methods in normal and disrupted conditions. Each algorithm was evaluated on the modified IEEE 13-bus system and a real power distribution system at Chattanooga, Tennessee, that encompasses more than 4500 buses. Each algorithms excels differently and may be suited for different scenarios depending on the condition, operations, and priorities of the power and communication systems.

24 POWER TRANSMISSION AND DISTRIBUTION↗