Search NASA⌕ Search

SEARCH · Search NASA

Results for “data distributions”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Four-dimensional phase space tomography from one-dimensional measurements of a hadron beam

In this paper, we use one-dimensional measurements to infer the four-dimensional phase space density of an accumulated proton beam in the Spallation Neutron Source (SNS) accelerator. The reconstruction was performed by maximizing the distribution’s entropy subject to the measurement constraints and thus represents the most conservative inference from the data. The reconstructed distribution reproduces the measured profiles down to the noise level, and simulations indicate that the problem is reasonably well constrained. Similar measurements could serve as benchmarks for beam dynamics simulations in the SNS or hadron accelerators.

43 PARTICLE ACCELERATORS↗

Super Resolution for Renewable Energy Resource Data With Wind From Reanalysis Data (Sup3rWind) and Application to Ukraine [Slides]

In this work we present a novel deep learning-based downscaling method, using generative adversarial networks (GANs), for generating high-resolution wind resource data from ECMWF Reanalysis v5 data (ERA5). We show that by training a GAN model on ERA5, as opposed to coarsened high-resolution data, we achieve results that are competitive with conventional dynamical downscaling. This GAN-based downscaling method additionally reduces computational costs over dynamical downscaling by two orders of magnitude. All GANs are trained on data sampled from CONUS, selected to provide a diverse sampling of terrain conditions, and validated on observational data along with data held out from training. This cross-validation shows low error and high correlations with observations and excellent agreement with hold out data across physical distributions. Our approach is finally used to downscale 30km hourly ERA5 to 2-km 5-minute wind data, for January 2000 through December 2023, at multiple hub heights, over Ukraine, Moldova, and part of Romania. Comparisons against observational data from Meteorological Assimilation Data Ingest System (MADIS) and multiple wind farms show the same level of performance as for CONUS validation. This 24 year data record is the first member of the "super resolution for renewable energy resource data with wind from reanalysis data" dataset (Sup3rWind).

17 WIND ENERGY↗

Distributed Wind Monitoring Best Practices

Accessible performance and operational data have been identified as a key enabler for distributed wind energy industry advancement. While utility-scale wind turbines benefit from reliable and continuous supervisory control and data acquisition (SCADA)-based monitoring platforms, monitoring of the U.S. fleet of distributed wind (DW) turbines has been more inconsistent, unreliable, and sometime difficult to access. Without fleet monitoring data, the industry will never understand and thus work to improve turbine under-performance and reliability issues. For the DW industry to scale up, attract investors, and boost credibility, fleetwide monitoring must be robust and reliable, select data must be made accessible to stakeholders, and the data must be in a format useful to users. To help move the industry toward a more standardized, accessible stream of monitoring data, this distributed wind monitoring best practices report attempts to cover topics including key monitoring channels, hardware, communication strategies, and accessibility. Strategic engagement with DW original equipment manufacturers (OEMs), service providers, lab and university researchers, testing organization, certification bodies, end users and solar photovoltaic (PV) monitoring experts has enabled a better understanding of the current state-of-the-art of monitoring and aided in articulating this set of best practices that will guide OEMs toward harmonized monitoring strategies, aimed at a future goal of achieving accessible performance and operational data for the entire fleet of U.S. distributed wind turbines.

17 WIND ENERGY↗

Data Center High-Temperature Liquid Cooling and Heat Reuse Techno-Economic Study: Preprint

Data centers are energy-intensive facilities with growing demands for efficiency and cost-effective operations. Smaller, more distributed edge inference data centers are expected to proliferate as AI applications require low latency closer to the user of AI tools, which presents a growing opportunity to explore the systems implications of liquid cooling on water and energy use. This study analyzes the implementation of high-temperature liquid cooling systems in a prototypical inference 1-MW data center and explores the potential for heat reuse across varying climates with a goal to optimize energy efficiency, reduce capital and operational costs, and identify opportunities for high-performance cooling and water use reduction infrastructure. This analysis evaluated configurations utilizing a peak day hourly sizing and systems performance spreadsheet to evaluate design and operational conditions from which component sizes, installed cost, operational cost, and performance metrics were determined for the Base case and the Elevated case. The techno-economic analysis included heat reuse applications across a range of heat recovery temperatures and heat rejection options. The analysis shows that high-temperature liquid cooling allows for improved energy efficiency, lower water consumption, and lower capital costs compared to traditional cooling approaches. Transitioning to elevated water inlet/outlet temperatures (50 degrees C/60 degrees C) eliminates the need for chillers, cooling towers, and heat recovery equipment in many scenarios across three distinct climate zones. This results in up to 75% capital cost savings for the cooling and heat recovery equipment, and with significantly reduced water consumption, especially in non-heat reuse applications. Heat generated from data centers can also be repurposed for space heating, domestic hot water, and other applications, and is most cost-effective when data center outlet temperatures exceed 55-60 degrees C.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Derivation of at-birth neutron energy distribution of americium-beryllium source based on ISO8529 spectral data

In many practical applications, the accuracy of neutron transport calculations depends on an estimate of the energy distribution provided in the source definition. While the energy distribution of neutrons emitted by spontaneous fission sources is well-defined, the spectra of neutrons produced in (α,n) reactions in compound mixtures, such as americium-beryllium (AmBe), depend on the microstructural properties of the source material, which may vary. Although the energy of neutrons produced in nuclear reactions can be calculated from first principles, such computations require detailed knowledge of the key characteristics of the source material, which are often unknown. Here, this study focuses on deriving the at-birth neutron energy distribution for an AmBe source by applying reverse transport calculations to the externally measured high-precision spectrum recommended in the latest revision of the international standard. The derived at-birth spectrum was validated through indirect energy-sensitive neutron emission rate measurements of AmBe sources calibrated at national primary metrology institutions. The resulting at-birth neutron energy distribution serves as essential input data for radiation transport modeling, providing a validated spectrum for broader experimental and computational applications.

AmBe↗

Search for neutron decay into an antineutrino and a neutral kaon in 0.401 megaton-years exposure of Super-Kamiokande

We searched for bound neutron decay via 𝑛 → $\bar{𝜈}$ +𝐾 0 predicted by the grand unified theories in 0.401 Mton·years exposure of all pure water phases in the Super-Kamiokande detector. About 4.4 times more data than in the previous search have been analyzed by a new method including a spectrum fit to kaon invariant mass distributions. No significant data excess has been observed in the signal regions. As a result of this analysis, we set a lower limit of 7.8 × 10 32 years on the neutron lifetime at a 90% confidence level.

Cherenkov detectors↗

Describing hadronization via histories and observables for Monte-Carlo event reweighting

We introduce a novel method for extracting a fragmentation model directly from experimental data without requiring an explicit parametric form, called Histories and Observables for Monte-Carlo Event Reweighting (HOMER), consisting of three steps: the training of a classifier between simulation and data, the inference of single fragmentation weights, and the calculation of the weight for the full hadronization chain. We illustrate the use of HOMER on a simplified hadronization problem, a q\bar{q} q q ‾ string fragmenting into pions, and extract a modified Lund string fragmentation function f(z) f ( z ) . We then demonstrate the use of HOMER on three types of experimental data: (i) binned distributions of high-level observables, (ii) unbinned event-by-event distributions of these observables, and (iii) full particle cloud information. After demonstrating that f(z) f ( z ) can be extracted from data (the inverse of hadronization), we also show that, at least in this limited setup, the fidelity of the extracted f(z) f ( z ) suffers only limited loss when moving from (i) to (ii) to (iii). Public code is available at https://gitlab.com/uchep/mlhad.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Conditional diffusion machine-learning framework for mapping valence electron distribution from convergent beam electron diffraction

Quantitative convergent beam electron diffraction (CBED) enables determination of aspherical valence electron distributions through refinement of low-order structure factors, which are highly sensitive to chemical bonding and charge density variations. However, conventional quantitative CBED (QCBED) requires solving a highly nonlinear inverse problem with many coupled parameters, and computationally intensive dynamical diffraction calculations, making it time-consuming and difficult to apply to complex systems. More broadly, reconstructing charge density and orbital electron distribution from diffraction data has long been a central challenge in both x-ray and electron crystallography. Here, in this study, we introduce an artificial-intelligence (AI)-based framework that replaces traditional refinement with a data-driven inverse solver. Using a large synthetic CBED dataset generated by Bloch-wave simulations, we train a conditional diffusion model to directly infer crystal structural parameters and multipole density formalism parameters, and hence valence electron distributions, from CBED patterns alone. By learning from forward simulations across realistic parameter space, the model effectively solves the inverse problem. Compared with direct regression approaches, the diffusion-based framework provides posterior parameter distributions for rigorous uncertainty quantification while preserving quantitative fidelity and reducing analysis time by orders of magnitude. By eliminating the need for external single-crystal x-ray diffraction data and complex nonlinear refinement, this approach enables practical, high-throughput, and in situ quantitative CBED, enabling real-time mapping of valence electron distributions and their correlation with functional responses in quantum and energy materials.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

An Efficient Storage-Driven Machine Learning Model for Performance in the Era of Multimodal Scientific Data

Scientific workflows are increasingly relying on machine learning (ML), simulation, and hybrid techniques to predict, understand, and optimize the behavior of complex experiments. High-performance computing has greatly improved researchers’ ability to acquire diverse data modalities in these workflows. Recent studies suggest that the performance of machine learning models can be improved by integrating data from various sources. Unfortunately, these workloads pose unprecedent pressure on the network storage to meet the demands associated with accessing these multimodal data. To mitigate the impact of intensive IO, we propose a solution that utilizes a multi-tier High-Performance Computing (HPC) distributed storage and data processing framework, placing computation where the data resides for better performance. By adopting this project, the scientific community will gain new opportunities to explore multimodal storage-driven possibilities, integrating multiple scientific data sources with advanced streaming frameworks. Additionally, our framework effectively utilizes computing resources and bridges the gaps identified by HPC experts. Our proposed approach tackles scalability and persistence challenges by leveraging native persistency, which has posed difficulties in traditional approaches. Furthermore, we seek to enhance fault-tolerance and load-balance of computations by leveraging real-time streaming in diverse scientific computing environments, thereby propelling advanced scientific computing research into the next generation.

97 MATHEMATICS AND COMPUTING↗

Elliptically-Contoured Tensor-variate Distributions with Application to Image Learning

Statistical analysis of tensor-valued data has largely used the tensor-variate normal (TVN) distribution that may be inadequate for data arising from distributions with heavier or lighter tails. We study a general family of elliptically contoured (EC) TV distributions and derive its characterizations, moments, marginal, and conditional distributions. We describe procedures for maximum likelihood estimation from data that are (1) uncorrelated draws from an EC distribution, (2) from a scale mixture of the TVN distribution, and (3) from an underlying but unknown EC distribution, for which we extend Tyler’s robust estimator. A detailed simulation study highlights the benefits of choosing an EC distribution over the TVN for heavier-tailed data. We develop TV classification rules using discriminant analysis and EC errors and show that they better predict cats and dogs from images in the Animal Faces-HQ dataset than the TVN-based rules. A novel tensor-on-tensor regression and TV analysis of variance (TANOVA) framework under EC errors is also demonstrated to better characterize gender, age, and ethnic origin than the usual TVN-based TANOVA in the celebrated labeled faces of the wild dataset.

97 MATHEMATICS AND COMPUTING↗

Explainable multi-fidelity Bayesian neural network for distribution system state estimation

Distribution System State Estimation (DSSE) is frequently constrained by limited real-time measurements, the uncertainties introduced by distributed energy resources, and the presence of bad data. To address them, this paper proposes an enhanced Multi-Fidelity Bayesian Neural Network (MFBNN) DSSE approach. A low-fidelity layer based on a Deep Neural Network (DNN) is first pre-trained on pseudo-measurement data to learn fundamental state features. Subsequently, a high-fidelity Bayesian Neural Network (BNN) layer leverages limited but high-quality real-time measurements to refine these features, thereby achieving accurate DSSE. Additionally, the deep SHapley Additive exPlanation (SHAP) is developed to quantify the influence of measurement data on DSSE through dual perspectives of global feature importance and local nodal contributions, establishing a hierarchical explainability framework for machine learning-based DSSE. Comparative studies conducted on the IEEE 13-bus system and a real-world 2135-node system from Dominion Energy demonstrate that the proposed method excels in estimation accuracy, even under situations of high noise levels, bad data, and missing data. Further comparisons with Weighted Least Squares (WLS) and other machine learning-based DSSE approaches verify that the proposed framework offers higher accuracy, improved interpretability, and enhanced robustness.

Bad data↗

Constellation: The autonomous control and data acquisition system for dynamic experimental setups

The operation of instruments and detectors in laboratory or beamline environments presents a complex challenge, requiring stable operation of multiple concurrent devices, often controlled by separate hardware and software solutions. These environments frequently undergo modifications, such as the inclusion of different auxiliary devices depending on the experiment or facility, adding further complexity. The successful management of such dynamic configurations demands a flexible and robust system capable of controlling data acquisition, monitoring experimental setups, enabling seamless reconfiguration, and integrating new devices with limited effort. This paper presents Constellation, a flexible and network-distributed control and data acquisition software framework tailored to laboratory and beamline environments, that addresses the limitations of existing solutions. The framework is designed with a focus on extensibility, providing a streamlined interface for instrument integration. It supports efficient system setup via network discovery mechanisms, promotes stability through autonomous operational features, and provides comprehensive documentation and supporting tools for operators and application developers such as controllers and logging interfaces. At the core of the architectural design is the autonomy of the individual components, called satellites, which can make independent decisions about their operation and communicate these decisions to other components. This paper introduces the design principles and framework architecture of Constellation, presents the available graphical user interfaces, shares insights from initial successful deployments, and provides an outlook on future developments and applications.

Autonomy↗

Probabilistic Hazard Assessment for Tornadoes, Straight-Line Wind, and Extreme Precipitation at the Savannah River Site

Recent data sets for three meteorological phenomena with the potential to inflict damage on SRS facilities – tornadoes, straight-line winds, and heavy precipitation – are analyzed using appropriate statistical techniques to estimate the occurrence probabilities for these events in the future. Summaries of the results for DOE-mandated return periods and comparisons to similar calculations performed in 2013 by Werth et al. (W2013) are given. Using tornado statistics for i) the combined states of Georgia and South Carolina, and ii) a 2⁰ square area surrounding SRS, we calculated the probability per year of any location at SRS being struck by a tornado (the ‘strike’ probability) and the probability that any point will experience winds above set thresholds. The strike probability was calculated to be 7.04E-4 (1 chance in 1420) per year and tornadic wind speeds for DOE mandated return periods of 50,000 years (corresponding to wind design category 3 (WDC-3), and 125,000 years (meeting WDC-4) (USDOE, 2016) were estimated to be 132 mph and 147 mph, respectively. By contrast, default tornado wind speeds taken from ANSI/ANS-2.3-2011 are somewhat higher: 161 mph for return periods of 50,000 years and 173 mph every 125,000 years (ANS, 2011). Although the ANS and the SRS evaluation used the same basic model (Ramsdell and Rishel, 2007), the region defined in ANS 2.3 that encompasses the SRS also includes areas of the Great Plains and lower Midwest, regions with much higher occurrence frequencies of strong tornadoes. The SRS straight-line wind values associated with various return periods were calculated by fitting existing wind data to a GEV1 distribution and extrapolating the values for any return period from the tail of that function. For the DOE mandated return periods, we expect straight-line winds of 117 mph every 2500 years (the required WDC-3 standard) and 125 mph every 6250 years (WDC-4) at any point within the SRS. These values are similar to those from the ANS-2.3-2011 report, which has wind speeds of 125mph and 133 mph for return periods of 2500 years and 6250 years, respectively. For extreme precipitation, we compared the fits of two different theoretical extreme-value distributions and applied the one that fit the data best for each of several accumulation periods. The DOE mandated 6-hr accumulated rainfall for return periods of 10,000 years (corresponding to precipitation design category 3 (PDC-3) and 25,000 years (PDC-4) were estimated as 9.1 inches and 10.1 inches, respectively. For the 24-hr rainfall return periods of 10,000 years and 25,000 years, total rainfall estimates were 12.02 inches and 13.17 inches, respectively, higher than comparable values provided in the W2013 report.

54 ENVIRONMENTAL SCIENCES↗

Fiber-Optic Sensing for Earthquake Hazards Research, Monitoring, and Early Warning

The use of fiber‐optic sensing systems in seismology has exploded in the past decade. Despite an ever‐growing library of ground‐breaking studies, questions remain about the potential of fiber‐optic sensing technologies as tools for advancing if not revolutionizing earthquake‐hazards‐related research, monitoring, and early warning systems. A working group convened to explore these topics; we comprehensively examined the application of fiber optics in various aspects of earthquake hazards, encompassing earthquake source processes, crustal imaging, data archiving, and technological challenges. There is great potential for fiber‐optic systems to advance earthquake monitoring and understanding, but to fully unlock their capabilities requires continued progress in key areas of research and development, including instrument testing and validation, increased dynamic range for applications focused on larger earthquakes, and continued improvement in subsurface and source imaging methods. A key current stumbling block results from the lack of clear data archiving requirements, and we propose an initial strategy that balances data volume requirements with preserving key data for a broad range of future studies. In addition, we demonstrate the potential for fiber‐optic sensing to impact monitoring efforts by documenting the data completeness in a number of long‐term experiments. Finally, we outline the features of a instrument testing facility that would enable progress toward reliable and standardized distributed acoustic sensing data. Overcoming these current obstacles would facilitate progress in fiber‐optic sensing and unlock its potential application to a broad range of earthquake hazard problems.

58 GEOSCIENCES↗

Confronting Large‐Eddy Simulations With Stereo Camera Data by Means of Reconstructed Hemispheric Cloud Size Distributions

High-resolution hemispheric camera images at a meteorological site in western Germany are used to analyze the multi-dimensional spatial characteristics of continental cumulus cloud fields, and to evaluate Large-Eddy Simulations on this aspect. Traditional non-hemispheric cloud-detecting instruments provide additional reference data. The main model-observation comparison focuses on cloud size distributions (CSDs), employing two methods: (a) directly using three-dimensional model fields, direct CSDs, and (b) using rendered hemispheric images of the model fields as produced by a camera simulator based on path-tracing. In the latter method, both the real and rendered images are used to three-dimensionally reconstruct the cloud fields, yielding hemispheric CSDs. Advantages of hemispheric comparisons over more classic approaches include (a) fair comparisons between model and data, and (b) full use of the enhanced resolutions and hemispheric spatial coverage of the camera imagery. Basic evaluation of the simulations demonstrates good agreement on thermodynamic structure and its diurnal cycle. Cloud heights and cloud cover are intercompared between the model, camera data and other instrumentation, providing insight into their structural differences. A consistent alignment is found between the hemispheric CSDs from both the model and the cameras. Power law fits reveal structurally lower exponents in hemispheric CSDs compared to non-hemispheric CSDs, which particularly caution against directly comparing hemispheric CSDs to non-hemispheric distributions. This result is robust for sample size and fitting method. These findings inform future use of hemispheric camera systems for studying cumulus cloud field morphology and model evaluation.

54 ENVIRONMENTAL SCIENCES↗

Microstructure and Mechanical Properties of Ni-based Alloys Fabricated by Laser Powder Bed Fusion

The Advanced Materials and Manufacturing Technologies (AMMT) program is aiming at the accelerated incorporation of new materials and manufacturing technologies into nuclear-related systems. Complex Ni-based components fabricated by laser powder bed fusion (LPBF) could enable operating temperatures at T > 700°C in aggressive environments such as molten salts or liquid metals. However, available mechanical properties data relevant to material qualification remains limited, in particular for Ni-based alloys routinely fabricated by LPBF such as IN718 (Ni- 19Cr-18Fe-5Nb-3Mo) and Haynes 282 (Ni-20Cr-10Co-8.5Mo-2.1Ti-1.5Al). Creep testing was conducted on LPBF 718 at 600°C and 650°C and on LPBF 282 at 750°C. finding that the creep strength of the two alloys was close to that of wrought counterparts. with lower ductility at rupture. Heat treatments were tailored to the LPBF-specific microstructure to achieve grain recrystallization and form strengthening γ' precipitates for LPBF 282 and γ' and γ" precipitates for LPBF 718. In-situ data generated during printing and ex-situ X-ray computed tomography (XCT) scans were used to correlate the creep properties of LPBF 282 to the material flaw distribution. In- situ data revealed that spatter particles are the potential causes for flaws formation in LPBF 282. with significant variation between rods based on their location on the build plate. XCT scans revealed the formation of a larger number of creep flaws after testing in the specimens with a higher initial flaw density. which led to a lower ductility for the specimen.

Dryepondt, Sebastien↗

The Draco Dwarf Spheroidal Galaxy in the First Year of Dark Energy Spectroscopic Instrument Data

We investigate the spatial distribution, kinematics, and metallicity of stars in the Draco dwarf spheroidal galaxy using data from the Dark Energy Spectroscopic Instrument (DESI). We identify 155 high-probability members of Draco using line-of-sight velocity and metallicity information derived from DESI spectroscopy along with Gaia Data Release 3 proper motions. We find a mean line-of-sight velocity of −290.62 ± 0.80 km s −1 with dispersion = $9.57$$^{+0.66}_{–0.62}$ km s −1 and mean metallicity [Fe/H] = −2.10 ± 0.04, consistent with previous results. We also find that Draco has a steep metallicity gradient within the half-light radius, and a metallicity gradient that flattens beyond the half-light radius. We identify eight high-probability members outside the King tidal radius, four of which we identify for the first time. These extratidal stars are not preferentially aligned along the orbit of Draco. We compute an average surface brightness of 34.02 mag arsec –2 within an elliptical annulus from the King tidal radius of 48'.1–81′.

79 ASTRONOMY AND ASTROPHYSICS↗

Graph-Learning-Assisted State and Event Tracking for Solar-Penetrated Power Grids with Heterogeneous Data Sources

Unlike transmission systems, distribution systems do not typically contain sufficient metering to enable real-time state estimation. The lack of sufficient real-time measurements prohibits accurate and timely monitoring of the state of distribution systems. As a result, control and optimal operation of distribution systems, especially those containing large numbers of renewable generation units are not possible without proper data and information about the current state of the system. The main motivation of this project is to address this shortcoming by developing an approach which provides “predicted” real-time measurements so that they can be used to execute a distribution system state estimator. Thus, the objective of the project is to make the distribution systems fully observable, such that the hosting capacity for solar generation can be accurately estimated, and unnecessary solar curtailments can be avoided. In order to accomplish this goal, the project investigated the use of a grid-model-informed machine learning (ML) tool which integrates heterogeneous data streams obtained from AMI meters, SCADA as well as PMU measurements and created synchronous measurement snapshots for the state estimator (SE); and developed a hybrid robust SE which provides not only accurate state estimates but also real-time feedback for the ML model refinement.

14 SOLAR ENERGY↗