Search NASA⌕ Search

SEARCH · Search NASA

Results for “data availability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 361 records · Page 20

Testing, Calibration, and UxS Integration of the Kromek GR1 Plus CZT Gamma Spectrometer

Collecting radiation measurements can be a hazardous task, especially in the presence of highly active sources. Various scenarios necessitate source search, classification, and quantification, often with limited or no a priori information. Activities such as disaster mitigation, emergency response, environmental monitoring, and site remediation may involve dangerous radioactive sources. In these situations, there is a pressing need for remote monitoring capabilities that protect human operators from potential harm and enhance adherence to the principle of “As Low as Reasonably Achievable” (ALARA) for radiation doses. The first step toward achieving remote radiation measurement capabilities is the remote operation of a radiation sensor. Once this milestone is reached, the next challenge is to integrate this remote sensing capability into suitable actuation agents, collectively referred to as uncrewed systems (UxS). These systems include familiar platforms such as robotic quadrupeds, aerial multirotor vehicles, and ground vehicles, any of which may be teleoperated, act autonomously, or utilize a combination of both. A critical factor in achieving remote radiation sensing is the availability of data from the appropriate sensor. Many commercially available radiation sensors have closed-source documentation for their communication protocols. Typical end-user products are often self-contained, handheld devices designed for manual measurement scenarios. While there are commercial off-the-shelf (COTS) integrations of radiation sensors with UxS available for purchase, these solutions are typically tailored for specific use cases and may not meet the requirements of different applications. This paper discusses efforts to remotely acquire radiation measurements from a small form-factor CZT gamma spectrometer. Sandia National Laboratories has successfully demonstrated the initial capability to integrate low size, weight, and power (SWaP) gamma spectroscopy into various UxS, alongside co-located GPS data logging and sensor calibration and qualification. With remote gamma spectroscopy achieved, the stage is set for UxS integration of this capability.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Utah FORGE: Wells 16A(78)-32 and 16B(78)-32 Stimulation Pressure and Circulation Data April, 2024

This dataset consists of all of the pressure pumping data for Utah FORGE Well16A78-32 fracturing stages, Well 16B78-32 fracturing stages, and a 9-hour circulation test which took place after all of the frac plugs were drilled out. Fracturing in the wells took place in stages from April 3rd to April 17th, while the circulation test occurred on April 27th, 2024. Data is in excel files, with ReadMe files for each of the .zip folders that detail what data is available. Also attached here is wellhead pressure data from Pason that is referred to in the ReadMe files.

15 GEOTHERMAL ENERGY↗

Bridging the Gap on Data and Analysis for Distribution System Planning: Information That Utilities Can Provide Regulators, State Energy Offices and Other Stakeholders

Electric utilities conduct planning annually to ensure their distribution system meets technical standards, policies, and regulations; addresses forecasted grid conditions; satisfies customer needs; and advances utility priorities. The plan identifies grid deficiencies, analyzes potential solutions, and prioritizes capital investments and other expenditures. About 20 U.S. states and jurisdictions require regulated utilities to file some type of distribution system plan with the public utility commission for review. Requirements for sharing distribution system data and analyses vary widely, from few specific requirements to a detailed list of information that must be provided. While utilities conduct extensive analysis to develop distribution system plans, in most jurisdictions regulators and stakeholders do not know what data are available and how the utility uses the data in planning and investing. This report aims to bridge the gap by increasing understanding of the types of data and analyses utilities employ to develop distribution system plans and how the information affects their decision-making. The report describes information that states and stakeholders can ask for related to 11 data categories: -Forecasting loads and distributed energy resources (DERs) -Scenario analysis -Worst-performing circuits -Asset management strategy -Hosting capacity analysis -Value of DERs -Grid needs assessment -Cost-effectiveness framework for investments -Distribution system investment strategy and implementation -Geotargeted programs -Non-wires alternatives procurements.

24 POWER TRANSMISSION AND DISTRIBUTION↗

New Particle Formation and Growth in the Houston Atmosphere During TRACER (Final Report)

From 2020-2025, researchers from UC Irvine, UC Riverside, and Colorado State University collaborated on a Department of Energy-funded project to understand how airborne particles form and grow in urban atmospheres, conducting an intensive field campaign in Houston, Texas during summer 2022. Using advanced instruments to measure gas-phase chemicals, particle composition, and a specialized chamber to study particle growth, the team discovered that sulfur-containing compounds from industrial and power plant emissions are the dominant driver of new particle formation in Houston, with particles typically forming locally in the city and growing as air moves away in the urban plume. The research revealed an important methodological insight: measurements from fixed ground stations can be misleading when interpreting how particles actually evolve as air masses move, which has significant implications for how scientists worldwide interpret atmospheric observations. These findings improve understanding of urban air quality and help reduce uncertainties in climate models, since these particles play critical roles in cloud formation and Earth's radiation balance, while also providing detailed information about ultrafine particle composition relevant to public health. The project trained three doctoral students, developed enhanced computer models for urban particle formation, and made all data publicly available through the DOE Atmospheric Radiation Measurement data archive for use by the broader scientific community.

54 ENVIRONMENTAL SCIENCES↗

Short-Term Electric Load Forecasting for a Residential Household in Alaska

Accurate short-term load forecasting at a fine scale is essential for demand response programs, peak shaving, and load-shedding strategies [1]. While traditionally, only aggregate short-term consumption data was available, advanced metering infrastructure (AMI) now provides data at the individual consumer level [1]. There is increasing interest in utilizing this data for short-term load forecasting (from an hour to a few days) to optimize grid operations. Electricity consumption in individual households is highly influenced by residents’ personal behaviors [2]. As a result, unlike aggregate loads, electrical power usage in single households often shows significant volatility, making meter-level load forecasting for individual users particularly challenging [3], [4]. Deep learning methods, with their strong ability to model nonlinear data, have become popular for improving the accuracy of household electricity consumption forecasting [4]. Notably, the Long ShortTerm Memory (LSTM) has attracted significant attention [5], [6].

42 ENGINEERING↗

Emerging anomaly detection techniques for electronic health records: A survey

Background Anomaly detection in electronic health records (EHRs) is a cornerstone of biomedical informatics, with direct implications for patient safety, clinical decision-making, and the prevention of healthcare fraud. Once guided primarily by simple rule-based methods, the field has advanced rapidly, driven by increased computing power, richer and more detailed health data, and the rise of machine learning and deep learning techniques. The objective of this paper is to provide a comprehensive overview of modern approaches to detecting anomalies in EHRs, outlining their strengths, limitations, and relevance to key healthcare challenges. We review traditional statistical methods alongside newer ML- and DL-based strategies and hybrid models, with particular attention to how these techniques support transparency and build clinical trust. Methods This paper presents a thorough and critical survey through systematic review (PRISMA-based) of the latest anomaly detection strategies in time-sequence data domains within electronic health record systems. Results We explore a broad spectrum of methodologies, including statistical models, supervised and unsupervised learning approaches, hybrid frameworks, and state-of-the-art ML-based techniques that collectively advance the precision and scalability of detecting anomalies in complex clinical datasets. In addition to mapping current capabilities, we address the enduring challenges that hinder widespread implementation and provide a forward-looking perspective on the future of anomaly detection in the data-rich landscape of modern healthcare. Summary The advancement in AI-based approaches is reported along with the basic principles of the individual approaches and their applicability. The increased availability of high-quality data, advancements in DL approaches, and enhanced computation power are leading to more frequent adaptation of DL-based approaches. Emerging DL-based approaches that have been adapted in other domains or recently applied in the EHR domain are also discussed in detail. Although DL-based approaches can improve model predictions by incorporating comorbidities, their application is limited in low-frequency data domains (e.g., when the total available data remains in the single digits). Therefore, the user must carefully consider the application based on data availability.

Anomaly detection↗

Uncovering hidden enhancers through unbiased in vivo testing

Chromatin signatures are widely used to identify tissue-specific in vivo enhancers, but their sensitivity and specificity remains unclear. Here we show that many developmental enhancers remain undetectable using currently available chromatin data. In an initial comparison of over 1200 developmental enhancers with tissue-matched chromatin data, 14% (n = 285) lacked canonical enhancer-associated chromatin signatures. To further assess the prevalence of enhancers missed by chromatin profiling approaches, we used a high-throughput transgenic enhancer assay to screen the regulatory landscapes of two key developmental genes at 5 kb resolution, spanning 1.3 Mb of mouse sequence in total. We observed that 23 of 88 (26%) in vivo enhancers discovered by this approach lacked enhancer-associated chromatin signatures in the respective tissue. Our findings suggest the existence of tens of thousands of enhancers that remain undiscovered by currently available chromatin data, underscoring the continued need for expanding resources for enhancer discovery.

Epigenomics↗

Prediction of Redox Potentials for Ac, Th, and Pa in Aqueous Solution

Density functional theory in conjunction with small core pseudopotentials and the associated basis sets was used to calculate potentials for multiple redox couples, covering a range of oxidation states for Ac (0 to III), Th (0 to IV), and Pa (0 to V) in aqueous solution. Solvation effects were incorporated using a supermolecule-continuum approach, with 30 water molecules representing two solvation shells, and the COSMO and SMD implicit solvation models. The calculated geometries for Ac(III), Th(IV), and Pa(V) were in reasonable agreement with the available experimental data. Using the COSMO model with the B3LYP functional, the calculated redox potentials were within ± 0.2 V from experiment for most redox couples. Several pathways were explored for the Pa(V/IV) redox couple for different forms of Pa(V) and Pa(IV). Most Pa(V/IV) redox couples have very similar potentials, ranging from 0 to -0.4 V up to a pH of 1.4. At pH = 1.4, the potentials shift to values that are more negative than -0.7 V, reflecting the growing unfavorable nature of the redox process at higher pH levels. The calculated values for An(III/II) potentials were consistent with prior estimates and the available experimental data. The predicted redox potentials for An(II/I) were highly negative, as expected. For An(I/0) potentials, Th and Pa exhibited positive values, contrasting with the negative values calculated for Ac. Furthermore, the An +m /An(0) potentials agreed better with the experimental data when using the COSMO solvation model as compared to the SMD model.

Chemical calculations↗

Quality Ranking of Unary Chloride Salt Property Data Included in MSTDB-TP

Molten salt reactor developers rely on thermal property data to design, license and operate the reactors. The Molten Salt Thermal Database-Thermophysical Properties (MSTDB-TP) was established under the DOE Nuclear Energy Advanced Modeling and Simulation (NEAMS) program and is managed by Oak Ridge National Laboratory to serve as a single source of thermophysical property values measured for a wide variety of molten salt systems for use by researchers, molten salt reactor developers, and regulators. These properties include density, viscosity and thermal diffusivity and conductivity. Published measurements of molten salt properties are lacking for many salts of interest and the data that are available are often inconsistent. This creates a challenge for MSR developers when determining which property values to use when designing their reactors. It is the purpose of this work to apply a consistent ranking system to all data entries that indicates the quality of property values listed in the database. These rankings will be the technical basis for down-selections by the database developers and alert users about the quality of the available property values. MSTDB-TP collects all available property data and indicates preferred data sets or correlations. However, all available data sets are included in the database. Quality assessments and rankings are being applied to data in MSTDB-TP to provide an indication of the quality of each data set independent of consistency with other data. Previous reports detailed the ranking system that was followed and assessments of unary fluoride data sets. Documentation of the quality of data in MSTDB-TP was continued by reviewing and assessing all available sources of density, viscosity and thermal diffusivity or conductivity values for unary chloride salts in MSTDB-TP V3.0 using the same criteria.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Reference Site Conditions for Floating Wind Arrays in the United States

Floating offshore wind farm design is highly site-specific, requiring detailed information about the specific conditions of a project area for realistic design studies. Unfortunately, publicly available site condition data for potential floating offshore wind project sites in the United States is scarce. To support U.S. offshore wind research, we developed reference site condition datasets, including metocean and seabed information, for four potential floating wind project areas in the U.S.: Humboldt Bay, Morro Bay, the Gulf of Maine, and the Gulf of Mexico. These datasets were compiled using publicly available data. Our metocean analysis, covering wind, waves, and surface currents, utilized measurement data from 2000 to 2020. Sources included the National Renewable Energy Laboratory’s National Offshore Wind Dataset for wind data, National Data Buoy Center buoys for wave data, and the High Frequency Radar Network for surface currents. These data were integrated into hourly time series used to compute extreme return periods up to 500 years, monthly statistics, and joint probability clusters for fatigue analysis. Soil conditions were evaluated using the usSEABED database and bathymetry grids were interpolated from the NCEI Digital Elevation Model Global Mosaic. In addition to providing curated reference site condition datasets for four U.S. areas, our assessment highlights the need for more publicly available metocean and soil condition data.

17 WIND ENERGY↗

CROCUS Low Cost All-in-One Weather Station AMB-001 Data Argonne National Laboratory Prairie Site

The Ambient Weather WS-2902D (AMB) is a low cost weather station that has become very useful for filling data gaps in harder to deploy locations. These low cost weather stations collect 13 second data, which is averaged to a five minute data output available to users through an API key. The data files contain measurements for precipitation, temperature, wind chill/heat index, relative humidity, dew point, UV index, solar radiation, wind speed, wind direction, wind gust, and with an external particulate matter 2.5 (PM 2.5) sensor. Having all of these measurements in one condense system allows for fast deploying and dense network capabilities. Three of the AMB weather stations were deployed at the Argonne Testbed for Multiscale Observational Science (ATMOS), a 20-acre prairie site at Argonne National Laboratory in Lemont, Illinois. The instruments are denoted by their three digit identifier (CMS-AMB-xxx) format. The data is presented as daily NetCDF (.nc) files, each containing approximately 24 hours of observations. Files follow the naming convention of: the project (CROCUS), location (atmos), instrument name (CMS-AMB-001), data level (raw, a1), and date (year, month, day). The NetCDF format can be accessed using common scientific software such as Python using xarray, netCDF4 or ACT-DOE.

54 ENVIRONMENTAL SCIENCES↗

CROCUS Low Cost All-in-One Weather Station AMB-002 Data Argonne National Laboratory Prairie Site

The Ambient Weather WS-2902D (AMB) is a low cost weather station that has become very useful for filling data gaps in harder to deploy locations. These low cost weather stations collect 13 second data, which is averaged to a five minute data output available to users through an API key. The data files contain measurements for precipitation, temperature, wind chill/heat index, relative humidity, dew point, UV index, solar radiation, wind speed, wind direction, wind gust, and with an external particulate matter 2.5 (PM 2.5) sensor. Having all of these measurements in one condense system allows for fast deploying and dense network capabilities. Three of the AMB weather stations were deployed at the Argonne Testbed for Multiscale Observational Science (ATMOS), a 20-acre prairie site at Argonne National Laboratory in Lemont, Illinois. The instruments are denoted by their three digit identifier (CMS-AMB-xxx) format. The data is presented as daily NetCDF (.nc) files, each containing approximately 24 hours of observations. Files follow the naming convention of: the project (CROCUS), location (atmos), instrument name (CMS-AMB-002), data level (raw, a1), and date (year, month, day). The NetCDF format can be accessed using common scientific software such as Python using xarray, netCDF4 or ACT-DOE.

54 ENVIRONMENTAL SCIENCES↗

Circularity Futures Workshop Series: Summary Report

The aim of this report is to synthesize key feedback received from the three-part Circularity Futures workshop series held in Spring 2024. The workshop series was conducted by the National Renewable Energy Laboratory (NREL) on behalf of U.S. Department of Energy, Office Energy Efficiency and Renewable Energy (EERE), and was broken into three workshops: Workshop 1 - Circularity Analysis Needs and Priorities; Workshop 2 - Circularity Metrics and Indicators; and Workshop 3 - Circularity Data. Together, the workshops focused on identifying the existing priorities and gaps in the circularity modeling space, understanding different stakeholders' use and interpretation of circularity metrics and indicators, identifying common data gaps and data quality challenges, and assessing the robustness of available solutions. The workshop series brought a diverse group of stakeholders - including representatives from U.S. government offices, national labs, nonprofit organizations, industry, and academia - to collect first-hand feedback on needs, priorities, challenges and opportunities in the circularity modeling and analysis space. The workshop discussions highlighted numerous common needs, priorities and challenges among the interviewed groups. Several topics were frequently discussed, including: 1) Circularity as a pathway for sustainable economic growth: While circularity is generally defined in terms of resource conservation and reducing wasteful disposal of materials, participants agreed that circular strategies should serve broader economic, environmental, and social goals. It is therefore crucial for circularity analysis to look beyond waste reduction and instead evaluate a variety of impact metrics such as cost savings, job creation, air quality, and pollutant emissions. Mutli-criteria decision-making frameworks may be useful for making sense of disparate metrics and evaluating tradeoffs between impact categories.; 2) Economic and social factors are not well understood: Underdevelopment of existing end-of-life (EOL) management infrastructure, inconsistent standardization codes and policy space in reusing recycled content, and suboptimal collection and sorting strategies collectively contribute to uncertainty about the economic potential of circular pathways. The latter observation is consistent among all technologies but more emphasized for renewable energy systems. Social impacts of circularity practices are less understood and less researched than other sustainability aspects.; 3) Inconsistent methods for assessing emerging technologies: LCA and TEA results vary widely depending on the assumptions made with regards to market adoption of new technologies. Emerging technologies suffer limited availability of data needed to conduct a robust circularity analysis. Yet, understanding projected impacts of proposed nascent technology is a key need for different stakeholder groups.; and 4) Lack of temporally and geospatially explicit data: There is a need for open data that represents variations in circularity technologies over time and location. The lack thereof leads to aggregated and potentially misrepresented results in circularity analysis. Sensitivity analyses should be included to verify whether options perceived as more sustainable align with real-world practices.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

An Overview of the Molten Salt Thermal Properties Database–Thermophysical, Version 4.0 (MSTDB-TP V.4.0)

A central repository of thermophysical and thermochemical properties of molten salt compositions of relevance to molten salt reactors (MSRs) is vital in supporting the broad community of MSR developers, who are at various stages of developing and deploying their reactor designs. In general, these MSR designs differ significantly from developer to developer (e.g., with respect to the hardness of the neutron spectra, level of fissile loading, target multicomponent temperatures and power levels, and moderating capabilities). Therefore, the fuel and coolant salts being considered vary greatly: they may be chlorides or fluorides, they utilize different actinides at different ratios, and the cations in the melt are selected based on perceived advantages and disadvantages. Considering the general need for thermal properties, and the vastness of the array of potential candidate salt mixtures, the Molten Salt Thermal Properties Database (MSTDB) was initiated in 2018 with the goal of providing thermophysical and thermochemical characterization of key molten salt compounds and mixtures across their temperature and compositional domains. The MSTDB is thus divided into the thermophysical arm (MSTDB-TP) and the thermochemical arm (MSTDB-TC). The MSTDB is an effort funded by the Department of Energy, Office of Nuclear Energy (DOE-NE) Nuclear Energy Advanced Modeling and Simulation (NEAMS) program, and the MSR Campaign. This report provides an overview of the MSTDB-TP v4.0 in terms of the data contained within, the state of the tools used to access the data, the availability of predictive models that leverage the raw data in the database, the preliminary status of developmental efforts that are currently underway, and an account of future goals for MSTDB-TP. The primary goal for the update from MSTDB-TP v.3.1 to v4.0 was the incorporation of surface tension data into the database; this property is important for thermal hydraulics modeling and species transport in other tools that have been developed under the NEAMS program. A breakdown of the surface tension data that have been added into MSTDB-TP v4.0 is provided herein, and the manner in which the quality of the data has been assessed is also documented. For MSTDB-TP v4.0, newly published thermophysical property data—primarily from collaborative experimental efforts under the MSR Campaign—have been incorporated into the database, and the resulting expansion is documented here. Because of the size to which MSTDB-TP has grown, the raw data format has now been recast into JavaScript Object Notation (JSON) format for easier connection with the MSTDB-TP application programming interface (API). Saline; the pre-existing comma-separated value (CSV) format has been deprecated but is still maintained, accessible, and up to date. As a final effort in packaging the MSTDB-TP v4.0 update, the graphical user interface (GUI) for MSTDB has been updated to allow full accessibility to the density and viscosity predictive models, which are based on Redlich-Kister expansions of MSTDB-TP raw data. Some other major aspects of this report, in terms of preliminary and future work, include: (1) documentation of the formalism and preliminary testing of a kinetic theory model that may act as a predictive model for thermal conductivity; (2) documentation of the candidate predictive models that may be considered in the future for surface tension, making use of the surface tension data now in MSTDB-TP v4.0; (3) a preliminary account of a data collection process that will enable the filling of additional gaps within MSTDB-TP, namely with data which have been collected computationally (e.g., through ab initio molecular dynamics).

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

FY24 Task 4: Studies of Phosphate and Fluoride Solubility for Dissolution of High Phosphate Tank Waste

Knowledge gaps have been identified in phosphate solubility in almost every single- and multi-component system, and particularly for aluminate in phosphate/hydroxide, where uncertainties in predicted values of aluminum solubility are greater than 50%. Solubility data of multicomponent, aqueous electrolytes containing sodium hydroxide, sodium fluoride, sodium phosphate, sodium nitrite, sodium nitrate, and dissolved gibbsite are also sparse. For example, for solutions of (i) sodium nitrate and sodium phosphate, data are only available at 30 and 50 °C, and for (ii) dissolved gibbsite in sodium hydroxide and sodium phosphate, only two data points are available at 20 and 40 °C. There is no data available on mixtures of sodium hydroxide, sodium phosphate, sodium fluoride, and dissolved gibbsite. These knowledge gaps were identified in a technical review of waste solubility data and the impact of dilution on solution stabilities and will be addressed in this work to predict aluminum hydroxide and sodium phosphate solubility in multicomponent electrolytes upon dilution, and upon variation of temperature. Results will provide the technical basis to develop accurate models for (i) gibbsite solubility and mass transfer of aluminum between solid and liquid forms following the sluicing of sludge and saltcake with water; and (ii) further blending of these suspensions with bismuth from bismuth phosphate waste, and zirconium and uranium left from the fuel decladding. This work will be essential to developing a disposition path for retrieval solutions from bismuth phosphate wastes. This effort would also support sludge washing to further reduce phosphate concentration if that process were to be added back to the flowsheet in the future.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Do we have globally representative data to understand soil processes?

Understanding and modeling soils and soil organic matter (SOM) are central to a variety of human needs, from food production to ecosystem management. Soil data have been collected for over a century, but the global spatial and process representativeness of soil data remains unclear. We assessed the representativeness of currently available soil data that could be used to understand a variety of SOM processes. We used 16 open-source soil databases and data from over 281,000 unique locations globally, categorizing the databases into three main data types necessary to understand SOM processes: soil carbon stocks and fluxes, mechanistic drivers of these stocks and fluxes, and soil carbon gain or loss potential. We found that stock and driver data have extensive global coverage. However, data on soil carbon gain or loss potential, particularly data describing change in soils over time such as time series data, are severely limited in their global coverage. We conclude that while significant strides have been made in measuring soil carbon stocks and fluxes, and their drivers, we are limited in global data related to changes in soils over time. Our recommendations for soil data generators are to ensure precise metadata reporting and prioritizing sampling in underrepresented areas like tropical, arctic, mountainous, wetland and arid regions. We also encourage designing revisit schemes that explicitly support change detection and reporting multi-modal datasets that can aid in model development. Targeted measurement of low coverage soil data types and regions is necessary for a range of applications including current and future biogeochemical predictions, and their management and policy implications.

carbon fluxes↗

Machine learning for domain transfer between simulated and experimental 2D X-ray diffraction patterns using generative adversarial networks

X-ray diffraction (XRD) is a well-established technique for analyzing materials at an atomic level. Dynamic compression experiments (DCE), in which materials are subject to extreme pressures, can provide fundamental understanding to pressure-induced phase transitions and compression of the crystal lattice. The analysis of XRD patterns from highly compressed samples is non-trivial given the sparsity of data, high experimental costs, and the fact that the data is often marred with X-ray background and other artifacts. While accurate computational frameworks exist, they solve the forward problem—from structures and orientations to XRD patterns. Solving the inverse problem for 2D experimental diffraction patterns is currently a complex manual process of matching and comparing experimentally observed patterns to computationally generated ones. Machine learning is a promising tool for automating the matching process but often requires data-intensive architectures. Here, in this study, we use a CycleGAN to translate the domain of limited experimental data to a domain in which there is readily available simulated data. This domain shift allows data-intensive machine learning models that have only been trained on simulated XRD patterns to be used in the analysis of experiments.

Brozak, Samantha Jean [Sandia National Laboratorie↗