Search NASASearch

SEARCH · Search NASA

Results for “Validation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

GeoCarb L2 Validation- Albedo Error Propagation into XCO2 in the Near Infrared

GeoCarb validation activities Pre-launch validation •Thermovac: Nov-Jan, 2022/3: Edwards/Dryden will move to Palo Alto Nov-March 2022/3 •Day-in-the life simulation-2022 •Simulations near Dallas, TX •Validation of TROPOMI in GeoCarb domain-2022 2GeoCarb post-launch GeoCarb will have gas precision of: •0.3% (1.2 ppm) for XCO2 •0.6% (11 ppb) for XCH4 •10% or 12 ppb for XCO GeoCarb SIF with precision of better than 0.75 W/m2/um/sr GeoCarb will utilize TCCON, GML/DOE aircraft, AirCore, and any validation site with systematic errors better than GeoCarb’s

GeoCarb

Validation of IMERG Oceanic Precipitation over Kwajalein

The integrated Multi-satellitE Retrievals for GPM (IMERG) Version V05B and V06B precipitation products from the Global Precipitation Measurement (GPM) mission are validated against ground-based observations from the Kwajalein Polarimetric S-band Weather Radar (KPOL) deployed at Kwajalein Atoll in the central Pacific Ocean. Such a validation is particularly important as comprehensive surface measurements over the oceans are practically infeasible, which hampers the identification of possible errors, and improvement of future versions of IMERG and other satellitebased retrieval algorithms. The V05B and V06B IMERG products are validated at their native 0.1°, 30 min resolution from 2014 to 2018 based on both volumetric and categorical metrics. This validation study indicates that precipitation rates from both IMERG V05B and V06B are underestimated with respect to radar surface estimates, but the underestimation is much reduced from V05B IMERG V06B outperforms V05B with reduced systematic bias and improved precipitation detectability. The IMERG performance is further traced back to its individual sensors and morphing-based algorithms. The overall underestimation in V05B is mainly driven by the negative relative biases from morphing-based algorithms which are largely corrected in V06B. Imagers perform generally better than sounders because of the usage of low-frequency channels in imagers which can better detect emission signals by the hydrometeors. Among imagers, the GPM Microwave Imager (GMI) and Advanced Microwave Scanning Radiometer Version 2 (AMSR2) are the best, followed by Special Sensor Microwave Imager/Sounder (SSMIS). Among sounders, the Microwave Humidity Sounder (MHS) is the best, followed by Advanced Technology Microwave Sounder (ATMS) and the Sounder for Atmospheric Profiling of Humidity in the Intertropics by Radiometry (SAPHIR) for V06B. Among all categories, morph-only and IR+morph only perform better than SAPHIR. SAPHIR shows the worst performance among all categories, likely due to its limited channel selection. It is envisaged that these results will improve our understanding of IMERG performance over oceans and aid in the improvement of future versions of IMERG.

KPOL

SPHINX: An SEP Model Validation Infrastructure developed through Community Challenges and the SEP Scoreboards

Solar Energetic Particle (SEP) events are interesting from a scientific perspective as they are the product of a broad set of physical processes from the corona out through the extent of the heliosphere, and provide insight into processes of particle acceleration and transport that are widely applicable in astrophysics. From the operations perspective, SEP events pose a radiation hazard for aviation, electronics in space, and human space exploration, in particular for missions outside of the Earth’s protective magnetosphere including to the Moon and Mars (Whitman et al 2022). For these reasons, SEP modelers have developed a rich and diverse set of models with a wide variety of aims. Some models probe the basic physics at the heart of particle acceleration and transport. Others produce fast statistical forecasts or employ disruptive new techniques like Machine Learning with the goal to assist end users in making operational decisions. To enable a consistent and quantitative understanding of SEP model performance, a generalized, automated validation infrastructure, called SPHINX, is being developed at NASA SRAG in close collaboration with NASA CCMC, NASA M2M, NOAA SWPC, and BIRA-IASB. This infrastructure has been built up through a multi-year community challenge. Starting in 2018 at the SHINE workshop, an effort was launched through SHINE, ISWAT, and ESWW to encourage quantitative, comprehensive, and consistent validation of SEP models. This effort has defined a set of challenge SEP events with the aim of generating quantitative comparisons between forecasts and observations and a set of challenge “non-events” to assess false alarms. In 2023, these challenge lists have been extended to statistically significant numbers with a prescribed set of rules for producing forecasts and supported through the dedicated SEPVAL working meetings. The participation of the research community has allowed the infrastructure to validate all the types of outputs being produced by SEP models. In parallel, the SPHINX code is being applied to real time forecasts submitted to the SEP Scoreboards, ensuring that the validation infrastructure can interpret forecasts produced in an operational scenario and provide metrics meaningful for operations. Upon completion, SPHINX and its interactive user interface, SPHINX-Web, will be made available for public use.

space weather

Pretest Simulations of a Supersonic Mixing and Combustion Validation Experiment to Assess Sensitivities

The reliance on CFD simulations to develop, design, and optimize scramjet systems (or components) has become commonplace. This reliance inevitably hinges on the ability of the computational analyst to quantify the level of confidence in their computational results. Unfortunately, nearly all the measured data available for this assessment comes from antiquated experimental datasets, or from tests that focused on the extraction of scramjet system (or component) performance. The objective of a CFD validation experiment is to quantify the predictive accuracy of one or more of the CFD physics submodels, implying that other uncertainties related to replicating the facility flow environment (e.g., knowledge of boundary conditions) must be minimized to the extent possible. This inevitably places stringent requirements on the quality and quantity of measurements taken to accurately specify inflow, outflow, and surface conditions for the CFD simulations; in addition to the measurements taken for the validation of physics submodels. This places additional demands on the experimental process above and beyond those for test article performance assessment. A recent high speed code credibility workshop series sponsored by AFRL identified a gap in existing validation data for fundamental assessments of turbulent mixing and combustion CFD closure models at scramjet engine relevant conditions. To address this gap, engineers at AFRL have designed a coaxial jet flame configuration that will be tested at two facilities (Research Cell 19 at the Air Force Research Lab, and at Purdue University). The effort described here documents pretest simulations of this validation experiment with the goal of fleshing out the extent of the facility flowpath that must be included to adequately reproduce the facility test section flow environment. The findings indicate that the flow around the support structure for the fuel injection centerbody upstream of the facility nozzle generates disturbances that persist throughout the nozzle expansion process; corrupting the azimuthal symmetry that was desired in the fuel/air mixing region of the test section. Simulations without this support structure maintained a high degree of azimuthal symmetry up until the fuel injection plane. However, even in this scenario the azimuthal symmetry was not maintained once the centerbody boundary layer transitioned to a wake flow downstream of the fuel injection plane.

CFD

Pretest Simulations of a Supersonic Mixing and Combustion Validation Experiment to Assess Sensitivities

The reliance on CFD simulations to develop, design, and optimize scramjet systems (or components) has become commonplace. This reliance inevitably hinges on the ability of the computational analyst to quantify the level of confidence in their computational results. Unfortunately, nearly all the measured data available for this assessment comes from antiquated experimental datasets, or from tests that focused on the extraction of scramjet system (or component) performance. The objective of a CFD validation experiment is to quantify the predictive accuracy of one or more of the CFD physics submodels, implying that other uncertainties related to replicating the facility flow environment (e.g., knowledge of boundary conditions) must be minimized to the extent possible. This inevitably places stringent requirements on the quality and quantity of measurements taken to accurately specify inflow, outflow, and surface conditions for the CFD simulations; in addition to the measurements taken for the validation of physics submodels. This places additional demands on the experimental process above and beyond those for test article performance assessment. A recent high speed code credibility workshop series sponsored by AFRL identified a gap in existing validation data for fundamental assessments of turbulent mixing and combustion CFD closure models at scramjet engine relevant conditions. To address this gap, engineers at AFRL have designed a coaxial jet flame configuration that will be tested at two facilities (Research Cell 19 at the Air Force Research Lab, and at Purdue University). The effort described here documents pretest simulations of this validation experiment with the goal of fleshing out the extent of the facility flowpath that must be included to adequately reproduce the facility test section flow environment. The findings indicate that the flow around the support structure for the fuel injection centerbody upstream of the facility nozzle generates disturbances that persist throughout the nozzle expansion process; corrupting the azimuthal symmetry that was desired in the fuel/air mixing region of the test section. Simulations without this support structure maintained a high degree of azimuthal symmetry up until the fuel injection plane. However, even in this scenario the azimuthal symmetry was not maintained once the centerbody boundary layer transitioned to a wake flow downstream of the fuel injection plane.

CFD

RELAP5-3D HTGR Validation Work at Idaho National Laboratory

Prismatic block-type high-temperature gas-cooled reactors (HTGRs) were built in the United States decades ago, and now advanced reactor vendors are seeking to deploy them again for a variety of applications. Deploying these reactors requires modelling and simulation tools that have been validated against conditions representative of the HTGR application. Idaho National Laboratory (INL) is leading the execution an of HTGR thermal hydraulics benchmark to accelerate the validation of thermal hydraulics modelling and simulation tools for these applications. That benchmark is based on a facility called the High Temperature Test Facility (HTTF). This work provides an overview of work conducted at INL over the last 2 years to validate RELAP5-3D against data from HTTF. This presentation shows results from multiple RELAP5-3D models and an HTTF experiment to assess the impact of certain modelling assumptions on results. The contents of this talk sit on the cutting edge of RELAP5-3D validation for HTGR analysis.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Modeling and Experimental Validation of a Direct-Contact Counter-Flow Fluidized Bed Heat Exchanger for Thermal Energy Storage (TES) Applications

Particle-based thermal energy storage (TES) systems are an emerging energy storage technology. The technological advances have reduced costs, making TES more competitive and reliable in the marketplace but an efficient and reliable operation is heavily dependent on coherent heat transfer between air to particles or vice versa. The particle-based TES technologies provide an intermediate system that can store energy for short (0-10 h), long (10-200 h) and seasonal (> 200 h) timescales. The TES systems store energy by converting electricity to thermal energy; electricity can be directly sourced intermittent generation technologies and/or the grid, helping manage peak loads and other mismatches in supply and demand. The overall efficiency of the TES system depends on the performance of system components (particle storage silos and particle transfer mechanism etc.). The particle heat exchanger is one of the key system components that affects the system efficiency. The pressurized fluidized bed heat exchanger (PFB HX) performance is challenging to predict due to the chaotic behavior of particle and fluid interaction. This research presents a computational study of a novel direct-contact, counter-flow and air-to-particles PFB HX, that contributes in advancing the particle-based long-duration TES technologies. For the current analysis an unsteady Eulerian-Eulerian CFD model was developed and validated against experiments performed at the National Laboratory of the Rockies for two particle sizes (600 ..mu..m and 825 ..mu..m ). Following validation, parametric simulations were conducted to evaluate the effects of interphase drag models (Syamlal-O'Brien and Gidaspow), particle size, bed height and the influence of a frictional-viscosity term on hydrodynamics and heat transfer between the air & particles. The key findings from the analysis are: (1) for the studied operating window Syamlal-O'Brien provides superior agreement with measured gas temperatures (errors generally < 10%) while Gidaspow shows large deviations for the coarse particle case; (2) model predictions are most sensitive in the lower 0.2 m above the air distributor where bubble initiation and local mixing dominate interphase heat transfer; (3) representation of the distributor (number of inlet ports) materially affects predicted local mixing and temperature stratification; and (4) the Eulerian-Eulerian framework reproduces bulk thermal trends but shows regime dependent limitations for coarse particles, motivating mesoscale informed closures for scale-up analysis for future studies. These results provide validated guidance for drag selection and distributor design in particle-based thermal energy storage applications. Collectively, the validated model and parametric results quantify key drivers of PHB-HX performance and provide practical guidance for design and optimization. The results provide confidence in the model predictability and provide a step forward to improve on heat exchange performance. The demonstrated performance and modeling approach support the deployment and further development of this novel PHB-HX concept for robust, particle-based long-duration thermal energy storage systems.

25 ENERGY STORAGE

Burst pressure models and validations for thick-walled pipelines containing corrosion defects

Corrosion is one major threat to pipeline integrity. Over the past decades, many corrosion models have been developed for determining the remaining strength of corroded pipelines, including ASME B31.G, Modified B31.G, LPC, PCORRC and their modified models. All these corrosion models are applicable only to large diameter, thin-walled pipelines with a diameter to wall thickness ratio D/t ≥ 20. In practice, many pipelines have a small diameter and thick wall with a D/t ratio < 20, and thus an adequate corrosion model is needed for assessing remaining strength for corroded thick-walled pipelines. This paper briefly reviews the theoretical burst pressure models for defect-free thin and thick-walled pipelines and four representative corrosion assessment models for thin-walled corroded pipelines. On this basis, two modified corrosion models are proposed to thick-walled pipelines in terms of the average shear stress yield theory. To verify the proposed corrosion models, comprehensive validations are performed. Numerical validations include the elastic-plastic finite element analysis to determine burst pressure for pipelines without and with corrosion defects and the model evaluation using a large dataset of available FEA results of burst pressure for machined defects. Experimental validations include a set of burst pressure tests for defect-free thick-walled pipes with different thicknesses and the model evaluation using one large burst dataset for machined defects with flat bottoms and another large dataset for real corrosion defects with curved river bottom profiles. Both numerical and experimental validations show that the proposed corrosion models can more accurately predict the remaining strength for corroded thin and thick-walled pipelines.

Pipeline

Application of research reactor environments for validation of the IRDFF-II dosimetry cross section library

Activation data from seven different reactor-based reference neutron fields are examined to provide enhanced validation evidence for the newly released IRDFF-II library. A least-squares based spectrum adjustment methodology is used and rigorous statistical metrics demonstrate the consistency of the set of IRDFF-II dosimetry cross sections. The use of these reactor-based neutron fields provides validation evidence for nine more reactions than have been addressed in the 252 Cf(sf) and 235 U(th) benchmark validation testing. The use of covers to shift their energy response increases the power of the response validation but makes it challenging to properly capture response correlations.

Griffin, Patrick

Monte Carlo toolkit for designing and validating step-range-filter spectrometer designs

Here, we present a Monte Carlo toolkit for validating step range filter (SRF) spectrometer designs. Geant4 is used to transport charged particles through the SRF filters to generate synthetic SRF data that include realistic CR-39 effects. Synthetic SRF spectra generated by this method inherently account for instrument response and allow for the quantification of SRF performance before shots. The usefulness of this toolkit is demonstrated through its application to a number of problems. A new broadband SRF for the ∼10 MeV wide 3He3He proton spectrum is validated, and an analysis method for analyzing 3He3He-p SRF data that accounts for instrument response is put forth. In addition, an SRF design for the compact recoil-proton spectrometer (CRS) on the Z-machine is validated. Finally, a new calibration technique for the DD-p SRF is proposed and validated.

Johnson, T. M. (ORCID:0000000193032949)

Automated annotation of scientific texts for ML-based keyphrase extraction and validation

Advanced omics technologies and facilities generate a wealth of valuable data daily; however, the data often lack the essential metadata required for researchers to find, curate, and search them effectively. The lack of metadata poses a significant challenge in the utilization of these data sets. Machine learning (ML)–based metadata extraction techniques have emerged as a potentially viable approach to automatically annotating scientific data sets with the metadata necessary for enabling effective search. Text labeling, usually performed manually, plays a crucial role in validating machine-extracted metadata. However, manual labeling is time-consuming and not always feasible; thus, there is a need to develop automated text labeling techniques in order to accelerate the process of scientific innovation. This need is particularly urgent in fields such as environmental genomics and microbiome science, which have historically received less attention in terms of metadata curation and creation of gold-standard text mining data sets. In this paper, we present two novel automated text labeling approaches for the validation of ML-generated metadata for unlabeled texts, with specific applications in environmental genomics. Our techniques show the potential of two new ways to leverage existing information that is only available for select documents within a corpus to validate ML models, which can then be used to describe the remaining documents in the corpus. The first technique exploits relationships between different types of data sources related to the same research study, such as publications and proposals. The second technique takes advantage of domain-specific controlled vocabularies or ontologies. In this paper, we detail applying these approaches in the context of environmental genomics research for ML-generated metadata validation. Our results show that the proposed label assignment approaches can generate both generic and highly specific text labels for the unlabeled texts, with up to 44% of the labels matching with those suggested by a ML keyword extraction algorithm.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION

Data-driven model validation for neutrino-nucleus cross section measurements

Neutrino-nucleus cross section measurements are needed to improve interaction modeling to meet the precision needs of neutrino experiments in efforts to measure oscillation parameters and search for physics beyond the Standard Model. We review the difficulties associated with modeling neutrino-nucleus interactions that lead to a dependence on event generators in oscillation analyses and cross section measurements alike. We then describe data-driven model validation techniques intended to address this model dependence. The method relies on utilizing various goodness-of-fit tests and the correlations between different observables and channels to probe the model for defects in the phase space relevant for the desired analysis. These techniques shed light on relevant mismodeling, allowing it to be detected before it begins to bias the cross section results. We compare more commonly used model validation methods which directly validate the model against alternative ones to these data-driven techniques and show their efficacy with fake data studies. These studies demonstrate that employing data-driven model validation in cross section measurements represents a reliable strategy to produce robust results that will stimulate the desired improvements to interaction modeling.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

SCEC/USGS Community Stress-Drop Validation Study: How Spectral Fitting Approaches Influence Measured Source Parameters

Spectral source parameters used to estimate an earthquake’s stress drop (⁠Δσ⁠) can vary significantly across measurement approaches. The Statewide California Earthquake Center/U.S. Geological Survey Community Stress‐Drop Validation Study was initiated to compare source parameter estimates, focusing initially on a dataset from the 2019 Ridgecrest earthquake sequence. As part of that validation effort, here we focus on one potential source of uncertainty: whether spectral fitting approaches alone, applied to a common set of spectra from the 2019 Ridgecrest sequence result in different source parameter estimates. By using a common set of benchmark spectra analyzed across a consistent frequency band of 1–40 Hz, we eliminate many sources of variability. A subgroup of validation study participants volunteered to estimate the low‐frequency displacement (⁠Ω 0 ⁠) and corner frequency (⁠ƒ c ⁠) by fitting a smooth function to benchmark displacement spectra. Participants used linear‐ or log‐sampled spectra, assumed a Brune or Boatwright spectral model, and applied different misfit criteria. We compare 17 approaches used to estimate ⁠Ω 0 ⁠, ƒ c ⁠, and Δσ for 54 earthquake spectra. Our results reveal that 35% of events have Δσ estimates within a factor of two, whereas others exhibit variations exceeding an order of magnitude. The variability in and can largely be attributed to whether a spectrum is consistent with the smooth function of an idealized simple crack model. The trade‐off between Ω 0 and ƒ c may be more pronounced when using linearly sampled spectra, as higher frequency spectral bumps control the fits. As expected, methods that assumed a Boatwright model tended to have lower Ω 0 and somewhat higher ƒ c compared to those assuming a Brune model, although resulting Δσ estimates are similar. Finally, when compared to the overall validation study results, the fitting approach alone may account for between 5% and 90% (25% on average) of the total variability in spectral Δσ⁠.

58 GEOSCIENCES

Development and Validation of Smart Building Technology Modules for Academic and Professional Education (Final Technical Report)

Smart building technologies can improve building energy efficiency and resilience, reduce carbon emissions, and provide load flexibility to the grid. However, in both college curricula and building professionals’ continuing education, there is a lack of systematic instruction on smart building technologies. Slipstream, partnering with Texas A&M University (TAMU), the Society of Building Science Educators (SBSE), and the National Institute of Building Sciences (NIBS), developed a semester-long smart building curriculum for college students and 16 training videos for building professionals and the general public. The education and training cover the drivers and benefits of smart building technologies, key building energy systems, the latest sensor technologies and IoT devices, and focus on topics related to smart building controls (i.e., energy management information systems, smart building control platforms, cybersecurity, grid-interactive-efficient buildings [GEBs], smart building control methods, and occupant-centric control). The smart building curriculum for college students was taught at TAMU in the Spring semester of 2024 as part of the validation process. Student feedback was collected and summarized in a validation report by TAMU. The curriculum material was also reviewed by SBSE faculty who are interested in teaching smart building technology-related courses. Suggestions on revisions and better adoption of the materials by other faculty across the architectural, engineering, and construction (AEC) domains were compiled in a distinct validation report by SBSE. The SBSE validation report was used to create structured subsets of the curriculum material for adoption at different levels in different sub-disciplines. These subsets are categorized and offered on the SBSE website (https://www.sbse.org/courses/Smart-Building-Technologies). The 16 training videos for building professionals and the general public were previewed by 17 industry experts, and feedback and suggested changes were incorporated into the final version of these videos. The videos are organized into a smart building technology training course and published on the Whole Building Design Guide website (https://www.wbdg.org/ce/doe/bto/sbtt), which is hosted by the National Institute of Building Sciences (NIBS). Project team members created marketing materials to promote the awareness of these free, publicly available education and training resources. Outreach and marketing activities included creating short promotional videos, building project webpages, making project announcements on social media, conducting an email campaign, and directly reaching out to faculties and building professionals. This report describes the project approach, provides outlines of the training materials, along with links to resources, and identifies lessons learned in creating the content. We also suggest ways to scale the instruction of smart building concepts to empower the workforce to accelerate the adoption of smart building technologies in the real world.

99 GENERAL AND MISCELLANEOUS

Final Design for Additional Thermal/Epithermal eXperiments (TEX) with Sodium Chloride Absorbers to Provide Validation Benchmarks for TerraPower

The first set of Thermal/Epithermal eXperiments (TEX) with chlorine absorbers (TEX-Cl) were executed in Q4FY24 and are in the process of being benchmarked for the ICSBEP. TEX-Cl builds upon the TEX-HEU baseline cases that were published in the 2022 International Criticality Safety Benchmark Evaluation Project (ICSBEP) Handbook. TEX-HEU, like TEX-Pu, was designed to be modular to allow for the incorporation of various absorbers and reflectors to test nuclear data and application case needs. For example, TEX-HEU with hafnium (TEX-Hf) utilizes hafnium plates as both absorbers and reflectors, depending on the tested configuration. A second set of chlorine experiments, dubbed More TEX-Cl, are laid out in this report to meet the needs of TerraPower for chlorine validation for their Molten Chloride Fast Reactor (MCFR) systems. TerraPower’s Molten Chloride Reactor Experiment (MCRE) and MCFR are fast molten salt reactors that utilize sodium chloride (NaCl) salt eutectics as the fuel and coolant. The MCRE eutectic is a mixture of NaCl and uranium trichloride (UCl 3 ). An abundant need for chlorine absorption validation has been expressed by multiple members of the community, including Y-12 (whose needs were addressed with the first set of experiments), LANL (whose needs were addressed with the Chlorine Worth Study (CWS)), TerraPower, institute de radioprotection et de sûreté nucléaire (IRSN), Savannah River Nuclear Solutions (SNRS), and others. Of the members who have expressed interest in this validation, most are interested in the fast neutron energy region, where the 35 Cl(n,p) reaction is most prominent. New 35 Cl(n,p) differential cross section measurements performed by LANL at LANCSE show substantial changes to the cross sections (Figure 1) and may be validated through these experiments as some configurations are optimally sensitive to this cross section.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

SCALE 6.3 Validation: Spent Nuclear Fuel

This report is the fifth volume in a series documenting the validation of SCALE 6.3 with ENDF/B-VII.1 libraries for nuclear criticality safety, reactor physics, radiation shielding, and spent nuclear fuel applications. This fifth volume, which focuses on validating SCALE capabilities that impact spent nuclear fuel applications, provides an update of the similar validation reported for SCALE 6.2.4. The experimental data used as basis for validation herein consist of measurement data for nuclide inventories and decay heat, including the following: 1. radiochemical assay (RCA) measurements of nuclides important to burnup credit, decay heat, and radiation shielding in 205 light-water reactor (LWR) spent nuclear fuel samples that cover burnups ranging up to 80 GWd/MTU and initial fuel enrichments up to 4.9% 235U; 2. full-assembly decay heat measurements for 236 LWR assemblies with initial fuel enrichments ranging up to 4% 235 U, assembly burnups of 5–51 GWd/MTU, and decay times after fuel discharge in the 2- to 27-year range (of importance to spent nuclear fuel storage, transportation, and disposal); and 3. pulse fission irradiations for fissionable materials at cooling times of interest to severe accident analyses (<10 5 s).

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Large Language Model for Validation, Optical Calibration, and Learning (VOCAL) Distributed Temperature Sensing Interface

Distributed temperature sensing (DTS) using fiber optic sensors (FOS) offers a promising method for temperature measurements in advanced reactors, such as sodium fast reactors and molten salt cooled reactors. To support the calibration and validation of DTS measurements, Argonne National Laboratory developed the Validation, Optical Calibration, and Learning (VOCAL) software package. This report describes the integration of a local large language model (LLM) with a retrieval-augmented generation (RAG) system into the VOCAL interface to serve as an interactive user assistant. The LLM framework enhances the VOCAL platform’s accessibility to users by explaining interface components, clarifying inputs and outputs, and answering user queries dynamically in real-time. The accuracy of the LLM assistant performance was evaluated with 20 queries regarding the interface and its parameters using experimental data from the Thermal Hydraulic Experimental Test Article (THETA) facility. Results demonstrate that the LLM achieved a 95% accuracy rate, with a BERTScore of 0.8816 and SBERT value of 0.7417. Furthermore, validation of the RAG system within the LLM framework showed optimal accuracy with k-values between 1 and 2 using the k-refinement convergence test. The prompt perturbation analysis demonstrated good initial consistency for the RAG system, exhibiting the highest accuracy under punctuation variations and the greatest sensitivity under query reordering. Notably, the model’s errors were limited to data retrieval failures rather than factual hallucinations, reinforcing its baseline reliability. The integration of LLM provides a highly accurate, userfriendly enhancement to the VOCAL platform without disrupting its core computational capabilities for FOS calibration and validation.

Hong, Evan

Statistical Validation of Multiple Related Data Sets—Case Study Using Interstellar Boundary Explorer Satellite Data

Abstract Space scientists often face the question of whether data collected by different instruments are measurements of the same source population. This paper proposes a statistical validation method for evaluating the agreement between such related data sets. It offers a detailed case study focused on validating a new data set from the Interstellar Boundary Explorer (IBEX) mission, which serves as a practical how-to guide for similar analyses. Since 2008, the IBEX satellite has been gathering data on heliospheric energetic neutral atoms (ENAs) while being exposed to various sources of background noise, such as cosmic rays and solar energetic particles. The IBEX mission initially released only a qualified triple-coincidence (qABC) data product, which was designed to provide observations of ENAs free of background contamination. Further measurements revealed that the qABC data were in fact susceptible to contamination, having relatively low ENA counts and high background rates. To mitigate this issue, the mission team recently considered releasing a certain qualified double-coincidence (qBC) data product, which has roughly twice the detection rate of the qABC data product. This paper presents a simulation-based validation of the new qBC data product against the already-released qABC data product. The results show that the qBCs can plausibly be said to be measuring the same source population as the qABCs up to an average absolute deviation of 3.6%. Visual diagnostics provide additional confirmation of source rate coherence across data products. The framework introduced here is general and can be applied to other validation problems both within and outside the field of space physics.

79 ASTRONOMY AND ASTROPHYSICS