Search NASASearch

SEARCH · Search NASA

Results for “text analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Topological Interpretability for Deep Learning

With the growing adoption of AI-based systems across everyday life, the need to understand their decision-making mechanisms is correspondingly increasing. The level at which we can trust the statistical inferences made from AI-based decision systems is an increasing concern, especially in high-risk systems such as criminal justice or medical diagnosis, where incorrect inferences may have tragic consequences. Despite their successes in providing solutions to problems involving real-world data, deep learning (DL) models cannot quantify the certainty of their predictions. These models are frequently quite confident, even when their solutions are incorrect. This work presents a method to infer prominent features in two DL classification models trained on clinical and non-clinical text by employing techniques from topological and geometric data analysis. We create a graph of a model's feature space and cluster the inputs into the graph's vertices by the similarity of features and prediction statistics. We then extract subgraphs demonstrating high-predictive accuracy for a given label. These subgraphs contain a wealth of information about features that the DL model has recognized as relevant to its decisions. We infer these features for a given label using a distance metric between probability measures, and demonstrate the stability of our method compared to the LIME and SHAP interpretability methods. This work establishes that we may gain insights into the decision mechanism of a DL model. This method allows us to ascertain if the model is making its decisions based on information germane to the problem or identifies extraneous patterns within the data.

Spannaus, Adam

Modification of Ni-20Cr corrosion dealloying behavior in molten fluorides via cold work induced plastic deformation

The corrosion dealloying behavior of cold-worked (CW) Ni20Cr alloy (wt%) was studied in molten LiF-NaF-KF (or FLiNaK) salts at 600 °C, equal to a homologous temperature (TH) of 0.52. Alloys were cold-rolled to achieve reductions of thickness of 10%, 30%, and 50% introducing plastic deformation and a high density of dislocations. Potentiostatic holds (Eapplied) were applied in two different electrode potential regimes. At 1.75VK+/K$${\text{V}}_{{\text{K}}^{+}/{\text{K}}}$$, Cr dealloying to Cr(II) and Cr(III) is predominant, while at 1.90 VK+/K$${\text{V}}_{{\text{K}}^{+}/{\text{K}}}$$, both Ni and Cr are oxidized in molten FLiNaK at 600 °C. In these potential regimes, dealloyed Ni20Cr displayed bicontinuous porosity within the grain interior and at grain boundaries, driven by the high driving force for Cr dissolution and sustained by defect mediated outward solid state diffusion of Cr in parallel with surface diffusion of Ni. The bicontinuous porous structure developed was observed to undergo further coarsening and densification of the Ni-rich ligaments at a higher electrode potential. The main effect of CW observed is the introduction of plastic deformation and dislocation substructures that serve as short-circuit paths for Cr solid state diffusion to surfaces exposed to FLiNaK. This modified the evolution of the bicontinuous porous structure which increased with CW substantially. Kinetic analysis reveals that the Cr dealloying at +1.75 VK+/K$${\text{V}}_{{\text{K}}^{+}/{\text{K}}}$$ and 1.90 VK+/K$${\text{V}}_{{\text{K}}^{+}/{\text{K}}}$$ is initially charge transfer controlled, except for the 50% CW condition at 1.90 VK+/K$${\text{V}}_{{\text{K}}^{+}/{\text{K}}}$$, where the process becomes limited by slow Cr defect mediated bulk diffusion. The rate determining factors are explored and compared to experimental results.

Chan, Ho Lun

A Model Based Approach to Extract Health Information from Textual Data

In current nuclear power plants (NPPs) a large amount of condition-based data is being generated and stored to assess and monitor component health and performance. The format of this data can be either numeric (e.g., pump vibration data) or textual (e.g., condition report which assess component health). While assessing component health from numeric data can be performed with a large variety of methods, the extraction of information from textual data still remains a challenge. Natural language processing (NLP) methods are starting to be deployed in current NPPs mainly to filter out incident reports (IRs) that are not safety related by employing supervised machine learning methods. However, these methods do not really provide the quantitative information that might be contained in IRs. This paper presents an approach to extract information from textual data (e.g., from IRs, maintenance reports) that is based on NLP data analytics methods coupled with model-based system engineer (MBSE) models. NLP methods are employed to perform syntactic and semantic analyses. Syntactic analysis analyzes the grammatical structure of a sentence; such analysis includes: part of speech (POS) tagging (i.e., identification of grammatic elements of each string - e.g., nouns, verbs), named entity recognition (i.e., identification of text entities - e.g., names, dates, events), and relation extraction (e.g., coreference resolution). On the other hand, semantic analysis is designed to analyze the logic structure of a sentence. Through a specific set of rules, our methods can identify whether a sentence contains health information of a component (e.g., degraded performance, anomaly behavior) or the causal relationship between two events (i.e., a cause-effect pair). An innovative element of our approach is that semantic analysis relies on MBSE models to identify links between textual elements. MBSE are diagrams designed to represent system and component dependencies (from both a form and functional point of view). In our approach, MBSE models emulate system engineer knowledge about component/system architecture. This paper presents in detail how the integration of NLP methods and MBSE models is performed. Few analysis examples focusing on centrifugal pumps are presented.

97 - MATHEMATICS AND COMPUTING

Libra

Libra is a Python package that extends support of its parent package, SQLAlchemy. Libra’s primary functionality supports the dynamic creation of object-oriented analogs of SQL tables from a variety of user-defined, text-based schema definition formats and provides quality control and analysis tools and methods

Spears, Brady

Mining Product Reviews for Important Product Features of Refurbished iPhones

Problem: Remanufacturers want to increase consumer interest in refurbished products, which motivates the need to understand which product features are important to buyers of refurbished products such as mobile phones. Research Questions: This study addresses two questions. First, which product features are most important for buyers of refurbished iPhones? Second, how do those preferences differ from the preferences of buyers of new iPhones? Methods: Online reviews of iPhones are obtained and converted into a document–term matrix. Using this text model, three subsets of features are identified using statistical analysis of frequency of mention: most frequent, average, and least frequent. A logistic regression (LR) model is then used to identify which features are most predictive of whether a review is for a new or refurbished phone. Results: Buyers of refurbished phones mention battery health, screen/display, shell condition, and brand significantly more often than other features. Directly contrasting reviews of refurbished versus new phones shows that shell condition, brand, speaker, and charger are found to be the most predictive product features indicated in reviews for refurbished phones. Of those, the shell condition is significantly more predictive than the others. Implications: The results identify product features that remanufacturers of iPhones can emphasize to increase customer demand.

Anisi, Atefeh

ARCH: Large-scale knowledge graph via aggregated narrative codified health records analysis

Objective: Electronic health record (EHR) systems contain a wealth of clinical data stored as both codified data and free-text narrative notes (NLP). The complexity of EHR presents challenges in feature representation, information extraction, and uncertainty quantification. Here, to address these challenges, we proposed an efficient Aggregated naRrative Codified Health (ARCH) records analysis to generate a large-scale knowledge graph (KG) for a comprehensive set of EHR codified and narrative features. Methods: Using data from 12.5 million Veterans Affairs patients, ARCH first derives embedding vectors and generates similarities along with associated p-values to measure the strength of relatedness between clinical features with statistical certainty quantification. Next, ARCH performs a sparse embedding regression to remove indirect linkage between features to build a sparse KG. Finally, ARCH was validated on various clinical tasks, including detecting known relationships between entity pairs, predicting drug side effects, disease phenotyping, as well as sub-typing Alzheimer’s disease patients. Results: ARCH produces high-quality clinical embeddings and KG for over 60,000 codified and narrative EHR concepts. The KG and embeddings are visualized in the R-shiny powered web-API.3 ARCH achieved high accuracy in detecting EHR concept relationships, with AUCs of 0.926 (codified) and 0.861 (NLP) for similar EHR concepts, and 0.810 (codified) and 0.843 (NLP) for related pairs. It detected drug side effects with a 0.723 AUC, which improved to 0.826 after fine-tuning. Using both codified and NLP features, the detection power increased significantly. Compared to other methods, ARCH has superior accuracy and enhances weakly supervised phenotyping algorithms’ performance. Notably, it successfully categorized Alzheimer’s patients into two subgroups with varying mortality rates. Conclusion: The proposed ARCH algorithm generates large-scale high-quality semantic representations and knowledge graph for both codified and NLP EHR features, useful for a wide range of predictive modeling tasks.

Electronic health records

Search for heavy resonances decaying into two Higgs bosons in the $\text {b}{\bar{\text {b}}} \tau ^{+} \tau ^{-}$ final state in proton–proton collisions at $\sqrt{s} = 13\,\text {Te}\hspace{-.08em}\text {V}$

A search is presented for massive narrow-width resonances in the mass range of 1–4.5 TeV , decaying into pairs of Higgs bosons (HH). The search uses proton–proton collision data at a center-of-mass energy of 13 TeV collected with the CMS detector at the CERN LHC during 2016–2018, corresponding to an integrated luminosity of 138 fb -1 . The analysis targets final states where one Higgs boson decays into a pair of bottom quarks and the other into a pair of tau leptons, ${\text {X}} \rightarrow {\text {HH}} \rightarrow \text {b}{\bar{\text {b}}}\,\tau ^{+}\tau ^{-}$. It uses a single large radius jet to reconstruct the ${\text {H}} \rightarrow \text {b}{\bar{\text {b}}}$ decay, while the ${\text {H}} \rightarrow \tau ^{+}\tau ^{-}$ decay products can either be contained within a single large radius jet or appear as two isolated tau leptons. The observed data are consistent with standard model background expectations. Upper limits at 95% confidence level are set on the production cross section for resonant HH production for masses between 1 and 4.5 TeV. This analysis sets the most sensitive limits to date on ${\text {X}} \rightarrow {\text {HH}} \rightarrow \text {b}\bar{\hbox {b}}\, \tau ^{+}\tau ^{-}$ decays in the mass range of 1.4–4.5 TeV.

Hayrapetyan, A. [Yerevan Physics Institute]

User’s Manual for RESRAD-RDD&IND Code Version 2: Vol. 2—User’s Guide for RESRAD-RDD&IND Code

Version 2.0 of the RESRAD-RDD&IND computer code is designed to support the implementation of protective action guides (PAGs) after a nuclear emergency incident including a radiological dispersal device (RDD) and/or an improvised nuclear device (IND) incident (EPA 2017). Eight different group types, addressing various decisions, are available for selection. The RESRAD-RDD&IND code calculates radiological doses, stay times, etc., for the selected group that the user wishes to focus on. (That is, the results for all the groups are not calculated simultaneously, and the input for those other groups do not matter, although some parameter values are shared between groups.) Version 2.0 has a user-friendly interface so that the RESRAD-RDD&IND code can be used with minimal training. For example, the user can select the major characteristics of the problem-event type, source term, and decision type from the left side of the interface and then calculate the results with the default assumptions for the exposure scenarios. More in-depth analysis would include specifying site-specific exposure scenario characteristics in the right side of the interface. The procedures for data entry and results viewing are self-explanatory. This is because common window maneuvering features and text instructions were incorporated in the interface design. General and context-specific help are available to aid users entering parameter values, as well. The RESRAD-RDD&IND computer code gives the user the option to select either an RDD or IND incident for analysis. For an RDD event analysis, 11 radionuclides (Am-241, Cf-252, Cm-244, Co-60, Cs-137, Ir-192, Po-210, Pu-238, Pu-239, Ra-226, and Sr-90) are included. These 11 radionuclides are the radionuclides most likely used for an RDD. More than 90 radionuclides can be selected for an IND event analysis. Initial default concentrations are provided for 44 radionuclides for a uranium-fueled IND event. These 44 radionuclides are those that would contribute significantly to the radiation dose associated with a uranium-fueled bomb detonation. The radionuclides generated from ingrowth of these 44 initial radionuclides are also automatically included in the analysis. Pu-239, Cs-134m, Ru-105, and Rb-89 and their progeny can be selected for analysis if they are detected and their concentrations are determined. This user’s guide, which is Volume 2 of the User’s Manual for RESRAD-RDD&IND Code Version 2, provides instructions to users on how to install the RESRAD-RDD&IND code, navigate the interface, and use the various features, including those discussed above, to set up an analysis and view/print the results in text outputs. Volume 1 of the User’s Manual for RESRAD-RDD&IND Code Version 2 (Yu et al. 2026), which contains descriptions of the methodology and theoretical basis for dose modeling and the mathematical equations implemented in the code, can be accessed and viewed through the Help menu in the code or can be downloaded from the RESRAD website (https://resrad.evs.anl.gov).

22 GENERAL STUDIES OF NUCLEAR REACTORS

TEMPEST3 surface runoff water chemistry and organic matter composition

Coastal flooding, driven by storm surges and sea level rise, can mobilize organic matter (OM) via runoff, while introducing compositionally distinct OM (e.g., estuarine OM) into the system. To understand event-scale OM dynamics, we monitored source waters and surface runoff during an ecosystem-scale field manipulation experiment, TEMPEST (Terrestrial Ecosystem Manipulation to Probe the Effects of Storm Treatments), in June 2024. The TEMPEST experiment is part of the COMPASS-FME (Coastal Observations, Mechanisms, and Predictions Across Systems and Scales – Field, Measurements, and Experiments) project and designed to investigate biogeochemical and ecological impacts of freshwater and seawater flooding on coastal terrestrial-aquatic interface ecosystems by simulating freshwater and seawater storm events in two 2000m2 coastal upland forest plots (freshwater and brackish seawater plots). The temporal coverage of this dataset is during the TEMPESTⅢ event (June 11-13, 2024). This dataset contains: - Surface runoff discharge measured by flumes - Sensor data (specific conductivity, salinity, dissolved oxygen, and temperature) - Particle size distribution - Total suspended sediment concentrations (TSS), particulate and dissolved organic carbon (POC, DOC) concentrations, total nitrogen and total dissolved nitrogen (TN, TDN) concentrations - Bulk particulate and dissolved OM compositions (stable C and N isotopes of particulates and optical measurements of chromophoric dissolved OM) - High resolution mass spectrometry analysis data - Water isotope data All data files are plain-text CSV (comma-separated value), and no special software is required to read them.

COMPASS-FME

Eucalyptus – An Analysis Suite for Fault Trees with Uncertainty Quantification

Eucalyptus is a novel code developed at Lawrence Livermore National Laboratory to incorporate uncertainty quantification into Fault Tree Analysis (FTA). This tool addresses the challenge of imperfect knowledge in “grey-box” systems by allowing analysts to incorporate and propagate uncertainty from component-level assessments to system-level effects. Eucalyptus facilitates a consistent evaluation of the impact of subject matter expert judgment and knowledge gaps on overall system response by Monte Carlo generation of possible system fault trees, sampling probabilities of the existence of subsystems and components. Here, the code supports the specification of fault trees through text and allows export to various formats, including auto-generated images, easing analysis and reducing errors. It has undergone extensive verification testing, demonstrating its reliability and readiness for deployment, and leverages on-node parallelism for rapid analysis. Example analyses are shown that include the identification of system failure paths and quantification of the value of further information about system components.

Fault Tree Analysis

From Text to Maps: LLM-Driven Extraction and Geotagging of Epidemiological Data

Epidemiological datasets are essential for public health analysis and decision-making, yet they remain scarce and often difficult to compile due to inconsistent data formats, language barriers, and evolving political boundaries. Traditional methods of creating such datasets involve extensive manual effort and are prone to errors in accurate location extraction. To address these challenges, we propose utilizing large language models (LLMs) to automate the extraction and geotagging of epidemiological data from textual documents. Our approach significantly reduces the manual effort required, limiting human intervention to validating a subset of records against text snippets and verifying the geotagging reasoning, as opposed to reviewing multiple entire documents manually to extract, clean, and geotag. Additionally, the LLMs identify information often overlooked by human annotators, further enhancing the dataset’s completeness. Our findings demonstrate that LLMs can be effectively used to semi-automate the extraction and geotagging of epidemiological data, offering several key advantages: (1) comprehensive information extraction with minimal risk of missing critical details; (2) minimal human intervention; (3) higher-resolution data with more precise geotagging; and (4) significantly reduced resource demands compared to traditional methods.

Harrod, Karly

DESI DR1 Ly$α$ forest: 3D full-shape analysis and cosmological constraints

We perform an analysis of the full shapes of Lyman-$α$ (Ly$α$) forest correlation functions measured from the first data release (DR1) of the Dark Energy Spectroscopic Instrument (DESI). Our analysis focuses on measuring the Alcock-Paczynski (AP) effect and the cosmic growth rate times the amplitude of matter fluctuations in spheres of $8$$h^{-1}\text{Mpc}$, $fσ_8$. We validate our measurements using two different sets of mocks, a series of data splits, and a large set of analysis variations, which were first performed blinded. Our analysis constrains the ratio $D_M/D_H(z_\mathrm{eff})=4.525\pm0.071$, where $D_H=c/H(z)$ is the Hubble distance, $D_M$ is the transverse comoving distance, and the effective redshift is $z_\mathrm{eff}=2.33$. This is a factor of $2.4$ tighter than the Baryon Acoustic Oscillation (BAO) constraint from the same data. When combining with Ly$α$ BAO constraints from DESI DR2, we obtain the ratios $D_H(z_\mathrm{eff})/r_d=8.646\pm0.077$ and $D_M(z_\mathrm{eff})/r_d=38.90\pm0.38$, where $r_d$ is the sound horizon at the drag epoch. We also measure $fσ_8(z_\mathrm{eff}) = 0.37\; ^{+0.055}_{-0.065} \,(\mathrm{stat})\, \pm 0.033 \,(\mathrm{sys})$, but we do not use it for cosmological inference due to difficulties in its validation with mocks. In $Λ$CDM, our measurements are consistent with both cosmic microwave background (CMB) and galaxy clustering constraints. Using a nucleosynthesis prior but no CMB anisotropy information, we measure the Hubble constant to be $H_0 = 68.3\pm 1.6\;\,{\rm km\,s^{-1}\,Mpc^{-1}}$ within $Λ$CDM. Finally, we show that Ly$α$ forest AP measurements can help improve constraints on the dark energy equation of state, and are expected to play an important role in upcoming DESI analyses.

Cuceu, Andrei [LBL, Berkeley; Chicago U., KICP] (O

Jackson, L., Johnson, M.B., Latrach, A., Grimes, D., Martinez, C., and Mclaughlin, J.F., 2024, Multidisciplinary geotechnical data collection, curation, and analysis for conformity with the regulatory framework for geologic carbon storage in Wyoming, USA: Geological Society of America Abstracts with Programs. Vol. 56, No. 5, 2024, doi: 10.1130/abs/2024AM-405024

Title: Multidisciplinary Geotechnical Data Collection, Curation, and Analysis for Conformity with the Regulatory Framework for Geologic Carbon Storage in Wyoming, USA. Text: Construction and operation of wells for geologic sequestration of carbon dioxide necessitate that they are permitted under the Environmental Protection Agency’s Underground Injection Control Class VI requirements. Class VI wells conform to stringent requirements to ensure long-term safety and integrity of the storage site and the protection of Underground Sources of Drinking Water. Entities pursuing Class VI permitting must provide comprehensive geologic site characterization, including regional geologic structure and stratigraphy, aquifer information, reservoir and confining unit geomechanical properties, geochemical analyses, assessment of trapping capacity and mechanisms, and a variety of other of multidisciplinary geotechnical data. The Wyoming Class VI Site Characterization Database Project is focused on developing a geologic site characterization database of geotechnical information, which has been compiled and verified from established, public databases/entities and scientific literature to expedite Class VI permitting in Sweetwater County within the Greater Green River Basin of southern Wyoming. The preliminary suite of compiled data from 14,000 wells includes 8,000 wells with logs and 7,250 wells with formation tops, ~70 wells with core data (e.g., X-Ray diffraction, petrographic, and petrophysical data), ~2,500 water analyses, ~740 seismic events data, and ~520 bottom-hole temperature measurements. Future work on—and stemming from—this project will include new core analyses, calculation and interpolation of subsurface temperature gradients, mechanical earth models, geochemical simulations, storage capacity estimation, stratigraphic column generation and correlation, and construction of subsurface maps. Finally, this work will help to inspire and facilitate subsurface data compilation and curation beyond Sweetwater County, Wyoming.

42 ENGINEERING

Constraints on anomalous Higgs boson couplings from its production and decay using the WW channel in proton–proton collisions at $\sqrt{s} = 13~\text {TeV}$

A study of the anomalous couplings of the Higgs boson to vector bosons, including ${\textit{CP}}$-violation effects, has been conducted using its production and decay in the WW channel. This analysis is performed on proton–proton collision data collected with the CMS detector at the CERN LHC during 2016–2018 at a center-of-mass energy of 13 TeV, and corresponds to an integrated luminosity of 138$\,\text {fb}^{-1}$. The different-flavor dilepton $({\textrm{e}} {{\upmu }})$ final state is analyzed, with dedicated categories targeting gluon fusion, electroweak vector boson fusion, and associated production with a W or Z boson. Kinematic information from associated jets is combined using matrix element techniques to increase the sensitivity to anomalous effects at the production vertex. A simultaneous measurement of four Higgs boson couplings to electroweak vector bosons is performed in the framework of a standard model effective field theory. All measurements are consistent with the expectations for the standard model Higgs boson and constraints are set on the fractional contribution of the anomalous couplings to the Higgs boson production cross section.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

Multicriteria Measures to Assess the Sustainability of Diets: A Systematic Review

Abstract Context Assessing the overall sustainability of a diet is a challenging undertaking requiring a holistic approach capable of addressing the multicriteria nature of this concept. Objective The aim was to identify and summarize the multicriteria measures used to assess the sustainability characteristics of diets reported at the individual level by healthy adults. Data Sources Articles were identified via PubMed, Scopus, and Web of Science. The search strategy consisted of key words and MeSH terms, and was concluded in September 2022, covering references in English, Spanish, and Portuguese. Data Extraction This systematic review followed the PRISMA guidelines. The search identified 5663 references, from which 1794 were duplicates. Two reviewers independently screened the titles and abstracts of each of the 3869 records and the full-text of the 144 references selected. Of these, 7 studies met the inclusion criteria. Data Analysis A total of 6 multicriteria measures were identified: 3 different Sustainable Diet Indices, the Quality Environmental Costs of Diet, the Quality Financial Costs of Diet, and the Environmental Impact of Diet. All of these incorporated a health/nutrition dimension, while the environmental and economic dimensions were the second and the third most integrated, respectively. A sociocultural sustainability dimension was included in only 1 of the measures. Conclusion Despite some methodological concerns in the development and validation process of the identified measures, their inclusion is considered indispensable in assessing the transition towards sustainable diets in future studies. Systematic Review Registration PROSPERO registration no. CRD42022358824.

Rei, Mariana (ORCID:0000000189453708)

A Solid State Zwitterionic Plastic Crystal with High Static Dielectric Constant

The dielectric data in Figure 3, Figure 4, Figure S6 of the published paper was extracted from 2EOIMTSA-BDS-DATA .txt file. This file can be directly opened using a text file editor. It can also be imported to Excel/ Origin for further plotting and analysis. The G' and G'' in Figure 3 of the publihsed paper was plotted from data in file 2EOImTSA-temperature-sweep.xlsx. This file can be directly opend using Excel. The details of DFT simulations mentioned in Figure 2, Figure 7, and Figure S9 of the published paper are included in the DFT.zip file.

Huang, Zitan [Pennsylvania State University]

Augmented Reality Technologies for Radiation Safety Training: A Systematic Review of Sensor Integration and Visualization Approaches

This paper presents a comprehensive systematic review examining the application of augmented reality (AR) and sensor technologies for visualizing ionizing radiation in virtual training environments. The review methodology involved systematic identification and analysis of the relevant literature based on predetermined criteria including publication type, year of publication, application domain, and technological approach. The literature search encompassed publications from 2011 to 2021 across four major academic databases: Web of Science, Google Scholar, IEEE Xplore, and Scopus. Through rigorous screening following PRISMA 2020 guidelines, 23 research articles met the inclusion criteria for detailed analysis. From 404 initial database records, 360 were excluded during title/abstract screening (primarily for lacking AR components, radiation focus, or training applications) and 4 during full-text assessment (all for lacking sensor integration). The findings reveal that AR-based ionizing radiation visualization has been successfully implemented across diverse domains, including nuclear facility operations, medical procedures, CERN research activities, and educational and monitoring applications. The analysis identified multiple dimensions of impact, encompassing distinct benefits, emerging opportunities, and implementation challenges associated with AR deployment for ionizing radiation training. Each of these dimensions is comprehensively examined and documented within this review. Additionally, this study identifies critical research gaps that currently limit the full potential of AR technology in supporting ionizing radiation training programs. These gaps are systematically analyzed and discussed to establish clear directions for future research endeavors in this emerging field.

61 - RADIATION PROTECTION AND DOSIMETRY

Debiasing Watermarks for Large Language Models via Maximal Coupling

Watermarking language models is essential for distinguishing between human and machine-generated text and thus maintaining the integrity and trustworthiness of digital communication. Here, we present a novel green/red list watermarking approach that partitions the token set into “green” and “red” lists, subtly increasing the generation probability for green tokens. To correct token distribution bias, our method employs maximal coupling, using a uniform coin flip to decide whether to apply bias correction, with the result embedded as a pseudorandom watermark signal. Theoretical analysis confirms this approach’s unbiased nature and robust detection capabilities. Experimental results show that it outperforms prior techniques by preserving text quality while maintaining high detectability, and it demonstrates resilience to targeted modifications aimed at improving text quality. This research provides a promising watermarking solution for language models, balancing effective detection with minimal impact on text quality.

97 MATHEMATICS AND COMPUTING