Search NASA⌕ Search

SEARCH · Search NASA

Results for “Human Performance”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Tractometry of the Human Connectome Project: resources and insights

The Human Connectome Project (HCP) has become a keystone dataset in human neuroscience, with a plethora of important applications in advancing brain imaging methods and an understanding of the human brain. We focused on tractometry of HCP diffusion-weighted MRI (dMRI) data. We used an open-source software library (pyAFQ; https://yeatmanlab.github.io/pyAFQ) to perform probabilistic tractography and delineate the major white matter pathways in the HCP subjects that have a complete dMRI acquisition (n = 1,041). We used diffusion kurtosis imaging (DKI) to model white matter microstructure in each voxel of the white matter, and extracted tract profiles of DKI-derived tissue properties along the length of the tracts. We explored the empirical properties of the data: first, we assessed the heritability of DKI tissue properties using the known genetic linkage of the large number of twin pairs sampled in HCP. Second, we tested the ability of tractometry to serve as the basis for predictive models of individual characteristics (e.g., age, crystallized/fluid intelligence, reading ability, etc.), compared to local connectome features. To facilitate the exploration of the dataset we created a new web-based visualization tool and use this tool to visualize the data in the HCP tractometry dataset. Finally, we used the HCP dataset as a test-bed for a new technological innovation: the TRX file-format for representation of dMRI-based streamlines. We released the processing outputs and tract profiles as a publicly available data resource through the AWS Open Data program's Open Neurodata repository. We found heritability as high as 0.9 for DKI-based metrics in some brain pathways. We also found that tractometry extracts as much useful information about individual differences as the local connectome method. We released a new web-based visualization tool for tractometry—“Tractoscope” (https://nrdg.github.io/tractoscope). We found that the TRX files require considerably less disk space-a crucial attribute for large datasets like HCP. In addition, TRX incorporates a specification for grouping streamlines, further simplifying tractometry analysis.

59 BASIC BIOLOGICAL SCIENCES↗

Next-Level Energy Management in Manufacturing: Facility-Level Energy Digital Twin Framework Based on Machine Learning and Automated Data Collection

This research introduces an energy prediction framework at the facility level supported by automated data collection and machine learning models. It investigates whether reducing the prediction time scale allows for applying more complex machine learning techniques and if those techniques improve the prediction accuracy. The primary advantages of this framework lie in its automation of the energy prediction process and its provision of real-time energy data suitable for use in energy dashboards or digital twins. A sitewide dataset was created by combining 15 min energy and daily production data of five shops—assembly, battery, body (electric), body (gas), and paint—from a globally recognized electric vehicle manufacturer. Various machine learning models were evaluated on daily, weekly, and monthly datasets, including, in increasingly complex order: naïve, simple linear regression, net regularized generalized linear regression, principal component regression, k-nearest neighbor, random forest, and Bayesian regularized neural network. Compared to the current state-of-the-art energy consumption prediction for the industrial facility level, this research investigates more complex models and smaller time intervals for higher accuracy. The findings revealed that the more complex monthly models require a minimum of a year and a half of data to operate, while weekly models demand a year of data to achieve improved accuracy. Daily models can operate with only six months of data but exhibit poor performance due to reduced prediction accuracy of production. Key challenges identified include access to reliable, high-quality energy and production data and the initial demand for human labor.

digital twin↗

DEVELOPMENT AND DEMONSTRATION TESTBED FOR THE REMOTE OPERATIONS AND MONITORING OF MICROREACTORS

The nuclear industry is rapidly developing many advanced-reactor concepts for near-term deployment in both traditional and non-traditional nuclear-powered applications. One such category of advanced reactor is the microreactor, a class of reactor with less than 20MWth power output, intended for applications where the economics or logistics of traditional power sources are difficult. This includes applications such as remote communities, mining sites, defense installations, or humanitarian and disaster-relief missions. One key enabling feature for the successful deployment of microreactors is a remote operations capability. Remote operations provide monitoring and control capabilities which can significantly reduce staffing costs by eliminating the need for licensed operators at each reactor facility and improve the economic viability for microreactor deployment. A remote concept of operations is not currently an established capability in the nuclear industry. In addition, no demonstration microreactor is expected to complete construction or go critical until at least 2026. This leaves two major capability gaps: the successful demonstration of a remote concept of operations for microreactors and a test bed suitable for said demonstration. Both gaps must be addressed in order to advance the remote concepts of nuclear operation and, more broadly, microreactors themselves from paper to reality. This paper aims to fill these gaps and describes a test bed that would support development and deployment of a remote concept of nuclear operations, initial experimental results from that test bed, and the application of the test bed and experimental results for a digital-twin-based remote concept of operations underdevelopment at Idaho National Laboratory (INL). The platform chosen as a remote concept of nuclear operations test bed is the Single Primary Heat Extraction and Removal Emulator, known as SPHERE, located at INL. SPHERE is a small-scale non-nuclear test bed that emulates thermal behavior of a microreactor. The small-scale and non-nuclear nature of SHPERE limit safety concerns associated with remote operations while still providing the physical response representative of a microreactor. A network connection was added to SPHERE that enables remote-monitoring capability. This allows for real-time data streaming to networked workstations, data historians, and human-machine interfaces (HMIs). These are all critical components in a remote concept of operations, thus providing a robust development and demonstration platform. An initial experiment was performed using the SPHERE remote operations testbed. This included running a comprehensively instrumented SPHERE through a series of steady-state and transient operating scenarios in both normal and abnormal operating conditions, all while streaming live test data to a remote HMI and data warehouse. This initial experiment served three purposes: (1) characterizing the response of SPHERE, (2) demonstrating the remote connection to SPHERE, and (3) providing a baseline data set for development of a digital-twin-based remote concept of operations that is under development at INL.

22 - GENERAL STUDIES OF NUCLEAR REACTORS↗

Evaluation of dried blood spot sampling for verification of exposure to chemical threat agents

Abstract Purpose Exposure to chemical threat agents (CTAs), including nerve agents, the vesicating agent sulfur mustard, and opioids, remains a significant threat to warfighter and civilian populations. Definitive analytical methods to verify exposure to CTAs require shipping refrigerated or frozen biomedical samples to reference laboratories for analysis. Logistical and financial burdens arise as the transport of biomedical samples is subject to strict restrictions and complex packaging, which, if done incorrectly, can lead to sample deterioration. The use of dried blood spot (DBS) sampling could provide operational improvements for collecting, storing, and shipping important forensic samples. Therefore, this effort focuses on developing DBS techniques with Mitra® 30-µL volumetric absorptive microsampling (VAMS®) devices for use in CTA exposure verification. Methods VAMS® devices were loaded and dried with human whole blood that was exposed to the metabolites pinacolyl methylphosphonic acid (PMPA), ethyl methylphosphonic acid (EMPA), 1,1’sulfonylbis[2-(methylsulfinyl)ethane] (SBMSE), norfentanyl, norcarfentanil, norsufentanil, and norlofentanil. Following extraction from the VAMS® devices, metabolites were detected using liquid chromatography-tandem mass spectrometry (LC–MS/MS). The methods were validated for performance by assessing sensitivity, precision, accuracy, and recovery. Results These methods were sensitive to 1 ng/mL for SBMSE, 0.5 ng/mL for PMPA, EMPA, and norfentanyl; 0.1 ng/mL for norlofentanil, and 0.05 ng/mL for norsufentanil and norcarfentanil. All methods met acceptable precision and accuracy criteria with favorable recovery. Conclusions These results demonstrated the utility of VAMS® in stabilizing human whole blood and show promise as an improved collection method for verification of exposure to various CTAs.

Toxicology↗

Aryl hydrocarbon receptor-dependent toxicity by retene requires metabolic competence

Polycyclic aromatic hydrocarbons (PAHs) are a class of organic compounds frequently detected in the environment with widely varying toxicities. Many PAHs activate the aryl hydrocarbon receptor (AHR), inducing the expression of a battery of genes, including xenobiotic metabolizing enzymes like cytochrome P450s (CYPs); however, not all PAHs act via this mechanism. We screened several parent and substituted PAHs in in vitro AHR activation assays to classify their unique activity. Retene (1-methyl-7-isopropylphenanthrene) displays Ahr2-dependent teratogenicity in zebrafish, but did not activate human AHR or zebrafish Ahr2, suggesting a retene metabolite activates Ahr2 in zebrafish to induce developmental toxicity. To investigate the role of metabolism in retene toxicity, studies were performed to determine the functional role of cyp1a, cyp1b1, and the microbiome in retene toxicity, identify the zebrafish window of susceptibility, and measure retene uptake, loss, and metabolite formation in vivo. Cyp1a-null fish were generated using CRISPR-Cas9. Cyp1a-null fish showed increased sensitivity to retene toxicity, whereas Cyp1b1-null fish were less susceptible, and microbiome elimination had no significant effect. Zebrafish required exposure to retene between 24 and 48 hours post fertilization (hpf) to exhibit toxicity. After static exposure, retene concentrations in zebrafish embryos increased until 24 hpf, peaked between 24 and 36 hpf, and decreased rapidly thereafter. We detected retene metabolites at 36 and 48 hpf, indicating metabolic onset preceding toxicity. This study highlights the value of combining molecular and systems biology approaches with mechanistic and predictive toxicology to interrogate the role of biotransformation in AHR-dependent toxicity.

59 BASIC BIOLOGICAL SCIENCES↗

A prototype cooling blanket for mitigating occupant overheating risk in a hot indoor environment: Modeling and assessments

Conventional ways of cooling a room or an entire house for occupant thermal comfort during summer consume a significant amount of energy and are vulnerable to overheating risk during power outages that lead to loss of cooling system operations. This study investigates a low-power cooling blanket, as a Personal Cooling System (PCS), that covers the upper human body for direct cooling during a five-day heat wave in a single-family house. A modeling framework is developed for evaluating the thermal and energy performance of the cooling blanket, which builds upon the co-simulation of three models: a house energy model, a personal thermal comfort model, and a cooling blanket model. Simulation results show that under the power outage scenario, the cooling blanket can greatly reduce the occupant heat stress with a reduction of daily hours of exceedance (discomfort hours defined as TSV>2) by up to 17.2 h (a 95.3 % improvement from the baseline power outage without the blanket). The cooling blanket, equipped with an innovative electrocaloric heat pump (COP as high as 10.1) consumes 6.31 W and can be operated by a portable battery for several days. The cooling blanket consumes only 0.28 % of the electricity of a central air-conditioning system running to provide cooling for the whole house during the five-day heatwave period. The findings justify further research of electrocaloric wearable PCS as low-power effective cooling to ensure thermal survivability of occupants during extreme indoor environments.

Electrocaloric heat pump↗

Circadian immunometabolic states impart a temporal response to SARS-CoV-2 spike proteins in mammalian macrophages

Circadian rhythms, the 24-hour cycles that tune organismal physiology to the daily rhythms of light and dark, optimally organize cellular processes such as metabolism and mitochondrial function. In mammals, macrophage functions are regulated by these 24-hour circadian rhythms such that the immunometabolic response is coordinated across the day, consolidating macrophage physiology into temporally distinct phases to time the cellular immune response. However, while it is known that there are time-of-day specific responses to stress in a macrophage, little has been done to determine if circadian regulation coordinates the response of a macrophage to real-world pathogens. Importantly, key proteins in the response to viral infection have been found to be under circadian control, and time of day of application is known to affect the efficacy of vaccinations, including in the case of the COVID-19 virus. Therefore, to investigate if the circadian regulation of macrophage physiology imparted a time-of-day response to viral exposure, we exposed primary mouse and human macrophages to the SARS-CoV-1 and CoV-2 spike proteins at different times over the circadian day. To establish a time-of-day effect, we performed a multi-omics analysis and in vitro tissue culture assays examining macrophage responses over circadian time. We found that, conserved across the species, the timing of spike protein exposure dictated two distinct temporal responses which were characterized by hallmarks of immunometabolic suppression and modest inflammatory activation. However, these responses were primarily influenced by central metabolic and mitochondrial changes and not by classical immune activation.

Circadian Biology↗

Robotic Assembly of SRF Cavity Pair

Superconducting Radio Frequency (SRF) cavities for particle accelerators require tight tolerances, ultrahigh vacuum, and strict cleanliness during assembly. As in the semiconductor industry, defects such as particles and residues are deleterious to performance, possibly rendering a cavity unfit for use. This problem is addressed primarily through cleanroom assembly during sensitive steps and rigorous chemical processing to prevent and remove such defects. Human workers are often the largest source of contamination, even with proper gowning and practices. The semiconductor industry has long integrated robots in cleanroom operation, but this has not occurred for SRF cavity production; SRF cavities, unlike wafers, are complex shapes, require more hands-on mechanical assembly, and are low-volume production items. At Jefferson Laboratory, a co-operative robot (cobot) has been setup to overcome these problems. Cobots are safe for use alongside human workers and can integrate new tools such as a 3D camera part detection and gripper for item manipulation. Therefore, cobots represent a promising avenue for reducing particulate generated during a variety of assembly tasks. A mockup of a cavity pair and coupler was setup and the cobot programmed to automatically pick up the coupler and place on the mating cavity flange. Particle counting methods were setup to measure human vs cobot assembly particulate generation inside a cavity mockup. Other potential uses will be discussed for further improving SRF cavity assembly steps, where a cobot can replace or assist a human operator, and what potential gains are expected.

Duzik, Adam [Thomas Jefferson National Accelerator↗

Systematic benchmarking demonstrates large language models have not reached the diagnostic accuracy of traditional rare-disease decision support tools

Large language models (LLMs) show promise in supporting differential diagnosis, but their performance is challenging to evaluate due to the unstructured nature of their responses, and their accuracy compared to existing diagnostic tools is not well characterized. To assess the current capabilities of LLMs to diagnose genetic diseases, we benchmarked these models on 5213 previously published case reports using the Phenopacket Schema, the Human Phenotype Ontology and Mondo disease ontology. Prompts generated from each phenopacket were sent to seven LLMs, including four generalist models and three LLMs specialized for medical applications. The same phenopackets were used as input to a widely used diagnostic tool, Exomiser, in phenotype-only mode. The best LLM ranked the correct diagnosis first in 23.6% of cases, whereas Exomiser did so in 35.5% of cases. While the performance of LLMs for supporting differential diagnosis has been improving, it has not reached the level of commonly used traditional bioinformatics tools. Future research is needed to determine the best approach to incorporate LLMs into diagnostic pipelines.

Reese, Justin T. [Lawrence Berkeley National Labor↗

Hierarchical semi-Markov models with duration-aware dynamics for activity sequences

Residential electricity demand at granular scales is driven by what people do and for how long. Accurately forecasting this demand for applications like microgrid management and demand response therefore requires generative models for activities that can produce realistic daily activity sequences, capturing both the timing and duration of human behavior. This paper develops a generative model of human activity sequences using nationally representative time-use diaries at a 10-min resolution. We use this model to quantify which demographic factors are most critical for improving predictive performance. We propose a hierarchical semi-Markov framework that addresses two key modeling challenges. First, a time-inhomogeneous Markov router learns the patterns of “which activity comes next.” Second, a semi-Markov hazard component explicitly models activity durations, capturing “how long” activities realistically last. To ensure statistical stability when data are sparse, the model pools information across related demographic groups and time blocks. The entire framework is trained and evaluated using survey design weights to ensure our findings are representative of the U.S. population. On a held-out test set, we demonstrate that explicitly modeling durations with the hazard component provides a substantial and statistically significant improvement over purely Markovian models. Furthermore, our analysis reveals a clear hierarchy of demographic factors: Sex, Day-Type, and Household Size provide the largest predictive gains, while Region and Season, though important for energy calculations, contribute little to predicting the activity sequence itself. The result is an interpretable and robust generator of synthetic activity traces, providing a high-fidelity foundation for downstream energy systems modeling.

24 POWER TRANSMISSION AND DISTRIBUTION↗

An agentic artificially intelligent X-ray scientist

Executing experimental tasks in both normal research laboratories and large-scale scientific facilities often requires extensive human supervision and remains a key challenge on the path to fully autonomous, artificial intelligence (AI)-driven science. Here we demonstrate a large language model-driven agent that autonomously performs X-ray sample alignment on a synchrotron beamline by planning actions, executing instrumental commands, interpreting observations and iterating towards experimental goals. Based on existing large language models with structured tool-use via the model context protocol, our AI X-ray scientist was guided and tested using an in-house-built virtual experimental setup that mirrors a six-circle diffractometer at an operational synchrotron beamline. The agentic workflow developed in the virtual environment was directly deployed on a real beamline, where it correctly identified reference reflections and determined the orientation matrix, an essential first step in any type of single-crystal scattering experiment. Our AI X-ray scientist responded effectively to unexpected experimental conditions, demonstrating adaptive problem-solving and readiness for addressing practical experimental situations. Our study provides a step towards autonomous operation across diverse experimental environments at large-scale scattering facilities.

Chen, Zhantao (ORCID:0000000319543868)↗

Platform Of Optimal Experiment Management

The platform of optimal experiment management, POEM, powered with automated machine learning to accelerate the discovery of optimal solutions, and automatically guide the design of experiments to be evaluated. POEM currently supports 1) random model explorations for experiment design, 2) sparse grid model explorations with Gaussian Polynomial Chaos surrogate model to accelerate experiment design ,3) time-dependent model sensitivity and uncertainty analysis to identify the importance features for experiment design, 4) model calibrations via Bayesian inference to integrate experiments to improve model performance, and 5) Bayesian optimization for optimal experimental design. In addition, POEM aims to simplify the process of experimental design for users, enabling them to analyze the data with minimal human intervention, and improving the technological output from research activities.

Wang, Congjian [Idaho National Laboratory (INL), I↗

Consistent performance of large language models in rare disease diagnosis across ten languages and 4917 cases

Background Large language models (LLMs) are increasingly used medicine for diverse applications including differential diagnostic support. The training data used to create LLMs such as the Generative Pretrained Transformer (GPT) predominantly consist of English-language texts, but LLMs could be used across the globe to support diagnostics if language barriers could be overcome. Initial pilot studies on the utility of LLMs for differential diagnosis in languages other than English have shown promise, but a large-scale assessment on the relative performance of these models in a variety of European and non-European languages on a comprehensive corpus of challenging rare-disease cases is lacking. Methods We created 4917 clinical vignettes using structured data captured with Human Phenotype Ontology (HPO) terms with the Global Alliance for Genomics and Health (GA4GH) Phenopacket Schema. These clinical vignettes span a total of 360 distinct genetic diseases with 2525 associated phenotypic features. We used translations of the Human Phenotype Ontology together with language-specific templates to generate prompts in English, Chinese, Czech, Dutch, French, German, Italian, Japanese, Spanish, and Turkish. We applied GPT-4o, version gpt-4o-2024-08-06, and the medically fine-tuned Meditron3-70B to the task of delivering a ranked differential diagnosis using a zero-shot prompt. An ontology-based approach with the Mondo disease ontology was used to map synonyms and to map disease subtypes to clinical diagnoses in order to automate evaluation of LLM responses. Findings For English, GPT-4o placed the correct diagnosis at the first rank 19.9% and within the top-3 ranks 27.0% of the time. In comparison, for the nine non-English languages tested here the correct diagnosis was placed at rank 1 between 16.9% and 20.6%, within top-3 between 25.4% and 28.6% of cases. The Meditron3 model placed the correct diagnosis within the first 3 ranks for 20.9% of cases in English and between 19.9% and 24.0% for the other nine languages. Interpretation The differential diagnostic performance of LLMs across a comprehensive corpus of rare-disease cases was largely consistent across the ten languages tested. This suggests that the utility of LLMs in clinical settings may extend to non-English clinical settings.

Artificial intelligence↗

Large language model evaluation for high–performance computing software development

We apply AI-assisted large language model (LLM) capabilities of GPT-3 targeting high-performance computing (HPC) kernels for (i) code generation, and (ii) auto-parallelization of serial code in C ++, Fortran, Python and Julia. Our scope includes the following fundamental numerical kernels: AXPY, GEMV, GEMM, SpMV, Jacobi Stencil, and CG, and language/programming models: (1) C++ (e.g., OpenMP [including offload], OpenACC, Kokkos, SyCL, CUDA, and HIP), (2) Fortran (e.g., OpenMP [including offload] and OpenACC), (3) Python (e.g., numpy, Numba, cuPy, and pyCUDA), and (4) Julia (e.g., Threads, CUDA.jl, AMDGPU.jl, and KernelAbstractions.jl). Kernel implementations are generated using GitHub Copilot capabilities powered by the GPT-based OpenAI Codex available in Visual Studio Code given simple + + prompt variants. To quantify and compare the generated results, we propose a proficiency metric around the initial 10 suggestions given for each prompt. For auto-parallelization, we use ChatGPT interactively giving simple prompts as in a dialogue with another human including simple “prompt engineering” follow ups. Results suggest that correct outputs for C++ correlate with the adoption and maturity of programming models. For example, OpenMP and CUDA score really high, whereas HIP is still lacking. We found that prompts from either a targeted language such as Fortran or the more general-purpose Python can benefit from adding language keywords, while Julia prompts perform acceptably well for its Threads and CUDA.jl programming models. Finally, we expect to provide an initial quantifiable point of reference for code generation in each programming model using a state-of-the-art LLM. Overall, understanding the convergence of LLMs, AI, and HPC is crucial due to its rapidly evolving nature and how it is redefining human-computer interactions.

97 MATHEMATICS AND COMPUTING↗

Improved Creep Testing Approach for Bentonite EBS

In the United States, the nuclear reactor contributes 18.7 percent of the nation’s total electricity generation and produces a significant amount of spent nuclear fuel (SNF) and high-level waste. Because SNF is being stored for longer periods than initially envisioned, the U.S. Department of Energy Office of Nuclear Energy, Office of Spent Fuel and Waste Science and Technology is assessing the technical performance of the SNF storage systems after extended durations. The concept of a sustainable geologic repository for SNF disposal relies heavily on the safety provided by multiple barriers, such as repository host rock, overlying rock formation, and human-made engineered barrier systems (EBS). One of the major components of EBS is bentonite-based buffer materials that isolate the nuclear-waste canisters from the host rock and fill the void left in the horizontally drilled boreholes. The resaturated bentonite has a swelling characteristic, which is useful for self-sealing the microcracks in the host rock and supporting the canister’s heavy weight. However, factors such as nonuniformity in the saturation, swelling pressure of bentonite, and the high temperature at the canister-bentonite interface may lead to instability of the waste canister inside the borehole. Therefore, evaluating the long-term deformation of bentonite-based material is necessary. RESPEC Company, LLC recently completed a preliminary study for the U.S. Department of Energy (DE SC0022804) that is attempting to understand the time-dependent deformation in the consolidated sand/bentonite (SB) specimens through triaxial creep experiments and improve the understanding of EBS performance under anticipated repository conditions. The project included developing a standard testing procedure for preparing the consolidated core specimen from a mixture of sand and bentonite in a 50:50 ratio by weight, fabricating a new pressure vessel in-house, modifying existing creep equipment to perform triaxial creep experiments on consolidated SB specimens, and performing six long-term triaxial creep experiments under low-deviatoric stress at two different temperature conditions, which is representative of hypothetical conditions that might be encountered in a geologic repository within a reasonable rate of success in completing creep experiments. The consolidated SB specimen was prepared using the novel specimen preparation procedure and had physical characteristics (e.g., moisture content and bulk density) reasonably comparable with the properties of bentonite-based buffer, which is proposed to be used in the EBS in the Swedish KBS-3H design and the Swiss design of SNF disposal in the horizontal drifts. The newly fabricated pressure vessel could withstand the confining pressure of up to 15 megapascals (MPa) while maintaining the test temperature of up to 90 degrees Celsius (°C) for long-term creep experiments. The preliminary results from the triaxial creep experiments suggested that the consolidated SB specimen may experience time-dependent deformation at a low-deviatoric stress state and room temperature condition. Based on the deformation recorded in the SB specimen from axial and radial linear variable differential transformers (LVDTs) and physical characteristic changes, it is apparent that the time-dependent deformation is a combination of consolidation and creep, which is influenced by the level of deviatoric stresses and temperature. The systematic creep testing of bentonite-based specimens helped develop a standard testing procedure applicable for analyzing the similar characteristics of shale or clay-like materials in different geologic conditions outside repository science, including in civil engineering and infrastructure projects, underground and open pit mining, and deep drilling.

58 GEOSCIENCES↗

Robustness of topological persistence in knowledge distillation for wearable sensor data

Topological data analysis (TDA) has shown great success in various applications involving wearable sensor data. However, there are difficulties in leveraging topological features in machine learning and wearable sensors because of the large time consumption and computational resources required to extract the features. To address this problem, knowledge distillation (KD) is utilized to generate a small model and accommodate topological features with persistence image (PI) representations from the raw time series data. Deploying topological knowledge in KD enables the student to achieve better performance compared to the one trained solely on raw time series data. However, it is not yet known if there are coherent characteristics for topological features in PI, which can aid in improving the performance during KD. In this paper, we investigate the suitability and challenges of utilizing topological features in KD for wearable sensor data, thereby contributing to the advancement of the field. Our study explores the impact of transferred topological features by comparing the Teacher-to-Student framework with Multiple Teachers-to-Student where teachers utilize both time series data and persistence images obtained by TDA as inputs. Additionally, we conduct a rigorous examination of topological knowledge effects by testing under various corruptions, knowledge types, and learning strategies in the context of human activity recognition tasks. Our analysis of topological features in KD presents the optimal strategy for incorporating these features. This study includes datasets of varying scales, window lengths, and activity classes, providing a comprehensive evaluation. Our results demonstrate that leveraging topological features in KD to enhance performance across databases.

97 MATHEMATICS AND COMPUTING↗

N-terminal domain swapping: A new paradigm for spermidine/spermine N -acetyltransferase (SSAT) protein structures?

Enterococcus faecalis is a multi-drug-resistant human pathogen that is found in a variety of environments and is challenging to treat. Under stress conditions, some bacteria regulate intracellular polyamine concentrations via polyamine acetyltransferases to reduce their toxicity. The E. faecalis genome encodes two polyamine acetyltransferases: PmvE and BltD. Both of these proteins belong to the Gcn5-related N-acetyltransferase (GNAT) superfamily. It is unclear why there are two enzymes with similar substrate specificities in this organism. To better understand the structure/function relationship of the E. faecalis BltD enzyme, we determined its crystal structure and performed additional assays to explore its oligomeric state and enzymatic activity. The goal was to determine whether there were structural or catalytic differences between this enzyme and other polyamine acetyltransferases that could explain this redundancy and be exploited for future development of targeted inhibitors for this important human pathogen. We found the BltD enzyme was structurally unique due to its N-terminal domain swapped dimer. However, this enzyme adopts a catalytically active monomer rather than dimer in solution. This indicates the crystal structure we obtained may represent a state that forms at high protein and salt concentrations and at low pH used during crystallization. The BltD dimer found in the crystal may represent a unique view of how an inhibitory peptide or molecule could be designed to occupy its active site. Additionally, this structure shows the extensive flexibility of the N-terminal portion of the E. faecalis BltD enzyme.

59 BASIC BIOLOGICAL SCIENCES↗

Archi: Agentic Operations at the CMS Experiment

We present Archi, an open-source, end-to-end framework for scientific collaborations that combines the systematic ingestion and organization of heterogeneous data sources with the deployment of configurable, private, and extensible agents that retrieve and reason over them. An instance of Archi has been deployed for the Computing Operations team of the CMS experiment at CERN's LHC since February 2026 as a support agent for technical operators, offering retrieval and analysis capabilities by combining documentation, historical data, and live monitoring systems. We evaluate the system on operator feedback and a question set collected from production usage, graded by human and automated panels. The system proves effective at operational tasks, resolving real-world queries posed by CMS operators. We also observe that locally-hosted, open-weight models perform competitively, enabling fully private management of sensitive data.

Lugato, Pietro [MIT; CERN]↗