Search NASASearch

SEARCH · Search NASA

Results for “Model context protocol”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Evolution of the Lunar Receiving Laboratory to the Astromaterial Sample Curation Facility: Technical Tensions Between Containment and Cleanliness, Between Particulate and Organic Cleanliness

The Lunar Receiving Laboratory (LRL) was planned and constructed in the 1960s to support the Apollo program in the context of landing on the Moon and safely returning humans. The enduring science return from that effort is a result of careful curation of planetary materials. Technical decisions for the first facility included sample handling environment (vacuum vs inert gas), and instruments for making basic sample assessment, but the most difficult decision, and most visible, was stringent biosafety vs ultra-clean sample handling. Biosafety required handling of samples in negative pressure gloveboxes and rooms for containment and use of sterilizing protocols and animal/plant models for hazard assessment. Ultra-clean sample handling worked best in positive pressure nitrogen environment gloveboxes in positive pressure rooms, using cleanable tools of tightly controlled composition. The requirements for these two objectives were so different, that the solution was to design and build a new facility for specific purpose of preserving the scientific integrity of the samples. The resulting Lunar Curatorial Facility was designed and constructed, from 1972-1979, with advice and oversight by a very active committee comprised of lunar sample scientists. The high precision analyses required for planetary science are enabled by stringent contamination control of trace elements in the materials and protocols of construction (e.g., trace element screening for paint and flooring materials) and the equipment used in sample handling and storage. As other astromaterials, especially small particles and atoms, were added to the collections curated, the technical tension between particulate cleanliness and organic cleanliness was addressed in more detail. Techniques for minimizing particulate contamination in sample handling environments use high efficiency air filtering techniques typically requiring organic sealants which offgas. Protocols for reducing adventitious carbon on sample handling surfaces often generate particles. Further work is needed to achieve both minimal particulate and adventitious carbon contamination. This paper will discuss these facility topics and others in the historical context of nearly 50 years' curation experience for lunar rocks and regolith, meteorites, cosmic dust, comet particles, solar wind atoms, and asteroid particles at Johnson Space Center.

Allton, J. H.

Selective Amnesia using Contrastive Subnet Erasure for Class Level Unlearning in Vision Models

We study concept-level forgetting in pretrained vision models: removing an entire semantic category so the system no longer recognizes that object in unseen images and contexts, rather than merely forgetting specific training examples. Prior work either applies blunt global projections or fine-tunes parameters, which can introduce collateral damage to unrelated features, add compute, and become unstable as forgetting strength increases. We introduce Contrastive Subnet Erasure (CSE), a training-free, encoder-centric edit that targets a compact set of channels most responsible for the class and attenuates them in a calibrated manner. The modification is algebraically folded into the subsequent layer, yielding no inference-time overhead and leaving task heads unchanged. To evaluate whether forgetting generalizes beyond the data used to specify the class, we introduce a cross dataset protocol in which the class is defined on a source dataset and performance is measured on a disjoint target dataset drawn from a different distribution with no shared images. This setup tests whether the model still fails to recognize the object when it looks different or appears in new scenes, and it helps avoid overfitting to patterns in the source dataset. Across CIFAR 10, CIFAR 100, and ImageNet under this protocol, CSE achieves stronger forgetting of the target class while better preserving non target utility than existing baselines in both single class and multi class settings. Overall, CSE provides a simple, stable, and deployment-ready mechanism for class-level unlearning in vision.

Kotevska, Olivera [ORNL] (ORCID:0000000316772243)

Harnessing land-atmosphere interactions to enhance subseasonal-to-seasonal predictability

2025 Advancing Understanding of Land-Atmosphere Interactions and Processes on S2S Predictability Workshop What: 227 registered workshop participants gathered in person (43%) and online (57%) to discuss state-of-the-art scientific understanding and modeling of land-atmosphere interactions and related processes in the context of subseasonal-to-seasonal (S2S) predictability. Topics covered sources of S2S predictability, land model initialization methods, model diagnosis and evaluation metrics, AI/ML analysis and applications, and coordination of future community multi-model S2S forecast focused experiments. To advance the science, this community workshop, organized by NSF NCAR, NOAA, NASA, and DOE, aimed to 1) identify process- and application-oriented metrics for assessing S2S prediction skill and 2) develop experimental protocols for coordinated experiments to isolate, quantify, and understand the role of land-atmosphere interactions in S2S predictability. When: June 16-18, 2025 Where: Boulder, CO, USA, and online.

Land-Atmosphere Interaction

A Survey Protocol to Assess Meaningfulness and Usefulness of Automated Topic Finding in the NASA Aviation Safety Reporting System

Context: The NASA Aviation Safety Reporting System (ASRS) is a voluntary confidential aviation safety reporting system. The ASRS receives reports from pilots, air traffic controllers, flight attendants and other involved in aviation operations. The reports are de-identified and coded by ASRS expert safety analysts and a short descriptive synopsis is written to describe the safety issue. The de-identified reports are then disseminated to the aviation community in a number of ways including entry into an online database, Safety Alert Bulletins and For Your Information Notices, and the CALLBACK newsletter. Key to these publications are the timely processing (de-identification, coding and summarization) of new reports, which is currently done by ASRS expert safety analysts. Thus, we believe topic modelling could decrease effort in ASRS, if topics are comprehensible. Aim: We propose a methodology to evaluate whether automated topic finding using topic modelling provides meaningful and useful topics. Method: We extend the total error survey methodology to evaluate user topic comprehension of machine learning outputs. To accomplish this we performed a literature review to identify existing methods and define a construct for topic comprehension, utilizing existing ASRS synopsis writing practices to more precisely define meaningfulness and usefulness. Results: A survey protocol was created that addresses the limitations of other survey protocols found in the literature review, which we found lacking in rationale and clear protocol definition. Conclusion: The surveying of user understanding in machine learning outputs presents challenges due to the explosion of parameters to control for and the lack of systematic approach presented in the literature. More reproducible work and survey protocols are needed in the literature and our work is one step towards that direction.

topic finding

The Integrated Medical Model: Statistical Forecasting of Risks to Crew Health and Mission Success

The Integrated Medical Model (IMM) helps capture and use organizational knowledge across the space medicine, training, operations, engineering, and research domains. The IMM uses this domain knowledge in the context of a mission and crew profile to forecast crew health and mission success risks. The IMM is most helpful in comparing the risk of two or more mission profiles, not as a tool for predicting absolute risk. The process of building the IMM adheres to Probability Risk Assessment (PRA) techniques described in NASA Procedural Requirement (NPR) 8705.5, and uses current evidence-based information to establish a defensible position for making decisions that help ensure crew health and mission success. The IMM quantitatively describes the following input parameters: 1) medical conditions and likelihood, 2) mission duration, 3) vehicle environment, 4) crew attributes (e.g. age, sex), 5) crew activities (e.g. EVA's, Lunar excursions), 6) diagnosis and treatment protocols (e.g. medical equipment, consumables pharmaceuticals), and 7) Crew Medical Officer (CMO) training effectiveness. It is worth reiterating that the IMM uses the data sets above as inputs. Many other risk management efforts stop at determining only likelihood. The IMM is unique in that it models not only likelihood, but risk mitigations, as well as subsequent clinical outcomes based on those mitigations. Once the mathematical relationships among the above parameters are established, the IMM uses a Monte Carlo simulation technique (a random sampling of the inputs as described by their statistical distribution) to determine the probable outcomes. Because the IMM is a stochastic model (i.e. the input parameters are represented by various statistical distributions depending on the data type), when the mission is simulated 10-50,000 times with a given set of medical capabilities (risk mitigations), a prediction of the most probable outcomes can be generated. For each mission, the IMM tracks which conditions occurred and decrements the pharmaceuticals and supplies required to diagnose and treat these medical conditions. If supplies are depleted, then the medical condition goes untreated, and crew and mission risk increase. The IMM currently models approximately 30 medical conditions. By the end of FY2008, the IMM will be modeling over 100 medical conditions, approximately 60 of which have been recorded to have occurred during short and long space missions.

Fitts, M. A.

Introduction to CAUSES: Description of Weather and Climate Models and Their Near‐Surface Temperature Errors in 5 Day Hindcasts near the Southern Great Plains

We introduce the Clouds Above the United States and Errors at the Surface (CAUSES) project with its aim of better understanding the physical processes leading to warm screen temperature biases over the American Midwest in many numerical models. In this first of four companion papers, 11 different models, from nine institutes, perform a series of 5 day hindcasts, each initialized from reanalyses. After describing the common experimental protocol and detailing each model configuration, a gridded temperature data set is derived from observations and used to show that all the models have a warm bias over parts of the Midwest. Additionally, a strong diurnal cycle in the screen temperature bias is found in most models. In some models the bias is largest around midday, while in others it is largest during the night. At the Department of Energy Atmospheric Radiation Measurement Southern Great Plains (SGP) site, the model biases are shown to extend several kilometers into the atmosphere. Finally, to provide context for the companion papers, in which observations from the SGP site are used to evaluate the different processes contributing to errors there, it is shown that there are numerous locations across the Midwest where the diurnal cycle of the error is highly correlated with the diurnal cycle of the error at SGP. This suggests that conclusions drawn from detailed evaluation of models using instruments located at SGP will be representative of errors that are prevalent over a larger spatial scale.

Morcrette, C. J.

Effect of Thermodynamic and Environmental Factors on Crystallization of DNA‐Origami Superlattices

The directed self‐assembly of nanoscale materials into ordered superlattices presents a powerful strategy for creating next‐generation materials with programmable mechanical, optical, and photonic properties. Deoxyribonucleic acid (DNA) origami has emerged as a versatile scaffold for encoding nanoscale geometry and guiding the crystallization of complex 3D architectures. However, a systematic understanding of the parameters that govern the efficiency and quality of superlattice formation remains limited. In this study, we utilize octahedral DNA nanoscale frames as a model system to investigate the relative influence of key factors, including buffer composition, ionic strength, frame concentration, and thermal annealing protocols, on the size, order, and reproducibility of the resulting superlattices. Our findings provide a quantitative framework to rationally optimize DNA‐based assembly pathways. Structural characterization via small‐angle x‐ray scattering (SAXS), scanning electron microscopy (SEM), and optical microscopy validates the quality and fidelity of the assembled lattices. Moreover, by templating these DNA frameworks into inorganic replicas, we establish general design principles that extend beyond biomolecular systems, providing a foundation for the synthesis of programmable materials in broader nanofabrication contexts.

77 NANOSCIENCE AND NANOTECHNOLOGY

UrbanScaping: Community Spatial Data Visualization & Analytics

Evaluating the electrification potential of buildings through retrofitting is crucial for reducing carbon emissions and the carbon footprint of built environments. This study leverages the Automatic Building Energy Modeling (AutoBEM) software, integrating the Model America database to create an urban context-based spatial analysis platform for community engagement and development. We selected Camp Hill Borough, PA, as a case study to analyze building-specific energy performance and evaluate the electrification potential of each building by switching to different Heating, Ventilation, and Air Conditioning (HVAC) systems and measurement components. The simulation results generated by the workflow provide retrofitting suggestions to help mitigate the carbon footprint as well as energy saving statistics of buildings. Additionally, the developed web-based interface serves as a community engagement platform, allowing residents to provide feedback and further develop interactive communication protocols. The outcomes of this project offer a baseline for community electrification planning and contribute to the design of low-carbon communities.

Chowdhury, Shovan [ORNL]

Modeling beta-adrenergic control of cardiac myocyte contractility in silico

The beta-adrenergic signaling pathway regulates cardiac myocyte contractility through a combination of feedforward and feedback mechanisms. We used systems analysis to investigate how the components and topology of this signaling network permit neurohormonal control of excitation-contraction coupling in the rat ventricular myocyte. A kinetic model integrating beta-adrenergic signaling with excitation-contraction coupling was formulated, and each subsystem was validated with independent biochemical and physiological measurements. Model analysis was used to investigate quantitatively the effects of specific molecular perturbations. 3-Fold overexpression of adenylyl cyclase in the model allowed an 85% higher rate of cyclic AMP synthesis than an equivalent overexpression of beta 1-adrenergic receptor, and manipulating the affinity of Gs alpha for adenylyl cyclase was a more potent regulator of cyclic AMP production. The model predicted that less than 40% of adenylyl cyclase molecules may be stimulated under maximal receptor activation, and an experimental protocol is suggested for validating this prediction. The model also predicted that the endogenous heat-stable protein kinase inhibitor may enhance basal cyclic AMP buffering by 68% and increasing the apparent Hill coefficient of protein kinase A activation from 1.0 to 2.0. Finally, phosphorylation of the L-type calcium channel and phospholamban were found sufficient to predict the dominant changes in myocyte contractility, including a 2.6x increase in systolic calcium (inotropy) and a 28% decrease in calcium half-relaxation time (lusitropy). By performing systems analysis, the consequences of molecular perturbations in the beta-adrenergic signaling network may be understood within the context of integrative cellular physiology.

Non-NASA Center

Development and implementation of high-throughput proteomic and metabolomics assays by using advanced chromatographic and mass spectrometric systems (CRADA Final Report)

The mission of this CRADA with Agilent was to couple powerful MS platforms (QQQ, IM-QTOFMS) with Agilent’s novel Ultra-High-Performance Liquid Chromatography (UHPLC) fast metabolomic workflows and perform ABF Machine Learning (ML) to generated datasets. Agilent transferred UHPLC methods to PNNL and LBNL and methods were implemented and demonstrated in both labs, achieving total acquisition times of < 10 min. Metabolites analyzed using Agilent’s shared methods included metabolites from central carbon metabolism, common across hosts, and metabolites unique to engineered strains. Standards were acquired in an UHPLC-Drift Tube Ion Mobility Mass Spectrometer (DTIMS) system for the first time within the context of ABF and methods were optimized based on Agilent’s protocols. Samples from ABF hosts Pseudomonas putida, Aspergillus pseudoterreus, Aspergillus niger and Rhodosporidium toruloides were analyzed using the UHPLC-DTIMS platform for a total of 276 runs. A data analysis workflow compatible with the Experimental Data Depot (EDD) and completely shareable was developed for the acquired UHPLC-DTIMS data. Samples were analyzed using a Data Independent Acquisition Approach (DIA), which for most of the standards provided more transitions therefore increasing detection confidence. Using the data acquired by PNNL, LBNL, and Agilent’s specifications from previous ML projects, SNL applied an ensemble ML strategy to pick the best performing model for automated LC-method selection. Finally, with the contribution of the participant labs and Agilent, SNL developed an Automated Method Selection (AMS) software tool to predict the best liquid chromatography method for analysis of any new molecules of interest. Samples with novel pathways and new metabolite targets of interest are generated at a high pace in the ABF. Overall, the project advanced rapid metabolomics by combining liquid chromatography, ion mobility spectrometry, and data-independent mass spectrometry with machine learning. This multidimensional approach uses retention time, collision cross-section, precursor mass, and fragment-ion information to distinguish chemically similar metabolites that can be difficult to resolve using conventional liquid- or gas-chromatography methods. The resulting workflow also provided automated metabolite-identification error estimates, addressing a recognized need for statistical confidence measures in metabolomics.

Petzold, Christopher [Lawrence Berkeley National L

PA51C-0786 Informing NASA's Indigenous Peoples Capacity Building Pilot Project Through Collaborative Storytelling and Systems Mapping

NASA's Capacity Building Program (CBP) aims to empower communities to use NASA Earth Observations for environmental decision-making. A new project, the Indigenous Peoples Pilot Project, focuses on building relationships across NASA and indigenous communities through remote sensing training, community engagement, and research opportunities. A recent workshop, held on the lands of the Red Cliff Band of Lake Superior Chippewa in Wisconsin, focused on understanding indigenous knowledge systems and comparisons with western science, and more specifically, NASA Earth Science. In order to meet the workshop goals, North American tribal members and NASA managers participated in storytelling and systems mapping. The storytelling portion focused on personal narratives to understand the barriers, challenges, and pathways to incite change in the context of western scientists working with tribal nations. This was based on the traditional model of oral history and allowed participants to share their unique experiences of failures and successes. The systems mapping portion of the workshop focused on finding leverage points within the "NASA Ecosystem" for creating sustained instrumental change. This included (1) dedicating time and resources to explore how to recognize western science and indigenous knowledge as equal, (2) supporting innovation in communication and knowledge system frameworks, and (3) identifying new partners, allies, and champions for the western science/indigenous relationship. These leverage points were then used to generate recommendations for NASA which included (1) an awareness phase (for NASA and for indigenous communities) for NASA/Indigenous work, (2) creating a cultural immersion for NASA managers, (3) developing a protocol for working with indigenous groups, (4) creating a NASA tribal liaison office, (5) acknowledging and accepting indigenous knowledge systems in funding solicitations and proposals, and (6) incorporation of indigenous knowledge into capacity building activities.

Mccullum, Amber Jean Kuss

Location, location: deciphering the significance of in-situ hydrogen analyses of Martian meteorite phases

Understanding how inner planets acquired their volatiles such as hydrogen (H) is fundamental to constrain models of solar system formation [e.g. 1]. One avenue of estimating the H content and isotopic characteristics of differentiated planets is to analyze the samples we have from them as meteorites. Ideally, the H content and isotopic characteristic of the mantle sources of these igneous rocks should give insight into the volatile origin of each planetary body. The mantle H signatures can be estimated from that of the parent melt, which in turn may be derived from that of the first crystallized phases. However, we will illustrate the processes that can modify the H content and D/H ratios of pyroxene, olivine and feldspar in selected Martian meteorites relative to those of their mantle sources, with two key protocols for SIMS analysis. The first key is to put each analysis in textural context ("location, location"). In particular, of prime importance is whether an analysis is done at the center or the edge of a mineral grain, close or far from a shock disturbed area, and in a mineral crystallized at the beginning or late in the differentiation sequence. Accompanying the H analyses done by SIMS with major and trace element data at the same analysis locations allows to constrain the history of crystallization, cooling, alteration and shock of the meteorite. For example, volcanic degassing can be evidenced by decreasing water contents and increasing D/H ratios from core to edge of nakhlite pyroxenes [2]. Traverses of H analyses in pyroxene and maskelynite in shergottite LAR 06319 provide examples of H contents and D/H ratios modified during degassing following shock [3]. The second key is to assess if the area analyzed by SIMS has shock generated damage of the mineral structure, as will be shown in pyroxene and olivine from shergottite RBT 04262. Olivine in shergottites have too high water contents to be explained with an igneous origin and their D/H ratios are more consistent with terrestrial alteration, as observed before [4]. A review of all these processes shows how careful one has to be prior to using H measurements in meteorites to infer the origin and amount of water in differentiated planetary interiors.

A.H. Peslier

Synthetic Scientific Image Generation with VAE, GAN, and Diffusion Model Architectures

Generative AI (genAI) has emerged as a powerful tool for synthesizing diverse and complex image data, offering new possibilities for scientific imaging applications. This review presents a comprehensive comparative analysis of leading generative architectures, ranging from Variational Autoencoders (VAEs) to Generative Adversarial Networks (GANs) on through to Diffusion Models, in the context of scientific image synthesis. We examine each model's foundational principles, recent architectural advancements, and practical trade-offs. Our evaluation, conducted on domain-specific datasets including microCT scans of rocks and composite fibers, as well as high-resolution images of plant roots, integrates both quantitative metrics (SSIM, LPIPS, FID, CLIPScore) and expert-driven qualitative assessments. Results show that GANs, particularly StyleGAN, produce images with high perceptual quality and structural coherence. Diffusion-based models for inpainting and image variation, such as DALL-E 2, delivered high realism and semantic alignment but generally struggled in balancing visual fidelity with scientific accuracy. Importantly, our findings reveal limitations of standard quantitative metrics in capturing scientific relevance, underscoring the need for domain-expert validation. We conclude by discussing key challenges such as model interpretability, computational cost, and verification protocols, and discuss future directions where generative AI can drive innovation in data augmentation, simulation, and hypothesis generation in scientific research.

Generative Adversarial Networks

Asi Nuclear Energy Sensors Data Portal Chatbot And Data Structuring Tool

The Idaho National Laboratory (INL) is advancing the development of an AI-powered chatbot and data structuring tool specifically designed to accelerate data mining processes for sensor-related information and seamlessly integrate the results into the ASI Sensors Data Portal (https://nes.energy.gov/). By doing so, the software aims to enhance the accessibility, usability, and organization of sensor data for nuclear energy applications. The software initial phase focuses on retrieving comprehensive datasets, prioritizing the past five years of publicly available information from the Office of Scientific and Technical Information (OSTI). These datasets will be meticulously processed to ensure compatibility, employing cleaning and preprocessing steps to eliminate irrelevant, incomplete, or corrupted information, thus establishing a robust foundation for subsequent AI use. The data will serve as the backbone for training an AI model and chatbot, which will act as an interactive tool enabling users to ask complex, context-specific questions and receive accurate, validated answers derived from constrained literature. In parallel, the project incorporates a data structuring process supported by AI to organize sensor information from multiple sources into a standardized format. This structured data will include detailed sensor specifications, such as measurement range, applications, accuracy, and operating conditions, generated and documented with AI. These specifications will be systematically integrated into the sensor portal. To maintain the highest levels of accuracy and relevance, all AI-generated outputs will be reviewed and validated by subject matter experts (SMEs), with additional fields or parameters added as needed. Future stages of the project aim to expand the dataset beyond OSTI to include other sources and potentially incorporate unclassified controlled information (UCI) with restricted access protocols to address security and confidentiality requirements.

Mapes, NormanJ. [Idaho National Laboratory (INL),

Thermal neutral format based on the step technology

The exchange of models is one of the most serious problems currently encountered in the practice of spacecraft thermal analysis. Essentially, the problem originates in the diversity of computing environments that are used across different sites, and the consequent proliferation of native tool formats. Furthermore, increasing pressure to reduce the development's life cycle time has originated a growing interest in the so-called spacecraft concurrent engineering. In this context, the realization of the interdependencies between different disciplines and the proper communication between them become critical issues. The use of a neutral format represents a step forward in addressing these problems. Such a means of communication is adopted by consensus. A neutral format is not directly tied to any specific tool and it is kept under stringent change control. Currently, most of the groups promoting exchange formats are contributing with their experience to STEP, the Standard for Exchange of Product Model Data, which is being developed under the auspices of the International Standards Organization (ISO 10303). This paper presents the different efforts made in Europe to provide the spacecraft thermal analysis community with a Thermal Neutral Format (TNF) based on STEP. Following an introduction with some background information, the paper presents the characteristics of the STEP standard. Later, the first efforts to produce a STEP Spacecraft Thermal Application Protocol are described. Finally, the paper presents the currently harmonized European activities that follow up and extend earlier work on the area.

Almazan, P. Planas

Benchmarking and Fidelity Response Theory of High-Fidelity Rydberg Entangling Gates

The fidelity of entangling operations is a key figure of merit in quantum information processing, especially in the context of quantum error correction. High-fidelity entangling gates in neutral atoms have seen remarkable advancement recently. A full understanding of error sources and their respective contributions to gate infidelity will enable the prediction of fundamental limits on quantum gates in neutral atom platforms with realistic experimental constraints. In this work, we implement the time-optimal Rydberg controlled-Z (CZ) gate, design a circuit to benchmark its fidelity, and achieve a fidelity, averaged over symmetric input states, of 0.9971 ( 5 ) , downward corrected for leakage error, which together with our recent work [Nature 634, 321–327 (2024)] forms a new state of the art for neutral atoms. The remaining infidelity is explained by an error model, consistent with our experimental results over a range of gate speeds, with varying contributions from different error sources. Further, we develop a fidelity response theory to efficiently predict infidelity from laser noise with nontrivial power spectral densities and derive scaling laws of infidelity with gate speed. Besides its capability of predicting gate fidelity, we also utilize the fidelity response theory to compare and optimize gate protocols, to learn laser frequency noise, and to study the noise response for quantum simulation tasks. Finally, we predict that a CZ gate fidelity of ≳ 0.999 is feasible with realistic experimental upgrades. Published by the American Physical Society 2025

Tsai, Richard Bing-Shiun (ORCID:0000000286758677)

Heterogeneous Multi-Domain Dataset Synthesis to Facilitate Privacy and Risk Assessments in Smart City IoT

The emergence of the Smart Cities paradigm and the rapid expansion and integration of Internet of Things (IoT) technologies within this context have created unprecedented opportunities for high-resolution behavioral analytics, urban optimization, and context-aware services. However, this same proliferation intensifies privacy risks, particularly those arising from cross-modal data linkage across heterogeneous sensing platforms. To address these challenges, this paper introduces a comprehensive, statistically grounded framework for generating synthetic, multimodal IoT datasets tailored to Smart City research. The framework produces behaviorally plausible synthetic data suitable for preliminary privacy risk assessment and as a benchmark for future re-identification studies, as well as for evaluating algorithms in mobility modeling, urban informatics, and privacy-enhancing technologies. As part of our approach, we formalize probabilistic methods for synthesizing three heterogeneous and operationally relevant data streams—cellular mobility traces, payment terminal transaction logs, and Smart Retail nutrition records—capturing the behaviors of a large number of synthetically generated urban residents over a 12-week period. The framework integrates spatially explicit merchant selection using K-Dimensional (KD)-tree nearest-neighbor algorithms, temporally correlated anchor-based mobility simulation reflective of daily urban rhythms, and dietary-constraint filtering to preserve ecological validity in consumption patterns. In total, the system generates approximately 116 million mobility pings, 5.4 million transactions, and 1.9 million itemized purchases, yielding a reproducible benchmark for evaluating multimodal analytics, privacy-preserving computation, and secure IoT data-sharing protocols. To show the validity of this dataset, the underlying distributions of these residents were successfully validated against reported distributions in published research. We present preliminary uniqueness and cross-modal linkage indicators; comprehensive re-identification benchmarking against specific attack algorithms is planned as future work. This framework can be easily adapted to various scenarios of interest in Smart Cities and other IoT applications. By aligning methodological rigor with the operational needs of Smart City ecosystems, this work fills critical gaps in synthetic data generation for privacy-sensitive domains, including intelligent transportation systems, urban health informatics, and next-generation digital commerce infrastructures.

IoT

Intelligent Integrated Health Management for a System of Systems

An intelligent integrated health management system (IIHMS) incorporates major improvements over prior such systems. The particular IIHMS is implemented for any system defined as a hierarchical distributed network of intelligent elements (HDNIE), comprising primarily: (1) an architecture (Figure 1), (2) intelligent elements, (3) a conceptual framework and taxonomy (Figure 2), and (4) and ontology that defines standards and protocols. Some definitions of terms are prerequisite to a further brief description of this innovation: A system-of-systems (SoS) is an engineering system that comprises multiple subsystems (e.g., a system of multiple possibly interacting flow subsystems that include pumps, valves, tanks, ducts, sensors, and the like); 'Intelligent' is used here in the sense of artificial intelligence. An intelligent element may be physical or virtual, it is network enabled, and it is able to manage data, information, and knowledge (DIaK) focused on determining its condition in the context of the entire SoS; As used here, 'health' signifies the functionality and/or structural integrity of an engineering system, subsystem, or process (leading to determination of the health of components); 'Process' can signify either a physical process in the usual sense of the word or an element into which functionally related sensors are grouped; 'Element' can signify a component (e.g., an actuator, a valve), a process, a controller, an actuator, a subsystem, or a system; The term Integrated System Health Management (ISHM) is used to describe a capability that focuses on determining the condition (health) of every element in a complex system (detect anomalies, diagnose causes, prognosis of future anomalies), and provide data, information, and knowledge (DIaK) not just data to control systems for safe and effective operation. A major novel aspect of the present development is the concept of intelligent integration. The purpose of intelligent integration, as defined and implemented in the present IIHMS, is to enable automated analysis of physical phenomena in imitation of human reasoning, including the use of qualitative methods. Intelligent integration is said to occur in a system in which all elements are intelligent and can acquire, maintain, and share knowledge and information. In the HDNIE of the present IIHMS, an SoS is represented as being operationally organized in a hierarchical-distributed format. The elements of the SoS are considered to be intelligent in that they determine their own conditions within an integrated scheme that involves consideration of data, information, knowledge bases, and methods that reside in all elements of the system. The conceptual framework of the HDNIE and the methodologies of implementing it enable the flow of information and knowledge among the elements so as to make possible the determination of the condition of each element. The necessary information and knowledge is made available to each affected element at the desired time, satisfying a need to prevent information overload while providing context-sensitive information at the proper level of detail. Provision of high-quality data is a central goal in designing this or any IIHMS. In pursuit of this goal, functionally related sensors are logically assigned to groups denoted processes. An aggregate of processes is considered to form a system. Alternatively or in addition to what has been said thus far, the HDNIE of this IIHMS can be regarded as consisting of a framework containing object models that encapsulate all elements of the system, their individual and relational knowledge bases, generic methods and procedures based on models of the applicable physics, and communication processes (Figure 2). The framework enables implementation of a paradigm inspired by how expert operators monitor the health of systems with the help of (1) DIaK from various sources, (2) software tools that assist in rapid visualization of the condition of the system, (3) analical software tools that assist in reasoning about the condition, (4) sharing of information via network communication hardware and software, and (5) software tools that aid in making decisions to remedy unacceptable conditions or improve performance.

Smith, Harvey