Search NASA⌕ Search

SEARCH · Search NASA

Results for “data analytics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 361 records · Page 20

Heterogeneous Multi-Domain Dataset Synthesis to Facilitate Privacy and Risk Assessments in Smart City IoT

The emergence of the Smart Cities paradigm and the rapid expansion and integration of Internet of Things (IoT) technologies within this context have created unprecedented opportunities for high-resolution behavioral analytics, urban optimization, and context-aware services. However, this same proliferation intensifies privacy risks, particularly those arising from cross-modal data linkage across heterogeneous sensing platforms. To address these challenges, this paper introduces a comprehensive, statistically grounded framework for generating synthetic, multimodal IoT datasets tailored to Smart City research. The framework produces behaviorally plausible synthetic data suitable for preliminary privacy risk assessment and as a benchmark for future re-identification studies, as well as for evaluating algorithms in mobility modeling, urban informatics, and privacy-enhancing technologies. As part of our approach, we formalize probabilistic methods for synthesizing three heterogeneous and operationally relevant data streams—cellular mobility traces, payment terminal transaction logs, and Smart Retail nutrition records—capturing the behaviors of a large number of synthetically generated urban residents over a 12-week period. The framework integrates spatially explicit merchant selection using K-Dimensional (KD)-tree nearest-neighbor algorithms, temporally correlated anchor-based mobility simulation reflective of daily urban rhythms, and dietary-constraint filtering to preserve ecological validity in consumption patterns. In total, the system generates approximately 116 million mobility pings, 5.4 million transactions, and 1.9 million itemized purchases, yielding a reproducible benchmark for evaluating multimodal analytics, privacy-preserving computation, and secure IoT data-sharing protocols. To show the validity of this dataset, the underlying distributions of these residents were successfully validated against reported distributions in published research. We present preliminary uniqueness and cross-modal linkage indicators; comprehensive re-identification benchmarking against specific attack algorithms is planned as future work. This framework can be easily adapted to various scenarios of interest in Smart Cities and other IoT applications. By aligning methodological rigor with the operational needs of Smart City ecosystems, this work fills critical gaps in synthetic data generation for privacy-sensitive domains, including intelligent transportation systems, urban health informatics, and next-generation digital commerce infrastructures.

IoT↗

Thermodynamic properties of “near-perfect” gas

When applied thermodynamics requires simple representations of thermodynamic quantities, one often finds quantities such as pressure and the thermal energy density expressed as a monomial linear product of density and temperature power laws. This is a simple generalization of the product of density and temperature used for the perfect gas that admits a broader range of thermodynamic behavior into analytic fluid calculations or provides an analytic form that is readily fit to tabulated equation-of-state data for an arbitrary material. This paper reviews the thermodynamic properties of this generalized perfect-gas model, treating it as a class of “near-perfect-gas” models that are defined and unified here, for the first time, through a shared Helmholtz free energy. This generalization from perfect to near-perfect-gas models preserves some perfect-gas properties (e.g., constant adiabatic exponents) but not others (e.g., constant specific heats) and is constrained by thermodynamic consistency. It is important to be aware of the properties of this class of near-perfect-gas models when using them as surrogates for real nonperfect materials that may have properties that cannot be captured by near-perfect-gas models.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Origin of anomalous magnetotransport in kagome superconductors 𝐴⁢V 3 ⁢Sb 5 (𝐴 = K,Rb,Cs)

Multiple anomalous features in electronic spectra of metals with a kagome lattice structure—van Hove singularities, Dirac points, and flat bands—imply that materials containing this structural motif may lie at a nexus of topological and correlated electron physics. Due to the prospects of such exceptional electronic behavior, the recent discovery of superconductivity coexisting with charge-density wave (CDW) order in the layered kagome metals 𝐴⁢V 3 ⁢Sb 5 (𝐴 = K,Rb,Cs) has attracted considerable attention. Notably, these archetypal kagome metals express unconventional magnetotransport behavior, including an unexpected linear-in-𝐻 diagonal resistivity at low fields, and an even more peculiar, nonmonotonic sign-changing behavior of the Hall resistivity, which has been speculated to arise from a chiral CDW. We argue here that this unusual magnetotransport derives not from such unconventional phenomena, but rather from the unique fermiology of the 𝐴⁢V 3 ⁢Sb 5 materials. Specifically, it is caused by a large, concave hexagonal Fermi surface sheet formed in the close proximity to the van Hove singularities, which is backfolded into a small hexagonal sheet and two large triangular sheets in the CDW state. We introduce and analyze a model of the electronic structure of these Fermi surface sheets that allows for a full analytical treatment within Boltzmann kinetic theory and that enables semi-quantitative fits of our transport data. Specifically, we find that the anomalous magnetotransport behavior is caused by the confluence of strong reduction of the Fermi velocity near the van Hove singularities located near the vertices of the hexagonal sheet and sharp corners in Fermi surface generated by the CDW reconstruction. In conclusion, our analytical approach not only explains the anomalous magnetotransport in the kagome superconductors but also can be extended to a variety of metallic systems hosting singular features in their Fermi surfaces.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

QCD Running Coupling in the Nonperturbative and Near-Perturbative Regimes

We use analytic continuation to extend the gauge-gravity duality nonperturbative description of the strong force coupling into the transition, near-perturbative, regime where perturbative effects become important. By excluding the unphysical region in coupling space from the flow of singularities in the complex plane, we derive a specific relation between the scales relevant at large and short distances; this relation is uniquely fixed by requiring maximal analyticity. The unified effective coupling model gives an accurate description of the data in the nonperturbative and the near-perturbative regions.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Protein Structure Inspired Discovery of a Novel Inducer of Anoikis in Human Melanoma

Drug discovery historically starts with an established function, either that of compounds or proteins. This can hamper discovery of novel therapeutics. As structure determines function, we hypothesized that unique 3D protein structures constitute primary data that can inform novel discovery. Using a computationally intensive physics-based analytical platform operating at supercomputing speeds, we probed a high-resolution protein X-ray crystallographic library developed by us. For each of the eight identified novel 3D structures, we analyzed binding of sixty million compounds. Top-ranking compounds were acquired and screened for efficacy against breast, prostate, colon, or lung cancer, and for toxicity on normal human bone marrow stem cells, both using eight-day colony formation assays. Effective and non-toxic compounds segregated to two pockets. One compound, Dxr2-017, exhibited selective anti-melanoma activity in the NCI-60 cell line screen. In eight-day assays, Dxr2-017 had an IC50 of 12 nM against melanoma cells, while concentrations over 2100-fold higher had minimal stem cell toxicity. Dxr2-017 induced anoikis, a unique form of programmed cell death in need of targeted therapeutics. Our findings demonstrate proof-of-concept that protein structures represent high-value primary data to support the discovery of novel acting therapeutics. This approach is widely applicable.

Oncology↗

ReVise: A Human-AI Interface for Incremental Algorithmic Recourse

The recent adoption of artificial intelligence in socio-technical systems raises concerns about the black-box nature of the resulting decisions in fields such as hiring, finance, admissions, etc. If data subjects—such as job applicants, loan applicants, and students—receive an unfavorable outcome, they may be interested in algorithmic recourse, which involves updating certain features to yield a more favorable result when re-evaluated by algorithmic decision-making. Unfortunately, when individuals do not fully understand the incremental steps needed to change their circumstances, they risk following misguided paths that can lead to significant, long-term adverse consequences. Existing recourse approaches focus exclusively on the final recourse goal but neglect the possible incremental steps to reach the goal with real-life constraints, user preferences, and model artifacts. To address this gap, we formulate a visual analytic workflow for incremental recourse planning in collaboration with AI/ML experts and contribute an interactive visualization interface that helps data subjects efficiently navigate the recourse alternatives and make an informed decision. We also present one of the many usage scenarios, developed during exploratory feedback sessions with twelve graduate students using a real-world dataset, which demonstrates that our approach can be instrumental for data subjects in choosing a suitable recourse path.

algorithmic recourse↗

Explainable AI for Multivariate Time Series Pattern Exploration: Latent Space Visual Analytics With Temporal Fusion Transformer and Variational Autoencoders in Power Grid Event Diagnosis

Detecting and analyzing complex patterns in multivariate time-series data is crucial for decision-making in urban and environmental system operations. However, challenges arise from the high dimensionality, intricate complexity, and interconnected nature of complex patterns, which hinder the understanding of their underlying physical processes. Existing AI methods often face limitations in interpretability, computational efficiency, and scalability, reducing their applicability in real-world scenarios. This paper proposes a novel visual analytics framework that integrates two generative AI models, Temporal Fusion Transformer (TFT) and Variational Autoencoders (VAEs), to reduce complex patterns into lower-dimensional latent spaces and visualize them in 2D using dimensionality reduction techniques such as PCA, t-SNE, and UMAP with DBSCAN. These visualizations, presented through coordinated and interactive views and tailored glyphs, enable intuitive exploration of complex multivariate temporal patterns, identifying patterns’ similarities and uncover their potential correlations for a better interpretability of the AI outputs. The framework is demonstrated through a case study on power grid signal data, where it identifies multi-label grid event signatures, including faults and anomalies with diverse root causes. Additionally, novel metrics and visualizations are introduced to validate the models and assess the performance, efficiency, and consistency of latent maps generated by VAE, which have been utilized in prior studies for latent space cartography and used as a benchmark in this study, and the emerging TFT architecture under various configurations. These analyses provide actionable insights for model parameter tuning and reliability improvements. Comparative results highlight that TFT achieves shorter run times and superior scalability to diverse time-series data shapes compared to VAE. This work advances fault diagnosis in multivariate time series, fostering explainable AI to support critical system operations.

Explainable AI↗

Analysis of Bis(trifluoromethylsulfonyl)imide Interactions with Metal Cations Through a Chemical Informatics Approach

Nominally weakly coordinating anions are useful for modulating the solubility and chemical properties of metal complexes, but identification and analysis of the systematics of the interactions of anions with cationic metal complexes has not received the attention it deserves. Here, a chemical informatics approach is demonstrated for identifying and quantitatively analyzing the ways that the bis(trifluoromethylsulfonyl)imide anion (TFSI) can interact with metal-containing species. An open access computer program (PyCIFTer) was developed to facilitate large-scale structural analysis of TFSI-containing species by utilization of experimental atomic coordinate data from single-crystal X-ray diffraction (XRD) studies obtained from the Cambridge Structural Database (CSD). PyCIFTer establishes a three-dimensional vector space from the raw atomic coordinates, generating acyclic, undirected graphs that are used to rapidly analyze the structural properties (bond lengths and angles) of TFSI in individual structures in sequential/batch fashion. The structures are sorted by PyCIFTer into groups based on pre-set and chemically sensible criteria, affording a comprehensive and systematic view of TFSI structural chemistry. This approach avoids tedious one-at-a-time interrogation of structures, a prospect unreasonable in this case, and many others of contemporary chemical relevance; there were over 1500 structures in the CSD containing TFSI as of November 2024. The results demonstrate that TFSI only rarely binds to cations in the solid state, favoring the formation of species in which TFSI is found in cations’ outer coordination spheres. The prospect of applying PyCIFTer to other moieties is also discussed. PyCIFTer is also schematically compared to the commercial CSD Python application programming interface (API). Taken together, this work demonstrates the usefulness of modular workflows for sequential/batch analysis of structural data from XRD, an approach that appears poised to accelerate the translation of legacy structural results into new chemical insights and hypotheses.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Continuum shock mixture models for Ni+Al multilayers: Individual layers and bulk equations of state

Continuum shock mixture models are reviewed and applied to determine the equations of state for five different compositions of Ni x Al y ⁠, as well as bulk Ni+Al reactive multilayers, by combining the fundamental property data for elemental nickel and aluminum. From the literature, we down-select and evaluate two analytical models for the mixture Hugoniot, i.e., the well-known method of kinetic energy averaging (KEA) and a recent model proposed by Jordan and Baer [J. Appl. Phys. 111, 083516 (2012)]. Fundamentally, the former method assumes pressure equilibrium, whereas the latter assumes a common particle velocity and mixture sound speed from compressible two-phase cavitating flows. Additionally, we construct thermodynamically complete equations of state by fitting Einstein oscillator series models for the specific heat at constant volume. Finally, the solid solution approximation is invoked for intermetallic compositions, which are not strictly physical mixtures. Overall, the KEA model provides a better fit to the available Ni x Al y and Ni+Al multilayer shock compression data; however, there are combinations of material properties where the performance of these two models is thought to be reversed. Moreover, the results of this work include the first analytical solution of Jordan–Baer that does not require numerical root finding, as well as proposed modifications to the Einstein oscillator series to incorporate some effects of local pressure–temperature equilibrium and reaction–diffusion. Future work is planned that will use these equations of state in mesoscale simulations to study shock-induced reaction in Ni+Al multilayers, and the intended application is illustrated with a brief 2D hydrocode example.

36 MATERIALS SCIENCE↗

Three-Dimensional Grid Visualization for Planning Activities: A Dubai Case Study

National Laboratory of the Rockies (NLR), in collaboration with the Dubai Electricity and Water Authority (DEWA) and Infra-X, has undertaken the Energy Visualization Analysis Project. The aim of this project is to enhance analytical and 3D visualization capabilities for distribution network planning and renewable energy integration. As modern grid continues to evolve with large-scale solar PV deployment and emerging distributed energy resources (DERs), the ability to effectively analyze, visualize, and communicate complex grid behaviors has become increasingly critical. The project focuses on developing empirical use cases based on real distribution feeder data and engineering workflows, ensuring the outcomes are directly aligned with operational environment. Through time-series power flow simulations and nodal hosting capacity analysis, the study quantifies the impacts of high PV penetration on voltage and thermal limits within representative 11 kV feeders. These analyses identify specific nodes and conditions where DER integration challenges arise. Furthermore, a Battery Energy Storage System (BESS) optimization algorithm was applied to determine the optimal size and placement of storage systems that can mitigate network constraints and enhance hosting capacity. The comparative results between base-case and BESS-augmented scenarios clearly demonstrate improvements in network stability and load management efficiency. In parallel, the NLR team developed an immersive 3D visualization framework, enabling interactive exploration of grid simulations using commodity head-mounted display (HMD) systems. This framework transforms conventional 2D simulation data into spatially intuitive visual environments - allowing engineers to analyze feeder conditions, PV hosting potential, and BESS effects in real time. This report represents the first foundational phase in establishing a visualization-driven analytical ecosystem. It provides a methodological foundation for data integration, visualization architecture, and simulation-based decision support, paving the way for large-scale adoption of immersive visualization across DEWA's Smart Grid Initiative, R&D activities, and future network resilience studies.

24 POWER TRANSMISSION AND DISTRIBUTION↗

A Hybrid Biophysical‐Machine Learning Framework for Diurnal Surface Energy Flux Estimation Using Proximal Sensing

Thermal infrared-based remote sensing of surface energy fluxes has traditionally relied on high spatial resolution satellite data with revisit frequencies on the order of weeks. In this study, we evaluate a biophysics-based analytical surface energy balance model for predicting latent energy ( LE ) and sensible heat ( H ) fluxes using proximal sensing observations. The Surface Temperature Initiated Closure (STIC1.2) model has been extensively validated across a wide range of spatial and temporal scales using various satellite-derived thermal infrared data sets. Here we extend this validation by applying STIC at sub-hourly temporal resolution over multiple growing seasons for four distinct agricultural systems. We further develop and evaluate novel STIC variants that incorporate machine learning (ML) techniques to eliminate the need for surface energy balance observations, specifically net radiation and soil heat flux, thereby enhancing model applicability in data-sparse settings. The integration of a ML component to estimate surface available energy is shown to have strong predictive performance for both LE (R 2 = 0.81–0.94) and H (R 2 = 0.46–0.72) across all agricultural systems examined here, demonstrating the potential of hybrid biophysical-machine learning approaches for surface energy balance modeling with minimal data requirements. This study concludes with a novel application of explainable machine learning (exML) to diagnose sources of model error. This exML framework attributes residual prediction errors to both model input variables and environmental drivers not explicitly included in the simulation experiments. This approach provides a new pathway for improving model design and integrating previously overlooked yet influential variables into future model iterations.

evapotranspiration↗

Measurements of Lund subjet multiplicities in 13 TeV proton-proton collisions with the ATLAS detector

This Letter presents a differential cross-section measurement of Lund subjet multiplicities, suitable for testing current and future parton shower Monte Carlo algorithms. This measurement is made in dijet events in 140 fb -1 of $\sqrt{s}$ =13 TeV proton–proton collision data collected with the ATLAS detector at CERN's Large Hadron Collider. The data are unfolded to account for acceptance and detector-related effects, and are then compared with several Monte Carlo models and to recent resummed analytical calculations. The experimental precision achieved in the measurement allows tests of higher-order effects in QCD predictions. Most predictions fail to accurately describe the measured data, particularly at large values of jet transverse momentum accessible at the Large Hadron Collider, indicating the measurement's utility as an input to future parton shower developments and other studies probing fundamental properties of QCD and the production of hadronic final states up to the TeV-scale.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Determination of nuclear PDFs using Markov chain Monte Carlo methods

Global QCD analyses of nuclear parton distribution functions (nPDFs) have traditionally relied on the Hessian method for uncertainty estimation. However, the inherent Gaussian approximation and reliance on local curvature often prove insufficient for nPDF fits, which are frequently characterized by limited data constraints and non-Gaussian likelihoods. In this paper, we present the first nPDF determination based on Markov Chain Monte Carlo (MCMC) techniques, implemented within the nCTEQ framework using an adaptive Metropolis-Hastings algorithm. The MCMC approach enables a direct mapping of the posterior distribution and reveals a highly nontrivial parameter-space structure, including multiple modes and pronounced non-Gaussian behavior, particularly for the valence PDFs. We perform the first single-nucleus global analysis of lead PDFs using exclusively lead data and compare it to a multi-nuclei fit employing a standard analytic A dependence. The inclusion of lighter nuclei reduces quark uncertainties and modifies the shape of the lead PDFs, while leaving the gluon distribution largely unaffected. A complementary Hessian analysis exposes systematic limitations of the Gaussian approximation. Our results demonstrate that MCMC methods provide a more reliable framework for uncertainty quantification in nPDF determinations.

Derakhshanian, N. [Institute of Nuclear Physics Po↗

Direct Measurement of Diffusion Coefficients: Evidence for Diffusive Stochastic Heating in Collisionless Plasmas

Open questions in collisionless plasma dissipation can be addressed using space-based observations in different astrophysical environments, with implications for both astrophysical and laboratory plasma systems. We study a low-𝛽, highly imbalanced, sub-Alfvénic stream observed by Parker Solar Probe (PSP) to identify and distinguish between signatures of stochastic heating (SH) and resonant heating (RH) by parallel ion cyclotron waves (∥-ICWs). Prior work studying this stream [Trevor A. Bowen et al., Stochastic heating in the sub-Alfvénic solar wind, Phys. Rev. Lett. 135, 255201 (2025)] showed that the SH rate, accounting for intermittency, matched the amplitude of the local energy transfer (LET) rate, while the RH rate did not. This comparison relied on a number of assumptions regarding the nature of the diffusive process and the calculation of the LET rate. We introduce a novel technique of inverting the proton guiding center equation to empirically measure velocity-space diffusion coefficients using three-dimensional proton velocity distribution functions, from the ion electrostatic analyzer (the Solar Probe Analyzer for Ions) on PSP. Measured diffusion coefficients are used to determine phase-space heating rates, leading to a calculation of a fully kinetic heating rate independent of assumptions made in prior work. We show that scale-dependent analytic expressions for SH via noncoherent fluctuations match the empirical measurements from PSP data, provided that we account for intermittency in the heating calculation. In contrast, the derived heating rates for SH that accounts for the effects of the helicity barrier and heating rates for RH via ∥-ICWs do not peak in the same region of velocity space as the empirical measurements, nor do they reach the required magnitude. Our approach provides novel methodology to uniquely identify and constrain heating processes in collisionless plasmas and shows evidence of a Fokker-Planck-like diffusive process in the near-Sun solar wind.

Plasma kinetic theory↗

Model-Free Control of Grid-Interactive Efficient Buildings Under Communication Time Delays

Grid-interactive efficient buildings (GEBs) have recently been used to enhance the reliability and stability of the electric grid through demand response (DR) programs. However, most existing DR control strategies require accurate modeling of the various building thermostatically controlled loads (TCLs) and are computationally expensive. To address these challenges, a model-free control (MFC)-based strategy has recently been introduced for coordinating and controlling GEBs. MFC is a data-enabled control strategy that is computationally efficient and does not require the analytical models of the various building equipment. In this paper, we numerically investigate the impact of communication time delays on the performance of MFC in maintaining the TCLs' temperatures within the desired comfort levels while meeting the assigned power allocation constraint.

Telsang, Bhagyashri [University of Tennessee, Knox↗

Boundary Effects in the Diffusion of New Products on Cartesian Networks

The Role of Boundaries in the Spreading of Solar Peer effects by neighbors play a key role in the spreading of residential solar. Thus, people are more likely to install a solar system on their roof if some of their neighbors have already done so. Because people who live near the municipality boundary have fewer neighbors, does this imply that they are less likely to adopt solar? In “Boundary Effects in the Diffusion of New Products on Cartesian Networks,” Fibich, Levin, and Gillingham analyze this problem analytically using the Bass model on two-dimensional networks and empirically using data on installations of solar systems. They show that boundaries have a significant impact on the adoption of residential units near the municipality boundary. Their effect on the aggregate adoption in the municipality, however, is negligible.

Business & Economics↗

HPC-FAIR: A Framework Managing Data and AI Models for Analyzing and Optimizing Scientific Applications

The increasing reliance on machine learning (ML) to analyze and optimize large-scale scientific applications on supercomputers faces a significant bottleneck: the lack of readily available, high-quality training datasets and the difficulty in reusing existing AI models. This project was motivated by the urgent need to address the “FAIR” principles (Findability, Accessibility, Interoperability, Reusability) for both training datasets and AI models in the high-performance computing (HPC) domain. The project developed HPC-FAIR, a high-performance computing data management framework designed to centralize HPC-related datasets and AI models within a unified hub. To ensure interoperability, the framework established a standardized representation and vocabulary (ontology) for both data and models. HPC-FAIR also implemented automated workflows to streamline data processing, model access, and benchmarking. Additionally, the project focused on optimizing data harnessing efficiency through advanced techniques like deep reuse and compression-based analytics.

97 MATHEMATICS AND COMPUTING↗