Search NASASearch

SEARCH · Search NASA

Results for “Distributed Data Parallelism”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

396 records · Page 8

Enhanced Preparation for Intelligent Cybermanufacturing Systems (EPICS)

Opportunities exist for realizing transformative advances in productivity and reductions in energy footprint through ubiquitous sensing in manufacturing environments. Enhanced Preparation for Intelligent Cybermanufacturing Systems (EPICS) is a 21-month (4 academic semesters, plus one summer) experience for graduate students that focuses on scaling the knowledge, understanding and leadership skills in the cyber manufacturing area. Masters students (8/year, 32 total) complete 2-year projects on industrially-driven project topics, rotating to internships in summer semester to work on scoping and implementation at project partners. Students complete academic training in embedded systems, process modeling, data science, and cloud-based systems design. Their projects are targeted toward sensor retrofit, process monitoring, root cause analysis, and sensor fusion.

Advanced Manufacturing

Detection and characterization of detached tidal dwarf galaxies

Tidal interactions between galaxies often give rise to tidal tails, which can harbor concentrations of stars and interstellar gas resembling dwarf galaxies. Some of these tidal dwarf galaxies (TDGs) have the potential to detach from their parent galaxies and become independent entities, but their long-term survival is uncertain. In this study, we conducted a search for detached TDGs associated with a sample of 39 interacting galaxy pairs in the local Universe using infrared, ultraviolet, and optical images. We employed IR colors and UV/optical/IR spectral energy distributions to identify potential interlopers, such as foreground stars or background quasars. Through spectroscopic observations using the Boller and Chivens spectrograph at San Pedro Mártir Observatory, we confirmed that six candidate TDGs are at the same redshift as their putative parent galaxy pairs. We identified and measured emission lines in the optical spectra and calculated nebular oxygen abundances, which range from log(O/H) = 8.10 ± 0.01 to 8.51 ± 0.02. We have serendipitously discovered an additional detached TDG candidate in Arp72 using available spectra from SDSS. Utilizing the photometric data and the CIGALE code for stellar population and dust emission fitting, we derived the stellar masses, stellar population ages, and stellar metallicities for these detached TDGs. Compared to standard mass-metallicity relations for dwarf galaxies, five of the seven candidates have higher than expected metallicities, confirming their tidal origins. One of the seven candidates remains unclear due to large uncertainties in metallicity, and another has stellar and nebular metallicities compatible with those of a preexisting dwarf galaxy. The latter object is relatively compact in the optical relative to its stellar mass, in contrast to the other candidate TDGs, which have large diameters for their stellar masses compared to most dwarf galaxies. The derived stellar population ages range from 100 Myr to 900 Myr, while the inferred stellar masses are between 2 × 10 6 M ⊙ and 8 × 10 7 M ⊙ . Four of the six TDGs are associated with the gas-rich M51-like pair Arp 72, one TDG is associated with a second M51-like pair Arp 86, and another is associated with Arp 65, an approximately equal mass pair. In spite of the relatively low stellar masses of these TDGs, they have survived for at least 100–900 Myrs, suggesting that they are stable and in dynamical equilibrium. We conclude that encounters with a relatively low-mass companion (1/10th–1/4th of the mass of the primary) can also produce long-lasting TDGs.

Astronomy & Astrophysics

CatTestHub: A benchmarking database of experimental heterogeneous catalysis for evaluating advanced materials

The ability to quantitatively compare newly evolving catalytic materials and technologies is hindered by the widespread availability of catalytic data collected in a consistent manner. While certain catalytic chemistries have been widely studied across decades of scientific research, quantitative comparisons based on literature information is hindered by variability in reaction conditions, types of reported data, and reporting procedures. Here, we present CatTestHub, an open-access database dedicated to benchmarking experimental heterogeneous catalysis data. Combining systematically reported catalytic activity data for selected probe chemistries, with relevant material characterization and reactor configuration information, the database provides a collection of catalytic benchmarks for distinct classes of active site functionality. Through key choices in data access, availability, and traceability, CatTestHub seeks to balance the fundamental information needs of chemical catalysis and the FAIR data design principles. Details of the database architecture and the means through which to navigate it are presented, highlighting examples of catalytic insights readily drawn from the available benchmarking data. In its current iteration, CatTestHub spans over 250 unique experimental data points, collected over 24 solid catalysts, that facilitated the turnover of 3 distinct catalytic chemistries. Here, a roadmap is presented through which to expand the open-access platform that serves as a community wide benchmark, primarily through continuous addition of kinetic information on select catalytic systems by members of the heterogeneous catalysis community at large.

Benchmark

SODAs: sparse optimization for the discovery of differential and algebraic equations

Differential-algebraic equations (DAEs) integrate ordinary differential equations (ODEs) with algebraic constraints, providing a fundamental framework for developing models of dynamical systems characterized by time-scale separation, conservation laws and physical constraints. While sparse optimization has revolutionized model development by allowing data-driven discovery of parsimonious models from a library of possible equations, existing approaches for dynamical systems assume DAEs can be reduced to ODEs by eliminating variables before model discovery. This assumption limits the applicability of such methods for DAE systems with unknown constraints and time scales. We introduce sparse optimization for differential-algebraic systems (SODAs), a data-driven method for the identification of DAEs in their explicit form. By discovering the algebraic and dynamic components sequentially without prior identification of the algebraic variables, this approach leads to a sequence of convex optimization problems. It has the advantage of discovering interpretable models that preserve the structure of the underlying physical system. To this end, SODAs improves since SODAs is singular numerical stability when handling high correlations between library terms, caused by near-perfect algebraic relationships, by iteratively refining the conditioning of the candidate library. We demonstrate the performance of our method on biological, mechanical and electrical systems, showcasing its robustness to noise in both simulated time series and real-time experimental data.

DAE

Data-Driven Discovery and Experimental Validation of Solvent Polarity Effects on Conjugated Polymer Solution-to-Film Assembly Pathways

Understanding how solvent properties influence the solution-to-film assembly of conjugated polymers remains a critical challenge due to the complex and intertwined nature of polymer–solvent interactions. In this study, we integrate a data-driven framework with experimental validation to identify key parameters influencing the assembly and performance of poly[2,5-(2-octyldodecyl)-3,6-diketopyrrolopyrrole-alt-5,5-(2,5-di(thien-2-yl)thieno[3,2-b]thiophene)] (DPP-DTT) in organic field-effect transistors (OFETs). A machine learning (ML) approach identified the normalized Reichardt polarity parameter (E T N ) as a significant descriptor correlated with DPP-DTT hole mobility (μ). Systematic DPP-DTT devices fabricated using solvents across a wide E T N range revealed that higher E T N solvents yield enhanced μ. To elucidate the structural origins of high μ, we conducted comprehensive analyses using UV–vis–NIR spectroscopy and grazing incidence wide angle X-ray scattering (GIWAXS) measurements. The results revealed that films processed from high E T N solvents exhibit reduced paracrystallinity. By analyzing the solution-state behavior using optical microscopy and solution WAXS, we revealed polymer solubility differences in the various solvents and associated distinct polymer assembly pathways, elucidating why the high E T N solvent produces long-range ordered films. Notably, the high E T N solvent shows a pronounced preference for liquid-crystal (LC)-mediated assembly, providing a mechanistic explanation for the enhanced structural order. Therefore, these results demonstrate that solvent polarity, as evaluated by E T N , serves as an important parameter that plays a significant role in the DPP-DTT assembly pathway and resultant solid-state morphology. This work provides a strategy for integrating data science with experiments to identify critical parameters associated with complex polymer systems and helps guide rational process design for high-performance organic electronics.

36 MATERIALS SCIENCE

Case Study of Integrating High-Temperature Heat Pump with LiBr-H2O Absorption Chiller for Data Center Liquid Cooling

Data centers (DCs) are physical infrastructures that support artificial intelligence workloads. The rapid growth of artificial intelligence is putting substantial pressure on the US power grid. Most of electricity consumed by IT equipment, accounting for 50%-60% of total DC power, ultimately becomes waste heat. This heat is dissipated by DC’s cooling facilities, accounting for an additional 30%-40% of total DC power. Recovering and repurposing this waste heat offers a significant opportunity to enhance energy efficiency and reduce operating costs of DCs. One potential pathway is converting heat to cold using thermal-driven absorption chillers, therefore, reducing the power consumption in DC cooling facilities. Existing studies mainly demonstrate the technical and economic feasibility of repurposing DC’s waste heat for cooling applications but provide limited technical details on how to integrate the thermal-driven absorption chillers with DC cooling systems. In addition, the low-grade waste heat available from DCs must be upgraded to higher temperatures suitable for absorption chillers. This paper presents a case study on integrating high-temperature heat pumps with a LiBr-H2O absorption chiller to use DC waste heat for cooling. A thermodynamic model of single-effect, LiBr-H2O absorption chiller and an empirical model of high-temperature heat pumps were built. The case study considers ASHRAE W17 liquid-cooled DC, with facility service water supplied at 17.0℃ and returned at 25.3℃. The thermal behaviors of absorption chiller components were predicted for the generation temperature ranging from 75.0℃ to 115.0℃. Based on the available waste heat in the integrated system, two waste heat recovery strategies were evaluated: a facility service water-based strategy and cooling water-based strategy. Results indicated that the cooling water-based strategy achieves higher Coefficient of Performance (COPs) than the facility service water-based strategy. The relatively low cooling COPs of single-effect LiBr-H2O absorption chillers could be offset by high heating COP of high temperature heat pumps. The maximum cooling COP of absorption chiller and the overall COP of integrated systems occur at lower generation temperatures, but these conditions also yield lower cooling capacities. In practice, system operation should balance the trade-off between the COP and cooling capacity

Wang, Pengtao [ORNL] (ORCID:0000000214713429)

System for controller area network payload decoding

A system for decoding an unknown automotive controller area network (“CAN”) message definitions. CAN data vehicle signal mappings are typically held in secret and varied by automotive model and year. Without knowledge of the mappings, the wealth of real-time vehicle data hidden in the automotive CAN packets is uninterpretable—impeding research, after-market tuning, efficiency and performance monitoring, fault diagnosis, and privacy-related technologies. This system can ascertain the CAN signals' boundaries (start bit and length), endianness (byte ordering), signedness (binary-to-integer encoding) from raw CAN data. This allows conversion of CAN data to time series. Interpreting the translated CAN data's physical meaning and finding a linear mapping to standard units (e.g., knowing the signal is speed and scaling values to represent units of miles per hour) can be achieved for many signals by leveraging diagnostic standards to obtain real-time measurements of in-vehicle systems. The system can be integrated into lightweight hardware enabling an OBD-II plugin for real-time in-vehicle CAN decoding or run on standard computers. The system can output a standard DBC file with the signal definition information.

Verma, Kiren E.

Integration of ultra-low coverage whole-genome sequences for reconstructing the evolutionary history of Galapagos giant tortoises

Genomic data from contemporary and historical samples often need to be coupled for evolutionary reconstructions of multitaxon complexes. However, the genetic data recovered from historical samples may result only in ultra-low coverage whole-genome sequences (ulcWGS; <0.15× depth), leading to inaccurate evolutionary inferences given a preponderance of missing data. Using the Galapagos giant tortoise radiation as a study system (Chelonoidis spp., composed of 13 extant and four extinct lineages), we assembled a novel methodological pipeline that removes potential noise introduced by the missing data and enhances the evolutionary signal from ulcWGS samples. We leveraged existing tools for phylogenomic placement (EPA-ng), population genomic structure (smartsnp) and admixture (Admixfrog, NGSadmix) to demonstrate that the evolutionary history of samples can be uncovered with sequencing depths as low as 0.008–0.139×. Importantly, these approaches do not use genotype imputation of the ulcWGS samples, which would require extensive reference datasets. Our application to two cases of extinct lineages of Galapagos giant tortoises, with and without references from the same lineage, demonstrates the general value of the approach. We confirm where the extinct lineages from San Cristóbal and Santa Fe islands fit into the Galapagos giant tortoise radiation, and that these lineages were evolutionarily distinct entities.

ancient DNA

Sequence-Structure–Property Relationships in Short-Chain Polyesters: How Primary Structure Governs Macroscopic Performance

While polymer properties are fundamentally linked to their nanostructure, the influence of monomer sequence remains less understood than stereochemical factors like tacticity. This study examines how sequence distribution affects the thermal behavior and morphology of homo- and copolyesters, specifically comparing polymers derived from constitutionally identical monomers but with varying degrees of sequence regularity depending on monomer structure or polymerization selectivity. Our findings show that increasing sequence defects progressively diminish thermal stability, crystallinity, melting temperatures, and morphological order. As new materials become more compositionally complex, this work underscores the importance of sequence control in the design of advanced polymers for emerging applications.

Bocharova, Vera [Oak Ridge National Laboratory (OR

Initial Feasibility Assessment and Broader Strategy Development for Fiber Integration via Advanced Manufacturing

Structural health monitoring is critical for ensuring the operational safety and cost-competitiveness of the developing advanced reactor designs. The extreme operational envelopes of these advanced architectures, having operating temperatures ranging 400°C–1,000°C and heightened displacement damage doses, render conventional commercially available piezoelectric transducers and resistive strain gauges unviable. Optical fiber sensors present an attractive solution for advanced radiation-hardened instrumentation due to their high thermal stability and distributed sensing capabilities. However, their deployment in embedded applications can be hindered by the severe thermomechanical strain driven by the coefficient of thermal expansion mismatch between fused silica glass and structural metal alloys like stainless steel (e.g., SS316L).

36 MATERIALS SCIENCE

DEPRECATED AI-Batt-OS (Autonomous Identification of Battery Life Models - Open Source) [SWR 21-17]

DEPRECATED. This repository was archived by the owner on Jun 30, 2026. It is now read-only. Open source implementation of some of the methods utilized by AI-Batt, a battery lifetime modeling and analysis toolkit provided by the National Laboratory of the Rockies (NLR). This software demonstrates the use of bi-level optimization and symbolic regression techniques to semi-autonomously identify algebraic models predicting the capacity fade of lithium-ion batteries during calendar aging. Modeling the degradation of batteries is a complex task, due to the difficulty in separating the time-dependent and time-independent factors impacting cell level degradation, across multiple data series with different numbers of measurements and/or data quality. Bi-level optimization enables model parameters to be optimized to either the entire data set or to individual data series, allowing statistical disambiguation of global behaviors (data series independent) and local behaviors (data series dependent). Symbolic regression is used to automatically search for optimal low-dimesional models predicting the variation of locally optimized parameters versus time-independent experimental variables from millions of possible models, resulting in a more accurate and repeatable model identification process than is possible by a manual search. The provided tools also implement cross-validation and bootstrap resampling schemes, empowering statistical model comparison/selection and quantification of model uncertainties. An example script replicates the results from the manuscript "Challenging Practices of Algebraic Battery Life Models through Statistical Validation and Model Identification via Machine-Learning", submitted to ECS. All code is written in MATLAB. Requires the Statistics and Machine Learning Toolbox. Contact Dr. Paul Gasper at Paul.Gasper@nlr.gov for any questions.

Gasper, Paul [National Renewable Energy Lab. (NREL

Large language model-driven database for thermoelectric materials

Thermoelectric materials have the ability to convert waste heat into electricity, offering a valuable solution for energy harvesting. However, their widespread use is hindered by low conversion efficiency, the reliance on expensive rare earth elements, and the environmental and regulatory concerns associated with lead-based materials. A fast and cost-effective way to identify highly efficient thermoelectric materials is through data-driven methods. These approaches rely on robust and comprehensive datasets to train models. Although there are several databases on thermoelectric materials, there is still a need to collect and integrate experimental data from peer-reviewed research articles to capture diverse compositions and properties of materials. Here, in this work, we developed a comprehensive database of 7,123 thermoelectric compounds, containing key information such as chemical composition, structural detail, seebeck coefficient, electrical and thermal conductivity, power factor, and figure of merit (ZT). We used the GPTArticleExtractor workflow, powered by large language models (LLM), to extract and curate data automatically from the scientific literature published in Elsevier journals. This process enabled the creation of a structured database that addresses the challenges of manual data collection. The open access database could stimulate data-driven research and advance thermoelectric material analysis and discovery.

Database

Electron temperature relations and the direct N, O, Ne, S, and Ar abundances of 49 959 star-forming galaxies in DESI data release 2

We present the largest direct-method abundance catalogue of galaxies to date, containing measurements of 49 959 star-forming galaxies at z<0.96 from DESI (Dark Energy Spectroscopic Instrument) data release 2. By directly measuring electron temperatures across multiple ionization zones, we provide constraints on a number of electron temperature relations. Using the temperature measurements, we derive reliable abundances for N, O, Ne, S, and Ar, and measure the evolution of abundances and abundance ratios of as a function of metallicity and other galaxy properties. Our measurements include direct oxygen abundances for 49 507 galaxies, leading to the discovery of the two most metal-poor galaxies in the nearby Universe, with oxygen abundances of 12+log⁡(O/H)=6.77−0.03+0.03 dex (1.2 per cent Z⊙⁠) and 12+log⁡(O/H)=6.81−0.04+0.04 dex (1.3 per cent Z⊙⁠). We identify a rare outlier population of 24 galaxies with high-N/O ratios at low metallicity, reminiscent of galaxy abundances observed in the early Universe. We find the Ne/O ratio is constant at low metallicity but increases gradually at 12+log(O/H)>8.105±0.004 dex. We show that the S/O and Ar/O abundance ratios are strongly correlated, consistent with the expected additional Type Ia enrichment channel for S and Ar. In this work, we present an initial survey of the key properties of the sample, with this data set serving as a foundation for extensive future work on galaxy abundances at low redshift.

Scholte, D. [Edinburgh U., Inst. Astron.] (ORCID:0

A physics informed bayesian optimization approach for material design: application to NiTi shape memory alloys

Abstract The design of materials and identification of optimal processing parameters constitute a complex and challenging task, necessitating efficient utilization of available data. Bayesian Optimization (BO) has gained popularity in materials design due to its ability to work with minimal data. However, many BO-based frameworks predominantly rely on statistical information, in the form of input-output data, and assume black-box objective functions. In practice, designers often possess knowledge of the underlying physical laws governing a material system, rendering the objective function not entirely black-box, as some information is partially observable. In this study, we propose a physics-informed BO approach that integrates physics-infused kernels to effectively leverage both statistical and physical information in the decision-making process. We demonstrate that this method significantly improves decision-making efficiency and enables more data-efficient BO. The applicability of this approach is showcased through the design of NiTi shape memory alloys, where the optimal processing parameters are identified to maximize the transformation temperature.

Chemistry

Enhancing cathode composites with conductive alignment synergy for solid-state batteries

Enhancing transport and chemomechanical properties in cathode composites is crucial for the performance of solid-state batteries. Our study introduces the filler-aligned structured thick (FAST) electrode, which notably improves mechanical strength and ionic/electronic conductivity in solid composite cathodes. The FAST electrode incorporates vertically aligned nanoconducting carbon nanotubes within an ion-conducting polymer electrolyte, creating a low-tortuosity electron/ion transport path while strengthening the electrode’s structure. This design not only mitigates recrystallization of the polymer electrolyte but also establishes a densified local electric field distribution and accelerates the migration of lithium ions. The FAST electrode showcases outstanding electrochemical performance with lithium iron phosphate as the active material, achieving a high capacity of 148.2 milliampere hours per gram at 0.2 C over 100 cycles with substantial material loading (49.3 milligrams per square centimeter). This innovative electrode design marks a remarkable stride in addressing the challenges of solid-state lithium metal batteries.

Science & Technology - Other Topics

Mesh-based multiphysics coupling acceleration for fusion neutronics through clustering for fusion blanket applications

Accurate modeling of particle transport within fusion blankets is essential for predicting performance metrics such as heat deposition and the tritium breeding ratio (TBR). However, high-fidelity coupling of thermal fluids from computational fluid dynamics (CFD) to neutronics simulations often incurs significant computational costs due to the complexity of surface intersection calculations in Monte Carlo codes. This paper presents an accelerated multiphysics coupling method for neutronics that utilizes hierarchical agglomerative clustering to map complex material property distributions to a neutronics model. Implemented within the fusion reactor design and assessment (FREDA) framework, the method leverages existing Python packages to automate the creation of clustered geometries for OpenMC. The approach is demonstrated on a sector model of an ARC-class tokamak with an immersion molten salt blanket, and an simple geometry with varying isotopic concentrations. Results show that the clustering method significantly reduces computational burden without compromising fidelity, providing a foundation for agile iteration of neutronics simulations involving multiple coupled material properties.

Bae, Jin Whan [ORNL] (ORCID:0000000326548907)

Commutative Algebra Modeling in Materials Science – A Case Study on Metal–Organic Frameworks (MOFs)

Metal-organic frameworks (MOFs) are a class of important crystalline and highly porous materials whose hierarchical geometry and chemistry hinder interpretable predictions in materials properties. Commutative algebra is a branch of abstract algebra that has been rarely applied in data and material sciences. We introduce the first ever commutative algebra modeling and prediction in materials science. Specifically, category-specific commutative algebra (CSCA) is proposed as a new framework for MOF representation and learning. It integrates element-based categorization with multiscale algebraic invariants to encode both local coordination motifs and global network organization of MOFs. These algebraically consistent, chemically aware representations enable compact, interpretable, and data efficient modeling of MOF properties such as Henry’s constants and uptake capacities for common gases. Compared to traditional geometric and graph-based approaches, CSCA achieves comparable or superior predictive accuracy while substantially improving interpretability and stability across data sets. By aligning commutative algebra with the chemical hierarchy, the CSCA establishes a rigorous and generalizable paradigm for understanding structure and property relationships in porous materials and provides a nonlinear algebra-based framework for data-driven material discovery.

Khaemba, Caleb S.

A Probabilistic Approach to Load Modeling for Central HVAC Systems in Large Commercial Buildings for Retrofit Decisions Under Uncertainty

Retrofitting central HVAC systems in large commercial buildings with advanced technologies like heat recovery chillers (HRCs) offers a significant opportunity to enhance energy efficiency. However, analyzing these retrofits is challenging with traditional whole-building simulation tools, which require intensive calibration and struggle to model innovative system configurations and controls. To overcome these limitations, this study proposes a load profilebased retrofit analysis framework that provides better decisions under uncertainty. The main focus of this paper is the development of a probabilistic load profile model that can be used in the framework by using exploratory data analysis (EDA) of measured building data to properly quantify its inherent variability. A non-parametric Gaussian Process (GP) model was employed to capture the time- and weather-dependent characteristics of the heating load while explicitly modeling its uncertainty. The model's effectiveness is demonstrated through strong predictive performance on unseen data and physically interpretable insights into load behavior. This data-driven, probabilistic load profile serves as a robust and flexible input for subsequent system simulations, enabling a more confident and statistically sound analysis of retrofit potential.

Ham, S W