Search NASA⌕ Search

SEARCH · Search NASA

Results for “data access”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 379 records · Page 21

Adaptable SEC‐SAXS data collection for higher quality structure analysis in solution

Abstract The two major challenges in synchrotron size‐exclusion chromatography coupled in‐line with small‐angle x‐ray scattering (SEC‐SAXS) experiments are the overlapping peaks in the elution profile and the fouling of radiation‐damaged materials on the walls of the sample cell. In recent years, many post‐experimental analyses techniques have been developed and applied to extract scattering profiles from these problematic SEC‐SAXS data. Here, we present three modes of data collection at the BioSAXS Beamline 4–2 of the Stanford Synchrotron Radiation Lightsource (SSRL BL4‐2). The first mode, the High‐Resolution mode, enables SEC‐SAXS data collection with excellent sample separation and virtually no additional peak broadening from the UHPLC UV detector to the x‐ray position by taking advantage of the low system dispersion of the UHPLC. The small bed volume of the analytical SEC column minimizes sample dilution in the column and facilitates data collection at higher sample concentrations with excellent sample economy equal to or even less than that of the conventional equilibrium SAXS method. Radiation damage problems during SEC‐SAXS data collection are evaded by additional cleaning of the sample cell after buffer data collection and avoidance of unnecessary exposures through the use of the x‐ray shutter control options, allowing sample data collection with a clean sample cell. Therefore, accurate background subtraction can be performed at a level equivalent to the conventional equilibrium SAXS method without requiring baseline correction, thereby leading to more reliable downstream structural analysis and quicker access to new science. The two other data collection modes, the High‐Throughput mode and the Co‐Flow mode, add agility to the planning and execution of experiments to efficiently achieve the user's scientific objectives at the SSRL BL4‐2.

Matsui, Tsutomu↗

FAIRmaterials: Ontology Tools with Data FAIRification in Development

The bilingual FAIRmaterials package simplifies the creation and visualization of materials and data science ontologies. FAIRmaterials, available in the Python and R languages, addresses the complexities associated with traditional ontology editors based on manual user input such as Protege with an intuitive workflow and easy-to-use templates, making it accessible to users both experienced and inexperienced with ontologies. The FAIRmaterials package is its ability to programatically convert simple and structured CSV inputs into rich, well-defined ontologies. This capability is designed to support the findability, accessibility, interoperability, and reusability (FAIR) of research data and serve as a tool in the process of data FAIRification. Its additional features, such as automated ontology merging, static visualizations, and comprehensive documentation for outputs extend its utility, making it a valuable tool for any researcher engaged in knowledge management.

Bradley, Alexander Harding [Case Western Reserve U↗

HPC-FAIR: A Framework Managing Data and AI Models for Analyzing and Optimizing Scientific Applications

The increasing reliance on machine learning (ML) to analyze and optimize large-scale scientific applications on supercomputers faces a significant bottleneck: the lack of readily available, high-quality training datasets and the difficulty in reusing existing AI models. This project was motivated by the urgent need to address the “FAIR” principles (Findability, Accessibility, Interoperability, Reusability) for both training datasets and AI models in the high-performance computing (HPC) domain. The project developed HPC-FAIR, a high-performance computing data management framework designed to centralize HPC-related datasets and AI models within a unified hub. To ensure interoperability, the framework established a standardized representation and vocabulary (ontology) for both data and models. HPC-FAIR also implemented automated workflows to streamline data processing, model access, and benchmarking. Additionally, the project focused on optimizing data harnessing efficiency through advanced techniques like deep reuse and compression-based analytics.

97 MATHEMATICS AND COMPUTING↗

Outcomes and Conclusions from the 2022 AM Bench Measurements, Challenge Problems, Modeling Submissions, and Conference

The Additive Manufacturing Benchmark Test Series (AM Bench) provides rigorous measurement data for validating additive manufacturing (AM) simulations for a broad range of AM technologies and material systems. AM Bench includes extensive in situ and ex situ measurements, simulation challenges for the AM modeling community, and a corresponding conference series. In 2022, the second round of AM Bench measurements, challenge problems, and conference were completed, focusing primarily upon laser powder bed fusion (LPBF) processing of metals, and both material extrusion processing and vat photopolymerization of polymers. In all, more than 100 people from 10 National Institute of Standards and Technology (NIST) divisions and 21 additional organizations were directly involved in the AM Bench 2022 measurements, data management, and conference organization. The international AM community submitted 138 sets of blind modeling simulations for comparison with the in situ and ex situ measurements, up from 46 submissions for the first round of AM Bench in 2018. Analysis of these submissions provides valuable insight into current AM modeling capabilities. The AM Bench data are permanently archived and freely accessible online. The AM Bench conference also hosted an embedded workshop on qualification and certification of AM materials and components.

36 MATERIALS SCIENCE↗

Machine learning in materials research: Developments over the last decade and challenges for the future

The number of studies that apply machine learning (ML) to materials science has been growing at a rate of approximately 1.67 times per year over the past decade. In this review, I examine this growth in various contexts. First, I present an analysis of the most commonly used tools (software, databases, materials science methods, and ML methods) used within papers that apply ML to materials science. The analysis demonstrates that despite the growth of deep learning techniques, the use of classical machine learning is still dominant as a whole. It also demonstrates how new research can effectively build upon past research, particular in the domain of ML models trained on density functional theory calculation data. Next, I present the progression of best scores as a function of time on the matbench materials science benchmark for formation enthalpy prediction. In particular, a dramatic improvement of 7 times reduction in error is obtained when progressing from feature-based methods that use conventional ML (random forest, support vector regression, etc.) to the use of graph neural network techniques. Finally, I provide views on future challenges and opportunities, focusing on data size and complexity, extrapolation, interpretation, access, and relevance.

36 MATERIALS SCIENCE↗

Overcoming barriers to improved decision-making for battery deployment in the clean energy transition

Decarbonization plans depend on the rapid, large-scale deployment of batteries to sufficiently decarbonize the electricity system and on-road transport. This can take many forms, shaped by technology, materials, and supply chain selection, which will have local and global environmental and social impacts. Current knowledge gaps limit the ability of decision-makers to make choices in facilitating battery deployment that minimizes or avoids unintended environmental and social consequences. These gaps include a lack of harmonized, accessible, and up-to-date data on manufacturing and supply chains and shortcomings within sustainability and social impact assessment methods, resulting in uncertainty that limits incorporation of research into policy making. These gaps can lead to unintended detrimental effects of large-scale battery deployment. To support decarbonization goals while minimizing negative environmental and social impacts, we elucidate current barriers to tracking how decision-making for large-scale battery deployment translates to environmental and social impacts and recommend steps to overcome them.

25 ENERGY STORAGE↗

Online energy consumption forecast for battery electric buses using a learning-free algebraic method

Accurately predicting the energy consumption plays a vital role in battery electric buses (BEBs) route planning and deployment. Based on the algebraic derivative estimation, we present a novel method to forecast the energy consumption in real time. In contrast to the mainstream machine-learning-based methods, the proposed method does not require access to the historical energy consumption data. It eliminates the time-consuming and computationally expensive offline training. Consequently, its prediction performance is not constrained by the quantity and quality of the training data. Moreover, the method can swiftly adapt to new situations not included in the previous driving cycles, which makes it especially suitable for emerging transport modes, e.g., on-demand transit services. In addition, its online execution only involves algebraic calculations, yielding superior calculation efficiency. Using real-world data, we comprehensively compare the performance of the proposed learning-free algebraic method with multiple representative machine-learning-based methods. Finally, the advantages and limitations of the proposed method are discussed in detail.

33 ADVANCED PROPULSION SYSTEMS↗

FatPlants: a comprehensive information system for lipid-related genes and metabolic pathways in plants

Abstract FatPlants, an open-access, web-based database, consolidates data, annotations, analysis results, and visualizations of lipid-related genes, proteins, and metabolic pathways in plants. Serving as a minable resource, FatPlants offers a user-friendly interface for facilitating studies into the regulation of plant lipid metabolism and supporting breeding efforts aimed at increasing crop oil content. This web resource, developed using data derived from our own research, curated from public resources, and gleaned from academic literature, comprises information on known fatty-acid-related proteins, genes, and pathways in multiple plants, with an emphasis on Glycine max, Arabidopsis thaliana, and Camelina sativa. Furthermore, the platform includes machine-learning based methods and navigation tools designed to aid in characterizing metabolic pathways and protein interactions. Comprehensive gene and protein information cards, a Basic Local Alignment Search Tool search function, similar structure search capacities from AphaFold, and ChatGPT-based query for protein information are additional features. Database URL: https://www.fatplants.net/

59 BASIC BIOLOGICAL SCIENCES↗

Microwave microscope studies of trapped vortex dynamics in superconductors

Trapped vortices in superconductors introduce residual resistance in superconducting radio-frequency (SRF) cavities and disrupt the operation of superconducting quantum and digital electronic circuits. Understanding the detailed dynamics of trapped vortices under oscillating magnetic fields is essential for advancing these technologies. We have developed a near-field magnetic microwave microscope to study the dynamics of a limited number of trapped vortices under the probe when stimulated by a localized rf magnetic field. By measuring the local second-harmonic response (𝑃 2⁢f ) at subfemto-Watt levels, we isolate signals exclusively arising from trapped vortices, excluding contributions from surface defects and Meissner screening currents. Toy models of niobium superconductor hosting vortex pinning sites are introduced and studied with time-dependent Ginzburg-Landau (TDGL) simulations of probe/sample interaction to better understand the measured second-harmonic response. The simulation results demonstrate that the second-harmonic response of trapped vortex motion under a localized rf magnetic field shares key features with the experimental data. Here, this measurement technique provides access to vortex dynamics at the micrometer scale, such as depinning events and spatially resolved pinning properties, as demonstrated in measurements on a niobium film with an antidot flux pinning array.

43 PARTICLE ACCELERATORS↗

XSub: Explanation-Driven Adversarial Attack against Blackbox Classifiers via Feature Substitution

Despite its significant benefits in enhancing the transparency and trustworthiness of artificial intelligence (AI) systems, explainable AI (XAI) can unintentionally provide adversaries with insights into blackbox models, increasing their vulnerability to various attacks. In this paper, we develop a novel explanation-driven adversarial attack against blackbox classifiers based on feature substitution, called XSub. The key idea of XSub is to strategically replace important features (identified via XAI) in the original sample with corresponding important features of a different label, thereby increasing the likelihood of the model misclassifying the perturbed sample. XSub only requires a minimal number of queries and can be easily extended to launch backdoor attacks in case the attacker has access to the model's training data. Our evaluation shows that XSub is not only effective and stealthy but also low-cost, showcasing its feasibility across a wide range of AI applications.

adversarial attack↗

NEPATEC2.0: NEPA Text Corpus v2.0

The National Environmental Policy Act of 1969, as amended (NEPA), is a major environmental law in the United States, requiring Federal agencies to consider and document potential environmental impacts before deciding on a proposed action. Modernization of NEPA and permitting processes faces significant challenges due to the lack of standardized formats and interoperable systems for organizing and sharing NEPA-related information across agencies. Much of the information gathered during NEPA reviews is written into documents such as categorical exclusions, environmental assessments, and environmental impact statements, then filed in predominately independent agency file stores that may or may not be publicly accessible. The application of metadata and data standards, such as those recommended by the Council on Environmental Quality (CEQ), to NEPA documents offers a shared vocabulary and structure for key entities like projects, processes, and documents that can streamline information exchange and enhance collaboration across systems. In this work, we publicly release NEPATEC2.0, an expanded corpus of NEPA documents with associated metadata. NEPATEC2.0 encompasses approximately 120,000 documents from 60,000 projects prepared by more than 60 different agencies. Modeled to align with CEQ metadata standards, NEPATEC2.0 promotes consistency in environmental reviews and supports the ongoing effort to modernize permitting technologies by facilitating more transparent, efficient, and data-driven decision-making. Importantly, NEPATEC2.0 demonstrates the possibilities and limitations of large language model-based prompting to extract information from NEPA documents at scale.

environmental review↗

ATcT — Active Thermochemical Tables Python Interface

SF-25-140 atct is a lightweight, Python client for the ATcT v1 API that enables programmatic access to high-accuracy thermochemical data and turnkey reaction-enthalpy analysis. The package implements full v1 endpoint coverage (species lookup by ATcT ID, name, formula, SMILES, InChI, CAS RN; covariance queries; health checks) with robust error handling, retries, and environment-based configuration for local/production endpoints. Beyond data retrieval, atct provides rigorously implemented reaction calculators that propagate uncertainties via either (i) a conventional independent-errors method (0 K or 298.15 K) or (ii) covariance-aware propagation using provided covariances at 298.15 K. Typed data classes ensure transparent, reproducible data structures and carry ATcT Thermochemical Network (TN) version identifiers for provenance. Dual import paths and comprehensive examples facilitate integration into research pipelines, enabling reproducible thermochemical calculations, automated validation, and downstream method development.

Bross, DavidHamilton [Argonne National Laboratory ↗

Coupling of high-resolution mass spectrometer and photosynthesis system for comprehensive leaf volatile metabolite profiling

Background Leaf-level biogenic volatile organic compounds (BVOCs) emissions represent a major source of organic gases in the atmosphere, influencing both climate and air quality. These emissions are strongly driven by environmental perturbations, which affect individual plant- to ecosystem-level processes. Uncovering all the BVOCs and understanding how their emissions respond to altered environmental conditions provide critical insights into vegetation-driven changes in atmospheric chemistry. We developed a tandem instrumentation setup that integrates a proton transfer reaction time-of-flight mass spectrometer (PTR-ToF-MS) with parts-per-trillion detection limits and a photosynthetic infrared gas exchange system for the untargeted survey of all the BVOCs. This novel system enables simultaneous, real-time monitoring of BVOC emissions and photosynthetic parameters at the leaf level, offering new opportunities to disentangle the physiological and environmental drivers of VOC release. Furthermore, we established the VOC Analysis and Processing Optimization Resource (VAPOR), an open-access software tool designed for rapid data post-processing and the analysis of the variability of hundreds of BVOCs. We assessed the performance of the tandem system under varying background conditions, using standard gas mixtures and a range of environmental factors. Results Blank emissions were substantially lower for major BVOCs (e.g., isoprene) compared to those observed in plant emissions. Despite this, the observation of background-level VOCs highlights the importance of routinely acquiring and accounting for blank measurements in analyses using the coupled instrumentation. Introduction of known VOC concentrations to the system demonstrated a linear response across different compounds with varying molecular compositions, indicating minimal gas loss regardless of chemical moieties within the coupled instrumentation. We applied the optimized system to investigate the physiological mechanisms driving BVOC emissions across different genotypes of poplar and pennycress. The high mass resolution capabilities of the PTR-ToF-MS, coupled with comprehensive VAPOR-driven data analysis, enabled the identification of several important BVOCs, including methanol and methanethiol; these BVOCs displayed substantial variation across pennycress genotypes and showed concentrations ~ 100–350% higher than the blank. Moreover, isoprene emissions varied significantly among poplar genotypes grown in different potting media. Conclusions Tandem instrumentation offers a powerful tool for profiling volatile molecular markers and elucidating their genetic and environmental underpinnings. This approach enhances our ability to predict BVOC emissions in response to genotype by environmental interactions and contributes to a deeper understanding of vegetation responses to environmental changes.

Biogenic volatile organic compounds↗

COMPASS-FME Synoptic Site Characterization

This dataset contains soil biogeochemical and physicochemical characterization data for the COMPASS-FME synoptic sites.This dataset also contains data for the paper Patel et al. 2025 "Transition zones at the changing coastal terrestrial-aquatic interface", https://doi.org/10.1029/2025JG008978.Coastal soils are a significant but highly uncertain component of global biogeochemical cycles. These systems experience unique spatial and temporal variability in biogeochemical processes, driven by wetland-to-upland gradients and hydrological fluctuations. We studied drivers of coastal soil variability (a) at regional scales and (b) across transects from upland forest to wetland, in two contrasting regions — Lake Erie, a freshwater lacustrine system, and Chesapeake Bay, a saltwater estuarine system. Salinity-related analytes were a key driver of soil variability, not just in the saltwater system, but surprisingly, also in the freshwater system. We had hypothesized linear trends in biogeochemical parameters along the TAI – however, contrary to expectations, transition soils were not consistently intermediate between upland and wetland endmembers; the non-monotonic trends of carbon, phosphorus, iron along our transects suggest that these are key analytes to study in our regions. Rapidly changing soil factors across coastal gradients provide insights into which soil processes may act as precursors to ecosystem shifts. Our comprehensive soil characterization across the coastal transects provides essential data for mechanistic modeling of ecosystem dynamics.The data are provided as processed, csv files. Raw data and processing scripts can be accessed on GitHub (https://github.com/COMPASS-DOE/cmps-soil_characterization).A note on the nomenclature: the experimental design represents three points along the coastal gradient -- upland, transition, and wetland. "wetland" is referred to as "marsh" in the corresponding paper. The two terms can be used interchangeably for the sites in this study.

54 ENVIRONMENTAL SCIENCES↗

Life Cycle Analysis of Natural Gas Extraction and Power Generation: U.S. 2020 Emissions Profile

This analysis expands upon previous life cycle analyses (LCAs) of natural gas systems performed by the National Energy Technology Laboratory (NETL). It provides a complete inventory of emissions to air and water, water consumption, and land use change. These environmental burdens are detailed for all supply chain steps from natural gas production through natural gas distribution. This package includes the report, the NETL Natural Gas Model, and appendices that include several Excel workbooks and a python script to provide transparent access to the calculations and resulting data. This is revision 1 of the 2024 study (published December 17, 2024 and updated on January 24, 2025) and corrects a modeling error in natural gas composition. See the errata on page 2 for more information.

03 NATURAL GAS↗

NEPATEC v2.0: Standardized Metadata and Text Corpus of National Environmental Policy Act Documents

The National Environmental Policy Act of 1969, as amended (NEPA), is a major environmental law in the United States, requiring Federal agencies to consider and document potential environmental impacts before deciding on a proposed action. Modernization of NEPA and permitting processes faces significant challenges due to the lack of standardized formats and interoperable systems for organizing and sharing NEPA-related information across agencies. Much of the information gathered during NEPA reviews is written into documents such as categorical exclusions, environmental assessments, and environmental impact statements, then filed in predominately independent agency file stores that may or may not be publicly accessible. The application of metadata and data standards, such as those recommended by the Council on Environmental Quality (CEQ), to NEPA documents offers a shared vocabulary and structure for key entities like projects, processes, and documents that can streamline information exchange and enhance collaboration across systems. In this work, we publicly release NEPATEC2.0, an expanded corpus of NEPA documents with associated metadata. NEPATEC2.0 encompasses approximately 120,000 documents from 60,000 projects prepared by more than 60 different agencies. Modeled to align with CEQ metadata standards, NEPATEC2.0 promotes consistency in environmental reviews and supports the ongoing effort to modernize permitting technologies by facilitating more transparent, efficient, and data-driven decision-making. Importantly, NEPATEC2.0 demonstrates the possibilities and limitations of large language model-based prompting to extract information from NEPA documents at scale.

54 ENVIRONMENTAL SCIENCES↗

ResStock Measure Documentation: Cold Climate Air-Source Heat Pump

This report is part of series describing a variety of different ResStock™ measures. "Measures" refers to energy efficiency retrofits that can be applied to buildings during modeling. This documentation covers the "Cold Climate Air-Source Heat Pump" measure upgrade methodology and briefly discusses key results. All results can be accessed on the ResStock Open Energy Data Initiative "End-Use Load Profiles for the U.S. Building Stock" data lake and on the data viewer at resstock.nlr.gov.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Network Security Challenges and Countermeasures for Software-Defined Smart Grids: A Survey

The rise of grid modernization has been prompted by the escalating demand for power, the deteriorating state of infrastructure, and the growing concern regarding the reliability of electric utilities. The smart grid encompasses recent advancements in electronics, technology, telecommunications, and computer capabilities. Smart grid telecommunication frameworks provide bidirectional communication to facilitate grid operations. Software-defined networking (SDN) is a proposed approach for monitoring and regulating telecommunication networks, which allows for enhanced visibility, control, and security in smart grid systems. Nevertheless, the integration of telecommunications infrastructure exposes smart grid networks to potential cyberattacks. Unauthorized individuals may exploit unauthorized access to intercept communications, introduce fabricated data into system measurements, overwhelm communication channels with false data packets, or attack centralized controllers to disable network control. An ongoing, thorough examination of cyber attacks and protection strategies for smart grid networks is essential due to the ever-changing nature of these threats. Previous surveys on smart grid security lack modern methodologies and, to the best of our knowledge, most, if not all, focus on only one sort of attack or protection. This survey examines the most recent security techniques, simultaneous multi-pronged cyber attacks, and defense utilities in order to address the challenges of future SDN smart grid research. The objective is to identify future research requirements, describe the existing security challenges, and highlight emerging threats and their potential impact on the deployment of software-defined smart grid (SD-SG).

24 POWER TRANSMISSION AND DISTRIBUTION↗