Search NASA⌕ Search

SEARCH · Search NASA

Results for “data usage”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Results from the last DD and DT JET campaigns in the framework of the EUROfusion Tokamak Exploitation Work Package activity

JET, the only tokamak capable of operating with deuterium–tritium (D–T) fuel (since TFTR was shutdown in 1999), has provided essential experimental data to support ITER and DEMO design and operation. Within the EUROfusion Tokamak Exploitation Work Package, JET completed its final campaigns (2022–2023), culminating in the third D–T campaign (DTE3). These experiments addressed key challenges in plasma scenarios, exhaust control, and tritium management under reactor-relevant conditions. Significant progress was achieved in demonstrating ITER-like integrated scenarios with impurity seeding, achieving partial divertor detachment and high confinement ($H_{98}(y,2)$ ≈ 0.85) at 3 MA in D–T plasmas. Advanced exhaust regimes such as quasi-continuous exhaust (QCE) and X-point radiator (XPR) were successfully achieved first in D–D and then extended to D–T operation, confirming their relevance for mixed isotope operation. Operational milestones included a new world record of 69 MJ fusion energy in tritium-rich hybrid plasmas and long-pulse H-mode operation up to 60 s, contributing with unique data to the CICLOP database. Physics studies focused on peeling-limited pedestals in support of ITER and improved understanding of edge stability and impurity screening in metallic environments. Extensive usage of the shattered pellet injector (SPI) on JET provided critical information for the design of the ITER disruption mitigation system (DMS). Real-time control systems for D/T ratio control and plasma exhaust were deployed and demonstrated in D–D and D–T, while energetic particle physics investigations unfolded the role of fast ions in turbulence suppression mechanisms. Comprehensive tritium retention studies using gas balance method, post-mortem analysis, and ITER-relevant laser induced desorption spectroscopy (LIDS) diagnostics provided essential input for tritium accountancy strategies. These results are validating the ITER operational concepts, inform DEMO design, and deliver critical experience in nuclear operation and scenario integration.

disruptions↗

ComStock Measure Documentation: Thermostat Setbacks During Unoccupied Periods

This report assesses the potential for nationwide adoption of thermostat setbacks in appropriate applications. Building on the 3-year End-Use Load Profiles project to calibrate and validate the U.S. Department of Energy’s ResStock™ and ComStock™ models, this work produces national datasets that enable cities, states, utilities, and other stakeholders to answer a broad range of questions regarding their commercial building stock. ComStock is a highly granular, bottom-up model that uses various data sources, statistical sampling methods, and advanced building energy simulations to estimate the annual sub-hourly energy consumption of the commercial building stock across the United States. The “baseline” model intends to represent the U.S. commercial building stock as it existed in 2018. The methodology of the baseline model is discussed in the ComStock Reference Documentation. The goal of this work is to develop energy efficiency and demand flexibility measures that cover market-ready technologies and study their mass-adoption impact on the baseline building stock. “Measures” refers to various “what-if” scenarios that can be applied to buildings. The results for the baseline and measure scenario simulations are published in public datasets that provide insights into building stock characteristics, operational behaviors, utility bill impacts, and annual and sub-hourly energy usage by fuel type and end use. This report describes the modeling methodology for a single ComStock measure scenario— Thermostat Setbacks During Unoccupied Periods—and briefly introduces key results. The full public dataset can be accessed on the ComStock data lake or via the Data Viewer at comstock.nrel.gov. The public dataset enables users to create custom aggregations of results for their use cases (e.g., filter to a specific county or building type).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Dataset: Breaking the barrier of human-annotated training data for machine-learning-aided plant research using aerial imagery

This dataset supports the implementation described in the manuscript "Breaking the Barrier of Human-Annotated Training Data for Machine-Learning-Aided Biological Research Using Aerial Imagery." It comprises UAV aerial imagery used to execute the code available at https://github.com/pixelvar79/GAN-Flowering-Detection-paper. For detailed information on dataset usage and instructions for implementing the code to reproduce the study, please refer to the GitHub repository.

generative and adversarial learning↗

Notice of Submittal – 2023 Radionuclide Air Emissions Report for Los Alamos National Laboratory

This report describes the emissions of airborne radionuclides from operations at Los Alamos National Laboratory (LANL) for calendar year 2023 and the resulting off-site dose from these emissions. This document fulfills the requirements established by the National Emissions Standards for Hazardous Air Pollutants in 40 CFR 61, Subpart H – Emissions of Radionuclides other than Radon from Department of Energy Facilities, commonly referred to as the Radionuclide NESHAP or Rad-NESHAP. Compliance with this regulation and preparation of this document is the responsibility of LANL’s Rad NESHAP compliance program, which is part of the Environmental Protection and Compliance (EPC) Division. The information in this report is required under the Clean Air Act and is being submitted to the U.S. Environmental Protection Agency (EPA) Headquarters and EPA Region 6. The highest effective dose equivalent (EDE) to an off-site member of the public was calculated using procedures specified by the EPA and described in this report. LANL’s EDE was 0.43 for 2023. The annual limit is 10 millirem per year, established by the EPA in 40 CFR 61 Subpart H. All measured air emissions are modeled to a single location, known as the Maximally Exposed Individual (MEI). During calendar year 2023, LANL continuously monitored radionuclide emissions at 28 “major” release points, or stacks. The Laboratory estimates emissions from an additional 59 “minor” release points using radionuclide usage source terms in lieu of stack monitoring. Also, LANL uses an EPA approved network of air samplers around the Laboratory perimeter to monitor ambient airborne levels of radionuclides. To provide data for dispersion modeling and dose assessment, LANL maintains and operates several meteorological monitoring towers. From these various systems, a comprehensive evaluation is conducted to calculate the MEI dose for the Laboratory. The MEI can be any member of the public at any off-site location where there is a residence, school, business, or office. In 2023, this MEI location was a business at 129 New Mexico State Road 4 (NM-4), located in the northern end of White Rock. The primary contributors to the off-site dose at this location are the ambient air data at that location combined with the collected potential emissions from unmonitored (minor) sources. Overall, the MEI dose in 2023 is similar to that which has been observed in recent years, and it remains well below the EPA’s 10 millirem per year limit. Doses reported to the EPA for the past 10 years are shown in Table E1.

54 ENVIRONMENTAL SCIENCES↗

Sandia Toolkit Manual (V.5.21.1)

This report provides documentation for the Sandia Toolkit (STK) modules. STK modules are intended to provide infrastructure that assists the development of computational engineering software such as finite-element analysis applications. STK includes modules for unstructured-mesh data structures, reading/writing mesh files, geometric proximity search, transfers, MPMD coupling support, and various other utilities. This document contains a chapter for each module, and each chapter contains overview descriptions and usage examples. Usage examples are primarily code listings which are generated from working test programs that are included in the STK code-base. A goal of this approach is to ensure that the usage examples will not fall out of date.

97 MATHEMATICS AND COMPUTING↗

A practical approach to using the Genomic Standards Consortium MIxS reporting standard for comparative genomics and metagenomics

Comparative analysis of (meta)genomes necessitates aggregation, integration, and synthesis of well-annotated data using standards. The Genomic Standards Consortium (GSC) collaborates with the research community to develop and maintain the Minimal Information about any (x) Sequence (MIxS) reporting standard for genomic data. To facilitate use of the GSC’s MIxS reporting standard, we provide a description of the structure and terminology, how to navigate ontologies for required terms in MIxS, and demonstrate practical usage through a soil metagenome example.

standards, metadata, genome, metagenome, schema, v↗

SALSA: a new versatile readout chip for MPGD detectors

The SALSA chip is a future readout ASIC foreseen for the MPGD detectors, developed in the framework of the EIC collider project, to equip the MPGD trackers of the EPIC experiment. It is designed to be versatile, to be adapted to other usages of MPGD detectors like TPC or photon detectors. It integrates a frontend block and an ADC for each of the 64 channels, associated to a configurable DSP processor meant to correct data and reduce the raw data flux to limit the output bandwidth. It will be compatible with the continuous readout foreseen for the EPIC DAQ, but will also work in a triggered environment. Several prototypes are already produced in order to qualify the different blocks of the chip, in particular the frontend, the ADC and the clock generation. The next 32-channel prototype is currently under development and is planed to be produced in 2025. In conclusion, the final prototype will be produced and tested from 2026 for a production of the SALSA chip at the horizon of 2027.

Data processing↗

DoCeph: DPU-Offloaded Messaging in Ceph for Reduced Host CPU Utilization

Ceph is a widely used distributed object store, but its messenger layer imposes substantial CPU overhead on the host. To address this limitation, we propose DoCeph, a DPU-offloaded storage architecture for Ceph that disaggregates the system by offloading the communication-intensive messaging component to the DPU while retaining the storage backend on the host. The DPU efficiently manages communication, using lightweight RPC for metadata operations and DMA for data transfer. Moreover, DoCeph introduces a pipelining technique that overlaps data transmission with buffer preparation, mitigating hardware-imposed transfer size limitations. We implemented DoCeph on a Ceph cluster with NVIDIA BlueField-3 DPUs. Evaluation results indicate that DoCeph cuts host CPU usage by up to 92% while sustaining stable throughput and providing larger performance benefits for object writes over 1 MB.

Park, Kuri [Sogang University]↗

Hot Droughts and Forest Tree Dynamics in the Amazon - Statistical Models, Scripts, Data, and Outputs

This package contains data, outputs, equations, and R scripts for analyses for manuscript entitled "Hot droughts in the Amazon: A window to a future hypertropical climate" by J. Chambers et al., in particular it contains statistical models and analyses for the INPA BIONTE tree mortality study. The Models folder contains details for all statistical models in PDF files. The Scripts folder contains the R scripts for Bayesian Hierarchical Models (two text files) and SEMs (one text file) are separate and reasonably annotated. All data associated with these scripts are in the data folder. The Data folder contains two of the three CSV files used for the analyses and are called by the R scripts. Two of them are part of published datasets (`BIONTE_mortality-rates.csv` from Lima et al. 2024, DOI:10.15486/ngt/1898910 and `SPEI.csv` from Pastorello et al. 2023 DOI:10.15486/ngt/1958257) and also provided in this package for convenience (please see the corresponding datasets for usage and citation terms). The third dataset (`BIONTE_gapfilled_wd.csv`) contains sensitive information and can be obtained by contacting the manuscript lead author. The Outputs folder contains the two output files that provide extra information about the analyses. The file `figuresFeb2025d.pdf` contains all the figures from the manuscript - captions are in the manuscript. The file `ChambersMS.pdf` contains primary results from Bayesian statistical models, regression analyses, and validation steps applied to the tree mortality data from the INPA experiments. The document includes visual summaries, model diagnostics, and leave-one-out (LOO) validation results. A breakdown of file contents can be found in the README file that is part of this package.

54 ENVIRONMENTAL SCIENCES↗

Predictive Model for Starlink Maritime Performance Using Multi-Horizon RandomForest

Low Earth orbit (LEO) satellite systems have become a crucial enabler of broadband access for maritime industries, where traditional networks are unavailable. However, the high mobility of LEO constellations and constantly changing weather conditions result in unpredictable link fluctuations, limiting the ability of maritime platforms to plan bandwidth usage proactively. To the best of our knowledge, no prior work has developed a short-term predictive model for maritime LEO connectivity using real experimental field measurements. This paper proposes a data-driven forecasting model that predicts future downlink throughput using multi-horizon RandomForest regression. The model is trained using real experimental coastal measurement data incorporating recent throughput history, network-layer indicators, and environmental variables. The proposed approach reduces mean absolute error by approximately 31% compared to a persistence baseline for 15-minute horizons. It maintains a measurable improvement at 30 minutes, despite increased stochasticity. These findings confirm that proactive bandwidth awareness is feasible on maritime platforms and can effectively support operational decisions such as adaptive streaming, routing, and resource scheduling. The performance gap between forecasting horizons also highlights the need for expanded offshore datasets to improve prediction robustness under harsher maritime environments.

97 MATHEMATICS AND COMPUTING↗

Visual Analytics of Multivariate Networks With Representation Learning and Composite Variable Construction

Multivariate networks are commonly found in real-world data-driven applications. Uncovering and understanding the relations of interest in multivariate networks is not a trivial task. This article presents a visual analytics workflow for studying multivariate networks to extract associations between different structural and semantic characteristics of the networks (e.g., what are the combinations of attributes largely relating to the density of a social network?). The workflow consists of a neural-network-based learning phase to classify the data based on the chosen input and output attributes, a dimensionality reduction and optimization phase to produce a simplified set of results for examination, and finally an interpreting phase conducted by the user through an interactive visualization interface. A key part of our design is a composite variable construction step that remodels nonlinear features obtained by neural networks into linear features that are intuitive to interpret. We demonstrate the capabilities of this workflow with multiple case studies on networks derived from social media usage and also evaluate the workflow with qualitative feedback from experts.

97 MATHEMATICS AND COMPUTING↗

ComStock Measure Documentation: Fan Static Pressure Reset for Multizone Variable Air Volume Systems

This report assesses the potential for nationwide adoption of a duct static pressure reset in MZ VAV systems in appropriate applications. Building on the 3-year End-Use Load Profiles project to calibrate and validate the U.S. Department of Energy’s ResStock™ and ComStock™ models, this work produces national datasets that enable cities, states, utilities, and other stakeholders to answer a broad range of questions regarding their commercial building stock. ComStock is a highly granular, bottom-up model that uses various data sources, statistical sampling methods, and advanced building energy simulations to estimate the annual sub-hourly energy consumption of the commercial building stock across the United States. The “baseline” model intends to represent the U.S. commercial building stock as it existed in 2018. The methodology of the baseline model is discussed in the ComStock Reference Documentation. The goal of this work is to develop energy efficiency and demand flexibility measures that cover market-ready technologies and study their mass-adoption impact on the baseline building stock. “Measures” refers to various “what-if” scenarios that can be applied to buildings. The results for the baseline and measure scenario simulations are published in public datasets that provide insights into building stock characteristics, operational behaviors, utility bill impacts, and annual and sub-hourly energy usage by fuel type and end use. This report describes the modeling methodology for a single ComStock measure scenario—Fan Static Pressure Reset for Multizone Variable Air Volume (VAV) Systems—and briefly introduces key results. The full public dataset can be accessed on the ComStock data lake or via the Data Viewer at comstock.nrel.gov. The public dataset enables users to create custom aggregations of results for their use case (e.g., filter to a specific county or building type).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Intelligent Experiments through Real-Time AI: Fast Data Processing and Autonomous Detector Control for High-Energy Nuclear Experiments

The aim of this project is to develop software and hardware for fast real-time data processing and autonomous detector control and calibration for the sPHENIX and the future EIC experiments. Below summarizes Georgia Tech team efforts in the past year: 1. We developed a real-time clustering algorithm and FPGA-based pipeline architecture for processing fired pixel data from ALPIDE sensors in sPHENIX experiments. Our Columnar Clustering Co-Design introduces a hardware-aware, stream-friendly approach that segments pixel data by column pairs using a Column Pair Clustering (CPC) strategy, followed by Cluster Stitching to merge adjacent subclusters. Implemented in Vitis HLS, the pipeline comprises five stages—read-in, subclustering, stitching, analysis, and write-out—connected by tagged HLS streams with custom end-of-event signaling for robust synchronization. We designed a pipelined dataflow model optimized for throughput, low latency, and minimal buffering, enabling scalable clustering across events of arbitrary size. Our system maintains spatial precision via center-of-mass and shape key extraction and efficiently handles edge cases such as fragmented or nested clusters. Compared against DBSCAN in both software and hardware, our approach demonstrates competitive performance under FPGA constraints. 2. We also conducted a comprehensive algorithm-to-hardware co-design of connected component analysis tailored for sPHENIX experiments, focusing on real-time, low-latency processing using FPGAs and High-Level Synthesis (HLS). Starting from a Python-based particle tracking pipeline, the team translated the core logic—graph traversal via DFS and Union-Find—into an HLS-compatible C++ model, replacing dynamic memory and recursion with static arrays and pipelined control flow. The final design includes a fully streamed and dataflow-compatible Union-Find kernel optimized across five iterations, incorporating loop pipelining, array partitioning, AXI/FIFO interface tuning, and function flattening. Experimental results show up to 14.8× speedup over the CPU baseline, reducing per-graph latency to 1.58 μs and demonstrating strong resource efficiency with only ~7k LUTs and zero BRAM usage. The design maintains functional correctness against the Python reference using a Python-based C-simulation framework and Mean Squared Error metrics. This work validates the potential of HLS-driven FPGA designs for edge-level HEP data acquisition, laying a scalable foundation for future integration with real-time detector pipelines and multi-graph processing systems.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

District-Scale Analysis of Electricity Load and Strategies to Improve Energy Reliability Using Prototype District Models

Projected increases in electricity demand in the U.S. highlight the urgent need for effective load management to ensure grid reliability. As the building sector accounts for approximately 75% of electricity usage, enhancing energy efficiency and flexibility in this sector is crucial. Adopting district-level approaches offers significant advantages over traditional individual building analyses by enabling shared infrastructure and economies of scale. To navigate the data and computational challenges associated with modeling energy at the district level, prototype district models have been proposed as holistic, system-level solutions that capture complex interactions within typical configurations. This study presents these models as a reference tool for analyzing district-scale energy systems across various climate zones in the U.S. Developed with input from stakeholders, these models integrate varied building characteristics, inter-building connections, and energy system interactions. A case study utilizing the Urban Edge prototype district model, implemented on the URBANopt™ platform, evaluates multiple demand scenarios and the impact of distributed energy resources such as fuel-fired backup generators, photovoltaic systems, and batteries. Findings suggest that while new electric systems can significantly reduce annual energy use, they may also elevate peak electricity loads, with a notable 43% increase in heating-dominant climate zone 5B. The optimal backup power solutions vary based on location, influenced by factors such as utility rates and incentives. For example, PV and batteries perform well in high-cost regions like New York City, while diesel backup generators are more suitable for backup needs in climate zone 3A, such as Atlanta. Thus, this research highlights the importance of prototype district models for future district-scale energy planning.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Investigating resource-efficient neutron/gamma classification ML models targeting eFPGAs

There has been considerable interest and resulting progress in implementing machine learning (ML) models in hardware over the last several years from the particle and nuclear physics communities. A big driver has been the release of the Python package, hls4ml, which has enabled porting models specified and trained using Python ML libraries to register transfer level (RTL) code. So far, the primary end targets have been commercial field-programmable gate arrays (FPGAs) or synthesized custom blocks on application specific integrated circuits (ASICs). However, recent developments in open-source embedded FPGA (eFPGA) frameworks now provide an alternate, more flexible pathway for implementing ML models in hardware. These customized eFPGA fabrics can be integrated as part of an overall chip design. In general, the decision between a fully custom, eFPGA, or commercial FPGA ML implementation will depend on the details of the end-use application. In this work, we explored the parameter space for eFPGA implementations of fully-connected neural network (fcNN) and boosted decision tree (BDT) models using the task of neutron/gamma classification with a specific focus on resource efficiency. We used data collected using an AmBe sealed source incident on Stilbene, which was optically coupled to an OnSemi J-series silicon photomultiplier (SiPM) to generate training and test data for this study. We investigated relevant input features and the effects of bit-resolution and sampling rate as well as trade-offs in hyperparameters for both ML architectures while tracking total resource usage. The performance metric used to track model performance was the calculated neutron efficiency at a gamma leakage of 10 -3 . The results of the study will be used to aid the specification of an eFPGA fabric, which will be integrated as part of a test chip.

47 OTHER INSTRUMENTATION↗

A Functional Survey of the Regulatory Landscape of Estrogen Receptor–Positive Breast Cancer Evolution

Abstract Only a handful of somatic alterations have been linked to endocrine therapy resistance in hormone-dependent breast cancer, potentially explaining ∼40% of relapses. If other mechanisms underlie the evolution of hormone-dependent breast cancer under adjuvant therapy is currently unknown. In this work, we employ functional genomics to dissect the contribution of cis-regulatory elements (CRE) to cancer evolution by focusing on 12 megabases of noncoding DNA, including clonal enhancers, gene promoters, and boundaries of topologically associating domains. Parallel epigenetic perturbation (CRISPRi) in vitro reveals context-dependent roles for many of these CREs, with a specific impact on dormancy entrance and endocrine therapy resistance. Profiling of CRE somatic alterations in a unique, longitudinal cohort of patients treated with endocrine therapies identifies a limited set of noncoding changes potentially involved in therapy resistance. Overall, our data uncover how endocrine therapies trigger the emergence of transient features which could ultimately be exploited to hinder the adaptive process. Significance: This study shows that cells adapting to endocrine therapies undergo changes in the usage or regulatory regions. Dormant cells are less vulnerable to regulatory perturbation but gain transient dependencies which can be exploited to decrease the formation of dormant persisters.

Oncology↗

Rural EVSE Planning and Analysis

The dataset includes detailed anonymized public charging station usage from several rural stations on the ChargePoint and Shell Recharge Solutions (formerly Greenlots) networks situated in and around Athens, Ohio, a rural Appalachian community. Both Level 2 and DC fast charging stations are represented. Historical data in the set date back to 2019; additional data will be uploaded semiannually until the project's completion in 2023. Each charging session recorded includes information on date and time, location, charging station level, session duration, energy delivered, and fuel savings.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Fast jet tagging with MLP-Mixers on FPGAs

We explore the innovative use of MLP-Mixer models for real-time jet tagging and establish their feasibility on resource-constrained hardware like FPGAs. MLP-Mixers excel in processing sequences of jet constituents, achieving state-of-the-art performance on datasets mimicking Large Hadron Collider conditions. By using advanced optimization techniques such as High-Granularity Quantization and Distributed Arithmetic, we achieve unprecedented efficiency. These models match or surpass the accuracy of previous architectures, reduce hardware resource usage by up to 97%, double the throughput, and half the latency. Additionally, non-permutation-invariant architectures enable smart feature prioritization and efficient FPGA deployment, setting a new benchmark for machine learning in real-time data processing at particle colliders.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗