Search NASA⌕ Search

SEARCH · Search NASA

Results for “data storage data management”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Managing negative values is reservoir inflow computation: A case study

Reservoir inflow is conventionally estimated using the water balance method, which involves the reservoir release and the change in storage during the period considered. As a result, the estimated inflow may sometimes be negative as the errors involved in each input variable build-up to the output. In our study, the fleet data was provided by the Tennessee Valley Authority (TVA) for their Norris Hydropower facility. Unlike the flow release data, which was readily accessible, the change in storage had to be calculated using the reservoir elevation and volume relationship. The original inflow estimates produced a wide range of negative values with large outliers, making it difficult to visualize the current trends. This paper describes a methodology to remove the negative values encountered during the inflow computation, and the results were analyzed by correlating with the nearby streamflow gaging stations.

Shibu, Asha↗

Fast, Controllable, and Modular Solid-State Circuit Breaker Design for Battery Management Systems

Electric power grid is experiencing a growing number of distributed and inertia-free generation resources. To facilitate the growing generation and load demands, and ensure stable operation, energy storage systems, especially behind-themeter-storage (BTMS), have emerged as a potential candidate. BTMS plays a vital role in the grid storage sector and supports high power charging for EVs. However, the potential of thermal runaway and associated safety concerns in the batteries can hamper their widespread adoption. In this work, we provide a solution for a fast and controllable discharge of a cell that was identified as a stressful/faulty, in the battery pack, and fast circuit breaking leveraging the Solid-State Circuit Breaker (SSCB) technology which provides active control over the cell connection as opposed to conventional passive solutions. Testing on the simulation platform successfully validated the concept, demonstrating its efficacy. Controlled discharge testing with 20Ah LiFePO4 Lithium Iron Phosphate (LFP) cells from 1C-10C current rate on the hardware prototype corroborated simulation results, demonstrating design feasibility and providing essential data for its performance and thermal characteristics, while also revealing limitations that inform areas for further optimization.

33 ADVANCED PROPULSION SYSTEMS↗

Unveiling the Potential of MeshGraphNets for Predicting Subsurface Evolution in Carbon Storage Projects

This is the conference paper accompanying an oral presentation “Unveiling the Potential of MeshGraphNets for Predicting Subsurface Evolution in Carbon Storage Projects” at the 17th International Conference on Greenhouse Gas Control Technologies GHGT-17 held in Calgary, Canada, October 20-24 , 2024. Carbon capture and storage (CCS) technology is critical for mitigating climate change but requires effective subsurface reservoir management to ensure safe containment of injected CO2. Accurate predictions of reservoir pressure and saturation are essential for assessing long-term CCS performance. Traditional numerical simulations, while effective, are computationally intensive, time-consuming, and constrained by data discretization. Previous work has shown the effectiveness of MeshGraphNets (MGN), a graph-based machine learning framework, as an innovative alternative for predicting reservoir behavior. MGN leverages graph neural networks (GNNs) and mesh representations to model complex geological formations, offering superior adaptability across different discretizations and reservoir configurations. Classic MGN implementations utilize an autoregressive technique to predict future behavior based on current predictions, but this technique is hampered by error accumulation over time. To enhance the model accuracy in time-series predictions, this study implemented a multi-step rollout strategy that integrates autoregressive predictions during training to stabilize prediction of saturation over time. Using the Illinois Basin – Decatur Project (IBDP) dataset, comprising 100 simulations of CO2 injection, pressure, and saturation changes, the framework demonstrated its ability to learn spatial dependencies and temporal dynamics. With inputs including permeabilities, porosities, and injection rates, MGN accurately predicted CO2 plume evolution over time, even with limited training data. Moreover, the addition of a multi-step rollout procedure during training improved the ability of MGN to predict stably over time by ~15%. This research positions MGN, enhanced with multi-step rollout capabilities, as a robust and efficient tool for CCS applications. It advances the field by enabling precise, computationally efficient predictions of reservoir behavior, providing a foundation for the broader adoption of machine learning frameworks in CCS and other geoscience domains.

Holcomb, Paul↗

Numerical Investigation of High Delta T Sensible Storage Integrated CO2 Heat Pump: Preprint

To assist building heating electrification, this paper numerically investigates a load flexible heat pump system for commercial buildings. The system consists of a CO2 vapor compression cycle, a sensible thermal storage tank, and an air handling unit. The thermal storage medium is inexpensive, non-toxic and stable anti-freeze solution (30% potassium acetate). The air handing unit has an indoor coil and a ventilation coil. The system can be used to manage building electric load. During peak hours, the heat pump is off and the hot solution water is discharged from the tank to heat up the indoor air and ventilation air. During the hour of charge, the heat pump delivers hot solution water to the tank and to the air. The tank can also stand by while the heat pump provides space heating directly. We selected a medium sized office building located in Minnesota as the representative building and used EnergyPlus to obtain its 24 hour load data. We designed three storage tank volumes assuming 50 degrees C, 65 degrees C and 80 degrees C tank temperatures to independently provide the building load for 4 hours in the morning. The higher the tank temperature, the smaller the required volume, and thus higher energy density. The effective energy density is 78 with an 80 degrees C tank, and 40 kWhth/m3 with 50 degrees C. We simulated the tank integrated heat pump performance subjected to the 24-hour building load profile and ambient data. The baseline is the same system without storage tank. There was a trade-off between the storage energy density and the charging COP. The charge hour COP was 2.77 to charge the tank to 80 degrees C, and 3.01 to 50 degrees C. The proposed system could shift building load from the peak hours (8:00 - 12:00) to off-business hour (23:00 - 7:00+1). It eliminated 100% compressor electricity use during the peak hours, and avoided a peak electric power of 34 kW. The 65 degrees C tank saved 9.5 kWhe (4%) considering all day operation, which was the best balance between energy density and the system operation efficiency among the three options.

CO2 heat pump↗

Dynamic CCS-EJ-SJ Database and Web Application - What's New

At the 2024 FECM/NETL Carbon Management Research Project Review Meeting, within the Carbon Transport and Storage Breakout Session 3, the presentation "Dynamic CCS-EJ-SJ Database and Web Application - What's New" highlights the critical tool designed to integrate environmental and social justice considerations into Carbon Capture and Storage (CCS) projects. Key features include an interactive dashboard for data access and visualization, which supports stakeholders in making informed decisions regarding CCS implementation, and updated data layers. The latest version enhances data integration and usability, providing a comprehensive resource for assessing the social and environmental impacts of CCS projects. There are 7 categories in the CCS EJSJ v2 database (released 03/31/2024): environmental justice, energy justice, economic justice, social justice, ecosystem assets, clean energy, and infrastructure. Most of the layers within each category have been updated in this version. As compared to the old database, there are 3 new categories in the v2 database: ecosystem assets, clean energy, and infrastructure.

Sharma, Maneesh↗

Risk-informed Hierarchical Control of Behind-the-Meter DERs with AMI Data Integration (Final Technical Report)

This project addresses several key barriers to implement the next generation demand response applications and provides a clear understanding of implementing hierarchical and standalone control using AMI data. Through this program, Eaton has developed and tested a meter-as-a-controller prototype with the help of other partners--- National Renewable Energy Laboratory (NREL), Electric Power Research Institute (EPRI), Pecan St Inc. (PSI), and Delaware Electric Cooperative (DEC). The controller can utilize residential controllable loads such as heating, ventilation, and air conditioner (HVAC), electric water heater and distributed energy resources like solar PV and battery energy storage systems for off-setting the demand that is required from the grid, thus providing reliable grid-services for demand reduction or peak shaving. The controller is also capable of coordinating the resources of the premises for better management and energy efficiency while meeting the comfort bound of the premises owner as quality-of-service. The development has been demonstrated in a three virtual-home setup at system performance lab of NREL with real appliances (HVAC, electric water heater, solar PV, and battery). The technology has also been proved through laboratory and field demonstration with successful interconnectivity (e.g., end-to-end communication and data exchange) between the residential appliances and utility through the RF network at Delaware Electric Co-op (DEC) in Delaware.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

From Depletion to Restoration: Lessons From Long‐Term Monitoring of Carbon Gains and Losses in Cropping Systems

As global atmospheric CO 2 rapidly approaches a key tipping point, there is an urgent need to implement strategies to reverse this pattern. A generally accepted understanding of carbon (C) in agricultural fields includes: (H1) substantial C loss occurs when natural vegetation is converted to crops, (H2) soils typically reach a steady-state C concentration under contemporary practices, and (H3) improved management or crop selection can enhance soil C stocks over time. Significant variability exists, but studies consistently show large C losses from agricultural ecosystems, supporting H1. Although steady-state C levels (H2) are commonly assumed, measuring C gains or losses in mature agroecosystems is challenging. Efforts to increase soil C storage (H3) have limited data due to the diversity of potential practices, compounded by substantial variability in soil C measurements. Here, long-term (7–17 year) ecosystem C flux data from diverse cropping systems revealed that conventionally tilled annual row crops (maize and soybean) act as significant long-term atmospheric C sources, challenging H2. Furthermore, conservation tillage practices reduced C losses compared with conventional tillage but showed minimal evidence for long-term ecosystem C storage, even after 20+ years. This indicates that no-till practices reduce C losses but imply that no soil C is added, challenging H3. By contrast, perennial Miscanthus × giganteus, Panicum virgatum, and restored tallgrass prairie systems store C at the ecosystem scale more effectively than minimally tilled annual row crops. Analysis over multiple years demonstrates significant ecosystem C storage with perennial crops, varying by species, starting in the first year of transition. These findings, although focused on one region, suggest that the assumptions of steady-state C levels and increased storage from conservation practices do not universally apply and that significant changes to agroecosystems are required to increase C storage.

59 BASIC BIOLOGICAL SCIENCES↗

Deploying and Operating CephFS for Scientific Applications at Fermilab

Fermilab has been running a Ceph cluster in production for several years to support high-throughput scientific computing. Our primary use case is CephFS, which serves interactive data analysis workloads, with growing interest in using RGW for scalable object storage of scientific datasets. In this talk, we'll share lessons learned from successfully deploying and maintaining our Ceph cluster with cephadm, including challenges faced, performance tuning, and operational practices. We'll also present custom tools we've developed to streamline monitoring and management and discuss how Ceph fits into our broader storage architecture for large-scale scientific research.

Peisker, Alison [Fermilab]↗

Product Defect Detection System: SYSM- 5620 Final Project

Retail sales is a growing market estimated to up to seven percent year over year. With this growing market there is also a trend in growing rate of retail returns, estimated just last year at $\$$850 billion. Retail stores must ensure that products available for purchase remain safe, undamaged, and acceptable to customers throughout their time in the store. This job exists regardless of the specific solution used because stores are always responsible for preventing damaged or defective products from reaching customers and when they fail to this is categorized under operation inefficiencies which accounts for an estimated $\$$12 billion in returns. When defective items remain on the sales floor, stores may experience increased returns, reduced customer satisfaction, loss of customer trust, and potential safety concerns depending on the product type. As a result, the core job to be done is to identify defective products quickly, remove them from the sales floor before they are purchased, and preserve useful information about the defect so that the store can improve its handling, stocking, and supplier coordination over time. The need for a more reliable process is especially important in high volume retail environments where employees manage large numbers of products across many aisles, shelves, and storage areas. In these settings, manual inspection alone can be inconsistent and difficult to sustain at the individual item level. At the same time, broader retail trends continue to emphasize operational efficiency, product visibility, and improved customer experience, creating an opportunity for more automated and data driven defect detection methods.

42 ENGINEERING↗

Cataloging Legacy Data from the Tritium Systems Test Assembly Program

The Tritium Systems Test Assembly (TSTA) at Los Alamos National Laboratory, operational from 1984 to 2001, was critical in advancing fusion fuel cycle technologies, including tritium storage, gas separation, and pumping. TSTA’s contributions, particularly in safe tritium operations, have influenced subsequent fusion projects. This paper discusses the ongoing effort to digitize and catalog TSTA’s historical data to create a searchable resource for the fusion research community. While the long-term objective is to develop a relational database for structured data management, the project remains in the early phase, with current efforts focused on scanning and indexing physical documents. Initial plans for database implementations are also presented, outlining key considerations for structure, query indexing, and standardization. As digitization progresses, future discussions will refine these implantation details to ensure an efficient and comprehensive system. This initiative aims to preserve critical legacy data, enhance the design of tritium system facilities, and support the next generation of fusion energy research.

42 ENGINEERING↗

BULKI-Store v0.3.2

BULKI-Store is a distributed object storage system optimized for high-performance computing environments. Built with a Rust core and Python bindings, it efficiently manages scientific and machine learning datasets across HPC clusters. The system employs a client-server architecture with MPI integration, enabling seamless scaling on supercomputers like Perlmutter. BULKI-Store's object-oriented approach provides intuitive data organization with rich metadata support, contrasting with traditional file-based solutions. Key optimizations include selective checkpoint loading, unified checkpoint files, and object chunking for large data transfers. For machine learning workloads, BULKI-Store offers advantages through fine-grained access patterns, dynamic data sharing between training instances, and reduced memory pressure. Memory management features include strategic Python GC calls, minimized data copies, and batch processing capabilities. The system leverages Rayon's thread pool for asynchronous data prefetching and supports multiple CPU architectures (ARM64, x86, AMD, RISC-V). By combining performance optimizations with developer-friendly APIs, BULKI-Store addresses the complex data management challenges of modern HPC applications while maintaining compatibility across heterogeneous computing environments.

Zhang, Wei [Lawrence Berkeley National Laboratory ↗

Fermilab 2025 Summer Internship Skills Learned from Working with Mu2e Tracker Team

As part of a DOE RENEW Grant, the undergraduate authors interned at Fermilab. They spent nine weeks over summer 2025 working on the tracker team for the Mu2e experiment. They gained practical experience in a multitude of new skills including electronics installation, testing, and repairs. Their work particularly increased productivity by helping repair the high voltage and calibration pre-amplifiers, over 20,000 of which are needed in the tracker. These components are very fragile and often break during the installation process. The experience also taught them several soft skills such as the importance of thoughtful data storage, record- keeping, and problem-solving skills. They learned the demands of a large, international collaboration and how to work in a team. The authors would like to acknowledge their advisor and PI of the grant, Dr. Christopher G. Fasano, and the Mu2e team lead by co- spokesperson Dr. Bob Bernstein and tracker L2 manager Dr. Brendan Kiburg.

de Zwart, Brontë↗

Development of a Multi-Robot System for Autonomous Inspection of Nuclear Waste Tank Pits

This paper introduces the overall design plan, development timeline, and preliminary progress of the Autonomous Pit Exploration System project. This project aims to develop an advanced multi-robot system for the efficient inspection of nuclear waste-storage tank pits. The project is structured into three phases: Phase 1 involves data collection and interface definition in collaboration with Hanford Site experts and university partners, focusing on tank riser geometry and hardware solutions. Phase 2 includes the selection of sensors and robot components, detailed mechanical design, and prototyping. Phase 3 integrates all components into a cohesive system managed by a master control package which also incorporates digital twin and surrogate models, and culminates in comprehensive testing and validation at a simulated tank pit at the Idaho National Laboratory. Additionally, the system’s communication design ensures coordinated operation through shared data, power, and control signals. For transportation and deployment, an electric vehicle (EV) is chosen to support the system for a full 10 h shift with better regulatory compliance for field deployment. A telescopic arm design is selected for its simple configuration and superior reach capability and controllability. Preliminary testing utilizes an educational robot to demonstrate the feasibility of splitting computational tasks between edge and cloud computers. Successful simultaneous localization and mapping (SLAM) tasks validate our distributed computing approach. More design considerations are also discussed, including radiation hardness assurance, SLAM performance, software transferability, and digital twinning strategies.

Nuclear waste management↗

2025 TEM Workshop

The TEM Data Management Workshop will take place on August 26 from 9 a.m. to 12 p.m. MT, and will be held virtually on TEAMS. The primary goal of this workshop is to engage NSUF users and stakeholders in discussions about the data needs for the utilization of AI and ML in the analysis of TEM data. Key topics to be covered include data storage, data sharing, data tagging, metadata inclusion, standardized data formats, data augmentation, and annotated training datasets. Additionally, the workshop will provide valuable insights into resources such as the Nuclear Research Data System (NRDS) for data storage and sharing, as well as open-source codes for data analysis.

Bachhav, Mukesh↗

Techno-Economic Analysis for the Addition of a Thermal Energy Storage System to a Central Plant

Increasing energy demand and rising peak loads present significant challenges for energy management in commercial and institutional settings. As climate change drives greater cooling needs, central plants must navigate the complex tradeoffs between operational efficiency, cost control, and grid stability. Thermal energy storage (TES) systems offer a viable solution by shifting energy consumption from peak to off-peak periods, thereby reducing peak demand, lowering utility expenses, and improving grid resilience. However, the success of TES implementation hinges on appropriate system sizing, effective control strategies, and alignment with local utility rate structures. This article presents a techno-economic analysis of integrating a chilled water TES system into the central plant at California State University, Dominguez Hills. Drawing on historical load profiles and utility tariffs, we assess three TES sizing approaches and their corresponding control strategies from both energy and economic perspectives. This article utilizes a model-based approach to assess the impact of TES sizing and control strategies on the techno-economic feasibility of integrating TES into an existing central plant. The models employed for this analysis were calibrated using 4 years of historical data. Here, the results demonstrated that utility tariffs and the campus's operational profiles dictate the most feasible sizing and control methods. The findings offer valuable insights for institutions and commercial building managers exploring sustainable energy solutions. By demonstrating how optimized TES strategies can improve operational efficiency while achieving financial savings, this study highlights the potential for TES to align performance with cost effectiveness in real-world applications.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Packages of Distributed Energy Technologies Demonstrating Demand Flexibility at Community Scale

The combination of increased electric load growth across all sectors, deferred electrical infrastructure investment, and other factors resulting in variable electric power supply, has created technical challenges to maintaining a resilient and reliable grid. Many federal, regional, and local efforts are in play to modernize the electric grid, including advancing building technologies and distributed energy resources (DERs) that are utilizing smarter controls to become responsive to both occupant and grid needs. This report reviews ten pilot projects demonstrating how groups of buildings combined with behind-the-meter (BTM) DERs such as electric vehicle (EV) charging, battery storage, flexible HVAC and domestic hot water systems, and photovoltaic systems can reliably and cost effectively provide grid services. Each of the ten pilot projects aim to deliver both energy efficiency and demand flexibility (DF) while supporting load growth. The ten demonstration teams are piloting flexible DER packages across diverse communities of residential and commercial buildings to address a variety of regional grid needs. The outcomes of these pilot projects will be used to inform future scaling through utility program development. This paper characterizes the ten teams, showcasing the decision-making process used by each group to develop their packages (Section 2), the grid services they plan to deliver (Section 3), the types of DER packages selected for deployment within building sectors (Section 4) and trends between building sector, DER types, and grid services In order to achieve community scale benefits, the pilot projects must utilize aggregated control mechanisms for coordinating buildings and DERs together. Several types of coordinated control architectures have evolved amongst the teams, influenced by use type, existing market conditions, and integration type. Three coordinated controls architectures have been characterized, highlighting their use cases, benefits, challenges, and tradeoffs in their design. These insights can aid utilities, control vendors, and developers in scaling community-level energy systems (Paul, 2024). Ultimately, the technology packages selected by the ten teams will be coordinated to provide power system services, also known as grid services. Insights from these demonstrations will be useful for grid operators, regulators, aggregators and other stakeholders as they look to deploy demand flexible resources as grid services in the future. The grid services that each team is targeting for demonstration are described in Section 3 and Section 4. Methods for evaluating the grid services have been described in the paper Metrics for Evaluating Grid Service Provision from Communities of Grid-interactive and Efficient Buildings and other DER (MacDonald, 2023). To identify technology packages for demonstration, Section 2 shows that project teams used a range of analysis approaches, including building energy modeling, AMI data analysis, cost-benefit frameworks, and utility pilot data. Some teams emphasized technical modeling to quantify grid impacts and demand reduction potential, while others prioritized economic evaluations, stakeholder input, or exploratory pilots to inform deployment decisions. This diversity reflects the need to tailor selection methods to project goals, available data, and organizational context. Section 5 discusses trends between the DER technologies deployed and the grid service provisions from each team. Residential buildings (multifamily and single family) lean towards technologies that enhance energy efficiency (e.g. weatherization upgrades, smart thermostats) and onsite power generation integration (e.g. solar PV). Commercial building demonstrations prioritize technologies that ensure operational reliability (e.g. battery storage) and centralized energy management systems and optimization solutions. Teams that are deploying controllable storage-based technologies are more likely to provide grid services that require a near real-time response. Teams incorporating load shifting technologies like smart thermostats with HEMs are likely to include energy markets participation and customer bill management offerings. Campus demonstrations are adopting diverse sets of DERs to emphasize renewable generation, paired with centralized control. This section also describes technologies that were considered during project planning but ultimately excluded from final deployment. These demonstrations reveal that effective DER package design should be tailored to building type, customer segment, and construction vintage. Multifamily buildings benefit from centralized HVAC upgrades and supervisory controls, while single-family homes are well-suited for individualized technologies like solar, storage, and smart home energy monitors. Commercial and campus settings prioritize EMIS integration and load optimization. New construction enables cost-effective integration of DER-ready infrastructure, whereas retrofits require deployments aligned with owner and tenant value streams. For utility program planners, early coordination with developers and building owners, paired with segmented and modular program offerings, can improve adoption, scalability, and grid impact.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Carbon Management Projects (CONNECT) Database and Explorer

Overview The Carbon Management Projects (CONNECT) Toolkit is an online exploratory visualization tool developed by the U.S. Department of Energy's (DOE) Office of Fossil Energy and Carbon Management (FECM) with support from other federal agencies such as the U.S. Environmental Protection Agency (EPA) and the U.S. Department of Transportation (DOT). It provides a single point of access to authoritative information on federal agency investment in a portfolio of research, development, and demonstration (RD&D) projects that have been publicly announced to advance technologies for point source carbon capture, carbon dioxide removal, transport, storage, and conversion, collectively referred to as carbon management. The RD&D programs covered in this tool are authorized by annual congressional appropriations ("Base Program") and the 2021 Infrastructure Investment and Jobs Act (IIJA). The tool also incorporates public information on other federal initiatives, such as the Regional Clean Hydrogen Hubs, and public information released by other government agencies, such as the Environmental Protection Agency's (EPA) and Primacy States’ Underground Injection Control Class VI permits and EPA’s facility level greenhouse gas (GHG) emissions. Developed in a geographic information system, the tool organizes carbon management projects into five groups based on the primary technology that a project aims to advance, each visually represented as a digital layer ("carbon management project layer"). Only federally funded projects are included, which can be awarded projects that are completed or ongoing, or projects that have been selected but are currently under negotiation. Project information can be viewed in the map or in the attribute table below it when turned on. In the map view, each project is displayed at either its host site (for field work), where available, or its performer site (project lead's location, further explained in the table below). Host sites and performer sites are represented in distinct icons. Several reference layers offer additional public information on infrastructural and natural resource environment for carbon management. These reference layers, combined with multiple geographical basemaps, enable users to visualize the carbon management project layers in context. Carbon management project information will be updated monthly based on feedback and information availability. Carbon management project layers Point Source Carbon Capture (PSC) This layer contains DOE-funded projects focused on capturing carbon dioxide (CO2) from power plants or industrial facilities. Carbon Dioxide Removal (CDR) This layer contains DOE-funded projects focused on capturing CO2 from the atmosphere, including direct air capture (DAC) and DAC hubs, direct ocean capture, enhanced mineralization, and biomass carbon removal and storage. For projects with multiple host sites, each of the sites are displayed individually with the project cost and cost sharing information representing the total for the entire project. Carbon Transport This layer contains DOE- and DOT-funded projects focused on CO2 transport. The Transport Research and Development sublayer contains projects that do not involve physical infrastructure; the Proposed Transport Corridor sublayer contains projects for which either a route for the transport infrastructure has been proposed or a general area for the transport infrastructure has been identified. Carbon Storage This layer contains DOE-funded key projects focused on CO2 storage. For projects with multiple field-work sites, each of the sites are displayed individually on the map with the project cost and cost sharing information representing the overall total for the entire project. Carbon Conversion This layer contains DOE-funded projects focused on converting CO2 into economically valuable products. Reference layers The following layers provide additional information in the geographic proximity of carbon management projects. Users should reference the original sources for more details (weblinks provided below and in pop-up windows on the map). Regional Clean Hydrogen Hub and Facility These layers illustrate the approximate areas of the Regional Clean Hydrogen Hubs announced by DOE's Office of Clean Energy Demonstrations (OCED) and the approximate locations of individual facilities that constitute the hubs (see "Where are the H2Hubs located?" on the webpage linked above). EPA Facility Level GHG Emissions (direct emitter) This layer shows direct CO2 emissions from stationary sources in 2022, using data extracted from EPA's Facility Level Information on GreenHouse gases Tool (FLIGHT). Captured and injected CO2 are not deducted from direct emitters’ total emissions. Contact EPA for additional details. Underground Injection Control Class VI permit/permit application This layer shows the locations of CO2 injection wells that are granted or in the process of applying for an Underground Injection Control Class VI permit by EPA or a Primacy State (currently Louisiana, North Dakota, and Wyoming). The URLs for the permits or permit applications are provided in the pop-up windows associated with the well locations. Contact EPA for additional details. Carbon Storage Resource This layer contains information on prospective CO2 storage resources in saline formations and oil and gas reservoirs provided by the National Carbon Sequestration Database and Geographic Information System (NATCARB) spatial database. Contact NETL for additional details. Existing CO2 pipeline This layer shows active CO2 pipelines based on information digitized from the map issued by the Pipeline and Hazardous Materials Safety Administration (PHMSA). Contact PHMSA for additional details.

Carbon Conversion↗

DEDUPKV: A Space-Efficient and High-Performance Key-Value Store via Fine-Grained Deduplication

Log-Structured Merge Tree (LSM-tree) based key-value stores excel in write-intensive environments but suffer from data duplication, consuming up to 49% of storage space in LSM-tree-based key-value store deployments. Traditional solutions like compression and coarse-grained file system-level deduplication introduce overhead or have limited effectiveness. In this study, we propose DedupKV, a fine-grained deduplication framework tailored for LSM-tree, maximizing data reduction efficiency while minimizing write stalls and read overheads. DedupKV features three key innovations: (1) FLUSH-integrated inline deduplication, which removes duplicates during memory-to-storage writes; (2) WAL file-based offline deduplication, repurposing write-ahead logs to avoid double writes; and (3) elastic execution, dynamically balancing inline and offline deduplication based on memory pressure and workload intensity. Additionally, dynamic granularity management reduces deduplication metadata overhead. We implemented these four ideas in RocksDB for the first time and conducted experiments in a Linux environment. Our evaluation shows that WAL file-based offline deduplication and DedupKV outperform BlobDB by 33% and 23%, respectively, in write-heavy workloads, while reducing write amplification by 1.2 ×, 2 ×, and 1.6 × for real KV datasets.

Jamil, Safdar [Sogang University]↗