Search NASA⌕ Search

SEARCH · Search NASA

Results for “false positive analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

31 records · Page 2

eDNAjoint: An R package for interpreting paired or semi‐paired environmental DNA and traditional survey data in a Bayesian framework

Abstract Environmental DNA (eDNA) sampling is increasingly used in surveys of species distribution as a potentially sensitive and efficient monitoring method. Yet access to modelling tools designed specifically for interpreting this new data type lags behind its ubiquity. While occupancy modelling software has dominated the analytical landscape for eDNA data analysis of single species, this type of model may not always be the most appropriate. The rate of eDNA detection often corresponds to species density, rather than just occupancy, and researchers often have access to observations from non‐genetic sampling methods at the same sites. To provide users access to a modelling framework designed to maximize the use of all available data, we developed an R package, eDNAjoint . The package provides an easy‐to‐use interface for fitting a ‘joint’ model that integrates data from paired or semi‐paired eDNA and traditional surveys in a Bayesian framework. The model can be used to estimate parameters like the probability of a false positive eDNA detection and mean catch rate at a site, and the package allows access to multiple model variations and Bayesian prior customization. Additional functionality can be used for model selection, summarising posteriors and comparing the relative sensitivities of the two survey methods. We demonstrate the use of eDNAjoint by fitting a variation of the model with site‐level covariates that scale the sensitivity of eDNA sampling relative to traditional sampling. The example workflow uses binary eDNA and seine count data for the endangered tidewater goby ( Eucyclogobius newberryi ) from a study by Schmelzle and Kinziger (2016). This use case includes a prior sensitivity analysis and an evaluation of the relationship between detection rates and environmental variables. eDNAjoint has the potential to greatly increase the range of users who will be able to rigorously analyse eDNA and traditional survey data in a Bayesian framework, understand if and how eDNA can improve monitoring practices, and gain confidence in the interpretability of eDNA data.

Keller, Abigail G. [Department of Environment Scie↗

Description of Pegethrix niliensis sp. nov., a Novel Cyanobacterium from the Nile River Basin, Egypt: A Polyphasic Analysis and Comparative Study of Related Genera in the Oculatellales Order

In this paper, we examine the filamentous cyanobacterial strain NILCB16 and describe it as a new species within the genus Pegethrix. The original population was sampled from a mat growing in an irrigation canal in the Nile River, Egypt. Initially classified under Plectonema or Planktolyngbya, the strain is a potential producer of the toxins microcystin and β-N-Methylamino-L-Alanine (BMAA). Additionally, we reviewed the taxonomic relationships between the Oculatellales genera. To describe the new species, we conducted a polyphasic study, encompassing 16S rRNA gene phylogenetic analyses performed using both Maximum Likelihood and Bayesian methods, sequence identity (p-distance) analysis, 16S-23S ITS secondary structures, and morphological and habitat comparisons. The phylogenetic analysis revealed that strain NILCB16 clustered within the Pegethrix clade with strong phylogenetic support, but in a distinct position from other species in the genus. The strain shared a maximum 16S rRNA gene identity of 97.3% with P. qiandaoensis and 96.1% with the type species, P. bostrychoides. Morphologically, NILCB16 can be differentiated from other species in the genus by its lack of false branching. Our phylogenetic analyses also show that Pegethrix, Cartusia, Elainella, and Maricoleus are clustered with strong phylogenetic support. They exhibit high 16S rRNA gene identity and are morphologically indistinguishable, suggesting they could potentially be merged into a single genus in the future.

Hentschke, Guilherme Scotta (ORCID:000000034396024↗

Identifying preferential flow from soil moisture time series: Review of methodologies

Abstract Identifying and quantifying preferential flow (PF) through soil—the rapid movement of water through spatially distinct pathways in the subsurface—is vital to understanding how the hydrologic cycle responds to climate, land cover, and anthropogenic changes. In recent decades, methods have been developed that use measured soil moisture time series to identify PF. Because they allow for continuous monitoring and are relatively easy to implement, these methods have become an important tool for recognizing when, where, and under what conditions PF occurs. The methods seek to identify a pattern or quantification that indicates the occurrence of PF. Most commonly, the chosen signature is either (1) a nonsequential response to infiltrated water, in which soil moisture responses do not occur in order of shallowest to deepest, or (2) a velocity criterion, in which newly infiltrated water is detected at depth earlier than is possible by nonpreferential flow processes. Alternative signatures have also been developed that have certain advantages but are less commonly utilized. Choosing among these possible signatures requires attention to their pertinent characteristics, including susceptibility to errors, possible bias toward false negatives or false positives, reliance on subjective judgments, and possible requirements for additional types of data. We review 77 studies that have applied such methods to highlight important information for readers who want to identify PF from soil moisture data and to inform those who aim to develop new methods or improve existing ones. Core Ideas Soil moisture data can be used to identify the occurrence of preferential flow (PF) and its initiating conditions. Various data‐analysis methods to identify PF differ in susceptibility to error, bias, and subjectivity. These methods can utilize vast amounts of data from soil moisture monitoring networks to develop understanding of when, where, and under what conditions PF occurs. Newly developed methods may lead to better accuracy and reliability, and reduce the need for subjective judgments. Plain Language Summary Preferential flow through soil occurs when a large amount of water is suddenly available, as during an intense storm. This type of flow moves rapidly through the soil in distinct narrow pathways rather than moving evenly throughout the body of soil, with major consequences for groundwater resources, ecosystems, spreading of contaminants, and other vital concerns. Methods of detecting preferential flow have been developed that utilize measurements of soil water content made by sensors installed at various depths. This measurement technology has been widely implemented, many locations now having datasets years in length, and various methods have been developed for using these to identify preferential flow. The various methods are based on different features in the soil moisture records and vary in their advantages and shortcomings. In this review, we explain and evaluate these methods, highlighting important information for their implementation to identify preferential flow from soil moisture data and for efforts to develop new methods or improve existing ones.

Nimmo, John R↗

Pushing the Limits of Pulse Shape Discrimination in a Large Liquid Xenon Detector

The LUX-ZEPLIN (LZ) experiment is a direct-detection dark matter experiment, optimized to search for weakly interacting massive particles (WIMPs) through WIMP-nucleon interactions. The main challenge in dark matter detection is differentiating between WIMP signals and background events. In LZ, the ratio of ionization to scintillation signals (charge-to-light) is the primary method for rejecting electronic recoil (ER) background. Pulse shape discrimination (PSD) offers a method for additional ER backgrounds rejection in liquid xenon detectors. In this paper, the discrimination power of PSD with the LZ experiment is discussed. To precisely characterize the scintillation pulse shape, an analysis framework is developed to reconstruct the detection time of individual photons. Using LZ calibration data, the photon-timing prompt fraction discriminator is optimized and achieves ER leakage as low as $15\%$. For specific background processes such as $^{124}$Xe double electron capture, the leakage is reduced further to about $5\%$. PSD is combined with charge-to-light to form two-factor discrimination (TFD). The optimized TFD performance is compared with the performance of the charge-to-light method, with the corresponding false positive rate reduced by up to a factor of two for large scintillation pulses. Finally, PSD and TFD are applied to data from LZ's WS2024 run and their performance is summarized.

Akerib, D. S. [SLAC; KIPAC, Menlo Park]↗

Pushing the limits of pulse shape discrimination in a large liquid xenon detector

Abstract The LUX-ZEPLIN (LZ) experiment is a direct-detection dark matter experiment, optimized to search for weakly interacting massive particles (WIMPs) through WIMP-nucleon interactions. The main challenge in dark matter detection is differentiating between WIMP signals and background events. In LZ, the ratio of ionization to scintillation signals (charge-to-light) is the primary method for rejecting electronic recoil (ER) background. Pulse shape discrimination (PSD) offers a method for additional ER backgrounds rejection in liquid xenon detectors. In this paper, the discrimination power of PSD with the LZ experiment is discussed. To precisely characterize the scintillation pulse shape, an analysis framework is developed to reconstruct the detection time of individual photons. Using LZ calibration data, the photon-timing tail fraction discriminator is optimized and achieves ER leakage as low as $$15\%$$ 15 % . For specific background processes such as $$^{124}$$ 124 Xe double electron capture, the leakage is reduced further to about $$5\%$$ 5 % . PSD is combined with charge-to-light to form two-factor discrimination (TFD). The optimized TFD performance is compared with the performance of the charge-to-light method, with the corresponding false positive rate reduced by up to a factor of two for large scintillation pulses. Finally, PSD and TFD are applied to data from LZ’s WS2024 run and their performance is summarized.

Akerib, D. S. [SLAC National Accelerator Laborator↗

Inferring precocial Chinook Salmon production through single‐parentage assignments

Abstract Objective Parentage analysis is a routine methodology in fisheries research, but study systems exist where it is impractical to sample both parents. The ability to reliably assign offspring to a single parent is beneficial in these situations. We applied single‐parentage assignments to a naturally spawning population of Chinook Salmon Oncorhynchus tshawytscha to quantify production of anadromous returns by unsampled precocial males. Methods We used an approach that focused on two important aspects of parentage analyses: (1) addressing the presence of family structure within the set of sampled parents and (2) controlling for false‐positive and false‐negative assignments. Result Results indicated that 30% of reproductively successful males were precocial males, which produced 20% of the returning anadromous offspring. Conclusion This study provides a framework for applying single‐parent assignments in a salmonid study system while explicitly addressing sources of assignment errors.

Steele, Craig A.↗

Characterisation and comparative analysis of mitochondrial genomes of false, yellow, black and blushing morels provide insights on their structure and evolution

Morchella species have considerable significance in terrestrial ecosystems, exhibiting a range of ecological lifestyles along the saprotrophism-to-symbiosis continuum. However, the mitochondrial genomes of these ascomycetous fungi have not been thoroughly studied, thereby impeding a comprehensive understanding of their genetic makeup and ecological role. In this study, we analysed the mitogenomes of 30 Morchellaceae species, including yellow, black, blushing and false morels. These mitogenomes are either circular or linear DNA molecules with lengths ranging from 217 to 565 kbp and GC content ranging from 38% to 48%. Fifteen core protein-coding genes, 28–37 tRNA genes and 3–8 rRNA genes were identified in these Morchellaceae mitogenomes. The gene order demonstrated a high level of conservation, with the cox1 gene consistently positioned adjacent to the rnS gene and cob gene flanked by apt genes. Some exceptions were observed, such as the rearrangement of atp6 and rps3 in Morchella importuna and the reversed order of atp6 and atp8 in certain morel mitogenomes. However, the arrangement of the tRNA genes remains conserved. We additionally investigated the distribution and phylogeny of homing endonuclease genes (HEGs) of the LAGLIDADG (LAGs) and GIY-YIG (GIYs) families. A total of 925 LAG and GIY sequences were detected, with individual species containing 19–48HEGs. These HEGs were primarily located in the cox1, cob, cox2 and nad5 introns and their presence and distribution displayed significant diversity amongst morel species. These elements significantly contribute to shaping their mitogenome diversity. Overall, this study provides novel insights into the phylogeny and evolution of the Morchellaceae.

59 BASIC BIOLOGICAL SCIENCES↗

Archetype-based Redshift Estimation for the Dark Energy Spectroscopic Instrument Survey

We present a computationally efficient galaxy archetype-based redshift estimation and spectral classification method for the Dark Energy Survey Instrument (DESI) survey. The DESI survey currently relies on a redshift fitter and spectral classifier using a linear combination of principal component analysis–derived templates, which is very efficient in processing large volumes of DESI spectra within a short time frame. However, this method occasionally yields unphysical model fits for galaxies and fails to adequately absorb calibration errors that may still be occasionally visible in the reduced spectra. Our proposed approach improves upon this existing method by refitting the spectra with carefully generated physical galaxy archetypes combined with additional terms designed to absorb data reduction defects and provide more physical models to the DESI spectra. We test our method on an extensive data set derived from the survey validation (SV) and Year 1 (Y1) data of DESI. Our findings indicate that the new method delivers marginally better redshift success for SV tiles while reducing catastrophic redshift failure by 10%–30%. At the same time, results from millions of targets from the main survey show that our model has relatively higher redshift success and purity rates (0.5%–0.8% higher) for galaxy targets while having similar success for QSOs. These improvements also demonstrate that the main DESI redshift pipeline is generally robust. Additionally, it reduces the false-positive redshift estimation by 5%–40% for sky fibers. We also discuss the generic nature of our method and how it can be extended to other large spectroscopic surveys, along with possible future improvements.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Predictive analytics of selections of russet potatoes

We explore the application of machine learning algorithms specifically to enhance the selection process of Russet potato (Solanum tuberosum L.) clones in breeding trials by predicting their suitability for advancement. This study addresses the challenge of efficiently identifying high-yield, disease-resistant, and climate-resilient potato varieties that meet processing industry standards. Leveraging manually collected data from trials in the state of Oregon, we investigate the potential of a wide variety of state-of-the-art binary classification models. The dataset includes 1086 clones, with data on 38 attributes recorded for each clone, focusing on yield, size, appearance, and frying characteristics, with several control varieties planted consistently across four Oregon regions from 2013 to 2021. We conduct a comprehensive analysis of the dataset that includes preprocessing, feature engineering, and imputation to address missing values. We focus on several key metrics such as accuracy, F1-score, and Matthews correlation coefficient (MCC) for model evaluation. The top-performing models, namely a feedforward neural network classifier (Neural Net), a histogram-based gradient boosting classifier (HGBC), and a support vector machine classifier (SVM), demonstrate consistent and significant results. To further validate our findings, we conducted a simulation study using the aims, data-generating mechanisms, estimands, methods, and performance measures (ADEMP) framework, simulating different data-generating scenarios to assess model robustness and performance through true positive, true negative, false positive, and false negative distributions, area under the receiver operating characteristic curve (AUC-ROC) and MCC. The simulation results highlight that non-linear models like SVM and HGBC consistently show higher AUC-ROC and MCC than logistic regression, thus outperforming the traditional linear model across various distributions, and emphasizing the importance of model selection and tuning in agricultural trials. Variable selection further enhances model performance and identifies influential features in predicting trial outcomes. The findings emphasize the potential of machine learning in streamlining the selection process for potato varieties, offering benefits such as increased efficiency, substantial cost savings, and judicious resource utilization. Our study contributes insights into precision agriculture and showcases the relevance of advanced technologies for informed decision-making in breeding programs.

60 APPLIED LIFE SCIENCES↗

CAHS: Context-Aware Homology Search

Protein homology search is foundational to bioinformatics: it supports annotation transfer, structure/function inference, and evolutionary analysis over rapidly expanding sequence repositories (e.g., UniProtKB). Profile hidden Markov models (pHMMs), as implemented in HMMER, remain the most widely trusted approach because they provide statistically calibrated E-values; however, their gap behavior is fixed once a profile is trained, despite biological evidence that insertion/deletion tolerance varies across flexible loops and intrinsically disordered regions. We present CAHS (Context-Aware Homology Search), a lightweight query-time adapter for pHMM search that incorporates learned and biologically motivated signals without changing HMMER's downstream search pipeline or its calibrated E-value reporting. Given a query sequence, CAHS computes per-residue representations from a protein language model and a disorder predictor, maps these to profile coordinates, and modulates only match-state transition rows (gap-open and gap-extension probabilities) while preserving Plan7 constraints. We comprehensively evaluate CAHS across six structurally diverse protein families and multi-domain architectures against a 570k-sequence target corpus. CAHS expands detection capability, retrieving thousands of additional remote homologs at relaxed thresholds by maintaining alignment quality through flexible regions. For multi-domain proteins, context-aware modulation resolves 94% of fragmented alignments. Crucially, CAHS preserves hit-set invariance at stringent operating points (E<10-10), demonstrating increased statistical confidence without inflating false positives. Furthermore, sharper statistical distinction between homologs and background noise during early filter stages yields up to a 3.87× acceleration in end-to-end wall-clock time on high-performance computing clusters. Overall, CAHS illustrates a practical AI-for-science design pattern: augmenting a trusted probabilistic model with query-specific learned signals to improve interpretable, reproducible inference in data-rich biology.

Bhattaram, Swethasree [Georgia Institute of Techno↗

Periodicity significance testing with null-signal templates: reassessment of PTF’s SMBH binary candidates

Periodograms are widely employed for identifying periodicity in time series data, yet they often struggle to accurately quantify the statistical significance of detected periodic signals when the data complexity precludes reliable simulations. We develop a data-driven approach to address this challenge by introducing a null-signal template (NST). The NST is created by carefully randomizing the period of each cycle in the periodogram template, rendering it non-periodic. It has the same frequentist properties as a periodic signal template, and we show with simulations that the distribution of false positives is the same as with the original periodic template, regardless of the underlying data. Thus, performing a periodicity search with the NST acts as an effective simulation of the null (no-signal) hypothesis, without having to simulate the noise properties of the data. We apply the NST method to the supermassive black hole binaries (SMBHB) search in the Palomar Transient Factory (PTF), where Charisi et al. had previously proposed 33 high signal-to-noise candidates utilizing simulations to quantify their significance. Our approach reveals that these simulations do not capture the complexity of the real data. There are no statistically significant periodic signal detections above the non-periodic background. To improve the search sensitivity, we introduce a Gaussian quadrature based algorithm for the Bayes Factor with correlated noise as a test statistic. We show with simulations that this improves sensitivity to true signals by more than an order of magnitude. However, the Bayes Factor approach also results in no statistically significant detections in the PTF data.

79 ASTRONOMY AND ASTROPHYSICS↗

Gamma-Ray and Cosmic Ray Muon Modalities for Cargo Inspection

Screening and inspection of cargo containers are two essential methods to nondestructively examine the contents of shipment. These methods enable the detection of illicit transportation of unauthorized materials such as nuclear and radioactive materials, explosives, drugs, and so on, typically at borders or secure facilities. Although high-energy X-ray transmission is a standard system and is widely used for cargo inspection, the inherent challenges of high false-positive rates and high attenuation factors necessitate the development of complementary techniques that can increase the detection efficiency and accuracy in large and dense materials. Gamma-rays, which possess higher penetration characteristics because of their high energy, offer an alternative nonintrusive modality for cargo scanning. They represent a promising inspection method when compared to X-rays for three reasons: (1) improved ability to detect nuclear and radioactive materials, (2) higher inspection throughput rates, and (3) lower false-positive rates. Currently, there are two main gamma-ray inspection techniques, active and passive interrogation. Active interrogation can be further grouped into (1) gamma-ray transmission imaging and (2) neutron-induced gamma-ray emission detection. Gamma-ray transmission imaging utilizes differences in material densities for mapping the shipment contents and detecting anomalies. It is analogous to the X-ray transmission method; however, the high-energy photons make it more difficult to shield against, which enables more efficient performance in large and dense material inspection. Neutron-induced gamma-ray emission inspection is designed for the detection of nuclear and radioactive material because those materials emit characteristic gamma-rays when they are activated by neutron absorption. On the other hand, passive interrogation techniques rely on high-efficiency detectors to detect radiation emitted from hidden special nuclear or other radioactive materials. Similar to passive interrogation, cosmic ray muon monitoring and imaging are relatively new techniques that do not require external radioactive sources. These techniques have received attention as a potential next-generation radiographic probe to identify illicit transportation of nuclear and radioactive materials in cargo containers. Cosmic ray muons have unique features, (1) much higher energies than X-rays or gamma-rays (on the order of 10−1—104 GeV), (2) enhanced penetration capability, and (3) natural occurrence, thereby eliminating the need for induced radiation sources. These features enable cosmic ray muons to be utilized for detection of special nuclear materials in high-background-noise environments. By analyzing incoming and outgoing muon trajectories, scattering angles, and energies, it has been shown that it would be possible to locate hidden and well-shielded materials in cargo containers via three-dimensional muon tomography images or signal analysis. Gamma-rays, cosmic ray muons, and other nonintrusive cargo inspection modalities are complementary to each other, allowing them to address various cargo inspection conditions (i.e., scanning time, cost, radiation exposure level, and types of target materials). This chapter presents a detailed review of the theoretical fundamentals and technical principles behind the current gamma-ray and cosmic ray muon modalities for cargo inspection. Additionally, critical assessments and suggestions for the future directions to advance the use of gamma and muon modalities are discussed.

Bae, Junghyun↗

CERF: IM3 Projected Western US Power Plant Locations

Overview The Capacity Expansion Regional Feasibility (CERF) model is an open-source geospatial python package that provides new power plant locations at a 1km resolution. The model ingests U.S. state or regional-scale electricity system capacity expansion plans, such as those produced by the Global Change Analysis Model (GCAM-USA), and identifies feasible, site-specific locations for individual new power plants (renewable and non-renewable). CERF combines high-resolution geospatial suitability analyses with an economic algorithm that selects individual plant siting locations based on grid interconnection costs and the locational marginal value of new generation. The model incorporates a wide range of dynamic constraints and opportunities, such as protected lands, population density, existing infrastructure, and water availability. This dataset provides CERF power plant siting results for IM3 Phase 2 simulations across eight different scenarios for the Western US through 2055. The scenarios include combinations of two Shared Socioeconomic Pathways (SSP3 and SSP5) with four high-resolution climate projections specific to the United States (see, https://tgw-data.msdlive.org/). These climate projections include "hotter" and "cooler" variants for two Representative Concentration Pathways (RCP4.5 and RCP8.5). The resulting eight simulations are: rcp45cooler_ssp3 rcp45cooler_ssp5 rcp45hotter_ssp3 rcp45hotter_ssp5 rcp85cooler_ssp3 rcp85cooler_ssp5 rcp85hotter_ssp3 rcp85hotter_ssp5 CERF siting results in this dataset correspond to capacity expansion plans in the GCAM-USA IM3 Phase 2 simulation data and are available for each of the above scenarios. Data Details Temporal Range: 2015-2055 in 5-year timesteps. Note that 2015 is the experiment base year and 2020 and beyond represent model simulation years. Spatial Range: Plant locations are provided for the eleven states in the Western US including Arizona, California, Colorado, Idaho, Montana, New Mexico, Nevada, Oregon, Utah, Washington, and Wyoming. Spatial Resolution: 1 km-squared, provided in x and y coordinates Geospatial Projection: Albers Equal Area Conic (ESRI:102003) File Type: csv The dataset contains subdirectories for each of the eight scenarios described in the overview. Each scenario folder contains two subfolders with the following information: 1. Power Plant Data This directory contains a single .csv file of power plant locations for both pre-existing (non-CERF sited plants in operation in 2015) and new (CERF-sited) power plants across the temporal range along with additional CERF model output parameters for CERF-sited plants. Plant with a siting year earlier than 2020 correspond to facilities that are operational leading into the first timestep CERF simulation. For a more detailed description of CERF model output parameters, see the CERF model documentation. Note that the cerf_plant_id parameter is unique within each scenario file but not across scenario files. Parameter Descriptions scenario - Name of scenario cerf_plant_id - Unique siting identifier cerf_sited - If True, indicates that plant was sited by CERF model. If False, indicates pre-existing facility region_name - Name of region (state) tech_id - Technology ID tech_name - Full generation technology name inclusive of cooling type (if applicable) and additional characteristics tech_simple - Simplified generation technology type unit_size_mw - Power plant unit size (MW) xcoord - X coordinate in the default CRS (meters) ycoord - Y coordinate in the default CRS (meters) index - Index position in the flattend 2D array buffer_in_km - Exclusion buffer around site (km) sited_year - Year of siting retirement_year - Year of retirement lmp_zone - Locational marginal price (LMP) zone ID locational_marginal_price_usd_per_mwh - Locational marginal price ($/MWh) generation_mwh_per_year - Generation output (MWh/yr) operating_cost_usd_per_year - Cost of plant operations ($/yr) net_operational_value - Net operational value based on LMP and and operating costs ($/yr) interconnection_cost - Cost of interconnection for transmission & gas pipeline (if applicable) net_locational_cost -- Difference of interconnection cost and operating value ($/yr) capacity_factor_fraction - Capacity factor (fraction) carbon_capture_rate_fraction - Carbon capture rate (fraction) fuel_co2_content_tons_per_btu - Fuel CO2 content (tons/Btu) fuel_price_usd_per_mmbtu - Fuel price ($/MMBtu) fuel_price_esc_rate_fraction - Fuel price escalation rate (fraction) heat_rate_btu_per_kWh - Heat rate (Btu/kWh) lifetime_yrs - Technology lifetime for annuity (years) operational_life_yrs - Operational lifetime for retirement (years) variable_om_usd_per_mwh - Variable operation and maintenance costs of yearly capacity use ($/MWh) variable_om_esc_rate_fraction - Variable operation and maintenance costs escalation rate (fraction) carbon_tax_usd_per_ton - Carbon tax ($/ton) carbon_tax_esc_rate_fraction - Carbon tax escalation rate (fraction) 2. Storage Data This directory contains information on new and pre-existing energy storage facilities operational in each timestep along with various storage operational parameters. The 2015 timestep provides pre-existing energy storage data and corresponds with facilities that are operational leading into the first model simulation timestep. Note that coordinates in the storage files correspond to the interconnection point on the grid (substation location), not individual energy storage locations. Energy storage is added in a cumulative process at each given interconnection point. That is, each individual file provides the total operational storage capacity interconnected to the specified substation for the given timestep, inclusive of previously installed storage at that location and new storage installed in that timestep at that location. Parameters scenario - Name of scenario timestep - Simulation timestep name - Unique storage identifier s_typ - Type of energy storage technology (battery or pumped storage hydro) s_node - Node ID of interconnecting substation xcoord - X coordinate in the default CRS (meters) ycoord - Y coordinate in the default CRS (meters) charge_rate - Maximum charge rate (power capacity) of storage system (MW) discharge_rate - Maximum discharge rate (power capacity) of storage system (MW) duration - Duration of storage system (hours) max_SoC - Allowed maximum state of charge (energy capacity) of storage system (MWh) min_SoC -Allowed minimum state of charge (energy capacity) of storage system (MWh) charge_eff - Efficiency of charge (fraction between 0 and 1) discharge_eff - Efficiency of discharge (fraction between 0 and 1) Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program.

CERF↗