Search NASA⌕ Search

SEARCH · Search NASA

Results for “toolkit”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 343 records · Page 19

RADAI: A Large-Scale Realistic Dataset for Radiation Detection Algorithm Development

Open, realistic datasets are essential for developing and benchmarking radiation detection algorithms, yet they remain scarce. The Radiological Anomaly Detection and Identification (RADAI) project was develop to create datasets that meet the training and testing needs for sophisticated radiation detection algorithms. The RADAI dataset is a large-scale synthetic resource that integrates high-fidelity Monte Carlo simulations with realistic urban scenarios to capture both background variability and source signatures. RADAI models construction-material NORM, people and vehicles, urban clutter, and dynamic environmental effects such as cosmic-ray and rain-induced transients, and they provide list-mode detector data with motion and response modeling suitable for algorithm training and evaluation. The RADAI project resulted in three publicly-released complementary datasets together with an online scoring portal for standardized performance assessment and an open software toolkit that supports data access, augmentation, model development, and evaluation. These resources enable reproducible comparisons across methods and promote rigorous studies at the scale required by contemporary machine learning. By grounding algorithm development in realistic, well-documented conditions, RADAI supports progress toward more robust detection, identification, and localization in complex urban environments.

Ghawaly, James M. [Division of Computer Science an↗

GASP: Gradient-Aware Shortest Path Algorithm for Boundary-Confined 2-Manifold Reeb Graph Visualization

Reeb graphs are an important tool for abstracting and representing the topological structure of a function defined on a manifold. We have identified three properties for faithfully representing Reeb graphs in a visualization: they should be constrained to the boundary, compact, and aligned with the function gradient. Existing algorithms for drawing Reeb graphs are agnostic to or violate these properties. In this paper, we introduce an algorithm to generate Reeb graph visualizations, called GASP, that is cognizant of these properties, thereby producing visualizations that are more representative of the underlying data. To demonstrate the improvements, the resulting Reeb graphs are evaluated both qualitatively and quantitatively against the geometric barycenter algorithm, using its implementation available in the Topology ToolKit (TTK), a widely adopted tool for calculating and visualizing Reeb graphs.

Rahman, Sefat [University of Utah]↗

Testing the Activation Analysis for Fusion in OpenMC

OpenMC is a community-developed Monte Carlo neutron and photon transport simulation code. It can perform fission simulations such as fixed-source, k-eigenvalue, and subcritical multiplication calculations on models built using either a constructive solid geometry or CAD representation. To explore the use of OpenMC for fusion activation analysis, a detailed model of the Fusion Neutronics Science Facility (FNSF) was first developed for comparisons against an existing SERPENT model. A 90-degree model of FNSF in Standard-Triangle-Language (STL) CAD format was converted to Constructive Solid Geometry (CSG) using each code's built-in functions, and the geometries were validated by ensuring no cells overlapped and no particles were lost during simulations. The neutron fluxes were calculated and compared for multiple components close to the plasma. The results show differences mostly below 1% in fluxes and averaged 8% for activity and decay heat. Here, the work described in this study tests the CAD-based geometry using the DagMC toolkit in OpenMC and compares the activation analysis of OpenMC to SERPENT code.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Simulation of Divertor Performance in ST40 Under Dynamic Double-Null Plasmas

A power fraction model was implemented for the simultaneous prediction of 3-D surface temperature evolution at all four divertor targets in near-double-null (DN) tokamak configurations, which is especially important for compact high-field devices that may not have the ability to dissipate large amounts of power on the high-field side. Evaluating the power-sharing between the four divertor strike points in a disconnected DN configuration is important for understanding the overall power balance, as well as for optimizing the power exhaust performance and prolonging the survivability of the plasma facing components (PFCs). This power-sharing is typically evaluated in terms of the separation between the primary and secondary separatrices at the outboard midplane, $\textit {dR}_{\text {sep}}$. The Heat flux Engineering Analysis Toolkit (HEAT) is coupled with Brunner’s power fraction model to simulate the deposited heat flux and resultant temperature change on 3-D divertor targets in a dynamic DN (DDN) pulse operation in ST40, a high-field spherical tokamak. The simulation results showed that with DDN operation, the operation time has significantly increased compared with single-null geometry configurations.

ST40↗

Turbo‐charging crop improvement: harnessing multiplex editing for polygenic trait engineering and beyond

Multiplex CRISPR editing has emerged as a transformative platform for plant genome engineering, enabling the simultaneous targeting of multiple genes, regulatory elements, or chromosomal regions. This approach is effective for dissecting gene family functions, addressing genetic redundancy, engineering polygenic traits, and accelerating trait stacking and de novo domestication. Its applications now extend beyond standard gene knockouts to include epigenetic and transcriptional regulation, chromosomal engineering, and transgene‐free editing. These capabilities are advancing crop improvement not only in annual species but also in more complex systems such as polyploids, undomesticated wild relatives, and species with long generation times. At the same time, multiplex editing presents technical challenges, including complex construct design and the need for robust, scalable mutation detection. We discuss current toolkits and recent innovations in vector architecture, such as promoter and scaffold engineering, that streamline workflows and enhance editing efficiency. High‐throughput sequencing technologies, including long‐read platforms, are improving the resolution of complex editing outcomes such as structural rearrangements—often missed by standard genotyping—when targeting repetitive or tandemly spaced loci. To fully realize the potential of multiplex genome engineering, there is growing demand for user‐friendly, synthetic biology‐compatible, and scalable computational workflows for gRNA design, construct assembly, and mutation analysis. Experimentally validated inducible or tissue‐specific promoters are also highly desirable for achieving spatiotemporal control. As these tools continue to evolve, multiplex CRISPR editing is poised to become a foundational technology of next‐generation crop improvement to address challenges in agriculture, sustainability, and climate resilience.

59 BASIC BIOLOGICAL SCIENCES↗

Black-box optimization of CT acquisition and reconstruction parameters: a reinforcement learning approach

Protocol optimization is critical in Computed Tomography (CT) for achieving desired diagnostic image quality while minimizing radiation dose. Due to the inter-effect of influencing CT parameters, traditional optimization methods rely on the testing of exhaustive combinations of these parameters. This poses a notable limitation due to the impracticality of exhaustive parameter testing. This study introduces a novel methodology leveraging Virtual Imaging Trials (VITs) and reinforcement learning to more efficiently optimize CT protocols. Computational phantoms with liver lesions were imaged using a validated CT simulator and reconstructed with a novel CT reconstruction Toolkit. The optimization parameter space included tube voltage, tube current, reconstruction kernel, slice thickness, and pixel size. The optimization process was done using a Proximal Policy Optimization (PPO) agent which was trained to maximize the Detectability Index (d’) of the liver lesion for each reconstructed image. Results showed that our reinforcement learning approach found the absolute maximum d’ across the test cases while requiring 79.7% fewer steps compared to an exhaustive search, demonstrating both accuracy and computational efficiency, offering a efficient and robust framework for CT protocol optimization. The flexibility of the proposed technique allows for use of varying image quality metrics as the objective metric to maximize for. Our findings highlight the advantages of combining VIT and reinforcement learning for CT protocol management.

Fenwick, David [Duke University Medical Center]↗

Tools for genetic engineering and gene expression control in Novosphingobium aromaticivorans and Rhodobacter sphaeroides

ABSTRACT Alphaproteobacteria have a variety of cellular and metabolic features that provide important insights into biological systems and enable biotechnologies. For example, some species are capable of converting plant biomass into valuable biofuels and bioproducts that have the potential to contribute to the sustainable bioeconomy. Among the Alphaproteobacteria, Novosphingobium aromaticivorans , Rhodobacter sphaeroides , and Zymomonas mobilis show promise as organisms that can be engineered to convert extracted plant lignin or sugars into bioproducts and biofuels. Genetic manipulation of these bacteria is needed to introduce engineered pathways and modulate expression of native genes with the goal of enhancing bioproduct output. Although recent work has expanded the genetic toolkit for Z. mobilis , N. aromaticivorans and R. sphaeroides still need facile, reliable approaches to deliver genetic payloads to the genome and to control gene expression. Here, we expand the platform of genetic tools for N. aromaticivorans and R. sphaeroides to address these issues. We demonstrate that Tn 7 transposition is an effective approach for introducing engineered DNA into the chromosome of N. aromaticivorans and R. sphaeroides . We screen a synthetic promoter library to identify isopropyl β-D-1-thiogalactopyranoside-inducible promoters with regulated activity in both organisms (up to ~15-fold induction in N. aromaticivorans and ~5-fold induction in R. sphaeroides ). Combining Tn 7 integration with promoters from our library, we establish CRISPR (Clustered Regularly Interspaced Short Palindromic Repeats) interference systems for N. aromaticivorans and R. sphaeroides (up to ~10-fold knockdown in N. aromaticivorans and R. sphaeroides ) that can target essential genes and modulate engineered pathways. We anticipate that these systems will greatly facilitate both genetic engineering and gene function discovery efforts in these species and other Alphaproteobacteria. IMPORTANCE It is important to increase our understanding of the microbial world to improve health, agriculture, the environment, and biotechnology. For example, building a sustainable bioeconomy depends on the efficient conversion of plant material to valuable biofuels and bioproducts by microbes. One limitation in this conversion process is that microbes with otherwise promising properties for conversion are challenging to genetically engineer. Here we report genetic tools for Novosphingobium aromaticivorans and Rhodobacter sphaeroides that add to the burgeoning set of tools available for genome engineering and gene expression in Alphaproteobacteria. Our approaches allow straightforward insertion of engineered pathways into the N. aromaticivorans or R. sphaeroides genome and control of gene expression by inducing genes with synthetic promoters or repressing genes using CRISPR interference. These tools can be used in future work to gain additional insight into these and other Alphaproteobacteria and to aid in optimizing yield of biofuels and bioproducts.

Hall, Ashley N.↗

PyOED: An Extensible Suite for Data Assimilation and Model-Constrained Optimal Design of Experiments

This article describes PyOED, a highly extensible scientific package that enables developing and testing model-constrained optimal experimental design (OED) for inverse problems. Specifically, PyOED aims to be a comprehensive Python toolkit for model-constrained OED. The package targets scientists and researchers interested in understanding the details of OED formulations and approaches. It is also meant to enable researchers to experiment with standard and innovative OED technologies with a wide range of test problems (e.g., simulation models). OED, inverse problems (e.g., Bayesian inversion), and data assimilation (DA) are closely related research fields, and their formulations overlap significantly. Thus, PyOED is continuously being expanded with a plethora of Bayesian inversion, DA, and OED methods as well as new scientific simulation models, observation error models, and observation operators. These pieces are added such that they can be permuted to enable testing OED methods in various settings of varying complexities. The PyOED core is completely written in Python and utilizes the inherent object-oriented capabilities; however, the current version of PyOED is meant to be extensible rather than scalable. Specifically, PyOED is developed to “enable rapid development and benchmarking of OED methods with minimal coding effort and to maximize code reutilization.” This article provides a brief description of the PyOED layout and philosophy and provides a set of exemplary test cases and tutorials to demonstrate the potential of the package.

97 MATHEMATICS AND COMPUTING↗

From Edge to HPC: Investigating Cross-Facility Data Streaming Architectures

In this paper, we investigate three cross-facility data streaming architectures, Direct Streaming (DTS), Proxied Streaming (PRS), and Managed Service Streaming (MSS). We examine their architectural variations in data flow paths and deployment feasibility, and detail their implementation using the Data Streaming to HPC (DS2HPC) architectural framework and the SciStream memory-to-memory streaming toolkit on the production-grade Advanced Computing Ecosystem (ACE) infrastructure at Oak Ridge Leadership Computing Facility (OLCF). We present a workflow-specific evaluation of these architectures using three synthetic workloads derived from the streaming characteristics of scientific workflows. Through simulated experiments, we measure streaming throughput, round-trip time, and overhead under work sharing, work sharing with feedback, and broadcast and gather messaging patterns commonly found in AI-HPC communication motifs. Our study shows that DTS offers a minimal-hop path, resulting in higher throughput and lower latency, whereas MSS provides greater deployment feasibility and scalability across multiple users but incurs significant overhead. PRS lies in between, offering a scalable architecture whose performance matches DTS in most cases.

George, Anjus [ORNL] (ORCID:0000000179737061)↗

SetGo: Metadata Readiness for Scientific AI Datasets

Scientific datasets intended for AI use require both computational readiness for model training and metadata readiness for discovery, sharing, and reuse. The Readiness Engine for Data Integration (REDI) addresses computational readiness, but no corresponding tool evaluates whether a dataset’s metadata are sufficiently complete, governed, and standards-compliant for publication and agent-based consumption. Existing FAIR assessors operate only on published repository records, and no single system covers FAIR compliance, licensing, provenance, governance, reproducibility, and catalog readiness together. We present SetGo, an open-source Python toolkit that assesses and repairs metadata readiness across these six dimensions before a dataset is published or archived. Applied to four scientific corpora, SetGo surfaces deficiencies that general-purpose tools do not detect: ERA5 climate metadata scores 4% on ACDD 1.3 compliance; materials datasets fail OPTIMADE species-definition requirements; and PDB-derived proteomics data carries licensing terms incompatible with standard SPDX identifiers. Guided enrichment raises overall FAIR scores from 52–57% to 81–91%, and a single setgo publish command pushes to Hugging Face Hub, CKAN, or OpenMetadata with ML Commons Croissant 1.0 metadata sidecars. To support interactive and automated workflows, SetGo integrates with coding agents powered by large language models (LLMs) through a /setgo skill that enables natural-language execution of the full assess–enrich–publish loop, with user involvement limited to supplying missing metadata values.

Wilkinson, Sean [ORNL] (ORCID:0000000214437479)↗

DEPRECATED AI-Batt-OS (Autonomous Identification of Battery Life Models - Open Source) [SWR 21-17]

DEPRECATED. This repository was archived by the owner on Jun 30, 2026. It is now read-only. Open source implementation of some of the methods utilized by AI-Batt, a battery lifetime modeling and analysis toolkit provided by the National Laboratory of the Rockies (NLR). This software demonstrates the use of bi-level optimization and symbolic regression techniques to semi-autonomously identify algebraic models predicting the capacity fade of lithium-ion batteries during calendar aging. Modeling the degradation of batteries is a complex task, due to the difficulty in separating the time-dependent and time-independent factors impacting cell level degradation, across multiple data series with different numbers of measurements and/or data quality. Bi-level optimization enables model parameters to be optimized to either the entire data set or to individual data series, allowing statistical disambiguation of global behaviors (data series independent) and local behaviors (data series dependent). Symbolic regression is used to automatically search for optimal low-dimesional models predicting the variation of locally optimized parameters versus time-independent experimental variables from millions of possible models, resulting in a more accurate and repeatable model identification process than is possible by a manual search. The provided tools also implement cross-validation and bootstrap resampling schemes, empowering statistical model comparison/selection and quantification of model uncertainties. An example script replicates the results from the manuscript "Challenging Practices of Algebraic Battery Life Models through Statistical Validation and Model Identification via Machine-Learning", submitted to ECS. All code is written in MATLAB. Requires the Statistics and Machine Learning Toolbox. Contact Dr. Paul Gasper at Paul.Gasper@nlr.gov for any questions.

Gasper, Paul [National Renewable Energy Lab. (NREL↗

Mbin v1.0

The Mbin software, is a software toolkit that implements the IMG metagenome binning pipeline. The software allows the user to process input metagenome contigs, and produces metagenome assembled genomes (metagenome bins) and valuation metrics per bin including completion and contamination estimates, quality assignment, predicted lineage and eukaryotic potential. It is currently packed as a portable docker container and provides the advantage of running the process of binning and analysis of the bins generated, using a suite of tools run sequentially with controls in place to capture errors and optional arguments to run a modified version depending on individual needs and capabilities.

Varghese, Neha↗

hFlux

hFlux is open source, lightweight, easy-to-use toolkit for simulation code developers working with magnetic fields to seamlessly check their intermediate results throughout the development of simulation codes.

Beznosov, Oleksii↗

BEAST

The Bilinear Ensemble Actuation Synthesis Toolkit (BEAST) is a computational platform that optimizes time-varying control signals to achieve a specified transfer of states governed by bilinear ensemble systems in which model parameters are subject to uncertainty.

Zlotnik, anatoly↗

Microreactor Optimization Using Simulation And Economics (mouse)

Microreactor Optimization Using Simulation and Economics (MOUSE) is a tool that integrates both nuclear microreactor design and reactor economics to provide comprehensive evaluations and optimizations. This tool enables stakeholders to explore the interplay between technical and economic variables, guiding them towards effective and competitive microreactor solutions. For the reactor core simulations, MOUSE leverages the OpenMC Monte Carlo Particle Transport Code to perform detailed core simulations for various microreactor designs. The included OpenMC models are 2D core designs of a Liquid Metal Thermal Microreactor (LMTR), a Gas-Cooled TRISO-Fueled Microreactor (GCMR), and a Heat Pipe Microreactor. Beyond core design, MOUSE includes simplified calculations for: - Calculating the masses of heat exchangers within the system. - Mechanical power of pumps. - Estimating the area occupied by various buildings within the nuclear plant. For the economic analysis, MOUSE provides detailed bottom-up cost estimates, encompassing a wide range of costs including preconstruction costs, direct costs, indirect costs, training costs, financial costs, operation & maintenance (O&M) costs, and fuel costs. These cost estimations are developed using data from the MARVEL project and additional literature sources, enabling the calculation of total capital costs and levelized cost of energy for both first-of-a-kind and nth-of-a-kind microreactors. MOUSE also enables analysis of the cost drivers and competitiveness in the electricity market. MOUSE allows users to modify a wide array of technical and economic parameters to evaluate different scenarios and their impacts. Examples of these parameters include: Fuels, coolants, or reflector materials Enrichment levels Control drum materials and geometry Fuel pin geometry and materials Moderator pin geometry and materials Reactor core and reflector dimensions Packing factor for the TRISO particles Nuclear reactor power and reactor burnup Number of sensors Shielding thickness Reactor vessel and guard vessel dimensions Operational staff requirements Number of emergency shutdowns Levelization period Interest rate Construction duration Since MOUSE is powered by the WATTS toolkit, it supports optimization studies, parametric analyses, and uncertainty calculations/propagation. The optimization techniques enable users to identify optimal design and economic configurations. The parametric analysis tools allow users to explore the sensitivity of various parameters, while uncertainty propagation helps quantify the impact of uncertainties on overall performance and cost. User Interface and Workflow: Currently, MOUSE is a command-line-based tool. Users can input various reactor design or economic parameters, modify the designs, run simulations, and visualize results through comprehensive data visualization and reporting capabilities. The typical workflow involves setting up the reactor model, defining economic parameters, running simulations, and analyzing the results to make informed decisions. By combining advanced design calculations with detailed economic modeling, MOUSE provides a robust framework for optimizing nuclear microreactor technologies, enhancing their competitiveness, and guiding stakeholders towards innovative and cost-effective solutions.

Hanna, Botros [Idaho National Laboratory (INL), Id↗

BAMBAM (The Behavior and Advanced Mobility Big Access Model) [SWR-25-123]

Access modeling toolkit for Rust built on the RouteE Compass energy-aware route planner. The Behavior and Advanced Mobility Big Access Model (BAMBAM) is a mobility research platform for scalable access modeling. The process begins with a grid defined at some spatial granularity (e.g., census block or 1 km grid) and a variety of travel configurations. For each grid cell and configuration, the platform executes constrained searches, uses the results to index points of interest (POI), and aggregates the findings to the grid level. It provides researchers access through R or Python running on HPC. The platform automates the import and merging of datasets from a variety of sources including data.gov (with automatic merging of Tiger/Lines geometries), OpenStreetMaps, OvertureMaps, and GTFS. It is built upon RouteE Compass, a scalable, energy-aware route planner written in Rust, extended to model multiple travel modes.

Fitzgerald, Robert [National Renewable Energy Labo↗

artdaq

The artdaq toolkit is a data-acquisition framework designed for high-energy physics experiments. It provides a flexible, reliable backbone for data transfers and has several locations where users can perform custom analysis tasks using the art framework.

Flumerfelt, EricL. [Fermi National Accelerator Lab↗