Search NASA⌕ Search

SEARCH · Search NASA

Results for “software tool”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 577 records · Page 32

From PINNs to PIKANs: recent advances in physics-informed machine learning

Physics-Informed Neural Networks (PINNs) have emerged as a key tool in Scientific Machine Learning since their introduction in 2017, enabling the efficient solution of ordinary and partial differential equations using sparse measurements. Over the past few years, significant advancements have been made in the training and optimization of PINNs, covering aspects such as network architectures, adaptive refinement, domain decomposition, and the use of adaptive weights and activation functions. A notable recent development is the Physics-Informed Kolmogorov-Arnold Networks (PIKANS), which leverage a representation model originally proposed by Kolmogorov in 1957, offering a promising alternative to traditional PINNs. In this review, we provide a comprehensive overview of the latest advancements in PINNs, focusing on improvements in network design, feature expansion, optimization techniques, uncertainty quantification, and theoretical insights. We also survey key applications across a range of fields, including biomedicine, fluid and solid mechanics, geophysics, dynamical systems, heat transfer, chemical engineering, and beyond. Lastly, we review computational frameworks and software tools developed by both academia and industry to support PINN research and applications.

Kolmogorov-Arnold networks↗

Automated Euler and Navier-Stokes Database Generation for a Glide-Back Booster

The past two decades have seen a sustained increase in the use of high fidelity Computational Fluid Dynamics (CFD) in basic research, aircraft design, and the analysis of post-design issues. As the fidelity of a CFD method increases, the number of cases that can be readily and affordably computed greatly diminishes. However, computer speeds now exceed 2 GHz, hundreds of processors are currently available and more affordable, and advances in parallel CFD algorithms scale more readily with large numbers of processors. All of these factors make it feasible to compute thousands of high fidelity cases. However, there still remains the overwhelming task of monitoring the solution process. This paper presents an approach to automate the CFD solution process. A new software tool, AeroDB, is used to compute thousands of Euler and Navier-Stokes solutions for a 2nd generation glide-back booster in one week. The solution process exploits a common job-submission grid environment, the NASA Information Power Grid (IPG), using 13 computers located at 4 different geographical sites. Process automation and web-based access to a MySql database greatly reduces the user workload, removing much of the tedium and tendency for user input errors. The AeroDB framework is shown. The user submits/deletes jobs, monitors AeroDB's progress, and retrieves data and plots via a web portal. Once a job is in the database, a job launcher uses an IPG resource broker to decide which computers are best suited to run the job. Job/code requirements, the number of CPUs free on a remote system, and queue lengths are some of the parameters the broker takes into account. The Globus software provides secure services for user authentication, remote shell execution, and secure file transfers over an open network. AeroDB automatically decides when a job is completed. Currently, the Cart3D unstructured flow solver is used for the Euler equations, and the Overflow structured overset flow solver is used for the Navier-Stokes equations. Other codes can be readily included into the AeroDB framework.

Chaderjian, Neal M.↗

User Interface Developed for Controls/CFD Interdisciplinary Research

The NASA Lewis Research Center, in conjunction with the University of Akron, is developing analytical methods and software tools to create a cross-discipline "bridge" between controls and computational fluid dynamics (CFD) technologies. Traditionally, the controls analyst has used simulations based on large lumping techniques to generate low-order linear models convenient for designing propulsion system controls. For complex, high-speed vehicles such as the High Speed Civil Transport (HSCT), simulations based on CFD methods are required to capture the relevant flow physics. The use of CFD should also help reduce the development time and costs associated with experimentally tuning the control system. The initial application for this research is the High Speed Civil Transport inlet control problem. A major aspect of this research is the development of a controls/CFD interface for non-CFD experts, to facilitate the interactive operation of CFD simulations and the extraction of reduced-order, time-accurate models from CFD results. A distributed computing approach for implementing the interface is being explored. Software being developed as part of the Integrated CFD and Experiments (ICE) project provides the basis for the operating environment, including run-time displays and information (data base) management. Message-passing software is used to communicate between the ICE system and the CFD simulation, which can reside on distributed, parallel computing systems. Initially, the one-dimensional Large-Perturbation Inlet (LAPIN) code is being used to simulate a High Speed Civil Transport type inlet. LAPIN can model real supersonic inlet features, including bleeds, bypasses, and variable geometry, such as translating or variable-ramp-angle centerbodies. Work is in progress to use parallel versions of the multidimensional NPARC code.

Source record↗

Control Strategies and Validation in the Hybrid Optimization and Performance Platform (HOPP)

The Hybrid Optimization and Performance Platform (HOPP) is a tool that simulates hybrid power plants in various configurations, and also calculates the financial feasibility of these plants. This report outlines an overview of HOPP and the energy storage dispatch strategies available. It then presents three case studies which demonstrate different applications of HOPP. The first case looks at the profitability of hybrid power plants in different locations in the USA. The second case examines the availability of hybrid power plants to provide energy reliability services. The third case presents a plant that produces both hydrogen and electricity, and demonstrates a dispatch strategy that chooses the most profitable energy vector based on price signals. The next section shows the validation of HOPP on operational data, using data from both unit-scale and utility-scale power plants. This validation process demonstrated that HOPP can simulate the power output of both wind and solar PV plants at both scales with comparable fidelity to an existing commercial software tool. Finally, HOPP is applied in a field test which applies an optimal dispatch strategy to a physical battery in a unit-scale hybrid plant at NREL. HOPP's optimal dispatch strategy, applied in a real-world setting, improved this hybrid plant's ability to meet a load signal while minimizing operational costs.

14 SOLAR ENERGY↗

NASA GRC ICME Schema for Materials Data Management: An Executive Summary

Integrated Computational Materials Engineering (ICME) has received a growing emphasis in attention due its potential impact on rapid material design, reduction in cost and time to market for new applications, and the promise of ‘fit-for-purpose’ materials coupled with recent advances in high performance computing and material characterization tools. However, for an organization to implement ICME practices for material discovery and design, a series of both technical and cultural challenges must be overcome to foster an environment that enables efficient, traceable, and predictive multiscale simulations of material behavior to enable virtual design of materials. In 2016, NASA sponsored a 2040 Vision study to define the potential 25-year future state required for integrated multiscale modeling of materials and systems to improve both the associated time and cost for aerospace and aeronautical innovation. The study envisions a cyber-physical-social ecosystem of experimentally validated computational models, tools, and techniques, along with the associated digital tapestry, that can enable rapid, optimized, ‘fit-for-purpose’ design of materials, components, and systems. A key requirement for such an ecosystem is the development of a robust information management system for materials across their full lifecycle, including material pedigree, experimental (real) and virtual (simulation) data, developed material models, and the implementation of models in engineering applications, such that process-structure-property-performance relationships can be established, thereby enabling the virtual design and optimization of materials. Such an information management system must be able to effectively capture: i) material information at each length scale; ii) test data and analysis; iii) associated material models; and iv) material and model deployment in engineering applications. These systems must also provide traceability between experimental and virtual representations of the material to ensure, when appropriate, the material digital twin is maintained. Additionally, this robust material information management system must be able to seamlessly connect with both commercial and an organization’s in-house software tools, be they analysis tools, other material databases, product lifecycle management (PLM) or simulation data management (SDM) tools, etc., such that automation of the design and analysis of a material across multiple length scales is possible. In this paper, an executive summary of the NASA GRC ICME Schema for materials information management is presented. The database best practices and schema design philosophy specifically for ICME materials data management and an overview description of each element in the schema is given, along with its associated role in an ICME workflow. Additionally, auxiliary tools that interact with the database and provide judicious automation with regards to importing, exporting, and analyzing materials data are presented. Such tools are critical to an ICME ecosystem, not only for their role in enabling optimization, but also in relieving users of tedious manual tasks, thus helping to promote adoption and combat the cultural challenges organizations face in enabling ICME.

Materials↗

PQML: Enabling the Predictive Reproducibility on NISQ Machines for Quantum ML Applications

Quantum computing represents a groundbreaking approach to high-performance computing. In recent years, quantum computers have progressed from single-qubit processors to systems boasting over 400 qubits. The presence of such a large number of qubits offers significant advantages, including enhanced computational speed—a capability beyond classical computing methods. However, the current stage of quantum computing is referred to as the noisy intermediate-scale quantum (NISQ) era. The existence of noise in this era presents challenges in testing quantum computing applications, leading to considerable variance in application results. Furthermore, the diverse noise characteristics observed across different machines exacerbate this issue, complicating the selection of the appropriate machine for application execution. In response to these challenges, we introduce our Predictive Quantum Machine Learning (PQML) tool. This tool is designed to predict outcomes when executing identical quantum machine learning applications—specifically, a critical suite of variational quantum algorithms—across various quantum computers during the NISQ era. This effort relies on data collected over a 12-month period. To the best of our knowledge, this study represents the first attempt to ensure reproducibility across quantum computers for complex circuits. Additionally, we have developed a model capable of forecasting the accuracy of quantum computers for variational quantum algorithms, with a particular emphasis on quantum machine learning as a case study.

Senapati, Priyabrata [Kent State University]↗

Use of New Communication Technologies to Change NASA Safety Culture: Incorporating the Use of Blogs as a Fundamental Communications Tool

The purpose of this paper is to explore an innovative approach to culture change at NASA that goes beyond reorganizations, management training, and a renewed emphasis on safety. Over the last five years, a technological social revolution has been emerging from the internet. Blogs (aka web logs) are transforming traditional communication and information sharing outlets away from established information sources such as the media. The Blogosphere has grown from zero blogs in 1999 to approximately 4.5 million as of November 2004 and is expected to double in 2005. Blogs have demonstrated incredible effectiveness and efficiency with regards to affecting major military and political events. Consequently, NASA should embrace the new information paradigm presented by blogging. NASA can derive exceptional benefits from the new technology as follows: 1) Personal blogs can overcome the silent safety culture by giving voice to concerns or questions that are not well understood or seemingly inconsequential to the NASA community at-large without the pressure of formally raising a potential false alarm. Since blogs can be open to Agency-wide participation, an incredible amount of resources from an extensive pool of experience can focus on a single issue, concern, or problem and quickly vetted, discussed and assessed for feasibility, significance, and criticality. The speed for which this could be obtained cannot be matched through any other process or procedure currently in use. 2) Through official NASA established blogs, lessons learned can be a real-time two way process that is formed and implemented from the ground level. Data mining of official NASA blogs and personal blogs of NASA personnel can identify hot button issues and concerns to senior management. 3) NASA blogs could function as a natural ombudsman for the NASA community. Through the recognition of issues being voiced by the community and taking a proactive stance on those issues, credibility within NASA Management can be restored. For NASA to harness the capabilities of blogs, NASA must develop an Agency-wide policy on blogging to encourage use and provide guidance. This policy should describe basic rules of conduct and content as well as a policy of non-retribution and/or anonymity. The Agency must provide sever space within their firewalls, provide appropriate software tools, and promote blogs in newsletters and official websites. By embracing the use of blogs, a potential pool of 19,000 experts could be available to address each posted safety issue, concern, problem, or question. Blogs could result in real NASA culture change.

Huls, Dale thomas↗

The XRISM Science Data Center: Optimizing the Scientific Return from a Unique X-ray Observatory

The X-Ray Imaging and Spectroscopy Mission, XRISM, is currently scheduled to launch in 2022 with the objective of building on the brief, but significant, successes of the ASTRO-H (Hitomi) mission in solving outstanding astrophysical questions using high resolution X-ray spectroscopy. The XRISM Science Operations Team (SOT) consists of the JAXA-led Science Operations Center (SOC) and NASA-led Science Data Center (SDC), which work together to optimize the scientific output from the Resolve high-resolution spectrometer and the Xtend wide-field imager through planning and scheduling of observations, processing and distribution of data, development and distribution of software tools and the calibration database (CaldB), support of ground and in-flight calibration, and support of XRISM users in their scientific investigations of the energetic universe. Here, we summarize the roles and responsibilities of the SDC and its current status and future plans. The Resolve instrument poses particular challenges due to its unprecedented combination of high spectral resolution and throughput, broad spectral coverage, and relatively small field-of-view and large pixel-size. We highlight those challenges and how they are being met.

XRISM↗

ISS Technology Demonstrations for Future Spaceflight Medical Systems

Throughout the history of human spaceflight, crewmembers have experienced various in-flight medical conditions including illness and injury. Planned missions to the Moon and Mars will require capabilities to maintain the health of future space travelers. Mass, power, and volume available in the vehicles and habitats for these missions will be severely constrained; resupply of resources will be limited or non-existent, as will opportunities for evacuation to Earth. Furthermore, ground-based support will be hampered by communication latencies and blackouts. These vehicle and mission constraints will necessitate a medical system that has been efficiently planned, providing on-board procedural guidance in addition to a variety of medical devices and consumable resources. Medical capabilities required for the diagnosis and treatment of potential medical conditions during future spaceflight missions may include real-time health monitoring, medical imaging, and biomarker analyses ( e.g., blood or urine). Terrestrial medicine shares these needs, thus many of these medical capabilities could likely be satisfied by Commercial-Off-The-Shelf (COTS) devices and methodologies; however, in some cases the unique space environment and increased mission duration will drive the need to modify technologies and the way care is provided. NASA’s Human Research Program (HRP) Exploration Medical Capability (ExMC) Element and Mars Campaign Office’s Exploration Medical Integrated Product Team (XMIPT) are working together to decrease medical risk during exploration missions. Flight-tested medical diagnostic and treatment technologies are necessary to effectively manage medical conditions relevant to exploration missions while meeting vehicle constraints, integrating with medical decision-support tools, and enabling increasingly Earth-independent operations. Several projects have leveraged the ISS as a testbed for exploration, including 1) i n- situ blood analysis, 2) medical inventory, 3) intravenous fluid generation, and 4) autonomous medical procedure guidance. Management of several in-flight medical conditions, such as bacterial and viral infections and acute radiation syndrome, is dramatically improved with ability to assess blood cell populations, electrolytes, and metabolites. I n December 2020 and January 2021 ExMC performed an ISS technology demonstration (Tech Demo) of the HemoCue® WBC DIFF analyzer (HemoCue, Brea, CA), a COTS device that was modified to enable functionality in a spaceflight environment. This Tech Demo marked the first time that hematology measurements were successfully performed real-time in microgravity. Also modified and demonstrated was the reusable Handheld Electrolytes and Lab Technology for Humans (rHEALTH) ONE analyzer (rHEALTH, Bedford, MA), which uses flow cytometry and sheath-based hydrodynamic focusing methodologies. The rHEALTH ONE ISS Tech Demo in May 2022 demonstrated test results obtained in-flight matched those on the ground. NASA currently relies on crew self-reporting to manage and maintain medical inventory on ISS.The ability to maintain an accurate inventory becomes more critical during long duration missions since the crew will need to be increasingly autonomous in finding and utilizing medical items, including those scenarios when alternative treatments need to be considered due to limited or no resupply. HRP’s Medical Consumables Tracking (MCT) project was developed by ZIN Technologies, Inc. (Cleveland, OH), and demonstrated real-time tracking of medical supplies aboard the ISS between December 2016 and July 2018. The MCT system design utilized Radio Frequency Identification Device (RFID) technology to perform automated inventory and was installed in the Crew Health Care System (CHeCS) Resupply Stowage Rack (RSR). The challenge of limited shelf life, exacerbated by the lack of resupply opportunities, affects a plethora of medical system components including consumables, pharmaceuticals, and intravenous (IV) fluid. In 2010, ExMC funded ZIN Technologies, Inc. (Cleveland, OH), to develop the Intravenous Fluid Generation (IVGEN) system. IV fluids were successfully generated with IVGEN using the potable water supply on ISS during ISS Expedition 23. The XMIPT is in the process of developing a miniaturized version of the original IVGEN hardware for a future Tech Demo aboard the ISS. Current ISS medical operations rely heavily on preflight training and real-time remote guidance, both of which become impractical or impossible for exploration missions. The primary goal of the Autonomous Medical Officer Support (AMOS) Software Tech Demos on ISS was to confirm telemedical proof-of-concept for autonomous medical imaging in an operational setting. This novel software tool shifts emphasis from preflight training and real-time remote guidance to in-flight just-in-time instruction, a new and necessary paradigm for crew medical autonomy. AMOS introduces a novel, streamlined skill management concept for exploration missions featuring comprehensive training and guidance modules for ultrasound examinations using the ISS Ultrasound 2 (a modified GE Vivid-q™; General Electric HealthCare, Chicago, IL). With no prior crew training or remote guidance, two Tech Demos on the ISS (April 2020 and June 2022) resulted in high quality, clinically useful image sets. We will provide a review of historical, current, and planned medical devices and technologies considered for inclusion within future spaceflight medical systems and summarize hardware development activities and medical device tech demos conducted on the ISS.

Astronaut health and performance↗

Bootstrapping the 3d Ising stress tensor

We compute observables of the critical 3d Ising model to high precision by applying the numerical conformal bootstrap to mixed correlators of the leading scalar operators σ and ϵ, and the stress tensor T μν . We obtain new precise determinations of scaling dimensions (∆ σ , ∆ ϵ ) = (0.518148806(24), 1.41262528(29)) as well as OPE coefficients involving σ, ϵ, and T μν . We also describe several improvements made along the way to algorithms and software tools for the numerical bootstrap.

Conformal and W Symmetry↗

Applying queueing theory to evaluate wait-time-savings of triage algorithms

Abstract In the past decade, artificial intelligence (AI) algorithms have made promising impacts in many areas of healthcare. One application is AI-enabled prioritization software known as computer-aided triage and notification (CADt). This type of software as a medical device is intended to prioritize reviews of radiological images with time-sensitive findings, thus shortening the waiting time for patients with these findings. While many CADt devices have been deployed into clinical workflows and have been shown to improve patient treatment and clinical outcomes, quantitative methods to evaluate the wait-time-savings from their deployment are not yet available. In this paper, we apply queueing theory methods to evaluate the wait-time-savings of a CADt by calculating the average waiting time per patient image without and with a CADt device being deployed. We study two workflow models with one or multiple radiologists (servers) for a range of AI diagnostic performances, radiologist’s reading rates, and patient image (customer) arrival rates. To evaluate the time-saving performance of a CADt, we use the difference in the mean waiting time between the diseased patient images in the with-CADt scenario and that in the without-CADt scenario as our performance metric. As part of this effort, we have developed and also share a software tool to simulate the radiology workflow around medical image interpretation, to verify theoretical results, and to provide confidence intervals for the performance metric we defined. We show quantitatively that a CADt triage device is more effective in a busy, short-staffed reading setting, which is consistent with our clinical intuition and simulation results. Although this work is motivated by the need for evaluating CADt devices, the evaluation methodology presented in this paper can be applied to assess the time-saving performance of other types of algorithms that prioritize a subset of customers based on binary outputs.

Thompson, Yee Lam Elim (ORCID:0000000196537707)↗

Machine learning in materials research: Developments over the last decade and challenges for the future

The number of studies that apply machine learning (ML) to materials science has been growing at a rate of approximately 1.67 times per year over the past decade. In this review, I examine this growth in various contexts. First, I present an analysis of the most commonly used tools (software, databases, materials science methods, and ML methods) used within papers that apply ML to materials science. The analysis demonstrates that despite the growth of deep learning techniques, the use of classical machine learning is still dominant as a whole. It also demonstrates how new research can effectively build upon past research, particular in the domain of ML models trained on density functional theory calculation data. Next, I present the progression of best scores as a function of time on the matbench materials science benchmark for formation enthalpy prediction. In particular, a dramatic improvement of 7 times reduction in error is obtained when progressing from feature-based methods that use conventional ML (random forest, support vector regression, etc.) to the use of graph neural network techniques. Finally, I provide views on future challenges and opportunities, focusing on data size and complexity, extrapolation, interpretation, access, and relevance.

36 MATERIALS SCIENCE↗

PDB-IHM: A System for Deposition, Curation, Validation, and Dissemination of Integrative Structures

Structures of many large biomolecular assemblies are now being determined using integrative approaches. In these approaches, information derived from multiple experimental and computational methods is combined to compute three-dimensional structures of multi-protein complexes and other macromolecular machines. A standalone prototype data resource for integrative structures called PDB-Dev was built, based on recommendations of the Integrative and Hybrid Methods (IHM) Task Force of the Worldwide Protein Data Bank (wwPDB). This effort included developing data standards and software tools for collecting, curating, validating, visualizing, archiving, and disseminating integrative structures that span diverse spatiotemporal scales and conformational states. Mechanisms have been created to validate integrative structures based on the experimental data underpinning them. Building upon this foundational framework, PDB-Dev has been further expanded to handle large dynamic macromolecular systems and integrative structures that combine, for example, experimental restraints with atomic coordinates computed by machine learning algorithms. Data standards and supporting tools have also been extended to capture information about biomolecular dynamics, such as conformational transitions and related kinetic data derived from biophysical methods. Recently, PDB-Dev was unified with the PDB archive and rebranded as PDB-IHM (pdb-ihm.org), further promoting FAIR (Findable, Accessible, Interoperable, and Reusable) principles of data stewardship for integrative structural biology.

IHMCIF↗

rcsb-api : Python Toolkit for Streamlining Access to RCSB Protein Data Bank APIs

The Protein Data Bank (PDB) was founded in 1971 as the first open-access digital data resource in biology to serve as the single global archive for three-dimensional (3D) macromolecular structure data. Current PDB holdings exceed 230,000 experimentally determined structures of proteins, nucleic acids, viruses, and macromolecular machines. The RCSB Protein Data Bank RCSB.org research-focused web portal facilitates search, analyses, and visualization of every PDB structure along with more than one million Computed Structure Models from AlphaFold DB and the ModelArchive. It is powered by a set of publicly available Application Programming Interfaces (APIs) that both support RCSB.org users and provide programmatic access to PDB data. Given the breadth and levels of granularity encompassed in this rich data collection, efficiently accessing the information programmatically may be challenging for new users. RCSB PDB has developed a Python software package, rcsb-api , that facilitates easy and efficient use of RCSB PDB APIs within a Python environment. This software tool is designed to streamline access to the extensive corpus of data housed within the PDB, enabling researchers to search, retrieve, and analyze 3D biostructure data seamlessly. Its use will accelerate research in structural biology, molecular biology and biochemistry, drug discovery, and bioinformatics by providing more efficient tools for data integration and analysis. The new toolkit is available on GitHub (github.com/rcsb/py-rcsb-api) and published to the public Python package repository (PyPI) to foster wider usage and support basic and applied research in fundamental biology, biomedicine, and the energy sciences.

FAIR principles↗

Key insights from US Department of Energy Better Plants workforce development bootcamps (2022–2025)

This study examines the effectiveness of the US Department of Energy’s Better Plants Program Bootcamps, which are designed to enhance participants’ technical skills in improving energy efficiency and optimizing operations in manufacturing facilities. Through the analysis of survey data collected from 529 participants across 9 bootcamps, the research investigates the motivations, benefits, and demographic trends of attendees. The findings reveal that skill acquisition and improvement are primary drivers for participation, with key benefits including hands-on training on diagnostic equipment and software tools, networking opportunities, and access to technical resources. The analysis shows strong participation from sectors characterized by high energy consumption and employment, such as chemical and transportation equipment manufacturing. Over 50% of participants have job titles that include “EHS” or “Energy” showing their key roles in leading energy efficiency and energy management efforts in manufacturing. Furthermore, the analysis highlights the distribution of participants across managerial, engineering, and technical roles, revealing a higher representation of managers and engineers. This observation suggests a need for targeted outreach to engage technicians, equipment operators, maintenance staff, and floor workers to ensure comprehensive workforce development. The post-bootcamp survey showed that the participants highly valued the opportunities for peer learning and idea exchange, and the benefits they gained from them. This research contributes to the advancement of manufacturing education by demonstrating the efficacy of specialized training in addressing critical industry challenges and fostering a more competent and empowered workforce.

Energy efficiency↗

Asymmetric errors

We present a procedure for handling asymmetric errors. Many results in particle physics are presented as values with different positive and negative errors, and there is no consistent procedure for handling them. We consider the difference between errors quoted, using pdfs and using likelihoods, and the difference between the rms spread of a measurement and the 68% central confidence region. We provide a comprehensive analysis of the possibilities, and software tools to enable their use.

Asymmetric↗

Abstraction hierarchy to define biofoundry workflows and operations for interoperable synthetic biology research and applications

Lack of standardization in biofoundries limits the scalability and efficiency of synthetic biology research. Here, we propose an abstraction hierarchy that organizes biofoundry activities into four interoperable levels: Project, Service/Capability, Workflow, and Unit Operation, effectively streamlining the Design‑Build‑Test‑Learn (DBTL) cycle. This framework enables more modular, flexible, and automated experimental workflows. It improves communication between researchers and systems, supports reproducibility, and facilitates better integration of software tools and artificial intelligence. Our approach lays the foundation for a globally interoperable biofoundry network, advancing collaborative synthetic biology and accelerating innovation in response to scientific and societal challenges.

Kim, Haseong↗

Flow matching meets biology and life science: a survey

Over the past decade, advances in generative modeling, such as generative adversarial networks, masked autoencoders, and diffusion models, have significantly transformed biological research and discovery, enabling breakthroughs in molecule design, protein generation, catalysis discovery, drug discovery, and beyond. At the same time, biological applications have served as valuable testbeds for evaluating the capabilities of generative models. Recently, flow matching has emerged as a powerful and efficient alternative to diffusion-based generative modeling, with growing interest in its application to problems in biology and life sciences. This paper presents the first comprehensive survey of recent developments in flow matching and its applications in biological domains. We begin by systematically reviewing the foundations and variants of flow matching, and then categorize its applications into three major areas: biological sequence modeling, molecule generation and design, and peptide and protein generation. For each, we provide an in-depth review of recent progress. We also summarize commonly used datasets and software tools, and conclude with a discussion of potential future directions.

59 BASIC BIOLOGICAL SCIENCES↗