Search NASASearch

SEARCH · Search NASA

Results for “Protein Design”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Flow matching meets biology and life science: a survey

Over the past decade, advances in generative modeling, such as generative adversarial networks, masked autoencoders, and diffusion models, have significantly transformed biological research and discovery, enabling breakthroughs in molecule design, protein generation, catalysis discovery, drug discovery, and beyond. At the same time, biological applications have served as valuable testbeds for evaluating the capabilities of generative models. Recently, flow matching has emerged as a powerful and efficient alternative to diffusion-based generative modeling, with growing interest in its application to problems in biology and life sciences. This paper presents the first comprehensive survey of recent developments in flow matching and its applications in biological domains. We begin by systematically reviewing the foundations and variants of flow matching, and then categorize its applications into three major areas: biological sequence modeling, molecule generation and design, and peptide and protein generation. For each, we provide an in-depth review of recent progress. We also summarize commonly used datasets and software tools, and conclude with a discussion of potential future directions.

59 BASIC BIOLOGICAL SCIENCES

Artificial intelligence methods for protein structure and interaction prediction: Recent advances and challenges

Recent advances in artificial intelligence have introduced novel methods for high-accuracy prediction of protein tertiary structures, protein complex structures, and interactions between proteins and other biomolecules, such as small molecules and nucleic acids. Such advancements are accelerating biomedical research and the development of new protein design and bioengineering methods among many other important biotechnology applications. Here, in this review, we outline the recent advances in protein-centric biomolecular structure and interaction prediction, highlight some major challenges in the field, and discuss potential directions to address them.

Morehead, Alex [Lawrence Berkeley National Laborat

Nanomolar Sensitivity Chirality Transfer from Designed Helical Repeat Proteins to Achiral CdS Nanorods

Bridging chirality across length scales with inorganic− organic hybrid materials is a rapidly expanding area of research. Here, we establish asymmetry at CdS nanorod (NR) interfaces using a designed helical repeat protein bearing four cysteine residues (DHR- 4Cys). Hydrophobic NRs are transferred into water with glycine, and then glycine is displaced by DHR-4Cys, leveraging the thiophilicity of cadmium. Circular dichroism (CD) in the visible, coincident with CdS electronic transitions, reveals a chiral DHR-4Cys:CdS interface. Here, the dissymmetry factor [g-factor = 4.5 × 10 −4 (short NRs) and 5.0 × 10 −4 (long NRs)] is weakly dependent on the NR length, and CD persists at nanomolar protein loadings. Additionally, control experiments demonstrate that DHR-4Cys:CdS NR chirality is dictated by the local coordination of Cys with no significant contribution from the chiral secondary structure of the protein (g-factors of short and long Cys:CdS NRs are 4.8 × 10 −4 and 4.0 × 10 −4 , respectively). Together with far-UV CD and transmission electron microscopy, which provide evidence of preserved protein structure, these results provide the first demonstration that a structurally defined protein can induce chirality in CdS nanocrystals while maintaining protein structure at biologically relevant concentrations.

Cadmium sulfide

Architector 2.0: Expanded Capabilities for Metal Complex Engineering

Automated three-dimensional molecular construction from two-dimensional graph representations is critical to high-throughput discovery eIorts. Software capabilities in this area have accelerated research across fields ranging from protein design and drug discovery to transition metal catalyst development. When Architector was first introduced, it uniquely enabled high-throughput, chemically relevant three-dimensional construction of f-element complexes. Since its introduction, Architector has been applied in large-scale computational campaigns, targeted studies in critical mineral extraction, and artificial intelligence-driven discovery eIorts.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Investigation of design principles for metal-binding and conductive protein assemblies

Throughout the lifetime of this initiative, including renewals, we focused on understanding the fundamental principles of protein-protein interface design that enable predictable and modular spatial and kinetic control of multi-component protein self-assembly in 1D, 2D, and 3D, including the interface with inorganic materials, small molecules, and metal ions. We designed individual protein components that bind specific metal ions, including REEs and transport ions across lipid membranes. We created helical 1D filaments of repeating units with programmed periodicity, pitch, and multi-component environmentally responsive self-assembling protein fibers. We showed that these filaments reversibly assemble and disassemble under specific pH conditions and created end-specific caps that independently tune the balance of attachment and detachment rates at each terminus of the filament. Using similar filaments, we succeeded in binding arrays of heme and chlorophyll molecules and assembling patterned helical coatings around carbon nanotubes in efforts to create de novo conductive nanowires. By arraying REE binding sites in a large circular tandem array with a repeat protein-based cyclic oligomer, we created a molecular scaffold for superradiance and paramagnetic quantum sensing. We created a range of one-component and two-component self-assembling 2D arrays and showed that when designed to engage cell receptors, these arrays can control cell behavior from outside the cell signal to inside the cell. We designed helical repeat proteins with variable lengths displaying charged residues in a pattern matched to the cation lattice of mica. achieved a range of ordered states with an epitaxial match to the underlying crystal lattice. We further applied the learned principles of protein-induced biomineralization to design proteins with an interface lattice matching CaCO 3 and guide the formation of specific crystal forms of CaCO 3 from solution, a significant advance toward the global need to manage carbon. In all cases of mineral lattice matching and biomineralization, we followed assembly using molecularly resolved in situ AFM imaging and extracted information about assembly pathways and energetics, applying deep learning to quantify the dynamics of protein self-organization. We developed techniques for using dynamic metal-dependent interfaces on protein nanopores for discriminatively sensing dilute REEs in solution and demonstrated the use of strong metal-binding interfaces to drive nanocage disassembly for conditional nanocompartmentalization applications. This grant supported 11 people, including Asim Bera, Evans Brackenbrough, Andrew Borst, Nikita Hanikel, Timothy Huddy, Emily Joyce, Alex Young-Seug Kang, Ryan Kibler, Joshua Morris Lubner, Harley Pyles, and Shuai Zhang. The research effort culminated in the production of published papers and theses. Electronic Thesis/Dissertation are distributed by ProQuest/UMI Dissertation Publishing and made available on an open access basis through UW Libraries ResearchWorks Service.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Unraveling design principles of protein landscapes in photosynthetic membranes in plant chloroplasts

The supramolecular organization of proteins within photosynthetic membranes is crucial for energy conversion in plants. Here, we introduce an analytical and computational pipeline that integrates high-resolution cryo–scanning electron microscopy, biochemical quantification, advanced Monte Carlo computer simulations, and statistical methods to elucidate the elusive protein landscapes of grana membranes in intact Arabidopsis leaves. Our integrated analysis challenges the prevailing view that particles on the exoplasmic fracture faces in freeze-fracture samples represent photosystem II exclusively. Instead, these particles also include cytochrome b 6 f complexes. Furthermore, our steric clash analysis demonstrates that stacked membranes contain a mixture of larger PSII supercomplexes (C 2 S 2 M 2 and C 2 S 2 ) in addition to a smaller complex (C 2 ). This suggests that in vivo PSII supercomplexes exist in an equilibrium distribution of differing sizes. Furthermore, we discovered that, although size exclusion effects govern the global protein arrangement, local packing exhibits orientational order indicative of lateral attractive protein-protein interactions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

De Novo Design of High‐Affinity Miniprotein Binders Targeting Francisella Tularensis Virulence Factor

Abstract Francisella tularensis poses considerable public health risk due to its high infectivity and potential for bioterrorism. Francisella‐like lipoprotein (Flpp3), a key virulence factor unique to Francisella, plays critical roles in infection and immune evasion, making it a promising target for therapeutic development. However, the lack of well‐defined binding pockets and structural information on native interactions has hindered structure‐guided ligand discovery against Flpp3. Here, we used a combination of physics‐based and deep‐learning methods to design high‐affinity miniprotein binders targeting two distinct sites on Flpp3. We identified four binders for site I with binding affinities ranging between 24–110 nM. For the second site, an initial binder showed a dissociation constant ( K D ) of 81 nM, and subsequent site saturation mutagenesis yielded variants with sub‐nanomolar affinities. Circular dichroism confirmed the topology of designed miniproteins. The X‐ray crystal structure of Flpp3 in complex with a site I binder is nearly identical to the design model (Cα root‐mean‐square deviation (RMSD): 0.9 Å). These designed miniproteins provide research tools to explore the roles of Flpp3 in tularemia and should enable the development of new therapeutic candidates.

Gokce‐Alpkilic, Gizem [Molecular Engineering and S

Ca X ML: Chemistry‐informed machine learning explains mutual changes between protein conformations and calcium ions in calcium‐binding proteins using structural and topological features

Proteins' flexibility is a feature in communicating changes in cell signaling instigated by binding with secondary messengers, such as calcium ions, associated with the coordination of muscle contraction, neurotransmitter release, and gene expression. When binding with the disordered parts of a protein, calcium ions must balance their charge states with the shape of calcium-binding proteins and their versatile pool of partners depending on the circumstances they transmit. Accurately determining the ionic charges of those ions is essential for understanding their role in such processes. However, it is unclear whether the limited experimental data available can be effectively used to train models to accurately predict the charges of calcium-binding protein variants. Here, we developed a chemistry-informed, machine-learning algorithm that implements a game theoretic approach to explain the output of a machine-learning model without the prerequisite of an excessively large database for high-performance prediction of atomic charges. We used the ab initio electronic structure data representing calcium ions and the structures of the disordered segments of calcium-binding peptides with surrounding water molecules to train several explainable models. Network theory was used to extract the topological features of atomic interactions in the structurally complex data dictated by the coordination chemistry of a calcium ion, a potent indicator of its charge state in protein. Our design created a computational tool of Ca X ML, which provided a framework of explainable machine learning model to annotate ionic charges of calcium ions in calcium-binding proteins in response to the chemical changes in an environment. Our framework will provide new insights into protein design for engineering functionality based on the limited size of scientific data in a genome space.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Exploring the binding properties and activities of ancestral expansins

Bacterial expansins are non-lytic proteins capable of loosening cellulose networks, offering promising applications in agriculture, biotechnology, and material science. Their ability to disrupt noncovalent interactions in biopolymer matrices such as cellulose and chitin positions them as valuable tools for upgrading abundant natural materials. However, their industrial use remains limited due to their relatively low wall-loosening activity compared to plant expansins. To address this limitation, we applied Ancestral Sequence Resurrection (ASR) to reconstruct and characterize ancient variants of the Bacillus subtilis expansin BsEXLX1. ASR is a powerful evolutionary tool that enables the inference and synthesis of ancestral proteins, allowing researchers to explore functional traits that may have been lost over time. This approach not only provides insights into protein evolution but also facilitates the design of proteins with enhanced properties, such as improved substrate affinity or structural stability. In this study, we combined biochemical and biophysical assays to evaluate the activity and binding behavior of ancestral expansins. Our results reveal that ancestral variants exhibit increased cellulose affinity, reduced binding to acidic polysaccharides, and greater salt resistance. Furthermore, these traits enhance their wall-loosening activity and demonstrate the utility of ASR in engineering surface-active proteins for industrial applications, particularly in biomass processing and cellulose modification.

09 BIOMASS FUELS

Computational design of highly signalling-active membrane receptors through solvent-mediated allosteric networks

Abstract Protein catalysis and allostery require the atomic-level orchestration and motion of residues and ligand, solvent and protein effector molecules. However, the ability to design protein activity through precise protein–solvent cooperative interactions has not yet been demonstrated. Here we report the design of 14 membrane receptors that catalyse G protein nucleotide exchange through diverse engineered allosteric pathways mediated by cooperative networks of intraprotein, protein–ligand and –solvent molecule interactions. Consistent with predictions, the designed protein activities correlated well with the level of plasticity of the networks at flexible transmembrane helical interfaces. Several designs displayed considerably enhanced thermostability and activity compared with related natural receptors. The most stable and active variant crystallized in an unforeseen signalling-active conformation, in excellent agreement with the design models. The allosteric network topologies of the best designs bear limited similarity to those of natural receptors and reveal an allosteric interaction space larger than previously inferred from natural proteins. The approach should prove useful for engineering proteins with novel complex protein binding, catalytic and signalling activities.

Chemistry

Assessing the potential of deep learning for protein–ligand docking

The effects of ligand binding on protein structures and their in vivo functions carry numerous implications for modern biomedical research and biotechnology development efforts such as drug discovery. Although several deep learning (DL) methods and benchmarks designed for protein–ligand docking have recently been introduced, so far no previous works have systematically studied the behaviour of the latest docking and structure prediction methods within the broadly applicable context of: (1) using predicted (apo) protein structures for docking (for example, for applicability to new proteins); (2) binding multiple (cofactor) ligands concurrently to a given target protein (for example, for enzyme design); and (3) having no previous knowledge of binding pockets (for example, for generalization to unknown pockets). To enable a deeper understanding of the real-world utility of docking methods, we introduce PoseBench, a comprehensive benchmark for broadly applicable protein–ligand docking. PoseBench enables researchers to rigorously and systematically evaluate DL methods for apo-to-holo protein–ligand docking and protein–ligand structure prediction using both primary ligand and multiligand benchmark datasets, the latter of which we introduce to the DL community. Empirically, using PoseBench, we find that: (1) DL cofolding methods generally outperform comparable conventional and DL docking baseline algorithms, but popular methods such as AlphaFold 3 are still challenged by prediction targets with new protein–ligand binding poses; (2) certain DL cofolding methods are highly sensitive to their input multiple sequence alignments, whereas others are not; and (3) DL methods struggle to strike a balance between structural accuracy and chemical specificity when predicting new or multiligand protein targets.

Morehead, Alex [Lawrence Berkeley National Laborat

Design of light- and chemically responsive protein assemblies through host-guest interactions

Host-guest (HG) interactions have been widely used to build responsive materials and molecular machines owing to their inherently dynamic nature, interaction specificity, and responsiveness to diverse stimuli. Here, in this work, we have set out to exploit these advantages of HG chemistry in the design of dynamic protein assemblies, using a C 4 symmetric protein, C98 RhuA, as a building block. We show that a C98 RhuA variant individually modified with β-cyclodextrin (βCD) (host) or azobenzene (guest) functionalities can specifically pair with each other to form highly ordered 1D and 2D assemblies. Association and dissociation of βCD RhuA- azo RhuA assemblies can be controlled by UV and visible light as well as by small-molecule modulators of βCD-azobenzene interactions. Kinetics analyses reveal that βCD RhuA- azo RhuA nanotubes assemble without a nucleation barrier, a highly unusual occurrence for helical supramolecular systems. Taken together, our findings provide a compelling example for achieving complex structural and dynamic outcomes in protein assembly through simple chemical design.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Protein Adhesion on Semi-Fluorinated Polystyrene Surfaces in Static and Dynamic Measurements

Reducing protein adhesion is a critical strategy in fouling-resistant material innovation, with broad applications spanning biomedical and healthcare devices, biosensors, industrial and environmental systems, and other important technological domains. Here, in this study, we elucidated protein adhesion behavior on polystyrene-based thin films by neutron reflectometry (NR) and quartz crystal microbalance with dissipation (QCM-D), using both lysozyme and bovine serum albumin (BSA) as model proteins. To this end, semifluorinated polystyrene thin films with gradient wettability and surface energy were fabricated through dry processing using plasma oxidation and gas-phase deposition. Although it is believed that a fully fluorinated alkyl chain offers extremely low surface energy, thus rejecting foulants, and has been used in many fouling-resistant surface designs, enhanced protein–surface interactions were observed consistently in NR and QCM-D results, due to the combined effects of surface morphology and chemistry. On the contrary, depositing shorter fluorinated silane onto a hydrophilic PS surface contributed to a more homogeneous nanoscale fluorine coating, resulting in less initial protein adsorption and improved surface recovery. Comparative analysis of proteins with different sizes on the nanopatterned semifluorinated surface revealed the influence of molecular characteristics on surface interactions. Lysozyme, being smaller and more compact, showed faster adsorption kinetics and higher surface coverage but largely reversible binding, whereas BSA, with its larger and more flexible structure, formed broader and more stable interfacial layers. This study fills the gap in understanding protein adhesion within the range of hydrophobicity (water contact angle ∼90°), as current strategies often associate with extreme hydrophilic and superhydrophobic surfaces due to hydration or low-surface-energy rejection mechanisms, respectively. It also provides in-depth insights into current combinatorial fouling-resistant surface design.

Yuan, Yue [Oak Ridge National Laboratory (ORNL), O

Highly multiplexed design of an allosteric transcription factor to sense new ligands

Allosteric transcription factors (aTF) regulate gene expression through conformational changes induced by small molecule binding. Although widely used as biosensors, aTFs have proven challenging to design for detecting new molecules because mutation of ligand-binding residues often disrupts allostery. Here, we develop Sensor-seq, a high-throughput platform to design and identify aTF biosensors that bind to non-native ligands. We screen a library of 17,737 variants of the aTF TtgR, a regulator of a multidrug exporter, against six non-native ligands of diverse chemical structures – four derivatives of the cancer therapeutic tamoxifen, the antimalarial drug quinine, and the opiate analog naltrexone – as well as two native flavonoid ligands, naringenin and phloretin. Sensor-seq identifies biosensors for each of these ligands with high dynamic range and diverse specificity profiles. The structure of a naltrexone-bound design shows shape-complementary methionine-aromatic interactions driving ligand specificity. To demonstrate practical utility, we develop cell-free detection systems for naltrexone and quinine. Sensor-seq enables rapid and scalable design of new biosensors, overcoming constraints of natural biosensors.

59 BASIC BIOLOGICAL SCIENCES

A generalized platform for artificial intelligence-powered autonomous enzyme engineering

Proteins are the molecular machines of life with numerous applications in energy, health, and sustainability. However, engineering proteins with desired functions for practical applications remains slow, expensive, and specialist-dependent. Here we report a generally applicable platform for autonomous enzyme engineering that integrates machine learning and large language models with biofoundry automation to eliminate the need for human intervention, judgement, and domain expertise. Requiring only an input protein sequence and a quantifiable way to measure fitness, this automated platform can be applied to engineer a wide array of proteins. As a proof of concept, we engineer Arabidopsis thaliana halide methyltransferase (AtHMT) for a 90-fold improvement in substrate preference and 16-fold improvement in ethyltransferase activity, along with developing a Yersinia mollaretii phytase (YmPhytase) variant with 26-fold improvement in activity at neutral pH. This is accomplished in four rounds over 4 weeks, while requiring construction and characterization of fewer than 500 variants for each enzyme. This platform for autonomous experimentation paves the way for rapid advancements across diverse industries, from medicine and biotechnology to renewable energy and sustainable chemistry.

59 BASIC BIOLOGICAL SCIENCES

Beyond Component Optimization: Systems Level Biodesign for Lanthanide Recovery

Global demand for lanthanides (Ln) is projected to rise sharply over the next decade, while geographically concentrated supply chains and the low concentrations and matrix complexity of secondary feedstocks limit the reach of conventional hydro- and pyrometallurgical separation. Engineered biological systems offer a selective, low-energy alternative, and component-level advances in Ln-binding proteins, AI-designed selective scaffolds, and cell-surface display platforms now rival synthetic chelators in affinity and selectivity. These components, however, remain functionally isolated. Currently, there are no engineered chassis coupling recognition, intracellular trafficking, accumulation, and controlled release into an end-to-end pipeline. Here, we outline how new biodesign strategies and chassis selection must move beyond bioleaching to encompass the full recovery pathway. Achieving this requires integrating AI/ML-guided design, genome-scale build tools, high-throughput phenotyping, and biophysical transport modeling within a Design–Build–Test–Learn cycle tuned to recognition, trafficking, accumulation, and release.

Biodesign