Search NASA⌕ Search

SEARCH · Search NASA

Results for “repeat protein”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Evolution and folding of repeat proteins

Repeat proteins are made with tandem copies of similar amino acid stretches that fold into elongated architectures. These proteins constitute excellent model systems to investigate how evolution relates to structure, folding, and function. Here, we propose a scheme to map evolutionary information at the sequence level to a coarse-grained model for repeat-protein folding and use it to investigate the folding of thousands of repeat proteins. We model the energetics by a combination of an inverse Potts-model scheme with an explicit mechanistic model of duplications and deletions of repeats to calculate the evolutionary parameters of the system at the single-residue level. These parameters are used to inform an Ising-like model that allows for the generation of folding curves, apparent domain emergence, and occupation of intermediate states that are highly compatible with experimental data in specific case studies. We analyzed the folding of thousands of natural Ankyrin repeat proteins and found that a multiplicity of folding mechanisms are possible. Fully cooperative all-or-none transitions are obtained for arrays with enough sequence-similar elements and strong interactions between them, while noncooperative element-by-element intermittent folding arose if the elements are dissimilar and the interactions between them are energetically weak. Additionally, we characterized nucleation-propagation and multidomain folding mechanisms. We show that the global stability and cooperativity of the repeating arrays can be predicted from simple sequence scores.

Ezequiel A. Galpern↗

Resolving the fine structure in the energy landscapes of repeat proteins

Ankyrin (ANK) repeat proteins are coded by tandem occurrences of patterns with around 33 amino acids. They often mediate protein–protein interactions in a diversity of biological systems. These proteins have an elongated non-globular shape and often display complex folding mechanisms. This work investigates the energy landscape of representative proteins of this class made up of 3, 4 and 6 ANK repeats using the energy-landscape visualisation method (ELViM). By combining biased and unbiased coarse-grained molecular dynamics AWSEM simulations that sample conformations along the folding trajectories with the ELViM structure-based phase space, one finds a three-dimensional representation of the globally funnelled energy surface. In this representation, it is possible to delineate distinct folding pathways. We show that ELViMs can project, in a natural way, the intricacies of the highly dimensional energy landscapes encoded by the highly symmetric ankyrin repeat proteins into useful low-dimensional representations. These projections can discriminate between multiplicities of specific parallel folding mechanisms that otherwise can be hidden in oversimplified depictions.

Murilo N. Sanches↗

De novo design of knotted tandem repeat proteins

De novo protein design methods can create proteins with folds not yet seen in nature. These methods largely focus on optimizing the compatibility between the designed sequence and the intended conformation, without explicit consideration of protein folding pathways. Deeply knotted proteins, whose topologies may introduce substantial barriers to folding, thus represent an interesting test case for protein design. Here we report our attempts to design proteins with trefoil (3 1 ) and pentafoil (5 1 ) knotted topologies. We extended previously described algorithms for tandem repeat protein design in order to construct deeply knotted backbones and matching designed repeat sequences (N = 3 repeats for the trefoil and N = 5 for the pentafoil). We confirmed the intended conformation for the trefoil design by X ray crystallography, and we report here on this protein’s structure, stability, and folding behaviour. The pentafoil design misfolded into an asymmetric structure (despite a 5-fold symmetric sequence); two of the four repeat-repeat units matched the designed backbone while the other two diverged to form local contacts, leading to a trefoil rather than pentafoil knotted topology. Our results also provide insights into the folding of knotted proteins.

59 BASIC BIOLOGICAL SCIENCES↗

Nanomolar Sensitivity Chirality Transfer from Designed Helical Repeat Proteins to Achiral CdS Nanorods

Bridging chirality across length scales with inorganic− organic hybrid materials is a rapidly expanding area of research. Here, we establish asymmetry at CdS nanorod (NR) interfaces using a designed helical repeat protein bearing four cysteine residues (DHR- 4Cys). Hydrophobic NRs are transferred into water with glycine, and then glycine is displaced by DHR-4Cys, leveraging the thiophilicity of cadmium. Circular dichroism (CD) in the visible, coincident with CdS electronic transitions, reveals a chiral DHR-4Cys:CdS interface. Here, the dissymmetry factor [g-factor = 4.5 × 10 −4 (short NRs) and 5.0 × 10 −4 (long NRs)] is weakly dependent on the NR length, and CD persists at nanomolar protein loadings. Additionally, control experiments demonstrate that DHR-4Cys:CdS NR chirality is dictated by the local coordination of Cys with no significant contribution from the chiral secondary structure of the protein (g-factors of short and long Cys:CdS NRs are 4.8 × 10 −4 and 4.0 × 10 −4 , respectively). Together with far-UV CD and transmission electron microscopy, which provide evidence of preserved protein structure, these results provide the first demonstration that a structurally defined protein can induce chirality in CdS nanocrystals while maintaining protein structure at biologically relevant concentrations.

Cadmium sulfide↗

Tetratricopeptide repeat protein SlREC2 positively regulates cold tolerance in tomato

Abstract Cold stress is a key environmental constraint that dramatically affects the growth, productivity, and quality of tomato (Solanum lycopersicum); however, the underlying molecular mechanisms of cold tolerance remain poorly understood. In this study, we identified REDUCED CHLOROPLAST COVERAGE 2 (SlREC2) encoding a tetratricopeptide repeat protein that positively regulates tomato cold tolerance. Disruption of SlREC2 largely reduced abscisic acid (ABA) levels, photoprotection, and the expression of C-REPEAT BINDING FACTOR (CBF)-pathway genes in tomato plants under cold stress. ABA deficiency in the notabilis (not) mutant, which carries a mutation in 9-CIS-EPOXYCAROTENOID DIOXYGENASE 1 (SlNCED1), strongly inhibited the cold tolerance of SlREC2-silenced plants and empty vector control plants and resulted in a similar phenotype. In addition, foliar application of ABA rescued the cold tolerance of SlREC2-silenced plants, which confirms that SlNCED1-mediated ABA accumulation is required for SlREC2-regulated cold tolerance. Strikingly, SlREC2 physically interacted with β-RING CAROTENE HYDROXYLASE 1b (SlBCH1b), a key regulatory enzyme in the xanthophyll cycle. Disruption of SlBCH1b severely impaired photoprotection, ABA accumulation, and CBF-pathway gene expression in tomato plants under cold stress. Taken together, this study reveals that SlREC2 interacts with SlBCH1b to enhance cold tolerance in tomato via integration of SlNCED1-mediated ABA accumulation, photoprotection, and the CBF-pathway, thus providing further genetic knowledge for breeding cold-resistant tomato varieties.

Zhang, Ying (ORCID:0000000169841164)↗

A Trichomonas vaginalis C2-XYPPX-repeat protein with a structured C2 domain displaying dampened flexibility upon binding calcium

C2 domains are ubiquitous membrane-binding modules of ∼130 residues in eukaryotes that are often associated with proteins involved in membrane trafficking and lipid modification. The genome of Trichomonas vaginalis, the most common, non-viral, sexually transmitted human pathogen, encodes eight genes that contain a N-terminal C2 module linked to a XYPPX-repeat domain of more than four XYPPX repeats (C2-XYPPX). While the function of the XYPPX-repeat domain remains unknown, its multiple association with C2 domains in T. vaginalis suggests it is important. Here, the C2 domain from one of these C2-XYPPX-repeat proteins, Tv-C2-1, was structurally and physically characterized using X-ray crystallography and NMR spectroscopy. The crystal structure for Tv-C2-1 shows that this domain shares a fold common to all C2 domains, a compact Greek-key motif composed of eight anti-parallel β-strands in the type-2 topology. An NMR chemical shift perturbation study with Ca 2+ showed that Tv-C2-1 bound two Ca 2+ atoms primarily via two loops (loop-1 and loop-3) on the predicted calcium binding face of the protein with K d s of 58.0 ± 0.1 μM and 232 ± 6 μM. Estimations of the overall rotational correlation time, τ c , in the apo (11.1 ns) and Ca 2+ -bound (9.2 ns) state suggests the protein becomes more compact upon Ca 2+ binding, consistent with a decrease in dynamics in loop-3 and marginally in loop-1 suggested by amide 15 N heteronuclear steady-state { 1 H}- 15 N NOEs. Showing Tv-C2-1 binds calcium and adopts a compact Greek-key motif structure, two primary features of C2 domains, suggests understanding the function of the XYPPX-repeat domain may be warranted.

NMR spectroscopy↗

Size and Structure of the Sequence Space of Repeat Proteins

The coding space of protein sequences is shaped by evolutionary constraints set by requirements of function and stability. We show that the coding space of a given protein family— the total number of sequences in that family—can be estimated using models of maximum entropy trained on multiple sequence alignments of naturally occurring amino acid sequences. We analyzed and calculated the size of three abundant repeat proteins families, whose members are large proteins made of many repetitions of conserved portions of *30 amino acids. While amino acid conservation at each position of the alignment explains most of the reduction of diversity relative to completely random sequences, we found that correlations between amino acid usage at different positions significantly impact that diversity. We quantified the impact of different types of correlations, functional and evolutionary, on sequence diversity. Analysis of the detailed structure of the coding space of the families revealed a rugged landscape, with many local energy minima of varying sizes with a hierarchical structure, reminiscent of frustrated energy landscapes of spin glass in physics. This clustered structure indicates a multiplicity of subtypes within each family and suggests new strategies for protein design.

Jacopo Marchi↗

Hallucination of closed repeat proteins containing central pockets

In pseudocyclic proteins, such as TIM barrels, β barrels, and some helical transmembrane channels, a single subunit is repeated in a cyclic pattern, giving rise to a central cavity that can serve as a pocket for ligand binding or enzymatic activity. Inspired by these proteins, we devised a deep-learning-based approach to broadly exploring the space of closed repeat proteins starting from only a specification of the repeat number and length. Biophysical data for 38 structurally diverse pseudocyclic designs produced in Escherichia coli are consistent with the design models, and the three crystal structures we were able to obtain are very close to the designed structures. Docking studies suggest the diversity of folds and central pockets provide effective starting points for designing small-molecule binders and enzymes.

59 BASIC BIOLOGICAL SCIENCES↗

High-density binding to Plasmodium falciparum circumsporozoite protein repeats by inhibitory antibody elicited in mouse with human immunoglobulin repertoire

Antibodies targeting the human malaria parasite Plasmodium falciparum circumsporozoite protein (PfCSP) can prevent infection and disease. PfCSP contains multiple central repeating NANP motifs; some of the most potent anti-infective antibodies against malaria bind to these repeats. Multiple antibodies can bind the repeating epitopes concurrently by engaging into homotypic Fab-Fab interactions, which results in the ordering of the otherwise largely disordered central repeat into a spiral. Here, we characterize IGHV3-33/IGKV1-5-encoded monoclonal antibody (mAb) 850 elicited by immunization of transgenic mice with human immunoglobulin loci. mAb 850 binds repeating NANP motifs with picomolar affinity, potently inhibits Plasmodium falciparum (Pf) in vitro and, when passively administered in a mouse challenge model, reduces liver burden to a similar extent as some of the most potent anti-PfCSP mAbs yet described. Like other IGHV3-33/IGKV1-5-encoded anti-NANP antibodies, mAb 850 primarily utilizes its HCDR3 and germline-encoded aromatic residues to recognize its core NANP motif. Biophysical and cryo-electron microscopy analyses reveal that up to 19 copies of Fab 850 can bind the PfCSP repeat simultaneously, and extensive homotypic interactions are observed between densely-packed PfCSP-bound Fabs to indirectly improve affinity to the antigen. Together, our study expands on the molecular understanding of repeat-induced homotypic interactions in the B cell response against PfCSP for potently protective mAbs against Pf infection.

59 BASIC BIOLOGICAL SCIENCES↗

Monomer-scale design of functional protein polymers using consensus repeat sequences

Protein-based polymers possess chemically defined sequences that can encode diverse properties and functions into a new class of biopolymeric materials. However, sequence variation that emerges from evolution can obscure the sequence–function relationships of naturally derived polymers. One strategy to clarify these relationships is to identify common sequences between proteins with similar functions. These conserved sequences often emerge from repeat proteins, and “consensus repeat sequences” provide a convenient platform for systematic investigations of biopolymer sequence–property relationships. In this review, we highlight recent approaches to engineer tunable polymeric materials using monomer-scale design of consensus repeat proteins. Here, we explore established and emerging protein-based materials with mechanical resilience, thermodynamic phase behavior, chemical responsiveness, biomolecular transport, and hierarchical structure. Overall, recent advances in the monomer-scale design of repetitive protein polymers present exciting fundamental and translational opportunities for polymer scientists and engineers.

36 MATERIALS SCIENCE↗

Tissue distribution and subcellular localization of the family of Kidney Ankyrin Repeat Domain (KANK) proteins

Kidney Ankyrin Repeat-containing Proteins (KANKs) comprise a family of four evolutionary conserved proteins (KANK1 to 4) that localize to the belt of mature focal adhesions (FAs) where they regulate integrin-mediated adhesion, actomyosin contractility, and link FAs to the cortical microtubule stabilization complex (CMSC). The human KANK proteins were first identified in kidney and have been associated with kidney cancer and nephrotic syndrome. Here, we report the distributions and subcellular localizations of the four Kank mRNAs and proteins in mouse tissues. We found that the KANK family members display distinct and rarely overlapping expression patterns. Whereas KANK1 is expressed at the basal side of epithelial cells of all tissues tested, KANK2 expression is mainly observed at the plasma membrane and/or cytoplasm of mesenchymal cells and KANK3 exclusively in vascular and lymphatic endothelial cells. KANK4 shows the least widespread expression pattern and when present, overlaps with KANK2 in contractile cells, such as smooth muscle cells and pericytes. Our findings show that KANKs are widely expressed in a cell type-specific manner, which suggests that they have cell- and tissue-specific functions.

60 APPLIED LIFE SCIENCES↗

De novo design of protein homodimers containing tunable symmetric protein pockets

Function follows form in biology, and the binding of small molecules requires proteins with pockets that match the shape of the ligand. For design of binding to symmetric ligands, protein homo-oligomers with matching symmetry are advantageous as each protein subunit can make identical interactions with the ligand. Here, we describe a general approach to designing hyperstable C2 symmetric proteins with pockets of diverse size and shape. We first designed repeat proteins that sample a continuum of curvatures but have low helical rise, then docked these into C2 symmetric homodimers to generate an extensive range of C2 symmetric cavities. We used this approach to design thousands of C2 symmetric homodimers, and characterized 101 of them experimentally. Of these, the geometry of 31 were confirmed by small angle X-ray scattering and 2 were shown by crystallographic analyses to be in close agreement with the computational design models. These scaffolds provide a rich set of starting points for binding a wide range of C2 symmetric compounds.

59 BASIC BIOLOGICAL SCIENCES↗

Genomic and phenotypic comparison of two variants of multidrug-resistant Salmonella enterica serovar Heidelberg isolated during the 2015–2017 multi-state outbreak in cattle

Salmonella enterica subspecies enterica serovar Heidelberg (Salmonella Heidelberg) has caused several multistate foodborne outbreaks in the United States, largely associated with the consumption of poultry. However, a 2015–2017 multidrug-resistant (MDR) Salmonella Heidelberg outbreak was linked to contact with dairy beef calves. Traceback investigations revealed calves infected with outbreak strains of Salmonella Heidelberg exhibited symptoms of disease frequently followed by death from septicemia. To investigate virulence characteristics of Salmonella Heidelberg as a pathogen in bovine, two variants with distinct pulse-field gel electrophoresis (PFGE) patterns that differed in morbidity and mortality during the multistate outbreak were genotypically and phenotypically characterized and compared. Strain SX 245 with PFGE pattern JF6X01.0523 was identified as a dominant and highly pathogenic variant causing high morbidity and mortality in affected calves, whereas strain SX 244 with PFGE pattern JF6X01.0590 was classified as a low pathogenic variant causing less morbidity and mortality. Comparison of whole-genome sequences determined that SX 245 lacked ~200 genes present in SX 244, including genes associated with the IncI1 plasmid and phages; SX 244 lacked eight genes present in SX 245 including a second YdiV Anti-FlhC(2)FlhD(4) factor, a lysin motif domain containing protein, and a pentapeptide repeat protein. RNA-sequencing revealed fimbriae-related, flagella-related, and chemotaxis genes had increased expression in SX 245 compared to SX 244. Furthermore, SX 245 displayed higher invasion of human and bovine epithelial cells than SX 244. These data suggest that the presence and up-regulation of genes involved in type 1 fimbriae production, flagellar regulation and biogenesis, and chemotaxis may play a role in the increased pathogenicity and host range expansion of the Salmonella Heidelberg isolates involved in the bovine-related outbreak.

59 BASIC BIOLOGICAL SCIENCES↗