Search NASASearch

SEARCH · Search NASA

Results for “Protein Structure, Secondary”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

PNNL-Predictive-Phenomics/ProCaliper

ProCaliper is a Python library that curates, organizes, and computes protein structure features in a way that easily interfaces with user-provided experimental data. It extracts or computes protein binding site, active site, charge, pLDDT (order/disorder), acid dissociation, protonation, solvent accessible surface area, disulfide bond distance, and protein secondary structure data using precomputed protein structures and publicly available databases. It provides a unified API for integrating additional residue-level data and for visualizing residue features in 3D.

Rozum, Jordan [Pacific Northwest National Lab]

Nanomolar Sensitivity Chirality Transfer from Designed Helical Repeat Proteins to Achiral CdS Nanorods

Bridging chirality across length scales with inorganic− organic hybrid materials is a rapidly expanding area of research. Here, we establish asymmetry at CdS nanorod (NR) interfaces using a designed helical repeat protein bearing four cysteine residues (DHR- 4Cys). Hydrophobic NRs are transferred into water with glycine, and then glycine is displaced by DHR-4Cys, leveraging the thiophilicity of cadmium. Circular dichroism (CD) in the visible, coincident with CdS electronic transitions, reveals a chiral DHR-4Cys:CdS interface. Here, the dissymmetry factor [g-factor = 4.5 × 10 −4 (short NRs) and 5.0 × 10 −4 (long NRs)] is weakly dependent on the NR length, and CD persists at nanomolar protein loadings. Additionally, control experiments demonstrate that DHR-4Cys:CdS NR chirality is dictated by the local coordination of Cys with no significant contribution from the chiral secondary structure of the protein (g-factors of short and long Cys:CdS NRs are 4.8 × 10 −4 and 4.0 × 10 −4 , respectively). Together with far-UV CD and transmission electron microscopy, which provide evidence of preserved protein structure, these results provide the first demonstration that a structurally defined protein can induce chirality in CdS nanocrystals while maintaining protein structure at biologically relevant concentrations.

Cadmium sulfide

Structural and functional insights into the interaction between the bacteriophage T4 DNA processing proteins gp32 and Dda

Abstract Bacteriophage T4 is a classic model system for studying the mechanisms of DNA processing. A key protein in T4 DNA processing is the gp32 single-stranded DNA-binding protein. gp32 has two key functions: it binds cooperatively to single-stranded DNA (ssDNA) to protect it from nucleases and remove regions of secondary structure, and it recruits proteins to initiate DNA processes including replication and repair. Dda is a T4 helicase recruited by gp32, and we purified and crystallized a gp32–Dda–ssDNA complex. The low-resolution structure revealed how the C-terminus of gp32 engages Dda. Analytical ultracentrifugation analyses were consistent with the crystal structure. An optimal Dda binding peptide from the gp32 C-terminus was identified using surface plasmon resonance. The crystal structure of the Dda–peptide complex was consistent with the corresponding interaction in the gp32–Dda–ssDNA structure. A Dda-dependent DNA unwinding assay supported the structural conclusions and confirmed that the bound gp32 sequesters the ssDNA generated by Dda. The structure of the gp32–Dda–ssDNA complex, together with the known structure of the gp32 body, reveals the entire ssDNA binding surface of gp32. gp32–Dda–ssDNA complexes in the crystal are connected by the N-terminal region of one gp32 binding to an adjacent gp32, and this provides key insights into this interaction.

Biochemistry & Molecular Biology

Secondary structure determines electron transport in peptides

Proteins play a key role in biological electron transport, but the structure–function relationships governing the electronic properties of peptides are not fully understood. Despite recent progress, understanding the link between peptide conformational flexibility, hierarchical structures, and electron transport pathways has been challenging. Here, we use single-molecule experiments, molecular dynamics (MD) simulations, nonequilibrium Green’s function-density functional theory (NEGF-DFT), and unsupervised machine learning to understand the role of secondary structure on electron transport in peptides. Our results reveal a two-state molecular conductance behavior for peptides across several different amino acid sequences. MD simulations and Gaussian mixture modeling are used to show that this two-state molecular conductance behavior arises due to the conformational flexibility of peptide backbones, with a high-conductance state arising due to a more defined secondary structure (beta turn or 3 10 helices) and a low-conductance state occurring for extended peptide structures. These results highlight the importance of helical conformations on electron transport in peptides. Conformer selection for the peptide structures is rationalized using principal component analysis of intramolecular hydrogen bonding distances along peptide backbones. Molecular conformations from MD simulations are used to model charge transport in NEGF-DFT calculations, and the results are in reasonable qualitative agreement with experiments. Projected density of states calculations and molecular orbital visualizations are further used to understand the role of amino acid side chains on transport. Overall, our results show that secondary structure plays a key role in electron transport in peptides, which provides broad avenues for understanding the electronic properties of proteins.

Science & Technology - Other Topics

Random heteropolymers as enzyme mimics

Despite successes in replicating the primary–secondary–tertiary structure hierarchy of protein, it remains elusive to synthetically materialize protein functions that are deeply rooted in their chemical, structural and dynamic heterogeneities. We propose that for polymers with backbone chemistries different from that of proteins, programming spatial and temporal projections of sidechains at the segmental level can be effective in replicating protein behaviours; and leveraging the rotational freedom of polymer can mitigate deficiencies in monomeric sequence specificity and achieve behaviour uniformity at the ensemble level. Here, guided by the active site analysis of about 1,300 metalloproteins, we design random heteropolymers (RHPs) as enzyme mimics based on one-pot synthesis. We introduce key monomers as the equivalents of the functional residues of protein and statistically modulate the chemical characteristics of key monomer-containing segments, such as segmental hydrophobicity. The resultant RHPs form pseudo-active sites that provide key monomers with protein-like microenvironments, co-localize substrates with catalytic or cofactor-binding sidechains and catalyse reactions such as oxidation and cyclization of citronellal with isopulegol/menthoglycol selectivity. This RHP design led to enzyme-like materials that can retain catalytic activity under non-biological conditions, are compatible with scalable processing and have expanded substrate scope, including environmentally long-lasting antibiotic tetracycline.

36 MATERIALS SCIENCE

Probing lanmodulin's mechanisms of rare-earth selectivity for protein-based bioseparations

Our BES Separation Science program project, DE-SC0021007, supported our efforts to begin to understand the mechanisms underlying selectivity of a novel class of lanthanide-binding proteins discovered by our laboratory, called lanmodulin (LanM), and to leverage these proteins for recovery and separations of trivalent rare earth elements (REEs) as well as of trivalent actinides. Overall, our work provides important insights into how higher-order (e.g., secondary, tertiary, and quaternary) protein structure modulates selectivity profiles of proteins that bind f-elements highly selectively. These results are important for advancing the concept of protein-based separations of REEs and, perhaps, of other critical minerals.

Lanmodulin, rare earth elements, protein-based met

Sequence-defined structural transitions by calcium-responsive proteins

Biopolymer sequences dictate their functions, and protein-based polymers are a promising platform to establish sequence–function relationships for novel biopolymers. To efficiently explore vast sequence spaces of natural proteins, sequence repetition is a common strategy to tune and amplify specific functions. This strategy is applied to repeats-in-toxin (RTX) proteins with calcium-responsive folding behavior, which stems from tandem repeats of the nonapeptide GGXGXDXUX in which X can be any amino acid and U is a hydrophobic amino acid. To determine the functional range of this nonapeptide, we modified a naturally occurring RTX protein that forms β-roll structures in the presence of calcium. Sequence modifications focused on calcium-binding turns within the repetitive region, including either global substitution of nonconserved residues or complete replacement with tandem repeats of a consensus nonapeptide GGAGXDTLY. Some sequence modifications disrupted the typical transition from intrinsically disordered random coils to folded β rolls, despite conservation of the underlying nonapeptide sequence. Proteins enriched with smaller, hydrophobic amino acids adopted secondary structures in the absence of calcium and underwent structural rearrangements in calcium-rich environments. In contrast, proteins with bulkier, hydrophilic amino acids maintained intrinsic disorder in the absence of calcium. In conclusion, these results indicate a significant role of nonconserved amino acids in calcium-responsive folding, thereby revealing a strategy to leverage sequences in the design of tunable, calcium-responsive biopolymers.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Viroid-like “obelisk” agents are widespread in the ocean and exceed the abundance of RNA viruses in the prokaryotic fraction

Abstract “Obelisks” are recently discovered ribonucleic acid (RNA) viroid-like elements present in diverse environments with no phylogenetic similarity to any known biological agent. obelisks were first identified in the human gut and in a commensal bacterium acting as a replicative host. They have a circular ∼1 kb RNA genome, rod-like secondary structures, and the encoding of a protein superfamily called “Oblins”. We performed a large-scale search of obelisks in the ocean using the Pebblescout program and the transcriptomic Sequence Archive Read databases, revealing the biogeography and abundance of these viroid-like RNA elements. We detected 55 obelisk genomes resulting in 35 marine clusters at the species level. These obelisks were detected in the prokaryotic fraction and to a lesser extent in the eukaryotic fraction, and distributed across all the oceans from surface to mesopelagic including the Arctic, and even in the coldest seawater of Earth beneath the Antarctic Ross Ice Shelf. The obelisk hallmark protein Oblin-1 confirmed by 3D models was found in various marine samples. Some of the detected marine obelisks harbor hammerhead self-cleaving ribozymes in both polarities. In the prokaryotic, but not the eukaryotic, fraction of the Tara Ocean dataset, relative abundance of obelisks calculated by transcriptomic fragment recruitment indicated that they are abundant in marine samples, reaching or even exceeding the relative abundance of the previously discovered uncultured RNA viruses. In conclusion, obelisks are abundant and widespread viroid-like elements that should be included in ocean biogeochemical models.

Environmental Sciences & Ecology

Structures of the mitochondrial single-stranded DNA binding protein with DNA and DNA polymerase γ

Abstract The mitochondrial single-stranded DNA (ssDNA) binding protein, mtSSB or SSBP1, binds to ssDNA to prevent secondary structures of DNA that could impede downstream replication or repair processes. Clinical mutations in the SSBP1 gene have been linked to a range of mitochondrial disorders affecting nearly all organs and systems. Yet, the molecular determinants governing the interaction between mtSSB and ssDNA have remained elusive. Similarly, the structural interaction between mtSSB and other replisome components, such as the mitochondrial DNA polymerase, Polγ, has been minimally explored. Here, we determined a 1.9-Å X-ray crystallography structure of the human mtSSB bound to ssDNA. This structure uncovered two distinct DNA binding sites, a low-affinity site and a high-affinity site, confirmed through site-directed mutagenesis. The high-affinity binding site encompasses a clinically relevant residue, R38, and a highly conserved DNA base stacking residue, W84. Employing cryo-electron microscopy, we confirmed the tetrameric assembly in solution and capture its interaction with Polγ. Finally, we derived a model depicting modes of ssDNA wrapping around mtSSB and a region within Polγ that mtSSB binds.

Biochemistry & Molecular Biology

Standardized Residue Numbering and Secondary Structure Nomenclature in the Class D β-Lactamases

Over 1370 class D β-lactamases are currently known, and they pose a serious threat to the effective treatment of many infectious diseases, particularly in some pathogenic bacteria where evolving carbapenemase activity has been reported. Detailed understanding of their molecular biology, enzymology, and structural biology are critically important, but the lack of a standardized residue numbering scheme and inconsistent secondary structure annotation has made comparative analyses sometimes difficult and cumbersome. Compounding this, in the post-AlphaFold world where we currently find ourselves, an extraordinary wealth of detailed structural information on these enzymes is literally at our fingertips; therefore it is vitally important that a standard numbering system is in place to facilitate the accurate and straightforward analysis of their structures. In conclusion, here we present a residue numbering and secondary structure scheme for the class D enzymes based on the sequence and structure of OXA-48 and apply it to test targets to demonstrate the ease with which it can be used.

59 BASIC BIOLOGICAL SCIENCES

Understanding the role of negative charge in the scaffold of an artificial enzyme for CO 2 hydrogenation on catalysis

Here, we have approached the construction of an artificial enzyme by employing a robust protein scaffold, lactococcal multidrug resistance regulator, LmrR, providing a structured secondary and outer coordination spheres around a molecular rhodium complex, [Rh I (P Et2 N gly P Et2 ) 2 ] - . Previously, we demonstrated a 2–3 fold increase in activity for one Rh-LmrR construct by introducing positive charge in the secondary coordination sphere. In this study, a series of variants was made through site-directed mutagenesis where the negative charge is located in the secondary sphere or outer coordination sphere, with additional variants made with increasingly negative charge in the outer coordination sphere while keeping a positive charge in the secondary sphere. Placing a negative charge in the secondary or outer coordination sphere demonstrates decreased activity by a factor of two compared to the wild-type Rh-LmrR. Interestingly, addition of positive charge in the secondary sphere, with the negatively charged outer coordination sphere restores activity. Vibrational and NMR spectroscopy suggest minimal changes to the electronic density at the rhodium center, regardless of inclusion of a negative or positive charge in the secondary sphere, suggesting another mechanism is impacting catalytic activity, explored in the discussion.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

The Factors Governing Metal Dependence of an Emergent Superfamily of Bimetallic Oxygenases

Metalloenzyme superfamilies are typically defined by their protein scaffolds and active sites. Owing to the high tunability of protein structures, members of a single superfamily can catalyze diverse reactions with the same metallocofactor. Some superfamilies, such as amidohydrolase-related dinuclear oxygenases (AROs), display further versatility by utilizing multiple metallocofactors. We have shown that certain AROs catalyze monooxygenation reactions with diiron, dimanganese, and/or mixed manganese−iron cofactors, but the molecular factors governing the selection of a particular cofactor remain unknown, and the extent of this superfamily in biology is unclear. Here, we report bioinformatic analyses that expand the ARO superfamily to approximately 17,000 unique UniProt sequences, far exceeding the number of previously characterized enzymes. Through the integration of structural, spectroscopic, and thermodynamic analyses of representative proteins with a bioinformatic pipeline that identifies key secondary- and tertiary-sphere residues, we can predict in silico the metal preference for the majority of reported ARO sequences. These annotations were validated via the characterization of multiple new AROs, including ones implicated in key oxidative steps of natural product biosyntheses. This study establishes the key structure−function relationships governing metal preferences in AROs and highlights their vastly underappreciated role in myriad biological processes.

Liu, Chang [University of California, Berkeley, CA

Ca X ML: Chemistry‐informed machine learning explains mutual changes between protein conformations and calcium ions in calcium‐binding proteins using structural and topological features

Proteins' flexibility is a feature in communicating changes in cell signaling instigated by binding with secondary messengers, such as calcium ions, associated with the coordination of muscle contraction, neurotransmitter release, and gene expression. When binding with the disordered parts of a protein, calcium ions must balance their charge states with the shape of calcium-binding proteins and their versatile pool of partners depending on the circumstances they transmit. Accurately determining the ionic charges of those ions is essential for understanding their role in such processes. However, it is unclear whether the limited experimental data available can be effectively used to train models to accurately predict the charges of calcium-binding protein variants. Here, we developed a chemistry-informed, machine-learning algorithm that implements a game theoretic approach to explain the output of a machine-learning model without the prerequisite of an excessively large database for high-performance prediction of atomic charges. We used the ab initio electronic structure data representing calcium ions and the structures of the disordered segments of calcium-binding peptides with surrounding water molecules to train several explainable models. Network theory was used to extract the topological features of atomic interactions in the structurally complex data dictated by the coordination chemistry of a calcium ion, a potent indicator of its charge state in protein. Our design created a computational tool of Ca X ML, which provided a framework of explainable machine learning model to annotate ionic charges of calcium ions in calcium-binding proteins in response to the chemical changes in an environment. Our framework will provide new insights into protein design for engineering functionality based on the limited size of scientific data in a genome space.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Energy metric prediction for double insertion mutants via the RoseNet deep learning framework

Studying the structural and functional implications of protein mutations is an important task in computational biology and bioinformatics. We leverage our previously proposed RoseNet neural network architecture to predict energy metrics of proteins with double amino acid insertions or deletions (InDels). We train models on previously generated benchmark datasets containing the exhaustive double InDel mutations for three proteins, as well as an additional three proteins for which ∼145k random mutants, each with two InDels, have been generated. We expand on our previous work by evaluating three additional proteins and analyzing domain features that impact the prediction capabilities of RoseNet. These features include InDels into secondary structures and the solvent accessible surface area (SASA) scores of the residues. We uncover further evidence to support that RoseNet has a higher proficiency of generalizing to unseen residue combinations than unseen insertion positions. We also observe that RoseNet produces higher-quality predictions when inserting into a β-sheet over an α-helix. Additionally, when the insertions fall in an area of high SASA, RoseNet often displays better performance than inserting into areas of low SASA.

59 BASIC BIOLOGICAL SCIENCES

Correlating Protein Dynamics and Catalytic Activity of a Model Hydrogenase Using Paramagnetic and Biological Nuclear Magnetic Resonance Spectroscopy

Rational catalyst design remains a significant challenge, with electronic structure, steric, and electrostatic effects known to contribute to activity. Recently, dynamics has been recognized as another factor that impacts catalysis, though identifying and predicting these effects has remained out of reach. Nickel-substituted rubredoxin (NiRd), a protein-based mimic of a hydrogenase enzyme, serves as a model catalytic system in which dynamics can be systematically investigated with respect to activity. While over 30 secondary-sphere mutants of NiRd have been shown to be catalytically active, no significant correlation was observed between the rates and catalytic overpotential or electronic structure, prompting questions about the protein-derived factors that modulate activity. Here, in this work, NMR spectroscopy was used to investigate the roles of substrate accessibility, protein dynamics, and protein stability in controlling catalysis. Significant paramagnetic effects from the nickel center (S = 1) isolate the methylene proton resonances of the metal-coordinating cysteine residues. The sensitivity of resonance positions and linewidths to local environment offers an opportunity to study dynamical molecular changes around the metal center with high resolution. Machine learning algorithms were employed to identify correlations between the catalytic activity and the paramagnetic NMR spectra. These analyses revealed spectroscopic features of specific cysteine protons that report on catalytic overpotential and increased turnover rates, which are further supported by the results obtained using high-field NMR techniques. Collectively, these studies indicate the potential for multifrequency NMR techniques to resolve key contributors to catalytic activity and highlight the importance of local and outer-sphere dynamics.

Protein Engineering

Matrix Metalloproteinases as Candidate Antigenic Determinants for Anti‐Tumor Autoantibodies in Human Ovarian Cancer: A Post Hoc Analysis

Circulating antibodies in patients with cancer can facilitate the identification of accessible epitopes on autoantigens expressed by tumors. To identify previously unrecognized protein targets in ovarian cancer, we computationally assessed a heptapeptide consensus motif (VPELGHE, flanked by two cysteine residues yielding a cyclic nonapeptide under oxidizing conditions) previously discovered via phage display-based epitope mapping of autoantibodies in patients. Eight proteins associated with ovarian cancer encompass amino acid sequences similar to the consensus motif and were, therefore, considered as candidate native autoantigens. Among these candidate targets, however, matrix metalloproteinase 14 (MMP14) demonstrates gene expression that is both high and negatively correlated with survival in ovarian cancer patient cohorts. MMP14 protein levels are also stable in tumor versus non-tumor tissues. Moreover, the corresponding heptapeptide mimic in MMP14 occurs within an α-helical secondary structural element observed in its catalytic domain. These findings demonstrate that a subset of patient-derived autoantibodies may interact with a previously unknown antigenic epitope found in MMP14 and other MMPs, thereby providing opportunities for the development of new targeted agents.

Biochemistry & Molecular Biology

Distinct binding conformations of epinephrine with α- and β-adrenergic receptors

Abstract Agonists targeting α 2 -adrenergic receptors (ARs) are used to treat diverse conditions, including hypertension, attention-deficit/hyperactivity disorder, pain, panic disorders, opioid and alcohol withdrawal symptoms, and cigarette cravings. These receptors transduce signals through heterotrimeric Gi proteins. Here, we elucidated cryo-EM structures that depict α 2A -AR in complex with Gi proteins, along with the endogenous agonist epinephrine or the synthetic agonist dexmedetomidine. Molecular dynamics simulations and functional studies reinforce the results of the structural revelations. Our investigation revealed that epinephrine exhibits different conformations when engaging with α-ARs and β-ARs. Furthermore, α 2A -AR and β 1 -AR (primarily coupled to Gs, with secondary associations to Gi) were compared and found to exhibit different interactions with Gi proteins. Notably, the stability of the epinephrine–α 2A -AR–Gi complex is greater than that of the dexmedetomidine–α 2A -AR–Gi complex. These findings substantiate and improve our knowledge on the intricate signaling mechanisms orchestrated by ARs and concurrently shed light on the regulation of α-ARs and β-ARs by epinephrine.

Biochemistry & Molecular Biology

Protein N -Glycans in Healthy and Sclerotic Glomeruli in Diabetic Kidney Disease

Diabetes is expected to directly affect renal glycosylation; yet to date, there has not been a comprehensive evaluation of alterations in N-glycan composition in the glomeruli of patients with diabetic kidney disease (DKD). Here, we used untargeted mass spectrometry imaging to identify N-glycan structures in healthy and sclerotic glomeruli in formalin-fixed paraffin-embedded sections from needle biopsies of five patients with DKD and three healthy kidney samples. Regional proteomics was performed on glomeruli from additional biopsies from the same patients to compare the abundances of enzymes involved in glycosylation. Secondary analysis of single-nucleus RNA sequencing (snRNAseq) data were used to inform on transcript levels of glycosylation machinery in different cell types and states. We detected 120 N-glycans, and among them, we identified 12 of these protein post-translated modifications that were significantly increased in glomeruli. All glomeruli-specific N-glycans contained an N-acetyllactosamine epitope. Five N-glycan structures were highly discriminant between sclerotic and healthy glomeruli. Sclerotic glomeruli had an additional set of glycans lacking fucose linked to their core, and they did not show tetra-antennary structures that were common in healthy glomeruli. Orthogonal omics analyses revealed lower protein abundance and lower gene expression involved in synthesizing fucosylated and branched N-glycans in sclerotic podocytes. In snRNAseq and regional proteomics analyses, we observed that genes and/or proteins involved in sialylation and N-acetyllactosamine synthesis were also downregulated in DKD glomeruli, but this alteration remained undetectable by our spatial N-glycomics assay. Integrative spatial glycomics, proteomics, and transcriptomics revealed protein N-glycosylation characteristic of sclerotic glomeruli in DKD.

60 APPLIED LIFE SCIENCES