Search NASA⌕ Search

SEARCH · Search NASA

Results for “Biological and medical sciences, Computer science”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Computationally restoring the potency of a clinical antibody against Omicron

The COVID-19 pandemic underscored the promise of monoclonal antibody-based prophylactic and therapeutic drugs and revealed how quickly viral escape can curtail effective options. When the SARS-CoV-2 Omicron variant emerged in 2021, many antibody drug products lost potency, including Evusheld and its constituent, cilgavimab. Cilgavimab, like its progenitor COV2-2130, is a class 3 antibody that is compatible with other antibodies in combination4 and is challenging to replace with existing approaches. Rapidly modifying such high-value antibodies to restore efficacy against emerging variants is a compelling mitigation strategy. We sought to redesign and renew the efficacy of COV2-2130 against Omicron BA.1 and BA.1.1 strains while maintaining efficacy against the dominant Delta variant. Here we show that our computationally redesigned antibody, 2130-1-0114-112, achieves this objective, simultaneously increases neutralization potency against Delta and subsequent variants of concern, and provides protection in vivo against the strains tested: WA1/2020, BA.1.1 and BA.5. Deep mutational scanning of tens of thousands of pseudovirus variants reveals that 2130-1-0114-112 improves broad potency without increasing escape liabilities. Our results suggest that computational approaches can optimize an antibody to target multiple escape variants, while simultaneously enriching potency. Our computational approach does not require experimental iterations or pre-existing binding data, thus enabling rapid response strategies to address escape variants or lessen escape vulnerabilities.

60 APPLIED LIFE SCIENCES↗

LDRD FY25 Program Overview

As Lawrence Livermore National Laboratory’s (LLNL’s) Laboratory Directed Research and Development (LDRD) program enters its fifth decade of leading-edge research and development, its impact and importance have never been stronger. The program continues to advance strategic investments in pioneering science, technology, and engineering, ensuring LLNL will be ready to deliver on our mission as it evolves over the coming decades. Investing in LDRD research, and the people who perform this critical work, gives LLNL the ability to sustain our role as a leader in the Department of Energy and National Nuclear Security Administration enterprise. The LDRD program enables high-risk, high-payoff research that anticipates emerging threats and future mission needs. By nurturing the ingenuity of the Lab’s greatest asset, its people, LDRD funding advances not only our research but also grows and nurtures our workforce: engaging future innovators with student mentoring, challenging postdoctoral researchers to apply their skills to support national security, and strengthening the leadership skills of early career staff. This annual report documents how LDRD investments advance LLNL’s science, technology, and engineering across our mission space. To assess LDRD’s impact we track both short and long-term metrics such as peer-reviewed publications, number of students, or professional fellows. In addition to reviewing these metrics, I encourage you to delve deeper into the breadth of science and technology that illustrate the strategic value of this research portfolio. For instance, a recent exploratory research project used advanced manufacturing to construct miniaturized three-dimensional ion traps for a quantum computer with reduced quantum error rates to enable applications that address national security missions and support basic science. Another project has delved into studying detonation by examining deflagration to enhance the safety and security of the nuclear weapons stockpile. LDRD researchers are also deploying AI agents on two of the world’s most powerful supercomputers to automate and accelerate inertial confinement fusion experiments. Other teams are delivering more accurate optical constants to enable improved validation for aluminum to advance atomic and molecular physics models. LDRD-driven discoveries of how metals deform under extreme conditions strengthen our ability to model and design materials for demanding national security environments. National security challenges are increasingly complex and continuously evolving. LDRD focuses our most innovative science and technology on these challenges, ensuring the Laboratory is developing creative, forward-leaning solutions for our nation and the world. The following pages feature highlights of published scientific advances, patents, and honors that stem from LDRD investments. As you read this report, I hope you will understand how these investments position the Laboratory, and our partners, to meet the demands of the decades ahead.

36 MATERIALS SCIENCE↗

WiDS Livermore Datathon 2025

The WiDS Datathon 2025 requires participants to build a model to predict both an individual’s sex and their ADHD diagnosis using functional brain imaging data of children and adolescents and their socio-demographic, emotions, and parenting information. The task is to create a multi-outcome model to predict two target variables: 1) ADHD (1=yes or 0=no) and 2) female (1=yes or 0=no).

59 BASIC BIOLOGICAL SCIENCES↗

HDBind: encoding of molecular structure with hyperdimensional binary representations

Traditional methods for identifying “hit” molecules from a large collection of potential drug-like candidates rely on biophysical theory to compute approximations to the Gibbs free energy of the binding interaction between the drug and its protein target. These approaches have a significant limitation in that they require exceptional computing capabilities for even relatively small collections of molecules. Increasingly large and complex state-of-the-art deep learning approaches have gained popularity with the promise to improve the productivity of drug design, notorious for its numerous failures. However, as deep learning models increase in their size and complexity, their acceleration at the hardware level becomes more challenging. Hyperdimensional Computing (HDC) has recently gained attention in the computer hardware community due to its algorithmic simplicity relative to deep learning approaches. The HDC learning paradigm, which represents data with high-dimension binary vectors, allows the use of low-precision binary vector arithmetic to create models of the data that can be learned without the need for the gradient-based optimization required in many conventional machine learning and deep learning methods. This algorithmic simplicity allows for acceleration in hardware that has been previously demonstrated in a range of application areas (computer vision, bioinformatics, mass spectrometery, remote sensing, edge devices, etc.). To the best of our knowledge, our work is the first to consider HDC for the task of fast and efficient screening of modern drug-like compound libraries. We also propose the first HDC graph-based encoding methods for molecular data, demonstrating consistent and substantial improvement over previous work. We compare our approaches to alternative approaches on the well-studied MoleculeNet dataset and the recently proposed LIT-PCBA dataset derived from high quality PubChem assays. We demonstrate our methods on multiple target hardware platforms, including Graphics Processing Units (GPUs) and Field Programmable Gate Arrays (FPGAs), showing at least an order of magnitude improvement in energy efficiency versus even our smallest neural network baseline model with a single hidden layer. Our work thus motivates further investigation into molecular representation learning to develop ultra-efficient pre-screening tools. We make our code publicly available at https://github.com/LLNL/hdbind.

59 BASIC BIOLOGICAL SCIENCES↗

Addressing the dynamic nature of reference data: a new nucleotide database for robust metagenomic classification

Accurate metagenomic classification relies on comprehensive, up-to-date, and validated reference databases. While the NCBI BLAST Nucleotide (nt) database, encompassing a vast collection of sequences from all domains of life, represents an invaluable resource, its massive size—currently exceeding 10 12 nucleotides—and exponential growth pose significant challenges for researchers seeking to maintain current nt-based indices for metagenomic classification. Recognizing that no current nt-based indices exist for the widely used Centrifuge classifier, and the last public version currently available was released in 2018, we addressed this critical gap by leveraging advanced high-performance computing resources. We present new Centrifuge-compatible nt databases, meticulously constructed using a novel pipeline incorporating different quality control measures, including reference decontamination and filtering. These measures demonstrably reduce spurious classifications, as shown through our reanalysis of published metagenomic data where Plasmodium annotations were dramatically reduced using our decontaminated database, highlighting how database quality can significantly impact research conclusions. Through temporal comparisons, we also reveal how our approach minimizes inconsistencies in taxonomic assignments stemming from asynchronous updates between public sequence and taxonomy databases. These discrepancies are particularly evident in taxa such as Listeria monocytogenes and Naegleria fowleri, where classification accuracy varied significantly across database versions. These new databases, made available as pre-built Centrifuge indexes, respond to the need for an open, robust, nt-based pipeline for taxonomic classification in metagenomics. Applications such as environmental metagenomics, forensics, and clinical metagenomics, which require comprehensive taxonomic coverage, will benefit from this resource. Our work highlights the importance of treating reference databases as dynamic entities, subject to ongoing quality control and validation akin to software development best practices. This approach is crucial for ensuring accuracy and reliability of metagenomic analysis, especially as databases continue to expand in size and complexity.

59 BASIC BIOLOGICAL SCIENCES↗

Generating Protein Structures for Pathway Discovery Using Deep Learning

Resolving the intricate details of biological phenomena at the molecular level is fundamentally limited by both length- and time scales that can be probed experimentally. Molecular dynamics (MD) simulations at various scales are powerful tools frequently employed to offer valuable biological insights beyond experimental resolution. However, while it is relatively simple to observe long-lived, stable configurations of, for example, proteins, at the required spatial resolution, simulating the more interesting rare transitions between such states often takes orders of magnitude longer than what is feasible even on the largest supercomputers available today. One common aspect of this challenge is pathway discovery, where the start and end states of a scientific phenomenon are known or can be approximated, but the mechanistic details in between are unknown. Here, we propose a representation-learning-based solution that uses interpolation and extrapolation in an abstract representation space to synthesize potential transition states, which are automatically validated using MD simulations. The new simulations of the synthesized transition states are subsequently incorporated into the representation learning, leading to an iterative framework for targeted path sampling. Our approach is demonstrated by recovering the transition of a RAS-RAF protein domain (CRD) from membrane-free to interacting with the membrane using coarse-grain MD simulations.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

SLAB: simultaneous labeling and binding affinity prediction for protein–ligand structures

Machine learning models are often used as scoring functions to predict the binding affinity of a protein–ligand complex. These models are trained with limited amounts of data with experimentally measured binding affinity values. A large number of compounds are labeled inactive through single-concentration screens without measuring binding affinities. These inactive compounds, along with the active ones, can be used to train binary classification models, while regression models are trained using compounds with binding affinities only. However, the classification and regression tasks are often handled separately, without sharing the learned feature representations. In this paper, we propose a novel model architecture that jointly performs regression and classification objectives, aiming to maximize data utilization and improve predictive performance by leveraging two complementary tasks. In our setup, the regression yields the binding affinity, whereas the classification task yields the label as active or inactive. We demonstrate our method using PDBbind, the standard 3D structure database, as well as a dataset of flavivirus protease compounds with binding affinity data. Our experiments show that the new joint training strategy improves the accuracy of the model, increasing applicability in various practical drug screening scenarios.

Biological and medical sciences↗

FY24 LLNL Laboratory Directed Research and Development (LDRD) Project Accomplishments

Large duration energy storage is the key to couple renewable energy generation with power supply. This project aimed to advance the durability of a low-cost and eco-friendly flow battery technology based on iron chemistry to speed up the technology readiness level rapidly and radically. A novel device which is called “an artificial kidney” was innovated through disruptive research to integrate with the flow battery for rebalancing capacity. Key results at 50 cm2 scale demonstrates that artificial kidney enabled retaining the storage capacity of the flow battery over 100 cycles. Comparison of outcome of this work showing 0% capacity degradation to the state-of-the-art technology corroborates that this novel system is a game changing innovation.

99 GENERAL AND MISCELLANEOUS↗

Free Energy and Flexibility Analysis of Autoinhibited Human BRAF

The RAF serine/threonine protein kinases function as direct effectors of RAS in the intracellular transmission of extracellular growth signals, and they are key targets for drug discovery, given the high incidence of oncogenic mutations in RAF and other components of this signaling pathway. In its inactive state, RAF is held in an autoinhibited conformation in the cytosol through a combination of intramolecular interactions and binding to a regulatory 14−3−3 protein dimer. Activation of RAF is initiated by its interaction with membrane-localized GTP-bound RAS, which induces conformational changes that release RAF from its autoinhibited state. However, the molecular mechanisms governing RAF activation remain incomplete, largely due to the challenges in experimentally capturing the intermediate conformational states in this process. To address this gap, we developed a comprehensive all-atom model of BRAF based on existing cryo-EM structures. Using this model, we performed extensive molecular dynamics simulations to evaluate the stability and free energy landscape of autoinhibited BRAF in solution. Our analysis reveals conformational flexibility within the autoinhibited complex, suggesting that this dynamic behavior may play a role in facilitating BRAF activation upon engagement with the membrane-bound RAS.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

GENTANGLE: integrated computational design of gene entanglements

The design of two overlapping genes in a microbial genome is an emerging technique for adding more reliable control mechanisms in engineered organisms for increased stability. The design of functional overlapping gene pairs is a challenging procedure, and computational design tools are used to improve the efficiency to deploy successful designs in genetically engineered systems. GENTANGLE (Gene Tuples ArraNGed in overLapping Elements) is a high-performance containerized pipeline for the computational design of two overlapping genes translated in different reading frames of the genome. This new software package can be used to design and test gene entanglements for microbial engineering projects using arbitrary sets of user-specified gene pairs.

59 BASIC BIOLOGICAL SCIENCES↗

Anisotropic interactions for continuum modeling of protein–membrane systems

In this work, a model for anisotropic interactions between proteins and cellular membranes is proposed for large-scale continuum simulations. The framework of the model is based on dynamic density functional theory, which provides a formalism to describe the lipid densities within the membrane as continuum fields while still maintaining the fidelity of the underlying molecular interactions. Within this framework, we extend recent results to include the anisotropic effects of protein–lipid interactions. As applications, we consider two membrane proteins of biological interest: a RAS–RAF complex tethered to the membrane and a membrane embedded G protein-coupled receptor. A strong qualitative and quantitative agreement is found between the numerical results and the corresponding molecular dynamics simulations. Combining the scope of continuum level simulations with the details from molecular level particle simulations enables research into protein–membrane behaviors at a more biologically relevant scale, which crucially can also be accessed via experiment.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗