Search NASA⌕ Search

SEARCH · Search NASA

Results for “feature annotation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Celestial Mapping System Videos

The Celestial Mapping System (CMS) is a software platform to generate virtual 3D globes for celestial bodies within our solar system. Multiple planetary data layers can be added to the virtual globe to provide visualization of high-resolution imagery and elevation data, which enables precise measurements, tools for analytical capabilities and a broad range of other functionalities to assist planetary scientists and mission planners. Third-party planetary data can be ingested into CMS with minimal effort. The present focus of CMS is on developing lunar mapping tools to provide features such as: 3D first person view with zoom and navigational capabilities, realistic terrain visualization based on LRO data, measurement tools, Apollo, CLPS and international mission landing site annotations, 3D Models, stereoscopic view, terrain profiling, line of sight analysis, sunlight shading and many more. The application has been developed to provide situational and domain awareness on the Lunar surface, planning capabilities for equipment placement and traverse path optimization.

Mapping↗

High-Resolution Tandem Mass Spectrometry-Based Analysis of Model Lignin–Iron Complexes: Novel Pipeline and Complex Structures

Understanding the chemical nature of soil organic carbon (SOC) with great potential to bind iron (Fe) minerals is critical for predicting the stability of SOC. Organic ligands of Fe are among the top candidates for SOCs able to strongly sorb on Fe minerals, but most of them are still molecularly uncharacterized. To shed insights into the chemical nature of organic ligands in soil and their fate, this study developed a protocol for identifying organic ligands using ultrahigh-performance liquid chromatography-high-resolution tandem mass spectrometry (UHPLC-HRMS/MS) and metabolomic tools. The protocol was used for investigating the Fe complexes formed by model compounds of lignin-derived organic ligands, namely, caffeic acid (CA), p-coumaric acid (CMA), vanillin (VNL), and cinnamic acid (CNA). Isotopologue analysis of 54/56 Fe was used to screen out the potential UHPLC-HRMS (m/z) features for complexes formed between organic ligands and Fe, with multiple features captured for CA, CMA, VNL, and CNA when 35/37 Cl isotopologue analysis was used as supplementary evidence for the complexes with Cl. MS/MS spectra, fragment analysis, and structure prediction with SIRIUS were used to annotate the structures of mono/bidentate mono/biligand complexes. The analysis determined the structures of monodentate and bidentate complexes of FeL x Cl y (L: organic ligand, x = 1–4, y = 0–3) formed by model compounds. The protocol developed in this study can be used to identify unknown organic ligands occurring in complex environmental samples and shed light on the molecular-level processes governing the stability of the SOC.

54 ENVIRONMENTAL SCIENCES↗

Shedding Light on Microbial Dark Matter with A Universal Language of Life

The majority of microbial genomes have yet to be cultured, and most proteins predicted from microbial genomes or sequenced from the environment cannot be functionally annotated. As a result, current computational approaches to describe microbial systems rely on incomplete reference databases that cannot adequately capture the full functional diversity of the microbial tree of life, limiting our ability to model high-level features of biological sequences. The scientific community needs a means to capture the functionally and evolutionarily relevant features underlying biology, independent of our incomplete reference databases. Such a model can form the basis for transfer learning tasks, enabling downstream applications in environmental microbiology, medicine, and bioengineering. Here we present LookingGlass, a deep learning model capturing a “universal language of life”. LookingGlass encodes contextually-aware, functionally and evolutionarily relevant representations of short DNA reads, distinguishing reads of disparate function, homology, and environmental origin. We demonstrate the ability of LookingGlass to be fine-tuned to perform a range of diverse tasks: to identify novel oxidoreductases, to predict enzyme optimal temperature, and to recognize the reading frames of DNA sequence fragments. LookingGlass is the first contextually-aware, general purpose pre-trained “biological language” representation model for short-read DNA sequences. LookingGlass enables functionally relevant representations of otherwise unknown and unannotated sequences, shedding light on the microbial dark matter that dominates life on Earth.

A Hoarfrost↗

An implementation of the programming structural synthesis system (PROSSS)

A particular implementation of the programming structural synthesis system (PROSSS) is described. This software system combines a state of the art optimization program, a production level structural analysis program, and user supplied, problem dependent interface programs. These programs are combined using standard command language features existing in modern computer operating systems. PROSSS is explained in general with respect to this implementation along with the steps for the preparation of the programs and input data. Each component of the system is described in detail with annotated listings for clarification. The components include options, procedures, programs and subroutines, and data files as they pertain to this implementation. An example exercising each option in this implementation to allow the user to anticipate the type of results that might be expected is presented.

Rogers, J. L., Jr.↗

Familiarization with LANDSAT imagery

Learning objectives of the activities provided include: (1) reading the annotation of a LANDSAT image; (2) becoming acquainted with the characteristics of 1:1,000,000 scale transparencies and prints of MSS images; (3) noting the general information visible in LANDSAT photo products; (4) observing changes of appearance of any ground feature or class in the black and white images made from the four MSS bands and the characteristic color of each class in color composites; (5) determining the degree to which a LANDSAT image meets map accuracy standards and can be fitted to map projections; (6) assessing the effects of LANDSAT enlargements and scale changes and of the limitations of satellite resolution relative to aerial photos; (7) observing the influence of time of acquisition (season) on a scene; (8) getting a feel for image quality as dependent on processing and photoreproduction; (9) appreciating the characteristics of the RBV and thermal band imagery obtained from LANDSAT-3; and (10) becoming familiar with certain attributes of adjacent LANDSAT images which permit them to be joined in mosaics and to be viewed in stereo.

Source record↗

Dataset for Leveraging CryoEM and AI-Driven Morphological Feature Analysis for Insights on Bacterial Structures

This repository hosts an AI-assisted image segmentation and analysis pipeline for Pantoea sp. YR343 cryo-electron microscopy (cryoEM) datasets. The workflow automates membrane thickness measurements, flagella detection, and field-of-view (FOV) screening from low-dose, high-resolution cryoEM micrographs eliminating the need for slow manual annotation. By integrating deep-learning based segmentation (YOLOv11) with quantitative post-processing, this toolkit provides a scalable and reproducible way to study bacterial morphology under hydrated, near-native conditions. The GitHub repository for AI-based tools for cryoEM bacteria ultrastructures can be found here: https://github.com/Sireesiru/Cryo-EM-Ultrastructures/tree/main

60 APPLIED LIFE SCIENCES↗

Geologic interpretation of Apollo 6 stereophotography from Baja California to west Texas

Excellent space photography of parts of the southwestern United States and northwestern Mexico was obtained during the unmanned Apollo 6 spaceflight. Two features of this photography made it useful for geologic interpretations: its vertical stereocoverage and its exposure under a relatively low angle of solar illumination through an unusually cloud-free and clear atmosphere. The structural patterns, which were topographically enhanced by the longer shadows, were annotated on the photographs, in order to analyze their trends with respect to the continental tectonic framework, and to attempt to correlate the pattern with known copper or other base metal deposits. The annotated fracture patterns showed the regional trends and their distribution. The area studied was a 100- to 105-mile swath of terrain covering a total land area of approximately 60,000 square statute miles. The coverage began from a point centered on Punta Colnett on the Pacific coast of Baja California and extended to the Sacramento Mountains of New Mexico and west Texas.

Gawarecki, S. J.↗

Film annotation system for a space experiment

This microprocessor system was designed to control and annotate a Nikon 35 mm camera for the purpose of obtaining photographs and data at predefined time intervals. The single STD BUSS interface card was designed in such a way as to allow it to be used in either a stand alone application with minimum features or installed in a STD BUSS computer allowing for maximum features. This control system also allows the exposure of twenty eight alpha/numeric characters across the bottom of each photograph. The data contains such information as camera identification, frame count, user defined text, and time to .01 second.

Browne, W. R.↗

Simple Math is Enough: Two Examples of Inferring Functional Associations from Genomic Data

Non-random features in the genomic data are usually biologically meaningful. The key is to choose the feature well. Having a p-value based score prioritizes the findings. If two proteins share a unusually large number of common interaction partners, they tend to be involved in the same biological process. We used this finding to predict the functions of 81 un-annotated proteins in yeast.

Liang, Shoudan↗

Building a FAIR data ecosystem for incorporating single-cell transcriptomics data into agricultural genome to phenome research

Introduction The agriculture genomics community has numerous data submission standards available, but the standards for describing and storing single-cell (SC, e.g., scRNA- seq) data are comparatively underdeveloped. Methods To bridge this gap, we leveraged recent advancements in human genomics infrastructure, such as the integration of the Human Cell Atlas Data Portal with Terra, a secure, scalable, open-source platform for biomedical researchers to access data, run analysis tools, and collaborate. In parallel, the Single Cell Expression Atlas at EMBL-EBI offers a comprehensive data ingestion portal for high-throughput sequencing datasets, including plants, protists, and animals (including humans). Developing data tools connecting these resources would offer significant advantages to the agricultural genomics community. The FAANG data portal at EMBL-EBI emphasizes delivering rich metadata and highly accurate and reliable annotation of farmed animals but is not computationally linked to either of these resources. Results Herein, we describe a pilot-scale project that determines whether the current FAANG metadata standards for livestock can be used to ingest scRNA-seq datasets into Terra in a manner consistent with HCA Data Portal standards. Importantly, rich scRNA-seq metadata can now be brokered through the FAANG data portal using a semi-automated process, thereby avoiding the need for substantial expert curation. We have further extended the functionality of this tool so that validated and ingested SC files within the HCA Data Portal are transferred to Terra for further analysis. In addition, we verified data ingestion into Terra, hosted on Azure, and demonstrated the use of a workflow to analyze the first ingested porcine scRNA-seq dataset. Additionally, we have also developed prototype tools to visualize the output of scRNA-seq analyses on genome browsers to compare gene expression patterns across tissues and cell populations. This JBrowse tool now features distinct tracks, showcasing PBMC scRNA-seq alongside two bulk RNA-seq experiments. Discussion We intend to further build upon these existing tools to construct a scientist-friendly data resource and analytical ecosystem based on Findable, Accessible, Interoperable, and Reusable (FAIR) SC principles to facilitate SC-level genomic analysis through data ingestion, storage, retrieval, re-use, visualization, and comparative annotation across agricultural species.

Genetics & Heredity↗

Ca X ML: Chemistry‐informed machine learning explains mutual changes between protein conformations and calcium ions in calcium‐binding proteins using structural and topological features

Proteins' flexibility is a feature in communicating changes in cell signaling instigated by binding with secondary messengers, such as calcium ions, associated with the coordination of muscle contraction, neurotransmitter release, and gene expression. When binding with the disordered parts of a protein, calcium ions must balance their charge states with the shape of calcium-binding proteins and their versatile pool of partners depending on the circumstances they transmit. Accurately determining the ionic charges of those ions is essential for understanding their role in such processes. However, it is unclear whether the limited experimental data available can be effectively used to train models to accurately predict the charges of calcium-binding protein variants. Here, we developed a chemistry-informed, machine-learning algorithm that implements a game theoretic approach to explain the output of a machine-learning model without the prerequisite of an excessively large database for high-performance prediction of atomic charges. We used the ab initio electronic structure data representing calcium ions and the structures of the disordered segments of calcium-binding peptides with surrounding water molecules to train several explainable models. Network theory was used to extract the topological features of atomic interactions in the structurally complex data dictated by the coordination chemistry of a calcium ion, a potent indicator of its charge state in protein. Our design created a computational tool of Ca X ML, which provided a framework of explainable machine learning model to annotate ionic charges of calcium ions in calcium-binding proteins in response to the chemical changes in an environment. Our framework will provide new insights into protein design for engineering functionality based on the limited size of scientific data in a genome space.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Considerations regarding the deployment of hypermedia at JSC

Electronic documents and systems are becoming the primary means of managing information for ground and space operations at NASA. These documents will utilize hypertext and hypermedia technologies to aid users in structuring and accessing information. Documents will be composed of static and dynamic data consisting of user-defined annotations and hypermedia links. The report consists of three major sections. First, it provides an overview of hypermedia and surveys the use of hypermedia throughout JSC. Second, it briefly describes a prototypical hypermedia system that was developed in conjunction with this work. This system was constructed to demonstrate various hypermedia features and to serve as a platform for supporting the electronic documentation needs for the MIDAS system developed by the Intelligent Systems Branch of the Automation and Robotics Division (Pac92). Third, it discusses emerging hypermedia technologies which have either been untapped by vendors or present significant challenges to the Agency.

Kacmar, Charles J.↗

An annotation system for 3D fluid flow visualization

Annotation is a key activity of data analysis. However, current systems for data analysis focus almost exclusively on visualization. We propose a system which integrates annotations into a visualization system. Annotations are embedded in 3D data space, using the Post-it metaphor. This embedding allows contextual-based information storage and retrieval, and facilitates information sharing in collaborative environments. We provide a traditional database filter and a Magic Lens filter to create specialized views of the data. The system has been customized for fluid flow applications, with features which allow users to store parameters of visualization tools and sketch 3D volumes.

Loughlin, Maria M.↗

Automated Bacterial Identification and Morphological Feature Analysis in Low‐Dose Cryo‐EM Using YOLOv11

Bacteria rapidly adapt to environmental cues through morphological and ultrastructural changes that correlate with physiology and behavior. Cryogenic transmission electron microscopy (cryo‐TEM) can capture these phenotypic changes in near‐native, vitrified states, but manual analysis of low‐dose micrographs is labor intensive and limits throughput. Here, we present an end‐to‐end workflow that combines low‐dose cryo‐TEM imaging with a YOLOv11‐based instance‐segmentation model to automatically identify bacteria and quantify key structural features directly from the micrographs. This workflow enables (i) robust bacterial localization and counting from low‐magnification atlas/montage images, (ii) automated measurements of cell‐envelope (outer–inner membrane) thickness and anisotropy from higher‐magnification views, and (iii) detection and quantification of bacteria–flagella interactions, including overlap length and curvature metrics for interacting versus noninteracting flagella. Using Pantoea sp. YR343 grown under distinct media conditions, we show that the automated measurements agree with manual annotations while substantially reducing analysis time. Together, these tools provide a practical framework for scalable bacterial identification and quantitative phenotyping in low‐dose cryo‐TEM datasets and establish a foundation for extending cryo‐TEM image analysis toward higher‐throughput studies of microbial heterogeneity and biointerfaces.

YOLOv11↗

Targeted genetic manipulation and yeast-like evolutionary genomics in the green alga Auxenochlorella

Auxenochlorella spp. are diploid oleaginous green algae whose streamlined genomes can be readily manipulated by homologous recombination, making them highly amenable to discovery research and bioengineering. Vegetatively diploid organisms experience specific evolutionary phenomena, including allodiploid hybridization, mitotic recombination, loss-of-heterozygosity, and aneuploidy; however, studies of these forces have largely focused on yeasts. Here, we present a telomere-to-telomere phased diploid genome assembly of Auxenochlorella UTEX 250-A (haploid length 22 Mb) and introduce a genetic toolkit for site-specific manipulation of the nuclear genome in multiple strains, featuring several selectable markers, inducible promoters, and fluorescent reporters for protein localization. UTEX 250-A is an allodiploid hybrid of Auxenochlorella protothecoides and Auxenochlorella symbiontica, two species differentiated by extensive chromosomal rearrangements. UTEX 250-A haplotypes are a mosaic of each parental species following mitotic recombination, and two chromosomes are trisomic. Loss-of-heterozygosity events are pervasive across Auxenochlorella and can evolve rapidly in the laboratory. High-quality structural annotation yielded ∼7,500 genes per haplotype. Auxenochlorella have experienced gene family loss and reduction, including core photosynthesis genes, and exhibit periodic adenine and cytosine methylation at promoters and gene bodies, respectively. Approximately 10% of genes, especially those involved in DNA repair and sex, overlap antisense long noncoding RNAs, which may participate in a regulatory mechanism. We demonstrate the utility of Auxenochlorella for fundamental research by knockout of a chlorophyll biosynthesis enzyme, and confirm one trisomy by allele-specific transformation. These results demonstrate the generality of several evolutionary forces associated with vegetative diploidy and provide a foundation for the use of Auxenochlorella as a reference organism.

CHL27↗

Constitutive down‐regulation of liguleless alleles in sorghum drives increased productivity and water use efficiency

Plant architecture influences the microenvironment throughout the canopy layer. Plants with a more erect leaf architecture allow for an increase in planting densities and allow more light to reach lower canopy leaves. This is predicted to increase crop carbon assimilation. Frictional resistance to wind reduces air movement in the lower canopy, resulting in higher humidity. By increasing the proportion of canopy photosynthesis in the more humid lower canopy, gains in the efficiency of water use might be expected, although this may be slightly offset by the more open erectophile form canopy. An anatomical feature in members of the Poaceae family that impacts leaf angle is the articulated junction of the sheath and blade, which also bares the ligule and auricles. Mutants, which lack ligules and auricles, show no articulation at this junction, resulting in leaves that are near vertical. In maize, these phenotypes termed liguleless result from null mutations of genes: ZmLG1 (Zm00001eb67740) and ZmLG2 (Zm00001eb147220). In sorghum, SbiRTx430.06G264300 (SbLG1) and SbiRTx430.03G392300 (SbLG2) are annotated as the respective maize homologues. A hair-pin element designed to down-regulate both SbLG1 and SbLG2 was introduced into the grain sorghum genotype RTx430. Derived transgenic events harbouring the hair-pin failed to develop ligules and displayed reduced leaf angles to the vertical, but less vertical than in null mutations. Under field settings, plots sown with these sorghum events having an erect architecture phenotype displayed an increase in photosynthesis in lower canopy levels, which led to increases in above-ground biomass and seed yield, without an increase in water use.

59 BASIC BIOLOGICAL SCIENCES↗

Data for "Constitutive Down-Regulation of Liguleless Alleles in Sorghum Drives Increased Productivity and Water Use Efficiency"

Plant architecture influences the microenvironment throughout the canopy layer. Plants with a more erect leaf architecture allow for an increase in planting densities and allow more light to reach lower canopy leaves. This is predicted to increase crop carbon assimilation. Frictional resistance to wind reduces air movement in the lower canopy, resulting in higher humidity. By increasing the proportion of canopy photosynthesis in the more humid lower canopy, gains in the efficiency of water use might be expected, although this may be slightly offset by the more open erectophile form canopy. An anatomical feature in members of the Poaceae family that impacts leaf angle is the articulated junction of the sheath and blade, which also bares the ligule and auricles. Mutants, which lack ligules and auricles, show no articulation at this junction, resulting in leaves that are near vertical. In maize, these phenotypes termed liguleless result from null mutations of genes: ZmLG1 (Zm00001eb67740) and ZmLG2 (Zm00001eb147220). In sorghum, SbiRTx430.06G264300 (SbLG1) and SbiRTx430.03G392300 (SbLG2) are annotated as the respective maize homologues. A hair-pin element designed to down-regulate both SbLG1 and SbLG2 was introduced into the grain sorghum genotype RTx430. Derived transgenic events harbouring the hair-pin failed to develop ligules and displayed reduced leaf angles to the vertical, but less vertical than in null mutations. Under field settings, plots sown with these sorghum events having an erect architecture phenotype displayed an increase in photosynthesis in lower canopy levels, which led to increases in above-ground biomass and seed yield, without an increase in water use.

Genome Engineering↗

Statistical Perspectives on Stratospheric Transport

Long-lived tropospheric source gases, such as nitrous oxide, enter the stratosphere through the tropical tropopause, are transported throughout the stratosphere by the Brewer-Dobson circulation, and are photochemically destroyed in the upper stratosphere. These chemical constituents, or "tracers" can be used to track mixing and transport by the stratospheric winds. Much of our understanding about the stratospheric circulation is based on large scale gradients and other spatial features in tracer fields constructed from satellite measurements. The point of view presented in this paper is different, but complementary, in that transport is described in terms of tracer probability distribution functions (PDFs). The PDF is computed from the measurements, and is proportional to the area occupied by tracer values in a given range. The flavor of this paper is tutorial, and the ideas are illustrated with several examples of transport-related phenomena, annotated with remarks that summarize the main point or suggest new directions. One example shows how the multimodal shape of the PDF gives information about the different branches of the circulation. Another example shows how the statistics of fluctuations from the most probable tracer value give insight into mixing between different regions of the atmosphere. Also included is an analysis of the time-dependence of the PDF during the onset and decline of the winter circulation, and a study of how "bursts" in the circulation are reflected in transient periods of rapid evolution of the PDF. The dependence of the statistics on location and time are also shown to be important for practical problems related to statistical robustness and satellite sampling. The examples illustrate how physically-based statistical analysis can shed some light on aspects of stratospheric transport that may not be obvious or quantifiable with other types of analyses. An important motivation for the work presented here is the need for synthesis of the large and growing database of observations of the atmosphere and the vast quantities of output generated by atmospheric models.

Sparling, L. C.↗