Search NASASearch

SEARCH · Search NASA

Results for “Vision”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

VISIONS: Remote Observations of a Spatially-Structured Filamentary Source of Energetic Neutral Atoms near the Polar Cap Boundary During an Auroral Substorm

We report initial results from the VISualizing Ion Outflow via Neutral atom imaging during a Substorm (VISIONS) rocket that flew through and near several regions of enhanced auroral activity and also sensed regions of ion outflow both remotely and directly. The observed neutral atom fluxes were largest at the lower energies and generally higher in the auroral zone than in the polar cap. In this paper, we focus on data from the latter half of the VISIONS trajectory when the rocket traversed the polar cap region. During this period, many of the energetic neutral atom spectra show a peak at 100 electronvolts. Spectra with peaks around 100 electronvolts are also observed in the Electrostatic Ion Analyzer (EIA) data consistent with these ions comprising the source population for the energetic neutral atoms. The EIA observations of this low energy population extend only over a few tens of kilometers. Furthermore, the directionality of the arriving energetic neutral atoms is consistent with either this spatially localized source of energetic ions extending from as low as about 300 kilometers up to above 600 kilometers or a larger source of energetic ions to the southwest.

Polar Cap

Technology for NASA's Planetary Science Vision 2050.

NASAs Planetary Science Division (PSD) initiated and sponsored a very successful community Workshop held from Feb. 27 to Mar. 1, 2017 at NASA Headquarters. The purpose of the Workshop was to develop a vision of planetary science research and exploration for the next three decades until 2050. This abstract summarizes some of the salient technology needs discussed during the three-day workshop and at a technology panel on the final day. It is not meant to be a final report on technology to achieve the science vision for 2050.

Vision

Improvements in Optical Surface Measurement Using Reflected Computer Vision Targets

Since 2021, NREL has been developing a system to measure heliostats by measuring the deflection of printed computer vision targets, called the Reflected Target Non-intrusive Assessment (ReTNA) [6], [7]. While this system will have lower resolution than a fringe deflectometry system, it has several important advantages that make it a complimentary technology: 2D surface slope measurement can be generated from a single image, it can operate in ambient lighting, target points can be directly located in 3D space with photogrammetry allowing for a non-flat target, and it's well suited to using a smaller target, and multiple images to measure larger optical surfaces. ReTNA has undergone several significant changes and improvements, described below. This talk will summarize new system layouts designed for commercial use, new computer vision algorithms used to automate the analysis process and validation campaigns for the ReTNA software.

computer vision

Advanced Fuels Campaign: Strategic Vision

The Advanced Fuels Campaign (AFC) is dedicated to propelling the United States to the forefront of nuclear fuel technology through a comprehensive strategic vision. The vision is structured around five key goals: driving U.S. leadership in nuclear fuel technology, expanding nuclear energy production from the existing fleet, completing the qualification basis for advanced reactor fuel technology, driving innovation in advanced nuclear fuel technology, and enabling a high-performing organization. By focusing on these goals, AFC aims to define and implement cutting-edge nuclear fuel technologies, support industry advancements, achieve critical fuel qualifications, foster innovative research, and ensure effective program management and stakeholder engagement. This strategic approach will solidify the United States’ position as a global leader in nuclear energy and fuel technology, while also preparing for the future of environmentally friendly and reliable nuclear power.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Advanced Fuels Campaign: Strategic Vision

The Advanced Fuels Campaign (AFC) is dedicated to propelling the United States to the forefront of nuclear fuel technology through a comprehensive strategic vision. The vision is structured around five key goals: driving U.S. leadership in nuclear fuel technology, expanding nuclear energy production from the existing fleet, completing the qualification basis for advanced reactor fuel technology, driving innovation in advanced nuclear fuel technology, and enabling a high-performing organization. By focusing on these goals, AFC aims to define and implement cutting-edge nuclear fuel technologies, support industry advancements, achieve critical fuel qualifications, foster innovative research, and ensure effective program management and stakeholder engagement. This strategic approach will solidify the United States’ position as a global leader in nuclear energy and fuel technology, while also preparing for the future of environmentally friendly and reliable nuclear power.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Automated Waterbox Inspection for Nuclear Power Plants Using Computer Vision - Based Change Detection

Nuclear power plant waterboxes require regular inspection for leaks, missing components, and structural damage during maintenance outages. Traditional manual inspection is time-consuming and poses safety risks from confined space entry. We developed an automated computer vision system for drone-based waterbox inspection in partnership with Florida Light and Power. Our approach uses feature detection and matching to identify critical changes between baseline and current inspection images, automatically flagging additions (leaks/debris), removals (missing plugs), and translations (displaced components) while compensating for drone movement and environmental variations. We systematically evaluated six feature matching methods, from classical approaches (SIFT+BF) to state-of-the-art neural networks (SuperPoint+SuperGlue), using both standard benchmarks (HPatches) and waterbox-specific validation with real-world augmentations. SuperPoint+SuperGlue achieved superior performance with 7.82 pixels RMSE and 100% success rate—2.8x better accuracy than our baseline. While the pre-trained model has commercial licensing restrictions for nuclear deployment, our findings validate this architecture for custom training. We implemented a real-time GUI demonstrating the SIFT+BF approach for immediate deployment, processing drone feeds at 30 FPS with color-coded change visualization. Future work includes training a custom SuperPoint+SuperGlue model on waterbox data and integrating Vision-Language Models for automated reporting and maintenance guidance.

46 - INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AN

Structural and functional dynamics of human cone cGMP-phosphodiesterase important for photopic vision

Cone cGMP-phosphodiesterase (PDE6) is the key effector enzyme for daylight vision, and its properties are critical for shaping distinct physiology of cone photoreceptors. We determined the structures of human cone PDE6C in various liganded states by single-particle cryo-EM that reveal essential functional dynamics and adaptations of the enzyme. Our analysis exposed the dynamic nature of PDE6C association with its regulatory γ-subunit (Pγ) which allows openings of the catalytic pocket in the absence of phototransduction signaling, thereby controlling photoreceptor noise and sensitivity. We demonstrate evolutionarily recent adaptations of PDE6C stemming from residue substitutions in the Pγ subunit and the noncatalytic cGMP binding site and influencing the Pγ dynamics in holoPDE6C. Thus, our structural analysis sheds light on the previously unrecognized molecular evolution of the effector enzyme in cones that advances adaptation for photopic vision.

Science & Technology - Other Topics

Spatial profiling of the interplay between cell type- and vision-dependent transcriptomic programs in the visual cortex

How early sensory experience during “critical periods” of postnatal life affects the organization of the mammalian neocortex at the resolution of neuronal cell types is poorly understood. We previously reported that the functional and molecular profiles of layer 2/3 (L2/3) cell types in the primary visual cortex (V1) are vision-dependent [S. Chenget al.,Cell185, 311–327.e24 (2022)]. Here, we characterize the spatial organization of L2/3 cell types with and without visual experience. Spatial transcriptomic profiling based on 500 genes recapitulates the zonation of L2/3 cell types along the pial–ventricular axis in V1. By applying multitasking theory, we suggest that the spatial zonation of L2/3 cell types is linked to the continuous nature of their gene expression profiles, which can be represented as a 2D manifold bounded by three archetypal cell types. By comparing normally reared and dark reared L2/3 cells, we show that visual deprivation-induced transcriptomic changes comprise two independent gene programs. The first, induced specifically in the visual cortex, includes immediate-early genes and genes associated with metabolic processes. It manifests as a change in cell state that is orthogonal to cell-type-specific gene expression programs. By contrast, the second program impacts L2/3 cell-type identity, regulating a subset of cell-type-specific genes and shifting the distribution of cells within the L2/3 cell-type manifold. Through an integrated analysis of spatial transcriptomics with single-nucleus RNA-seq data, we describe how vision patterns cortical L2/3 cell types during the critical period.

Science & Technology - Other Topics

Uncertainty quantification of fireball features extracted from nuclear test films using computer vision

Films from the US’s historic nuclear testing era comprise the only extensive collection of imagery depicting high-yield detonations. These films offer unique insights into the characteristics of flows occurring on scales that are difficult to replicate experimentally, and they are a valuable source of data for the validation of models used to describe nuclear detonations. In recent work, we implemented modern computer vision and machine learning techniques to extract features of the fireball following nuclear detonation. With a training dataset of fireball films, we fine-tuned a You Only Look Once 11 (YOLO11) model to detect and track the fireball. Applied to a video, the outer bounding box produced in each frame by YOLO11 is used as an input prompt to Meta’s Segment Anything Model 2 (SAM2), which is shown to accurately predict the boundary of the fireball over time with high resolution. These state-of-the-art computer vision foundation models exhibit impressive visual accuracy in their results but lack an output of values that robustly quantify uncertainty in scientific applications. In this paper, we develop procedures for uncertainty quantification of extracted fireball features. We outline the application of a parallel attention mechanism to calculate uncertainty ranges that complement and better pose model validation data. This higher quality fireball validation data may serve to improve prognostic models describing nuclear detonations in support of nuclear forensic and emergency response activities.

Khristy, Joel [ORNL] (ORCID:0000000209963060)

ObstacleSense: Low-Power Neuromorphic Vision for Corridor Obstacle Awareness in Low-Level ADAS

The automotive industry’s pursuit of Level 5 autonomy is constrained by substantial perception-compute power requirements, often reaching 1, 000 + watts in full autonomy stacks. Reducing this energy burden requires rethinking perception not only at the high-end autonomy level, but also at the foundational Advanced Driver Assistance Systems (ADAS) level where low-power, safety-critical sensing can have broad impact. Neuromorphic vision provides a promising starting point: HD Dynamic Vision Sensors (DVS) can operate below 100 mW at the sensor level by reporting only asynchronous brightness changes. However, low-power sensing alone is insufficient if downstream perception reintroduces dense, energy-intensive computation. In particular, many event-driven object-detection pipelines still rely on CNN backbones, while purely spiking alternatives often trade away accuracy or ignore deployment constraints. We introduce ObstacleSense, a highly compact, CNN-free hybrid ANN–SNN framework for Level 0–1 forward-corridor obstacle awareness. Instead of performing full-scene object detection with a convolutional feature backbone, ObstacleSense targets the safety-critical question of whether the ego corridor is occupied and how far the nearest obstacle is. The architecture combines polarity-conditioned event encoding, lightweight temporal spiking dynamics, axial spatial mixing, and coarse-to-fine range estimation within a regular fixed-grid compute pattern. This design avoids the dense CNN backbone commonly used in event-based detection while maintaining a small state footprint suitable for eventual small-FPGA deployment. Before hardware mapping, we evaluate the software implementation using a model-side power proxy derived from MACs, weight and activation traffic, and spiking state updates under shared FP16 assumptions. On simulated CARLA event corpora, the deployment-oriented model achieves 0.9464 objectness F1, 0.9978 grid-level mAP, and 0.8987 m distance Mean Absolute Error at an estimated 1.92 mW proxy cost, while maintaining performance on unseen generalization test sequences.

Johnson-Scott, Zac [ORNL]

Lost in OCR Translation? Vision-Based Approaches to Robust Document Retrieval

Code for ‘Lost in OCR Translation?’: robust document retrieval under degradation. Compares OCR-based, vision-only, and hybrid pipelines; includes SambaNova LLaMA Vision OCR, Nougat, and ViDoRe baselines. Provides QA data generation, RAG evaluation, and metrics (Levenshtein, nDCG@k, Recall@k, EM/F1) with reproducible scripts. Includes dataset guides

Bhattarai, Manish [Los Alamos National Labs]

Catalyst-Vision (PEM Catalyst Layer Image Analysis Tool) [SWR-25-100]

Catalyst-Vision (PEM Catalyst Layer Image Analysis Tool) provides an advanced Python-based tool, primarily designed for use in a Jupyter/Colab notebook, for the quantitative morphological analysis of pre-segmented shapes. While developed for analyzing PEM catalyst layers from microscopy, its methodology is suitable for characterizing any grayscale object provided on a uniform white background. The tool uses a robust computer vision pipeline based on the Euclidean Distance Transform and skeletonization to accurately measure local thickness and tortuosity, providing a comprehensive characterization of an object's geometry and internal texture. If you find this code useful, please cite our preprint as: Chan, Ai-Lin and Hayden, Steven and Harvey, Steven P. and Smeaton, Michelle and Okrucky, Caleb and Watt, John and Ulična, Soňa and Spurgeon, Steven and Jungjohann, Katherine and Alia, Shaun, Mechanism-informed breakdown: understanding degradation by controlling voltage hold patterns in PEM water electrolyzers. Preprint (2025).

Spurgeon, Steven [National Laboratory of the Rocki

Sequence length scaling in vision transformers for scientific images on frontier

Vision Transformers (ViTs) are pivotal for foundational models in scientific imagery, including Earth science applications, due to their capability to process large sequence lengths. While transformers for text have inspired scaling sequence lengths in ViTs, adapting these for ViTs introduces unique challenges. We develop distributed sequence parallelism for ViTs, enabling them to handle up to 1M tokens. Our approach, leveraging DeepSpeed-Ulysses and Long-Sequence-Segmentation with model sharding, is the first to apply sequence parallelism in ViT training, achieving a 94% batch scaling efficiency on 2,048 AMD-MI250X GPUs. Evaluating sequence parallelism in ViTs, particularly in models up to 10B parameters, highlighted substantial bottlenecks. We countered these with hybrid sequence, pipeline, and flash attention strategies, to scale beyond single GPU memory limits. Our method significantly enhances climate modeling accuracy by 20% in temperature predictions, marking the first training of a vision transformer model to convergence with a sequence length of 188K tokens, using full self-attention.

Tsaris, Aristeidis (aris) [ORNL] (ORCID:0000000277

Visioning Energy: Science Fiction Author-Energy Researcher Collaboration Workshop Recap

The Visioning Energy: Science Fiction Author-Energy Researcher Collaboration Workshop brought together speculative fiction authors and NREL researchers to examine possible scenarios of the future. The goals of the workshop were to support out-of-the-box thinking and creative future visioning for the researchers and to provide authors with insights into the latest clean energy technologies and their potential. This report outlines the proceedings of the workshop and insights from the collaborative brainstorming activity, highlighting themes and opportunities for further exploration.

29 ENERGY PLANNING, POLICY, AND ECONOMY

Integrating Machine-learning-assisted Computer Vision with RICH System

Developments in artificial intelligence have vastly expanded the capabilities of robots. Currently, the Spallation Neutron Source (SNS) beamlines at Oak Ridge National Lab (ORNL) have robotic sample loaders to increase the efficiency of running experiments. However, they require retraining if anything about the situation changes, e.g., where the samples are, and cannot notice if errors occur. So, the viability of using computer vision and machine learning to enhance these sample loaders’ functionality was investigated. In this project, the RICH system with a Dobot CR3 6-axis robot present at the VULCAN beamline assisted by an Intel Realsense D435i camera, a unique camera that enables convenient translation of 2D pixel coordinates to 3D world points, was programmed to load ceramic crucibles into a thermogravimetric analyzer (TGA) furnace. An algorithm was constructed in Python with three major phases planned: (1) obtaining a sample, (2) moving it to the target location, and then (3) bringing the sample back to its original location once the experiment finished. In the first phase, the algorithm would dynamically detect sample locations using ArUco markers to recognize the samples’ general location and a custom-trained yolov5 object detection model to locate the crucibles’ centers. Afterward, the robot would be directed to pick up samples based on the crucibles’ calculated positions. In the second phase, the robot would move the sample to a secondary point, reorient its grip, and place the sample at the target location. In the final phase, the robot would determine whether the sample was intact and would bring it back to its original place if it was or raise an alarm. Using this algorithm, the robot was able to pick up different types of crucibles at varying positions. These results indicate that integrating machine-learning-assisted computer vision with robotic sample loaders can result in effective autonomous detection of samples.

97 MATHEMATICS AND COMPUTING

Permeability Prediction Using Vision Transformers

Accurate permeability predictions remain pivotal for understanding fluid flow in porous media, influencing crucial operations across petroleum engineering, hydrogeology, and related fields. Traditional approaches, while robust, often grapple with the inherent heterogeneity of reservoir rocks. With the advent of deep learning, convolutional neural networks (CNNs) have emerged as potent tools in image-based permeability estimation, capitalizing on micro-CT scans and digital rock imagery. This paper introduces a novel paradigm, employing vision transformers (ViTs)—a recent advancement in computer vision—for this crucial task. ViTs, which segment images into fixed-sized patches and process them through transformer architectures, present a promising alternative to CNNs. We present a methodology for implementing ViTs for permeability prediction, its results on diverse rock samples, and a comparison against conventional CNNs. The prediction results suggest that, with adequate training data, ViTs can match or surpass the predictive accuracy of CNNs, especially in rocks exhibiting significant heterogeneity. This study underscores the potential of ViTs as an innovative tool in permeability prediction, paving the way for further research and integration into mainstream reservoir characterization workflows.

58 GEOSCIENCES

ESPPU INPUT: C$^3$ within the "Linear Collider Vision"

The Linear Collider Vision calls for a Linear Collider Facility with a physics reach from a Higgs Factory to the TeV-scale with $e^+e^{-}$ collisions. One of the technologies under consideration for the accelerator is a cold-copper distributed-coupling linac capable of achieving high gradient. This technology is being pursued by the C$^3$ collaboration to understand its applicability to future colliders and broader scientific applications. In this input we share the baseline parameters for a C$^3$ Higgs-factory and the energy reach of up to 3 TeV in the 33 km tunnel foreseen under the Linear Collider Vision. Recent results, near-term plans and future R&D needs are highlighted.

43 PARTICLE ACCELERATORS

Effect of elimination of nitrogen and/or hypoxia or restricted visual environment on color vision and range of accommodation

The effects upon range of accommodation and color vision of reduced atmospheric pressure, at partial and complete elimination of nitrogen, of hypoxia, and of exposure for varying periods of time to restricted visual environment, have been studied alone or in various combinations. Measurements were made on the electroretinogram, the electrooculogram, and the diameter of the retinal vessels as an indicator of blood flow to the retina at the time of total elimination of nitrogen. An objective method was used to test range of accommodation. In the color vision test the flicker colors of a Benham's top were matched with a colorimeter.

Wolbarsht, M. L.