Search NASA⌕ Search

SEARCH · Search NASA

Results for “vision transformers”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Pictorial communication in virtual and real environments

Papers about the communication between human users and machines in real and synthetic environments are presented. Individual topics addressed include: pictorial communication, distortions in memory for visual displays, cartography and map displays, efficiency of graphical perception, volumetric visualization of 3D data, spatial displays to increase pilot situational awareness, teleoperation of land vehicles, computer graphics system for visualizing spacecraft in orbit, visual display aid for orbital maneuvering, multiaxis control in telemanipulation and vehicle guidance, visual enhancements in pick-and-place tasks, target axis effects under transformed visual-motor mappings, adapting to variable prismatic displacement. Also discussed are: spatial vision within egocentric and exocentric frames of reference, sensory conflict in motion sickness, interactions of form and orientation, perception of geometrical structure from congruence, prediction of three-dimensionality across continuous surfaces, effects of viewpoint in the virtual space of pictures, visual slant underestimation, spatial constraints of stereopsis in video displays, stereoscopic stance perception, paradoxical monocular stereopsis and perspective vergence. (No individual items are abstracted in this volume)

Ellis, Stephen R.↗

Convergent Aeronautics Solutions Project

NASA is committed to transforming our aviation system to best meet demands and opportunities of the future. With a vision of safe, efficient, flexible, and environmentally sustainable air transportation, the NASA Aeronautics Research Mission Directorate is conducting research and development to address future needs of the aviation community, the Nation, and the world. While our NASA Aeronautics vision and strategy reaches into the next 25 years and beyond, we recognize that our vision and strategy must be responsive to new discoveries and emerging markets. For this reason, we are empowering our research community to redefine the future of aviation by dreaming up convergent/transformative ideas and studying if those ideas are possible. By modeling the new NASA Aeronautics' Convergent Aeronautics Solutions (CAS) Project after the venture capital community, we created opportunities for teams of intrapreneurs to mature their ideas into concepts through rapid feasibility studies. The CAS Project expects teams to consider the complexities and potential benefits of multi-disciplinary solutions and to leverage technology advances from outside the field of aeronautics. We also expect teams to explore their concepts in a rapid, iterative manner that allows them to learn and adjust their research approach. Within a year or two, teams are responsible for reporting on the feasibility of their concept. The findings inform NASA Aeronautics strategic planning and further investment. This presentation will give an overview of CAS.

Transformative↗

Localization Using Visual Odometry and a Single Downward-Pointing Camera

Stereo imaging is a technique commonly employed for vision-based navigation. For such applications, two images are acquired from different vantage points and then compared using transformations to extract depth information. The technique is commonly used in robotics for obstacle avoidance or for Simultaneous Localization And Mapping, (SLAM). Yet, the process requires a number of image processing steps and therefore tends to be CPU-intensive, which limits the real-time data rate and use in power-limited applications. Evaluated here is a technique where a monocular camera is used for vision-based odometry. In this work, an optical flow technique with feature recognition is performed to generate odometry measurements. The visual odometry sensor measurements are intended to be used as control inputs or measurements in a sensor fusion algorithm using low-cost MEMS based inertial sensors to provide improved localization information. Presented here are visual odometry results which demonstrate the challenges associated with using ground-pointing cameras for visual odometry. The focus is for rover-based robotic applications for localization within GPS-denied environments.

Swank, Aaron J.↗

MIT-NASA Workshop: Transformational Technologies

As a space faring nation, we are at a critical juncture in the evolution of space exploration. NASA has announced its Vision for Space Exploration, a vision of returning humans to the Moon, sending robots and eventually humans to Mars, and exploring the outer solar system via automated spacecraft. However, mission concepts have become increasingly complex, with the potential to yield a wealth of scientific knowledge. Meanwhile, there are significant resource challenges to be met. Launch costs remain a barrier to routine space flight; the ever-changing fiscal and political environments can wreak havoc on mission planning; and technologies are constantly improving, and systems that were state of the art when a program began can quickly become outmoded before a mission is even launched. This Conference Publication describes the workshop and featured presentations by world-class experts presenting leading-edge technologies and applications in the areas of power and propulsion; communications; automation, robotics, computing, and intelligent systems; and transformational techniques for space activities. Workshops such as this one provide an excellent medium for capturing the broadest possible array of insights and expertise, learning from researchers in universities, national laboratories, NASA field Centers, and industry to help better our future in space.

Mankins, J. C.↗

Solving The Space Weather Problem: A 15+ Year Roadmap to Revolutionize Space Weather Research, Protect NASA Space Assets, and Enable Robust Operations

The White Paper (WP) describes a roadmap to address the Space Weather (SpWx) problem. It presents a strategic vision of how a community-wide effort could be organized and implemented to enable transformative advancement in SpWx research and ultimately, in applications. We envision a ‘system-of-systems’—an integrated web of SpWx stations and state-ofthe-art modeling facilities to enable a transformative advance in SpWx nowcasting and forecasting (Figure 1). The Space Weather Aggregated Network of Systems (SWANS) will enable space situational awareness for end-users invested in spaceflight operations, infrastructure risk mitigation, and future human endeavors in space exploration while profoundly transforming Heliophysics research.

A Vourrlidas↗

Computational models of human vision with applications

Perceptual problems in aeronautics were studied. The mechanism by which color constancy is achieved in human vision was examined. A computable algorithm was developed to model the arrangement of retinal cones in spatial vision. The spatial frequency spectra are similar to the spectra of actual cone mosaics. The Hartley transform as a tool of image processing was evaluated and it is suggested that it could be used in signal processing applications, GR image processing.

Wandell, B. A.↗

National Campaign Development Test Executive Summary

NASA’s vision for Advanced Air Mobility (AAM) is to provide safe, sustainable, accessible, and affordable aviation for transformational local and intraregional missions and includes the transportation of passengers and cargo as well as aerial work missions, such as infrastructure inspection or search and rescue operations. NASA’s technical expertise, intergovernmental relationships, and high level of public trust will help this technology come to market concurrent with infrastructure readiness, public acceptance, and constructive regulation. By advising and integrating disparate AAM efforts across the country and working collaboratively with the FAA, the National Campaign (NC) objective is to motivate industry progress and support the development of policy, regulatory, and technical standards in a manner that best ensures public safety and benefit to the American people. NC began with the Dry Run and Developmental Test (NC-DT), which served as a pathfinder for collaboration and direct involvement with industry partners in integrated simulation exercises and flight tests. NC-DT culminated with an acoustics-gathering flight test in September 2021 with industry partner Joby’s prototype S4 2.0 air vehicle, which delivered the first foundational baseline of noise levels present in Electric Vertical Takeoff and Landing (eVTOL) vehicles. Through the DT phase, the NC team built and tested the airspace and range infrastructure while assessing the readiness level of industry partners leading up to future NC events. The purpose of this paper is to provide an overview of NC-DT.

National Campaign↗

Database Integrity Monitoring for Synthetic Vision Systems Using Machine Vision and SHADE

In an effort to increase situational awareness, the aviation industry is investigating technologies that allow pilots to visualize what is outside of the aircraft during periods of low-visibility. One of these technologies, referred to as Synthetic Vision Systems (SVS), provides the pilot with real-time computer-generated images of obstacles, terrain features, runways, and other aircraft regardless of weather conditions. To help ensure the integrity of such systems, methods of verifying the accuracy of synthetically-derived display elements using onboard remote sensing technologies are under investigation. One such method is based on a shadow detection and extraction (SHADE) algorithm that transforms computer-generated digital elevation data into a reference domain that enables direct comparison with radar measurements. This paper describes machine vision techniques for making this comparison and discusses preliminary results from application to actual flight data.

Cooper, Eric G.↗

Adventures in Astronomical Time Series Analysis

Welcome to a tour of the volatile, highly active Universe — in stark contrast to earlier serene ``clockwork’’ visions. Innovative data analysis techniques have illuminated explosive physical processes animating these systems. Examples include a Fourier transform suited to the irregular sampling characteristic of much astronomical data, but time domain techniques will be emphasized for these applications: gamma-ray activity in the Crab Nebula, gamma-ray bursts, active galactic nuclei, and gravitational waves. I hope this talk will change some of the ways you carry out statistical data analysis.

Jeffrey D Scargle↗

Automated Registration of Multi-Mode Nondestructive Evaluation Data

Registration techniques play a central role in applications of image processing to computer vision, medical imaging, and automatic target tracking. Feature-based techniques such as scale-invariant feature transform (SIFT) and speeded up robust features (SURF) are commonly used to register images derived from a single modality. However, SIFT and SURF struggle to register images from different modalities because the features tend to manifest rather differently and at sometimes very different length-scales. The most successful methods that have been developed to register multi-modal data use information-theoretic approaches. These methods play a key part in nondestructive evaluation scenarios where data that is collected by sensors of different modalities must be registered to be fused. In this paper, automated registration based on normalized mutual information is applied to align data derived from ultrasonic and radiographic inspections of (i) additively manufactured titanium alloy test coupons, and (ii) thin, lithium metal pouch-cell batteries. The quality of the registration is quantified in terms of computational resources and spatial accuracy. In the first case the X-ray computed tomography (XCT) data is captured on a region corresponding to a small subset of the ultrasonic data, while in the case of the lithium batteries the digital radiography (DR) captures a larger region of interest than the ultrasonic data. In both cases the radiographic data resolution is much higher than for ultrasound, but interestingly, in both cases the accuracy of the registration is approximately equal to two-to-three-pixel lengths in the ultrasonic images.

Nondestructive Evaluation↗

Visual detection of spatial contrast patterns: evaluation of five simple models

The ModelFest Phase One dataset is a collection of luminance contrast thresholds for 43 two-dimensional monochromatic spatial patterns confined to an area of approximately two by two degrees. These data were collected by a collaboration among twelve laboratories, and were designed to provide a common database for calibration and testing of spatial vision models. Here I report fits of the ModelFest data with five models: Peak Contrast, Contrast Energy, Generalized Energy, a Gabor Channels model, and a Discrete Cosine Transform model. The Gabor Channels model provides the best fit, though the other, simpler models, with the exception of Peak Contrast, provide remarkably good fits as well. Though there are clear individual differences, regularities in the data suggest the possibility of constructing a standard observer for spatial vision. c2000 Optical Society of America.

NASA Discipline Space Human Factors↗

Optical See-Through Head Mounted Display Direct Linear Transformation Calibration Robustness in the Presence of User Alignment Noise

Augmented Reality (AR) is a technique by which computer generated signals synthesize impressions that are made to coexist with the surrounding real world as perceived by the user. Human smell, taste, touch and hearing can all be augmented, but most commonly AR refers to the human vision being overlaid with information otherwise not readily available to the user. A correct calibration is important on an application level, ensuring that e.g. data labels are presented at correct locations, but also on a system level to enable display techniques such as stereoscopy to function properly [SOURCE]. Thus, vital to AR, calibration methodology is an important research area. While great achievements already have been made, there are some properties in current calibration methods for augmenting vision which do not translate from its traditional use in automated cameras calibration to its use with a human operator. This paper uses a Monte Carlo simulation of a standard direct linear transformation camera calibration to investigate how user introduced head orientation noise affects the parameter estimation during a calibration procedure of an optical see-through head mounted display.

Axholt, Magnus↗

Real-time optical multiple object recognition and tracking system and method

System for optically recognizing and tracking a plurality of objects within a field of vision. Laser (46) produces a coherent beam (48). Beam splitter (24) splits the beam into object (26) and reference (28) beams. Beam expanders (50) and collimators (52) transform the beams (26, 28) into coherent collimated light beams (26', 28'). A two-dimensional SLM (54), disposed in the object beam (26'), modulates the object beam with optical information as a function of signals from a first camera (16) which develops X and Y signals reflecting the contents of its field of vision. A hololens (38), positioned in the object beam (26') subsequent to the modulator (54), focuses the object beam at a plurality of focal points (42). A planar transparency-forming film (32), disposed with the focal points on an exposable surface, forms a multiple position interference filter (62) upon exposure of the surface and development processing of the film (32). A reflector (53) directing the reference beam (28') onto the film (32), exposes the surface, with images focused by the hololens (38), to form interference patterns on the surface. There is apparatus (16', 64) for sensing and indicating light passage through respective ones of the positions of the filter (62), whereby recognition of objects corresponding to respective ones of the positions of the filter (62) is affected. For tracking, apparatus (64) focuses light passing through the filter (62) onto a matrix of CCD's in a second camera (16') to form a two-dimensional display of the recognized objects.

Chao, Tien-Hsin↗

An Automated Classification Technique for Detecting Defects in Battery Cells

Battery cell defect classification is primarily done manually by a human conducting a visual inspection to determine if the battery cell is acceptable for a particular use or device. Human visual inspection is a time consuming task when compared to an inspection process conducted by a machine vision system. Human inspection is also subject to human error and fatigue over time. We present a machine vision technique that can be used to automatically identify defective sections of battery cells via a morphological feature-based classifier using an adaptive two-dimensional fast Fourier transformation technique. The initial area of interest is automatically classified as either an anode or cathode cell view as well as classified as an acceptable or a defective battery cell. Each battery cell is labeled and cataloged for comparison and analysis. The result is the implementation of an automated machine vision technique that provides a highly repeatable and reproducible method of identifying and quantifying defects in battery cells.

McDowell, Mark↗

Advances in image compression and automatic target recognition; Proceedings of the Meeting, Orlando, FL, Mar. 30, 31, 1989

Various papers on image compression and automatic target recognition are presented. Individual topics addressed include: target cluster detection in cluttered SAR imagery, model-based target recognition using laser radar imagery, Smart Sensor front-end processor for feature extraction of images, object attitude estimation and tracking from a single video sensor, symmetry detection in human vision, analysis of high resolution aerial images for object detection, obscured object recognition for an ATR application, neural networks for adaptive shape tracking, statistical mechanics and pattern recognition, detection of cylinders in aerial range images, moving object tracking using local windows, new transform method for image data compression, quad-tree product vector quantization of images, predictive trellis encoding of imagery, reduced generalized chain code for contour description, compact architecture for a real-time vision system, use of human visibility functions in segmentation coding, color texture analysis and synthesis using Gibbs random fields.

Tescher, Andrew G.↗

Effectively Transforming IMC Flight into VMC Flight: An SVS Case Study

A flight-test experiment was conducted using the NASA LaRC Cessna 206 aircraft. Four primary flight and navigation display concepts, including baseline and Synthetic Vision System (SVS) concepts, were evaluated in the local area of Roanoke Virginia Airport, flying visual and instrument approach procedures. A total of 19 pilots, from 3 pilot groups reflecting the diverse piloting skills of the GA population, served as evaluation pilots. Multi-variable Discriminant Analysis was applied to three carefully selected and markedly different operating conditions with conventional instrumentation to provide an extension of traditional analysis methods as well as provide an assessment of the effectiveness of SVS displays to effectively transform IMC flight into VMC flight.

Glaab, Louis J.↗

Onboard Image Registration from Invariant Features

This paper describes a feature-based image registration technique that is potentially well-suited for onboard deployment. The overall goal is to provide a fast, robust method for dynamically combining observations from multiple platforms into sensors webs that respond quickly to short-lived events and provide rich observations of objects that evolve in space and time. The approach, which has enjoyed considerable success in mainstream computer vision applications, uses invariant SIFT descriptors extracted at image interest points together with the RANSAC algorithm to robustly estimate transformation parameters that relate one image to another. Experimental results for two satellite image registration tasks are presented: (1) automatic registration of images from the MODIS instrument on Terra to the MODIS instrument on Aqua and (2) automatic stabilization of a multi-day sequence of GOES-West images collected during the October 2007 Southern California wildfires.

descriptors↗

Machine Learning Airport Surface Model

Future needs of the National Airspace System require decision support tools to adopt a service-oriented architecture in alignment with the FAA’s vision for an Info-Centric NAS. To achieve this, many existing systems will need to undergo a digital transformation from a monolithic decision support tool to a service-oriented architecture where individual services are exposed through well defined Application Programming Interfaces (APIs). To enable this transformation, NASA has developed the Digital Information Platform as a cloud based foundation for development of aviation services with a special focus towards Artificial Intelligence and Machine Learning (ML) services. This paper describes the work required for the transformation of NASA’s legacy surface management system to a real-time ML based decision support system deployed in the cloud. Details of the Machine Learning Operations (MLOps) infrastructure and best practices are described which enabled the end-toend lifecycle management of ML within an integrated software system. Validation results are provided from an operational field evaluation where performance was benchmarked against the legacy approach.

Jeremy Coupe↗