Search NASA⌕ Search

SEARCH · Search NASA

Results for “Support for Usability Evaluation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

115 records · Page 7

Development of a Safety Hazards Risk Assessment Tool for Uncrewed Aircraft System Traffic Management during Preflight Planning

Tremendous growth in the uncrewed and remotely piloted vehicle market is expected in low-altitude, uncontrolled airspace, resulting in potential decreases in safety without systems that support monitoring, assessing, and mitigating risk. At NASA, the System-Wide Safety (SWS) project has been developing a suite of data-driven tools to predict hazards so that the potential risks that these hazards pose can be mitigated. Services to predict various hazards have been developed, including battery capacity, proximity to static obstacles, population risks, global positioning system signal strength, radio frequency spectrum interference risk, and vertiport congestion. These services can monitor hazards along a flight path and if any risks posed by these hazards exceed a threshold, the uncrewed aircraft system (UAS) fleet manager can be alerted to mitigate the risk by modifying the flight path, changing the scheduled departure or arrival times, and/or diverting the vehicle to an alternate vertiport. These services were originally developed to monitor and assess risks during flight, but they have been adapted to assess hazard risks prior to departure so that a fleet manager can evaluate the potential risks for a fleet of UAS along their planned flight paths. These services have been integrated into a prototype tool called the Supplemental Data Service Provider-Consolidated Dashboard (SDSP-CD), developed at NASA Ames Research Center. The tool consists of a dashboard which provides a comprehensive overview for a number of risks and a map display that shows the details of the hazards along each flight’s path. Based on the findings from three previous studies, the SDSP-CD has been updated with new design elements and functions. In this paper, we describe lessons learned from the previous studies, changes made to the interface, and the feedback received during a follow-up usability study. Overall, participants reported that there is a substantial benefit of having a fleet manager use a consolidated dashboard to assess hazards for the vehicles in their fleet and to provide situational awareness to potential risks so that they can be mitigated prior to flight. Once the SDSP-CD matures, it will need to be integrated into flight and mission planning tools. Some initial thoughts on how this integration should be accomplished are shared in this paper. Finally, the functional differences between preflight vs. in-flight risk assessment and the differences in fleet manager vs. UAS pilot roles that may require different information and user interactions are discussed.

preflight↗

Modified Advanced Crew Escape Suit Intravehicular Activity Suit for Extravehicular Activity Mobility Evaluations

The use of an intravehicular activity (IVA) suit for a spacewalk or extravehicular activity (EVA) was evaluated for mobility and usability in the Neutral Buoyancy Laboratory (NBL) environment at the Sonny Carter Training Facility near NASA Johnson Space Center in Houston, Texas. The Space Shuttle Advanced Crew Escape Suit was modified to integrate with the Orion spacecraft. The first several missions of the Orion Multi-Purpose Crew Vehicle will not have mass available to carry an EVA-specific suit; therefore, any EVA required will have to be performed by the Modified Advanced Crew Escape Suit (MACES). Since the MACES was not designed with EVA in mind, it was unknown what mobility the suit would be able to provide for an EVA or whether a person could perform useful tasks for an extended time inside the pressurized suit. The suit was evaluated in multiple NBL runs by a variety of subjects, including crewmembers with significant EVA experience. Various functional mobility tasks performed included: translation, body positioning, tool carrying, body stabilization, equipment handling, and tool usage. Hardware configurations included with and without Thermal Micrometeoroid Garment, suit with IVA gloves and suit with EVA gloves. Most tasks were completed on International Space Station mock-ups with existing EVA tools. Some limited tasks were completed with prototype tools on a simulated rocky surface. Major findings include: demonstrating the ability to weigh-out the suit, understanding the need to have subjects perform multiple runs prior to getting feedback, determining critical sizing factors, and need for adjusting suit work envelope. Early testing demonstrated the feasibility of EVA's limited duration and limited scope. Further testing is required with more flight-like tasking and constraints to validate these early results. If the suit is used for EVA, it will require mission-specific modifications for umbilical management or Primary Life Support System integration, safety tether attachment, and tool interfaces. These evaluations are continuing through calendar year 2014.

Watson, Richard D.↗

Methods for semi-automated indexing for high precision information retrieval

OBJECTIVE: To evaluate a new system, ISAID (Internet-based Semi-automated Indexing of Documents), and to generate textbook indexes that are more detailed and more useful to readers. DESIGN: Pilot evaluation: simple, nonrandomized trial comparing ISAID with manual indexing methods. Methods evaluation: randomized, cross-over trial comparing three versions of ISAID and usability survey. PARTICIPANTS: Pilot evaluation: two physicians. Methods evaluation: twelve physicians, each of whom used three different versions of the system for a total of 36 indexing sessions. MEASUREMENTS: Total index term tuples generated per document per minute (TPM), with and without adjustment for concordance with other subjects; inter-indexer consistency; ratings of the usability of the ISAID indexing system. RESULTS: Compared with manual methods, ISAID decreased indexing times greatly. Using three versions of ISAID, inter-indexer consistency ranged from 15% to 65% with a mean of 41%, 31%, and 40% for each of three documents. Subjects using the full version of ISAID were faster (average TPM: 5.6) and had higher rates of concordant index generation. There were substantial learning effects, despite our use of a training/run-in phase. Subjects using the full version of ISAID were much faster by the third indexing session (average TPM: 9.1). There was a statistically significant increase in three-subject concordant indexing rate using the full version of ISAID during the second indexing session (p < 0.05). SUMMARY: Users of the ISAID indexing system create complex, precise, and accurate indexing for full-text documents much faster than users of manual methods. Furthermore, the natural language processing methods that ISAID uses to suggest indexes contributes substantially to increased indexing speed and accuracy.

Evaluation Studies↗

Mitigating Headward Fluid Shifts with Venoconstrictive Thigh Cuffs during Spaceflight

Venoconstrictive thigh cuffs (VTC) are a mechanical countermeasure capable of attenuating the spaceflight induced headward fluid shift, and thus may be a viable spaceflight associated neuro-ocular syndrome (SANS) countermeasure. Crewmembers can use VTC to mitigate the headward fluid shift to aid in adapting to spaceflight. However, data are needed to determine if VTC affect ocular structures. PURPOSE The purpose of this study is to determine the efficacy of long duration use of VTC application to mitigate the spaceflight-induced headward fluid shift. We hypothesize that a VTC countermeasure will temporarily reverse the headward fluid shift and attenuate spaceflight-induced changes of internal jugular vein (IJV) cross-sectional area, IJV pressure, stroke volume, cardiac output, intraocular pressure (IOP), and optic nerve head and retinal morphology. METHODS This study will evaluate the effectiveness of VTC countermeasure application on the headward fluid shift, as well as cardiovascular and ocular variables. VTC during spaceflight will be worn for an extended duration (up to 6 hours) with data collected at three time points (30 minutes, 3 hours, and 6 hours) to characterize the temporal profile of key fluid shift outcome measures of the vascular fluid shift, IOP, and ocular structure changes. Ten astronauts will be recruited to participate and will be studied before and during approximately 180-day International Space Station (ISS) spaceflight missions. Baseline data collection will occur approximately 90-days before launch (Figure). Preflight, each leg of the crewmember will be measured to determine the VTC cuff size and a cuff fit check session will occur prior to the preflight baseline ground imaging to verify the appropriate fit measured via a surface contact pressure. Prior to donning the VTC, baseline measures without VTC will be collected seated, supine, and supine followed by data collection with the VTC. The baseline data collection will include ultrasound (IJV area and pressure, stroke volume and cardiac output), brachial blood pressure and heart rate, optical coherence tomography (OCT) imaging (total retinal thickness and choroid thickness), and IOP. An inflight cuff fit check session, same as the preflight fit check, will occur prior to VTC use on ISS. The inflight VTC experiment will be conducted early (FD45) and late (R-45) to determine if mission duration affects VTC fit and the efficacy of fluid redistribution. A system usability scale comfort questionnaire will be included in each VTC session to capture feedback from crewmembers regarding the comfort and usability of the VTC. SCIENTIFIC & MISSION IMPACT Results will narrow knowledge gaps described in the Human Research Roadmap to mitigate the headward fluid shift during spaceflight and help NASA to 1) determine the efficacy of extended use VTC application to mitigate the spaceflight-induced headward fluid shift and 2) further the understanding of the use of VTC on vascular fluid shifts, IOP, and ocular structure during spaceflight. Supported by NASA Human Research Program Directed Research. Figure. Detailed Testing Schedule. Baseline data collection at L-90 will be performed in a randomized order of position.

J V Jasien↗

Technology Development and Advanced Planning for Curation of Returned Mars Samples

NASA Johnson Space Center (JSC) curates extraterrestrial samples, providing the international science community with lunar rock and soil returned by the Apollo astronauts, meteorites collected in Antarctica, cosmic dust collected in the stratosphere, and hardware exposed to the space environment. Curation comprises initial characterization of new samples, preparation and allocation of samples for research, and clean, secure long-term storage. The foundations of this effort are the specialized cleanrooms (class 10 to 10,000) for each of the four types of materials, the supporting facilities, and the people, many of whom have been doing detailed work in clean environments for decades. JSC is also preparing to curate the next generation of extraterrestrial samples. These include samples collected from the solar wind, a comet, and an asteroid. Early planning and R\&D are underway to support post-mission sample handling and curation of samples returned from Mars. One of the strong scientific reasons for returning samples from Mars is to search for evidence of current or past life in the samples. Because of the remote possibility that the samples may contain life forms that are hazardous to the terrestrial biosphere, the National Research Council has recommended that all samples returned from Mars be kept under strict biological containment until tests show that they can safely be released to other laboratories. It is possible that Mars samples may contain only scarce or subtle traces of life or prebiotic chemistry that could readily be overwhelmed by terrestrial contamination . Thus, the facilities used to contain, process, and analyze samples from Mars must have a combination of high-level biocontainment and organic / inorganic chemical cleanliness that is unprecedented. JSC has been conducting feasibility studies and developing designs for a sample receiving facility that would offer biocontainment at least the equivalent of current maximum containment BSL-4 (BioSafety Level 4) laboratories, while simultaneously maintaining cleanliness levels equaling those of state-of-the-art cleanrooms. Unique requirements for the processing of Mars samples have inspired a program to develop handling techniques that are much more precise and reliable than the approach (currently used for lunar samples) of employing gloved human hands in nitrogen-filled gloveboxes. Individual samples from Mars are expected to be much smaller than lunar samples, the total mass of samples returned by each mission being 0.5- 1 kg, compared with many tens of kg of lunar samples returned by each of the six Apollo missions. Smaller samples require much more of the processing to be done under microscopic observation. In addition, the requirements for cleanliness and high-level containment would be difficult to satisfy while using traditional gloveboxes. JSC has constructed a laboratory to test concepts and technologies important to future sample curation. The Advanced Curation Laboratory includes a new-generation glovebox equipped with a robotic arm to evaluate the usability of robotic and teleoperated systems to perform curatorial tasks. The laboratory also contains equipment for precision cleaning and the measurement of trace organic contamination.

Lindstrom, David J.↗

Using a Large Language Model as a Building Block to Generate Usable Validation and Verification Suite for OpenMP

In the HPC area, both hardware and software move quickly. Often new hardware is developed and deployed, the corresponding software stack, including compilers and other tools, are under active development while leading edge software developers are working to port and tune their applications, all at the same time. While the software ecosystem is in flux, one of the key challenges for users is obtaining insight into the state of implementation of key features in the programming languages and models their applications are using – whether they have been implemented, and whether the implementation conforms to the specification, especially for newly implemented features (less tested by widespread use). OpenMP is one of the most prominent shared memory programming models used for on-node programming in HPC. With the shift towards accelerators (such as GPUs and FPGAs) and heterogeneous programming OpenMP features are getting more complex. It is natural to ask whether generative AI approaches, and large language models (LLMs) in particular, can help in producing validation and verification test suites to allow users better and faster insights into the availability and correctness of OpenMP features of interest. In this work, we explore the use of ChatGPT-4 to generate a suite of tests for OpenMP features. We have chosen a set of directives and clauses, a total of 78 combinations, which first appeared in OpenMP 3.0 (released in May 2008) but are also relevant for accelerators. We prompted ChatGPT to generate tests in the C and Fortran languages, for both host (CPU) and device (accelerator). On the Summit super-computer using the GNU implementation, we found that, of the 78 generated tests 67 C tests and 43 Fortran tests compiled successfully and fewer than those executed to completion. On further analysis we show that not all generated tests are valid. We document the process, results, and provide detailed analysis regarding the quality of tests generated. With the aim of providing input to a production quality validation and verification suite, we manually implement the corrections required to make the tests valid according to the current OpenMP specification. We quantify this effort as small, medium, or large, and record the lines of code changed to correct the invalid tests. With the corrected tests we validate recent implementations from HPE, AMD, and GNU on the Frontier supercomputer. Our experiment and subsequent analysis show that although LLMs are capable of producing HPC specific codes, they are limited by their understanding of the deeper semantics and restrictions of programming models such as OpenMP. Unsurprisingly more commonly used features have better support, while some OpenMP 3.0 directives such as sections and tasking are not universally supported on accelerators. We demonstrate that successful compilation and execution to completion are inadequate metrics for evaluating generated code and that, at this time, commodity LLMs require expert intervention for code verification. This points to gaps in the training data that is currently available for HPC. We demonstrate that with "small" effort 37% of generated invalid C tests and 63% of generated invalid Fortran tests could be corrected. This improves productivity of test generation as we circumvent writing from scratch and the common programming errors associated with it.

Pophale, Swaroop [ORNL] (ORCID:0000000185446367)↗

Human-Robot Interaction Directed Research Project

Human-robot interaction (HRI) is a discipline investigating the factors affecting the interactions between humans and robots. It is important to evaluate how the design of interfaces and command modalities affect the human's ability to perform tasks accurately, efficiently, and effectively when working with a robot. By understanding the effects of interface design on human performance, workload, and situation awareness, interfaces can be developed to appropriately support the human in performing tasks with minimal errors and with appropriate interaction time and effort. Thus, the results of research on human-robot interfaces have direct implications for the design of robotic systems. This DRP concentrates on three areas associated with interfaces and command modalities in HRI which are applicable to NASA robot systems: 1) Video Overlays, 2) Camera Views, and 3) Command Modalities. The first study focused on video overlays that investigated how Augmented Reality (AR) symbology can be added to the human-robot interface to improve teleoperation performance. Three types of AR symbology were explored in this study, command guidance (CG), situation guidance (SG), and both (SCG). CG symbology gives operators explicit instructions on what commands to input, whereas SG symbology gives operators implicit cues so that operators can infer the input commands. The combination of CG and SG provided operators with explicit and implicit cues allowing the operator to choose which symbology to utilize. The objective of the study was to understand how AR symbology affects the human operator's ability to align a robot arm to a target using a flight stick and the ability to allocate attention between the symbology and external views of the world. The study evaluated the effects type of symbology (CG and SG) has on operator tasks performance and attention allocation during teleoperation of a robot arm. The second study expanded on the first study by evaluating the effects of the type of navigational guidance (CG and SG) on operator task performance and attention allocation during teleoperation of a robot arm through uplinked commands. Although this study complements the first study on navigational guidance with hand controllers, it is a separate investigation due to the distinction in intended operators (i.e., crewmembers versus ground-operators). A third study looked at superimposed and integrated overlays for teleoperation of a mobile robot using a hand controller. When AR is superimposed on the external world, it appears to be fixed onto the display and internal to the operators' workstation. Unlike superimposed overlays, integrated overlays often appear as three-dimensional objects and move as if part of the external world. Studies conducted in the aviation domain show that integrated overlays can improve situation awareness and reduce the amount of deviation from the optimal path. The purpose of the study was to investigate whether these results apply to HRI tasks, such as navigation with a mobile robot. HRP GAPS This HRI research contributes to closure of HRP gaps by providing information on how display and control characteristics - those related to guidance, feedback, and command modalities - affect operator performance. The overarching goals are to improve interface usability, reduce operator error, and develop candidate guidelines to design effective human-robot interfaces.

Sandor, Aniko↗