Search NASASearch

SEARCH · Search NASA

Results for “test protocol”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Efficiently improving the performance of noisy quantum computers

Using near-term quantum computers to achieve a quantum advantage requires efficient strategies to improve the performance of the noisy quantum devices presently available. We develop and experimentally validate two efficient error mitigation protocols named "Noiseless Output Extrapolation" and "Pauli Error Cancellation" that can drastically enhance the performance of quantum circuits composed of noisy cycles of gates. By combining popular mitigation strategies such as probabilistic error cancellation and noise amplification with efficient noise reconstruction methods, our protocols can mitigate a wide range of noise processes that do not satisfy the assumptions underlying existing mitigation protocols, including non-local and gate-dependent processes. We test our protocols on a four-qubit superconducting processor at the Advanced Quantum Testbed. We observe significant improvements in the performance of both structured and random circuits, with up to 86 % improvement in variation distance over the unmitigated outputs. Our experiments demonstrate the effectiveness of our protocols, as well as their practicality for current hardware platforms.

97 MATHEMATICS AND COMPUTING

NASA'S Standard Measures During Bed Rest: Adaptations in the Cardiovascular System

Bed rest is a well-accepted analog of space flight that has been used extensively to investigate physiological adaptations in a larger number of subjects in a shorter amount of time than can be studied with space flight and without the confounding effects associated with normal mission operations. However, comparison across studies of different bed rest durations, between sexes, and between various countermeasure protocols have been hampered by dissimilarities in bed rest conditions, measurement protocols, and testing schedules. To address these concerns, NASA instituted standard bed rest conditions and standard measures for all physiological disciplines participating in studies conducted at the Flight Analogs Research Unit (FARU) at the University of Texas-Medical Branch. Investigators for individual studies employed their own targeted study protocols to address specific hypothesis-driven questions, but standard measures tests were conducted within these studies on a non-interference basis to maximize data availability while reducing the need to implement multiple bed rest studies to understand the effects of a specific countermeasure. When possible, bed rest standard measures protocols were similar to tests nominally used for medically-required measures or research protocols conducted before and after Space Shuttle and International Space Station missions. Specifically, bed rest standard measures for the cardiovascular system implemented before, during, and after bed rest at the FARU included plasma volume (carbon monoxide rebreathing), cardiac mass and function (2D, 3D and Doppler echocardiography), and orthostatic tolerance testing (15- or 30-minutes of 80 degree head-up tilt). Results to-date indicate that when countermeasures are not employed, plasma volume decreases and the incidence of presyncope during head-up tilt is more frequent even after short-duration bed rest while reductions in cardiac function and mass are progressive as bed rest duration increases. Additionally, while plasma volume loss can be corrected and cardiac mass can be prevented with properly applied countermeasures, orthostatic tolerance is more difficult to protect when supine exercise is the only countermeasure. Similar results have been observed after space flight. Plasma volume, cardiac chamber volume, and orthostatic tolerance recover relatively quickly with resumption of ambulation and normal activity levels after bed rest but restoration of cardiac mass is prolonged.

Lee, Stuart M. C.

Quality Assurance Program Plan for SFR Metallic Fuel Data Qualification

This document contains an evaluation of the applicability of the current Quality Assurance Standards from the American Society of Mechanical Engineers Standard NQA-1 (NQA-1) criteria and identifies and describes the quality assurance process(es) by which attributes of historical, analytical, and other data associated with sodium-cooled fast reactor [SFR] metallic fuel will be evaluated. This process is being instituted to facilitate validation of data to the extent that such data may be used to support future licensing efforts associated with advanced reactor designs. The initial data to be evaluated under this program were generated during the US Integral Fast Reactor program between 1984-1994, where the data include, but are not limited to, research and development data and associated documents, test plans and associated protocols, operations and test data, technical reports, and information associated with past United States Nuclear Regulatory Commission reviews of SFR designs. It is recognized that managing the data generated by large research and development projects presents a significant challenge for retaining data integrity and availability. American Society of Mechanical Engineers Standard NQA-1 (NQA-1) 2008/2009a provides appropriate requirements for this plan.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Shuttle bioresearch laboratory breadboard simulations

Laboratory breadboard simulations (Tests I and II) were conducted to test concepts and assess problems associated with bioresearch support equipment, facilities, and operational integration for conducting manned earth orbital Shuttle missions. This paper describes Test I and discusses the major observations made in Test II. The tests emphasized candidate experiment protocols and requirements: Test I for biological research and Test II for crew members (simulated), subhuman primates, and radioisotope tracer studies on lower organisms. The procedures and approaches developed for these simulation activities could form the basis for Spacelab simulations and developing preflight integration, testing, and logistics of flight payloads.

Taketa, S. T.

Determination of Material Properties Near the Glass Transition Temperature for an Isogrid Boom

Experiments were performed and results obtained to determine the temperature dependence of the modulus of elasticity for a thermoplastic isogrid tube. The isogrid tube was subjected to axial tensile loads of 0-100 lbf and strain was measured at room and elevated temperatures of 100, 120, 140, 160, 180, 190, and 200 F. These were based on tube manufacturer specifying an incorrect glass transition temperature of 210 F. Two protocols were used. For the first protocol the tube was brought to temperature and a tensile test performed. The tube was allowed to cool between tests. For the second protocol the tube was ramped to the desired test temperature and held. A tensile test was performed and the tube temperature ramped to the next test temperature. The second protocol spanned the entire test range. The strain rate was constant at 0.008 in/min. Room temperature tests resulted in the determination of an average modulus of 2.34 x 106 Psi. The modulus decreased above 100 F. At 140 F the modulus had decreased by 7.26%. The two test protocols showed good agreement below 160 F. At this point the glass transition temperature had been exceeded. The two protocols were not repeated because the tube failed.

Blandino, Joseph R.

IPHE Regulations Codes and Standards Working Group - Type IV COPV Round Robin Testing

This manuscript presents the results of a multi-lateral international activity intended to understand how to execute a cycle stress test as specified in a chosen standard (GTR, SAE, ISO, EIHP...). The purpose of this work was to establish a harmonized test method protocol to ensure that the same results would be achieved regardless of the testing facility. It was found that accurate temperature measurement of the working fluid is necessary to ensure the test conditions remain within the tolerances specified. Continuous operation is possible with adequate cooling of the working fluid but this becomes more demanding if the cycle frequency increases. Recommendations for future test system design and operation are presented.

Maes, M.

Results of the Field Test and Defining Sensormotor Fitness for Duty Standards

The primary goals of the Field Test were (1) to determine functional abilities associated with long-duration space flight crews beginning as soon after landing as possible (< 2 hours), and (2) to characterize the time course of recovery with additional follow-up measurement sessions within 24 hours after landing. The NASA and Russian teams have collected data on a total of 39 different United States Orbital Segment (USOS) and Russian crewmembers, with 9 Russian crewmembers being tested twice (total of 48 tests). Eighteen subjects (7 Russian, 11 USOS) completed a reduced Field Test (pilot) protocol and 30 subjects (19 Russian and 11 USOS) participated in the full Field Test. This presentation will focus on a subset of the measures that will be used in the upcoming ground study to determine sensorimotor fitness for duty standards, namely: tandem walk, obstacle walk, eye-hand coordination task, finger-to-nose task, and Computerized Dynamic Posturography (CDP). Exploration-class missions including Artemis, Gateway, and beyond will require a new level of autonomy around periods of gravitational transition, where sensorimotor disturbances increase. The operational support that is available upon return to Earth including rescue teams, medical interventions, and the ability to rest as needed will not be available after landing on the lunar or Martian surface. Because of this, there is a need to define fitness for duty standards that will help inform crew capabilities during and soon after gravitational transitions. Due to the new requirements of the exploration environment, we must utilize a set of exploration field measures, for which no previous spaceflight data exists to define fitness for duty standards. A Sensorimotor Adaptation Analog (SAA) that can provide different levels of acute disorientation through combined vestibular, visual, and proprioceptive disruptions will be used to increase the range of performance in exploration field measures, simulating the moderate-to-severe performance decrements observed in spaceflight. The levels of SAA will be titrated and validated by comparison to gold standard measures that have a wealth of spaceflight data at different time points during recovery. These data will be pulled directly from the results of Field Test and help ensure the range of performance while being exposed to SAA mimics the range of performance from pre-flight to immediately postflight in Field Test subjects. Specifically, we will be using the tandem walk, obstacle walk, eye-hand coordination task, finger-to-nose task, and CDP as the gold standard measures. Referencing this Field Test data in the form of gold standard measures will also help us characterize and contextualize how each magnitude of SAA disorientation compares to recovery from long-term microgravity exposure.

M J F Rosenberg

Design of a 2-Hour Prebreathe Protocol for Space Walks (EVAs) from the International Space Station (ISS)

The majority of extravehicular activities (EVAs) performed from the shuttle use a 10.2 psi staged decompression. The International Space Station (ISS) will operate at 14.7 psi, requiring crews to "campout" in the airlock at 10.2 psi. The constraints associated with campout (crew isolation, oxygen usage, and waste management), provided the rationale to develop a 2-hour prebreathe protocol from 14.7 psi. Previous studies on the affect of microgravity and exercise during prebreathe suggested the feasibility of this approach. Various combinations of adynamia (nonwalking subjects), prebreathe exercise doses, and space suit donning options (10.2 vs. 14.7 psi) were analyzed against timeline and consumable constraints. Prospective decompression sickness (DCS) and venous gas emboli (VGE) accept/reject criteria were defined from statistical analysis of historical DCS data, combined with risk management of DCS under ISS mission circumstances. Maximum operational DCS levels were defined based on protecting for EVA capability with two crew members at 95% confidence, throughout ISS lifetime (within the constraints of NASA DCS disposition policy JPG 1800.3). The accept / reject limits were adjusted for greater safety (including Grade IV VGE criteria) based on analysis of related medical factors. Monte-Carlo simulation was performed to design a closed sequential, multi-center laboratory trial, including the capability of rejecting the primary protocol and testing at least one alternate exercise dose, within the 2-hour prebreathe. The 2-hour protocol incorporates 0, breathing for 5 0 min at 14.7 psi, including 10 min dual cycle ergometry at 75%VO(2max). It requires an additional 30 minO2breathing during depress from 14.7 to 10.2 psi, followed by a 30-60 min suit donning break at 10.2 psi/26.5% O2. It concludes with a 40 min in-suit O2 prebreathe. The protocol would be accepted for operations, if the incidence of DCS was less than 15% and Grade IV VGE less than 20%, both at 95% confidence. The above protocol and accept/reject limits were implemented in a multi-center study.

Gernhardt, M. L.

Human Trials of a 2-Hour Prebreathe Protocol

We evaluate 2-hour prebreathe protocols combining simulated microgravity and exercise during prebreathe with the objective of validating a protocol for use on International Space Station (ISS). The protocol was tested with four different exercise doses during prebreathe in a multi-center trial involving three laboratories. Subject selection, Doppler monitoring techniques for venous gas emboli (VGE), test termination criteria, and definitions of decompression sickness (DCS) were standardized in all laboratories. The Phase II protocol met the accept criteria for a prebreathe procedure for use by astronauts during assembly and maintenance of the ISS Dual-cycle ergometry or light exercise individually was not sufficient to protect against DCS at acceptable levels. The combination of both was successful.

Butler, Bruce D.

Viability of Small Dimension Crew Quarters for Surface Habitation

It is possible that in the next twenty years NASA may fly crew quarters on twice as many spacecraft as it has in the past fifty years. In short, US experience with spacecraft crew quarters is limited and with few available standards to guide their design there is significant uncertainty facing spacecraft currently in development, several of which are also subject to substantial mass and volume challenges. Those spacecraft developments will face considerable pressure to minimize crew quarters size, including those intended for use on the lunar surface. Given that a crew quarters is the only space a crew member can call his or her own during missions that can last weeks to years in duration, providing an appropriate volume is especially important. This is even more critical when one considers the reality that all crew quarters flown to date have been smaller than minimum standards for US jail cells. This research will categorize functional capabilities of crew quarters and explore physical and virtual prototypes of small crew quarters that have attempted to include these capabilities. The Exploration Atmospheres Test at NASA Johnson Space Center represents the first opportunity to collect multi-day test data on crew quarters of this size in a gravitational environment. Intended to validate exploration prebreathe protocols, this test will house eight people inside a vacuum chamber that has been outfitted as a habitat prototype for twelve days. In addition to their other test activity, the crew will evaluate the acceptability of their crew quarters. This data will aid in establishing design guidelines for crew quarters in both short and long duration missions beyond low Earth orbit.

Crew Quarters

Crew report

A 56-day chamber simulation of Skylab was successfully completed. The atmosphere (5 psi, 70 percent oxygen, 30 percent nitrogen, 5 mm carbon dioxide) and medical features including a 21-day pre- and 18-day post-test medical protocols were closely simulated. No apparent crew health problems were induced by the atmosphere, semiclosed environment, or other test features; and no appreciable crew degradation appeared over this period. The chamber and associated systems performed without major problems.

Bobko, K. J.

Spaceflight Autonomous Multigenerational Microbial Sequencer (SAMMS) in Support of Plant-Growth Systems

As the National Aeronautics and Space Association (NASA) begins to pursue long-duration space flights, they will need to be able to provide astronauts with a nutritious and reliable food source. To meet the administration’s goal of traveling to the Moon and Mars, astronauts will need to begin to grow their own food in space. To protect their food source, extensive monitoring will occur to test for the effects of a space flight environment (e.g., radiation) as well as for early pathogen and disease detection. Genomic sequencing allows for both concerns to be tested on a regular basis. However, NASA’s current sequencer is unable to process plant tissues. Therefore, a novel method for plant DNA extraction using microneedle (MN) patches that will be able to feed into NASA’s existing system, but also require minimal human input is proposed. To support this, the design was broken down into four components (1) MN patch fabrication (2) MN patch extraction, (3) automated sampling motion control, and (4) a processing module. The MN patch is fabricated using a custom mold with conically shaped needles. The mold is filled with Polyvinyl alcohol (PVA) solution and placed in a vacuum desiccator. The mold is left in the vacuum overnight until the patch is dry and ready for use. The protocol was tested with varying pressures, drying times, volume amounts, and preparation methods to determine if highquality needles can be produced. A MN is a method of DNA extraction where the patch is applied to a leaf, the needles penetrate the leaf, breaking the rigid plant cell wall to isolate the DNA. A protocol for this method of extraction was tested to ensure the patch could produce the needed yield and purity. The tests varied by the number of patches, number of applications, and plant type. To automate the MN extraction method, motion control will utilize two separate axis tables which move in the x and y directions. The y-axis table will have an end effector that fits a MN patch and will have the ability to apply the patch to the leaf sample. This end effector will also act as a lid for a downstream processing module. The other axis will position the leaf sample and processing container so that the patch can be applied accurately. The Joint Comprehensive Sequencing System (JCSS) module integrates all the components together. The output of this module feeds into the NASA Charged Information-Storage Polymer Preparation System (CHIPPS) for genomic sequencing. The extraction module operates using a series of syringes and tubing to pump the varying reagents needed for the extraction protocol. The results of the study proved that MN patches are a viable method of DNA extraction. While fabrication of high-quality needles was unsuccessful, the protocol was able to be further developed using centrifugation. The integrated design between the motion control and the JCSS enabled the potential for automation with a complete conceptual design and prototype. Future research and development for this study would include (1) further testing for fabrication (2) expanding the range of plant species compatible with the MN patch, and (3) building a working prototype for the integrated system.

Peter Ling

Mars Sample Return: Risk Management & Sample Safety Assessment

Returning samples from Mars has long been a major planetary science objective due to the high scientific value and transformative potential of the resulting data. An exciting dimension of this objective is the potential for the detection of ancient microbiological life, and the possibility of improving our understanding of the evolution of habitable environments on Mars and the development of life on Earth. To ensure that returned samples meet stringent planetary protection requirements and do not expose Earth to potential biohazards, the joint NASA/ESA Sample Receiving Project (SRP) assembled the Sample Safety Assessment Protocol Tiger Team (SSAP-TT). Members were recruited with the specific goal of creating a multi-disciplinary and internationally distributed team of experts in their respective fields across the federal government, academia, and private industry. This team was chartered with reassessing previous sample safety assessment strategies, defining what constitutes a biological hazard, developing a protocol to test for potential biohazards, and establishing a statistical framework to determine if samples are “safe” for release. The team developed a three-step protocol, supported by a Bayesian statistical framework, to assess whether returned samples contain potential biohazards that could present a risk to Earth’s biosphere. Initial conclusions indicated that an effective and comprehensive safety assessment protocol is feasible using modern techniques and does not require an excessive amount of sample consumption or traditional microbiological detection methodology. Herein, we will present an overview of the MSR SRP, the proposed safety assessment protocol, and how aspects of this novel approach can be applied to biological assessment in healthcare product manufacturing practices.

Alvin L Smith

Hardware Testing and Implementation of RapidIO Protocol on the ISAAC iBoard for Use in the NEXUS Testbed

I am assisting the NEXUS (NEXtbUS) and ISAAC (Instrument ShAred Artifact for Computing) teams in achieving a highly reusable, highly configurable FPGA (Field Programmable Gate Array) system which uses a unified, high-speed bus standard. One of my tasks is to verify that all new ISAAC iBoards features are functioning as expected by using my previous builds to test these features and resolving any errors found. My other task is to investigate and implement the RapidIO protocol and to implement it onto the new ISAAC iBoards to be used in the NEXUS testbed. Completing these tasks will demonstrate the potential of both NEXUS and ISAAC technology. This will allow others to see the power in making use of a unified, modular system.

RapidIO Protocol

Characterization of Low Frequency Auditory Filters

The purpose of this study is to characterize auditory filters at low frequencies, defined as below about 100 Hz. Three experiments were designed and executed. They were conducted in the Exterior Effects Room at the NASA Langley Research Center, a psychoacoustic facility designed for presentation of aircraft flyover sounds to groups of test subjects. The first experiment measured 36 subjects’ hearing threshold for pure tones (at 25, 31.5, 40, 50, 63 and 80 Hz) in “quiet” conditions. The subjects, male and female, had a wide age range. This experiment allowed the performance of the test facility to be assessed and also provided screened test subjects for participation in subsequent experiments. The second and third experiments used 20 and 10 test subjects, respectively, and measured psychophysical tuning curves (PTCs) that describe auditory filters with center frequencies of approximately 63 and 50 Hz. The latter is assumed to be the lowest (bottom) auditory filter; thus, sounds at frequencies below about 50 Hz are perceived via the lower skirt of this lowest filter. All experiments used an adaptive, three-alternative forced-choice test procedure using either variable level tones or variable level, narrowband noise maskers. Measured PTCs were found to be very similar to other recently published data, both in terms of mean values and intersubject variation, despite different experimental protocols, different test facilities, and a wide range in subjects’ age.

Rafaelof, Menachem

Sizing Single Cantilever Beam Specimens for Characterizing Facesheet/Core Peel Debonding in Sandwich Structure

This technical publication details part of an effort focused on the development of a standardized facesheet/core peel debonding test procedure. The purpose of the test is to characterize facesheet/core peel in sandwich structure, accomplished through the measurement of the critical strain energy release rate associated with the debonding process. Following an examination of previously developed tests and a recent evaluation of a selection of these methods, a single cantilever beam (SCB) specimen was identified as being a promising candidate for establishing such a standardized test procedure. The objective of the work described here was to begin development of a protocol for conducting a SCB test that will render the procedure suitable for standardization. To this end, a sizing methodology was developed to ensure appropriate SCB specimen dimensions are selected for a given sandwich system. Application of this method to actual sandwich systems yielded SCB specimen dimensions that would be practical for use. This study resulted in the development of a practical SCB specimen sizing method, which should be well-suited for incorporation into a standardized testing protocol.

Ratcliffe, James G.

AI Model Benchmarking for Nonproliferation Applications: Steel Thread Benchmarking Task Force Technical Report (Rev. 2)

Steel Thread is a NA-22 venture that seeks to build trustworthy, reliable AI models that can be used in a wide variety of nonproliferation tasks. A key aspect of building these models is developing appropriate benchmarks and evaluation methods, which will enable the venture to identify and adapt models to provide the most value in the nonproliferation domain. Benchmarks must be relevant to key tasks in this domain, such as question answering, information retrieval, document summarization and classification, consensus analysis, and image and data analysis. This report 1) provides an overview of benchmark design, evaluation, and challenges; 2) reviews a variety of open benchmarks, with a focus on language models and tasks; and 3) identifies benchmarks that are most relevant to Steel Thread. This report is intended to serve as a basis for further efforts to classify and evaluate benchmarks and their correlation with success on nonproliferation-specific tasks. The Steel Thread venture has defined benchmarks to be a particular combination of a dataset (or datasets) and a metric (or metrics) conceptualized as representing one or more specific tasks or sets of abilities for a specific modality. It is adopted by a research community as a shared framework for comparing methods.1 It includes 1) Data: Labeled (a designated subset not used for training, which could be all the data), 2) Metric: A way to quantify performance, 3) Task/Ability: What the benchmark is testing, 4) Protocol: A structured and repeatable evaluation process, 5) Baseline/Reference Model: For comparison; could be statistical, rule-based, SME-derived, or another model, and 6) Maintenance Plan: to update with new information over time; important for long-term utility. For further clarity, the definition includes what a benchmark, in this context, is not. It is not a corpus of training data, specific to a model (it is intended to apply to a range of models), a universal evaluation of performance, a guarantee that the ‘top’ model on the leaderboard will be the best fit for every specific use case, an all-encompassing proof of a model’s universal quality, nor is it a one-size-fits-all measure of success. It does not cover every real-world constraint (like operational, ethical, or cost considerations), a systems integration test, or a unit test. This definition was inspired by and resulted from discussions within the Steel Thread Benchmarking Task Force. This group was formed to define what we would mean as a benchmark within Steel Thread but persisted as the need to develop a thorough understanding of the large and expanding existing benchmarking space. This technical report is a result of the group’s divide and conquer approach to exploring this space. The release of benchmarks might not be progressing as quickly as model development, but it is moving very fast, as many benchmarks quickly become saturated, when state-of-the-art models score so close to the benchmark’s ceiling that their results are virtually indistinguishable. At that point, the test no longer differentiates between new systems, so researchers usually stop reporting scores as the benchmark no longer informs about improvements from the next generation of models. In the OpenAI announcement of GPT-5, they reported results on six flagship public benchmarks (AIME 2025, SWE-bench Verified, Aider Polyglot, MMMU, HealthBench Hard, GPQA) but the full system-card covers roughly thirty-five separate evaluations, comprising hundreds of test task items in total. There have been some efforts to summarize benchmarks in specific fields, like for text-to-image generation, but these surveys have had a narrow methodology scope. Therefore, a comprehensive survey of all benchmarks or even all benchmarks that could be relevant to Steel Thread is outside of the scope of this report. We chose some specific benchmarks to investigate in detail.

97 MATHEMATICS AND COMPUTING