Search NASA⌕ Search

SEARCH · Search NASA

Results for “Computer systems performance”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 649 records · Page 36

Versatile, Fast Computer Core

Versatile computer core serves as state-of-the-art component and tool for development of computing systems required to process data rapidly, particularly as part of control tasks involving relatively large volumes of input and output data. Exploits new technology to enhance performance needed in flight computers. Computing and other equipment specific to flight system added around this core to develop flight systems rapidly without incurring time and monetary costs of designing new core.

Ross, Douglas↗

The complex structured singular value

A tutorial introduction to the complex structured singular value (mu) is presented, with an emphasis on the mathematical aspects of mu. The mu-based methods discussed here have been useful for analyzing the performance and robustness properties of linear feedback systems. Several tests for robust stability and performance with computable bounds for transfer functions and their state space realizations are compared, and a simple synthesis problem is studied. Uncertain systems are represented using linear fractional transformations which naturally unify the frequency-domain and state space methods.

Packard, A.↗

Biocybernetic Adaptation Strategies: Machine Awareness of Human Engagement for Improved Operational Performance

Human operators interacting with machines or computers continually adapt to the needs of the system ideally resulting in optimal performance. In some cases, however, deteriorated performance is an outcome. Adaptation to the situation is a strength expected of the human operator which is often accomplished by the human through self-regulation of mental state. Adaptation is at the core of the human operator's activity, and research has demonstrated that the implementation of a feedback loop can enhance this natural skill to improve training and human/machine interaction. Biocybernetic adaptation involves a “loop upon a loop,” which may be visualized as a superimposed loop which senses a physiological signal and influences the operator’s task at some point. Biocybernetic adaptation in, for example, physiologically adaptive automation employs the “steering” sense of “cybernetic,” and serves a transitory adaptive purpose – to better serve the human operator by more fully representing their responses to the sys- tem. The adaptation process usually makes use of an assessment of transient cog- nitive state to steer a functional aspect of a system that is external to the operator’s physiology from which the state assessment is derived. Therefore, the objective of this paper is to detail the structure of biocybernetic systems regarding the level of engagement of interest for adaptive systems, their processing pipeline, and the adaptation strategies employed for training purposes, in an effort to pave the way towards machine awareness of human state for self-regulation and improved operational performance.

Stephens, Chad↗

Analysis and interpretation of arterial sounds using a small clinical computer system

A small mobile bed-side computer system is described that is capable of performing phonoangiographic analyses as well as many other common data analysis tasks in a hospital. The clinical application of phonoangiography is found to be greatly facilitated by the computer-provided availability of data acquisition and analysis capabilities.

Dewey, C. F., Jr.↗

Autonomous Integrated Receive System (AIRS) requirements definition. Volume 2: Design and development

Functional requirements and specifications are defined for an autonomous integrated receive system (AIRS) to be used as an improvement in the current tracking and data relay satellite system (TDRSS), and as a receiving system in the future tracking and data acquisition system (TDAS). The AIRS provides improved acquisition, tracking, bit error rate (BER), RFI mitigation techniques, and data operations performance compared to the current TDRSS ground segment receive system. A computer model of the AIRS is used to provide simulation results predicting the performance of AIRS. Cost and technology assessments are included.

Chie, C. M.↗

Real-time optical multiple-object recognition and tracking demonstration: A friendly challenge to the digital field

Researchers demonstrated the first optical multiple object tracking system. The system is capable of simultaneous tracking of multiple objects, each with independent movements in real-time, limited only to the TV frame rate (30 msec). In order to perform a similar tracking operation, a large computer system and very complex software would be needed. Although researchers have demonstrated the tracking of only 3 objects, the system capacity can easily be expanded by 2 orders of magnitude.

Chao, Tien-Hsin↗

A Hands-On Curriculum for Training in HPC Cluster Deployment and Management

This paper presents the design, methodology, and outcomes of the High-Performance Computing Technologies (HPCT) course, a hands-on training program focused on the system-side of HPC cluster deployment and administration. Delivered as part of the Master in High Performance Computing (MHPC) program, the course introduces students to key concepts in cluster configuration, including networking, software stack provisioning, job scheduling, and monitoring. Initially taught in person, the course was transitioned to an online format during the COVID-19 pandemic. This shift led to the development of openly available instructional material and a flipped-classroom approach that continues to support both in-person and hybrid delivery. All course materials are publicly available at www.hpc.temple.edu/mhpc/hpc-technology/index.html. By documenting the structure, infrastructure, and evolution of HPCT, this paper offers a model for accessible HPC system training that supports workforce development in computational science.

Posada Correa, Fernando [ORNL] (ORCID:000000022565↗

Fabrication and Testing of Solid-Solution Strengthened Corrosion Resistant Alloys For Service in Molten Fluoride Environments

The demand for higher system thermal efficiencies requires the operation of power generation cycles and heat conversion systems at progressively higher temperatures. As the system operating temperature increases, existing materials may not provide adequate mechanical properties or environmental compatibility or both. There is an increasing commercial interest in the development and deployment of liquid-fueled Molten Salt Reactors (MSRs). Hastelloy®N, the highest performing candidate MSR structural alloy, is not capable of operations at temperatures above 700°C, thus limiting the performance of these systems. Using an Integrated Computational Materials Engineering (ICME)-approach and small laboratory scale heats, ORNL developed a class of patented alloys covered by U.S. Patent 9, 435, 011 B2, “Creep-resistant, Cobalt-free alloys for high temperature, liquid-salt heat exchanger systems,” similar to Hastelloy®N in that they are primarily solid solution strengthened. In contrast to precipitation strengthened alloys, the microstructure of solid solution alloys and hence the high temperature mechanical properties are stable for extended periods of time allowing long reactor operating life. The new alloys have shown to possess good resistance to liquid fluorides at temperatures up to 850°C and have significantly improved creep properties when compared to Hastelloy®N. The purpose of the CRADA project was for ORNL to collaborate with Haynes International- a materials producer, MetalTek International- a foundry, and Kairos Power – an advanced reactor developer – to scale-up selected alloys, evaluate their properties, and identify one solid solution strengthened alloy that can meet the property requirements for the reactor being developed by Kairos Power and other similar liquid fluoride-salt cooled reactors. As part of the project, eight alloys were down-selected and fabricated in larger industrial scale heats by Haynes International. Resistance to molten salt was evaluated in flowing FLiNaK and FLiBe by Kairos Power using their Rotating Cage Loop (RCL) system. Accounting for iron deposition during these tests, the new alloys displayed very low net mass change showing excellent corrosion performance in molten salt. Creep properties evaluated at ORNL were found to be better than that of Hastelloy®N and 316 stainless steel. Long-term stabilities of the alloys evaluated by Haynes International showed that these alloys have excellent thermal stability in the temperature range 704.4-815.6°C, with the change in strength and ductility being less than 10-15% after a 4000-hour exposure at 815.6°C. Autogenously Gas Tungsten Arc Welding (GTAW) welded samples showed less than 10% change in yield strength / ultimate tensile strength / total elongation compared to the basemetal, indicating that the alloys have excellent weldability. Three parts were successfully investment-cast using one alloy with very little voiding showing feasibility of fabricating parts using the casting process. This project enabled extensive interaction between the material producer Haynes International, casting supplier MetalTek, and reactor developer Kairos Power. This facilitated testing of materials and components produced using the newly developed alloys by the end-user. This allowed the generation of critical dataset required for down-selection of a few promising alloys for further development. This data is also currently being shared with other reactor designers for them to evaluate the suitability of this alloy for their reactor design. The availability of this alloy will ultimately enable the design and development and deployment of MSRs with increased temperature of operation and thus, improved efficiencies.

99 GENERAL AND MISCELLANEOUS↗

Fabrication and Testing of Solid-Solution Strengthened Corrosion Resistant Alloys For Service in Molten Fluoride Environments

The demand for higher system thermal efficiencies requires the operation of power generation cycles and heat conversion systems at progressively higher temperatures. As the system operating temperature increases, existing materials may not provide adequate mechanical properties or environmental compatibility or both. There is an increasing commercial interest in the development and deployment of liquid-fueled Molten Salt Reactors (MSRs). Hastelloy®N, the highest performing candidate MSR structural alloy, is not capable of operations at temperatures above 700°C, thus limiting the performance of these systems. Using an Integrated Computational Materials Engineering (ICME)-approach and small laboratory scale heats, ORNL developed a class of patented alloys covered by U.S. Patent 9,435,011 B2, “Creep-resistant, Cobalt-free alloys for high temperature, liquid-salt heat exchanger systems,” similar to Hastelloy®N in that they are primarily solid solution strengthened. In contrast to precipitation strengthened alloys, the microstructure of solid solution alloys and hence the high temperature mechanical properties are stable for extended periods of time allowing long reactor operating life. The new alloys have shown to possess good resistance to liquid fluorides at temperatures up to 850°C and have significantly improved creep properties when compared to Hastelloy®N. The purpose of the CRADA project was for ORNL to collaborate with Haynes International- a materials producer, MetalTek International- a foundry, and Kairos Power – an advanced reactor developer – to scale-up selected alloys, evaluate their properties, and identify one solid solution strengthened alloy that can meet the property requirements for the reactor being developed by Kairos Power and other similar liquid fluoride-salt cooled reactors. As part of the project, eight alloys were down-selected and fabricated in larger industrial scale heats by Haynes International. Resistance to molten salt was evaluated in flowing FLiNaK and FLiBe by Kairos Power using their Rotating Cage Loop (RCL) system. Accounting for iron deposition during these tests, the new alloys displayed very low net mass change showing excellent corrosion performance in molten salt. Creep properties evaluated at ORNL were found to be better than that of Hastelloy®N and 316 stainless steel. Long-term stabilities of the alloys evaluated by Haynes International showed that these alloys have excellent thermal stability in the temperature range 704.4-815.6°C, with the change in strength and ductility being less than 10-15% after a 4000-hour exposure at 815.6°C. Autogenously Gas Tungsten Arc Welding (GTAW) welded samples showed less than 10% change in yield strength / ultimate tensile strength / total elongation compared to the basemetal, indicating that the alloys have excellent weldability. Three parts were successfully investment-cast using one alloy with very little voiding showing feasibility of fabricating parts using the casting process. This project enabled extensive interaction between the material producer Haynes International, casting supplier MetalTek, and reactor developer Kairos Power. This facilitated testing of materials and components produced using the newly developed alloys by the end-user. This allowed the generation of critical dataset required for down-selection of a few promising alloys for further development. This data is also currently being shared with other reactor designers for them to evaluate the suitability of this alloy for their reactor design. The availability of this alloy will ultimately enable the design and development and deployment of MSRs with increased temperature of operation and thus, improved efficiencies.

36 MATERIALS SCIENCE↗

IRIS: Exploring Performance Scaling of the Intelligent Runtime System and its Dynamic Scheduling Policies

High-Performance Computing is becoming increasingly heterogeneous, relying on a diverse mix of hardware to achieve good performance. Paradoxically, current drivers and frameworks for these devices typically require separate languages and implementations for each vendor. Furthermore, there are few tools and little support to schedule codes between these devices in a truly heterogeneous manner-partly because of this fragmentation between vendors and the languages each supports. To overcome both limitations, the Intelligent Runtime System (IRIS) was developed. It allows a common task abstraction to automatically be shared among contemporary vendors and is run from a single host-side API. At runtime, IRIS queries the host system and registers which frameworks and drivers are available, these determine which kernels can be used by the scheduler-CPUs via OpenMP, Nvidia GPUs (CUDA), AMD GPUs (HIP), and Intel and Xilinx FPGAs with OpenCL. IRIS enables tasks to be scheduled to any heterogeneous device and resolves to the appropriate kernel binary at runtimeit only uses the devices supported by the system on which it is run. IRIS supports single-task and graph-based expressions of dependencies of tasks. Additionally, IRIS features a range of dynamic scheduling policies, allowing complex chains of tasks and interactions to be executed, relieving the programmer/user from considering the system to assign tasks to devices optimally. This paper presents the peak performance attainable by IRIS over a range of systems-each with different numbers and types of accelerator devices, it highlights the flexibility of IRIS since these devices are truly heterogeneous, relying on different backends (drivers, frameworks, and languages) which historically required unique implementations to utilize them. We then use this peak performance as a baseline to compare increasingly complex chains of tasks (with increasingly complex task dependencies) and evaluate how IRIS copes. Finally, we consider the performance of different IRIS scheduling policies on this range of task graphs.

Johnston, Beau↗

Performance assessment of aero-assisted orbital transfer vehicles

Aero-assisted orbital transfer vehicles are analyzed. The aerodynamic characteristics over the flight profile and three- and six-degree-of-freedom performance analyses were determined. The important results, to date, are: (1) the aerodynamic preliminary analysis system, an interactive computer program, used to predict the aerodynamics (performance, stability, and control) for these vehicles; (2) the performance capability, e.g., maximum inclination change, maximum heating rate, and maximum sensed acceleration, can be determined using continuum aerodynamics only; (3) guidance schemes can be developed that allow for errors in atmospheric density prediction, mispredicted trim angle of attack, and off-nominal atmospheric interface conditions, even for vehicles with a low lift-to-drag ratio; and (4) multiple pass trajectories can be used to reduce the maximum heating rate.

Powell, R. W.↗

Improved neutron activation prediction code system development

Two integrated neutron activation prediction code systems have been developed by modifying and integrating existing computer programs to perform the necessary computations to determine neutron induced activation gamma ray doses and dose rates in complex geometries. Each of the two systems is comprised of three computational modules. The first program module computes the spatial and energy distribution of the neutron flux from an input source and prepares input data for the second program which performs the reaction rate, decay chain and activation gamma source calculations. A third module then accepts input prepared by the second program to compute the cumulative gamma doses and/or dose rates at specified detector locations in complex, three-dimensional geometries.

Saqui, R. M.↗

A parallel-vector algorithm for rapid structural analysis on high-performance computers

A fast, accurate Choleski method for the solution of symmetric systems of linear equations is presented. This direct method is based on a variable-band storage scheme and takes advantage of column heights to reduce the number of operations in the Choleski factorization. The method employs parallel computation in the outermost DO-loop and vector computation via the loop unrolling technique in the innermost DO-loop. The method avoids computations with zeros outside the column heights, and as an option, zeros inside the band. The close relationship between Choleski and Gauss elimination methods is examined. The minor changes required to convert the Choleski code to a Gauss code to solve non-positive-definite symmetric systems of equations are identified. The results for two large scale structural analyses performed on supercomputers, demonstrate the accuracy and speed of the method.

Storaasli, Olaf O.↗

A parallel-vector algorithm for rapid structural analysis on high-performance computers

A fast, accurate Choleski method for the solution of symmetric systems of linear equations is presented. This direct method is based on a variable-band storage scheme and takes advantage of column heights to reduce the number of operations in the Choleski factorization. The method employs parallel computation in the outermost DO-loop and vector computation via the 'loop unrolling' technique in the innermost DO-loop. The method avoids computations with zeros outside the column heights, and as an option, zeros inside the band. The close relationship between Choleski and Gauss elimination methods is examined. The minor changes required to convert the Choleski code to a Gauss code to solve non-positive-definite symmetric systems of equations are identified. The results for two large-scale structural analyses performed on supercomputers, demonstrate the accuracy and speed of the method.

Storaasli, Olaf O.↗

Resilience of the Electric Grid Through Trustable IoT-Coordinated Assets

The electricity grid has evolved from a physical system to a cyberphysical system with digital devices that perform measurement, control, communication, computation, and actuation. The increased penetration of distributed energy resources (DERs) including renewable generation, flexible loads, and storage provides extraordinary opportunities for improvements in efficiency and sustainability. However, they can introduce new vulnerabilities in the form of cyberattacks, which can cause significant challenges in ensuring grid resilience. We propose a framework in this paper for achieving grid resilience through suitably coordinated assets including a network of Internet of Things devices. A local electricity market is proposed to identify trustable assets and carry out this coordination. Situational Awareness (SA) of locally available DERs with the ability to inject power or reduce consumption is enabled by the market, together with a monitoring procedure for their trustability and commitment. With this SA, we show that a variety of cyberattacks can be mitigated using local trustable resources without stressing the bulk grid. Multiple demonstrations are carried out using a high-fidelity cosimulation platform, real-time hardware-in-the-loop validation, and a utility-friendly simulator.

distributed energy resources↗

Digital Radar-Signal Processors Implemented in FPGAs

High-performance digital electronic circuits for onboard processing of return signals in an airborne precipitation- measuring radar system have been implemented in commercially available field-programmable gate arrays (FPGAs). Previously, it was standard practice to downlink the radar-return data to a ground station for postprocessing a costly practice that prevents the nearly-real-time use of the data for automated targeting. In principle, the onboard processing could be performed by a system of about 20 personal- computer-type microprocessors; relative to such a system, the present FPGA-based processor is much smaller and consumes much less power. Alternatively, the onboard processing could be performed by an application-specific integrated circuit (ASIC), but in comparison with an ASIC implementation, the present FPGA implementation offers the advantages of (1) greater flexibility for research applications like the present one and (2) lower cost in the small production volumes typical of research applications. The generation and processing of signals in the airborne precipitation measuring radar system in question involves the following especially notable steps: The system utilizes a total of four channels two carrier frequencies and two polarizations at each frequency. The system uses pulse compression: that is, the transmitted pulse is spread out in time and the received echo of the pulse is processed with a matched filter to despread it. The return signal is band-limited and digitally demodulated to a complex baseband signal that, for each pulse, comprises a large number of samples. Each complex pair of samples (denoted a range gate in radar terminology) is associated with a numerical index that corresponds to a specific time offset from the beginning of the radar pulse, so that each such pair represents the energy reflected from a specific range. This energy and the average echo power are computed. The phase of each range bin is compared to the previous echo by complex conjugate multiplication to obtain the mean Doppler shift (and hence the mean and variance of the velocity of precipitation) of the echo at that range.

Berkun, Andrew↗

Implementing Access to Data Distributed on Many Processors

A reference architecture is defined for an object-oriented implementation of domains, arrays, and distributions written in the programming language Chapel. This technology primarily addresses domains that contain arrays that have regular index sets with the low-level implementation details being beyond the scope of this discussion. What is defined is a complete set of object-oriented operators that allows one to perform data distributions for domain arrays involving regular arithmetic index sets. What is unique is that these operators allow for the arbitrary regions of the arrays to be fragmented and distributed across multiple processors with a single point of access giving the programmer the illusion that all the elements are collocated on a single processor. Today's massively parallel High Productivity Computing Systems (HPCS) are characterized by a modular structure, with a large number of processing and memory units connected by a high-speed network. Locality of access as well as load balancing are primary concerns in these systems that are typically used for high-performance scientific computation. Data distributions address these issues by providing a range of methods for spreading large data sets across the components of a system. Over the past two decades, many languages, systems, tools, and libraries have been developed for the support of distributions. Since the performance of data parallel applications is directly influenced by the distribution strategy, users often resort to low-level programming models that allow fine-tuning of the distribution aspects affecting performance, but, at the same time, are tedious and error-prone. This technology presents a reusable design of a data-distribution framework for data parallel high-performance applications. Distributions are a means to express locality in systems composed of large numbers of processor and memory components connected by a network. Since distributions have a great effect on the performance of applications, it is important that the distribution strategy is flexible, so its behavior can change depending on the needs of the application. At the same time, high productivity concerns require that the user be shielded from error-prone, tedious details such as communication and synchronization.

James, Mark↗

Regen: An object layout regenerator on large-scale production HPC systems

This article proposes an object layout regenerator called Regen which regenerates and removes the object layout dynamically to improve the read performance of applications. Regen first detects frequent access patterns from the I/O requests of the applications. Second, Regen reorganizes the objects and regenerates or preallocates new object layouts according to the identified access patterns. Finally, Regen removes or reuses the obsolete or regenerated object layouts as necessary. As a result, Regen accelerates access to objects by providing a flexible object layout. We implement Regen as a framework on top of Proactive Data Container (PDC) and evaluate it on Cori supercomputer, a production-scale HPC system, by using realistic HPC I/O benchmarks. The experimental results show that Regen improves the I/O performance by up to 16.92 × compared with an existing system.

Distributed file system↗