Search NASA⌕ Search

SEARCH · Search NASA

Results for “subset selections”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Aided Active Learning (AAL) for Enhanced Critical Heat Flux Prediction

Accurate prediction of critical heat flux (CHF) is crucial for the safe and efficient operation of nuclear reactors. Traditional CHF modeling methods often require extensive experimental data, which are hard to obtain. This study introduces the Aided Active Learning (AAL) framework, which strategically minimizes data requirements without sacrificing model accuracy. Unlike conventional Active Learning (AL), AAL introduces an additional step of randomly selecting a subset from the sample pool before applying the query strategy. To evaluate the performance of AAL, two query strategies—uncertainty-based sampling and error-reduction sampling—were evaluated across the following models: random forest (RF), feedforward neural network (FNN), and variational feedforward neural network (vFNN). The proposed framework demonstrated that AAL effectively reduces the number of training samples needed to achieve comparable predictive accuracy. For the RF model, AL required only 710 samples to achieve an R2 score of 0.98, as compared to the 4,785 samples needed by random sampling. Similarly, the FNN model achieved the same R2 score with just 355 samples when using AL, a significant improvement over the 825 samples required by random sampling. In case of uncertainty-based sampling strategy, vFNN attained an R2 of 0.98 with 3,420 samples, reducing the sample requirement by 47% relative to the 6,440 samples needed for random sampling. Its performance suggests that larger training data are required to fully leverage its uncertainty quantification capabilities.

22 - GENERAL STUDIES OF NUCLEAR REACTORS↗

An assessment of LANDSAT data acquisition history on identification and area estimation of corn and soybeans

Multitemporally registered LANDSAT MSS data from four acquisitions during the 1978 growing season were used in classification of eight sample segments in Iowa and Indiana. The results illustrate that use of LANDSAT acquisition when corn has tasseled is critical, as this is the optimum time for separation of corn and soybeans. An early season acquisition when the summer crops appear as bare soil can be beneficial in reducing the confusion between these two crops and other cover types. A subset of one visible and one infrared band from each date was found to produce results not significantly different from the use of all bands. Selection of a subset of these bands may also be feasible for multitemporal analysis.

Hixson, M. M.↗

Evaluation of criteria for selecting the spectral attributes of digital LANDSAT MSS imagery for discriminating lithological units in the lower Curaca River Valley, Bahia

The use of spectral attributes criteria was investigated, based on measures of statistical distance of separability between thematic classes in MSS digital LANDSAT imagery, in order to select the best subsets of channels in composite colors for the detection and discrimination of lithological units in the lower valley of Curaca River, State of Bahia, Brazil. Three situations were investigated: (1) selection of the three best channels, considering all of the original bands (channels 4, 5, 6, and 7); (2) selection of the three best bands, considering the six MSS band-ratios (channels 4/5, 4/6. 4/7, 5/6, 5/7, and 6/7); and (3) selection of the three best bands in a hybrid approach (the four original bands and the six ratios). A visual analysis was done on color composite images using the selected sets. Results show that the hybrid product (bands 4, 5/7, and 7 with green, blue, and red respectively) and the Normal Color Composite (bands 4, 5, and 7 with blue, green, and red colors respectively) had the best performance.

Paradella, W. R.↗

Comparative Ergonomic Evaluation of Spacesuit and Space Vehicle Design

With the advent of the latest human spaceflight objectives, a series of prototype architectures for a new launch and reentry spacesuit that would be suited to the new mission goals. Four prototype suits were evaluated to compare their performance and enable the selection of the preferred suit components and designs. A consolidated approach to testing was taken: concurrently collecting suit mobility data, seat-suit-vehicle interface clearances, and qualitative assessments of suit performance within the volume of a Multi-Purpose Crew Vehicle mockup. It was necessary to maintain high fidelity in a mockup and use advanced motion-capture technologies in order to achieve the objectives of the study. These seemingly mutually exclusive goals were accommodated with the construction of an optically transparent and fully adjustable frame mockup. The construction of the mockup was such that it could be dimensionally validated rapidly with the motioncapture system. This paper describes the method used to create a space vehicle mockup compatible with use of an optical motion-capture system, the consolidated approach for evaluating spacesuits in action, and a way to use the complex data set resulting from a limited number of test subjects to generate hardware requirements for an entire population. Kinematics, hardware clearance, anthropometry (suited and unsuited), and subjective feedback data were recorded on 15 unsuited and 5 suited subjects. Unsuited subjects were selected chiefly based on their anthropometry in an attempt to find subjects who fell within predefined criteria for medium male, large male, and small female subjects. The suited subjects were selected as a subset of the unsuited medium male subjects and were tested in both unpressurized and pressurized conditions. The prototype spacesuits were each fabricated in a single size to accommodate an approximately average-sized male, so select findings from the suit testing were systematically extrapolated to the extremes of the population to anticipate likely problem areas. This extrapolation was achieved by first comparing suited subjects performance with their unsuited performance, and then applying the results to the entire range of the population. The use of a transparent space vehicle mockup enabled the collection of large amounts of data during human-in-the-loop testing. Mobility data revealed that most of the tested spacesuits had sufficient ranges of motion for the selected tasks to be performed successfully. A suited subject's inability to perform a task most often stemmed from a combination of poor field of view in a seated position, poor dexterity of the pressurized gloves, or from suit/vehicle interface issues. Seat ingress and egress testing showed that problems with anthropometric accommodation did not exclusively occur with the largest or smallest subjects, but also with specific combinations of measurements that led to narrower seat ingress/egress clearance.

England, Scott↗

Combinatorial Reasoning: Selecting Reasons in Generative AI Pipelines via Combinatorial Optimization

Recent Large Language Models (LLMs) have demonstrated impressive capabilities at tasks that require human intelligence and are a significant step towards human-like artificial intelligence (AI). Yet the performance of LLMs at reasoning tasks have been subpar and the reasoning capability of LLMs is a matter of significant debate. While it has been shown that the choice of the prompting technique to the LLM can alter its performance on a multitude of tasks, including reasoning, the best performing techniques require human-made prompts with the knowledge of the tasks at hand. We introduce a framework for what we call Combinatorial Reasoning (CR), a fully-automated prompting method, where reasons are sampled from an LLM pipeline and mapped into a Quadratic Unconstrained Binary Optimization (QUBO) problem. The framework investigates whether QUBO solutions can be profitably used to select a useful subset of the reasons to construct a Chain-of-Thought style prompt. We explore the acceleration of CR with specialized solvers. We also investigate the performance of simpler zero-shot strategies such as linear majority rule or random selection of reasons. Our preliminary study indicates that coupling a combinatorial solver to generative AI pipelines is an interesting avenue for AI reasoning and elucidates design principles for future CR methods.

combinatorial reasoning↗

Selecting Data from a Star Catalog

MCDUMP is a computer program that selects data from the SKYMAP SKY2000 Master Star Catalog a database about 150 MB in size, stored on a computer hard drive. The database describes about 300,000 stars, each by means of a 500-byte entry. MCDUMP reads all 300,000 entries, then generates an output file that comprises a subset of entries selected according to one or more criteria entered by the user. Examples of criteria that could be entered include: location in a selected portion of the sky; constancy or a specified degree of variability of brightness; absence of nearby, bright companion stars; a particular surface temperature; and brightness sufficient to enable detection by a specified astronomical instrument. The output of MCDUMP can be in the form of either a single 520-column file or multiple files that contain fewer columns to facilitate printing. MCDUMP has been configured and tested for use under the HP-UX 10.20 operating system (a Hewlett-Packard version of the UNIX operating system). It should also be possible to adapt MCDUMP to other versions of UNIX.

Tracewell, David A.↗

Self-Organizing-Map Program for Analyzing Multivariate Data

SOM_VIS is a computer program for analysis and display of multidimensional sets of Earth-image data typified by the data acquired by the Multi-angle Imaging Spectro-Radiometer [MISR (a spaceborne instrument)]. In SOM_VIS, an enhanced self-organizing-map (SOM) algorithm is first used to project a multidimensional set of data into a nonuniform three-dimensional lattice structure. The lattice structure is mapped to a color space to obtain a color map for an image. The Voronoi cell-refinement algorithm is used to map the SOM lattice structure to various levels of color resolution. The final result is a false-color image in which similar colors represent similar characteristics across all its data dimensions. SOM_VIS provides a control panel for selection of a subset of suitably preprocessed MISR radiance data, and a control panel for choosing parameters to run SOM training. SOM_VIS also includes a component for displaying the false-color SOM image, a color map for the trained SOM lattice, a plot showing an original input vector in 36 dimensions of a selected pixel from the SOM image, the SOM vector that represents the input vector, and the Euclidean distance between the two vectors.

Li, P. Peggy↗

Dimensionality Reduction Through Classifier Ensembles

In data mining, one often needs to analyze datasets with a very large number of attributes. Performing machine learning directly on such data sets is often impractical because of extensive run times, excessive complexity of the fitted model (often leading to overfitting), and the well-known "curse of dimensionality." In practice, to avoid such problems, feature selection and/or extraction are often used to reduce data dimensionality prior to the learning step. However, existing feature selection/extraction algorithms either evaluate features by their effectiveness across the entire data set or simply disregard class information altogether (e.g., principal component analysis). Furthermore, feature extraction algorithms such as principal components analysis create new features that are often meaningless to human users. In this article, we present input decimation, a method that provides "feature subsets" that are selected for their ability to discriminate among the classes. These features are subsequently used in ensembles of classifiers, yielding results superior to single classifiers, ensembles that use the full set of features, and ensembles based on principal component analysis on both real and synthetic datasets.

Oza, Nikunj C.↗

A statistical analysis of the broadband 0.1 to 3.5 keV spectral properties of X-ray-selected active galactic nuclei

We survey the broadband spectral properties of approximately 500 X-ray-selected active galactic nuclei (AGNs) observed with the Einstein Observatory. Included in this survey are the approximately 450 AGNs in the Extended Medium Sensitivity Survey (EMSS) of Gioia et al. (1990) and the approximately 50 AGNs in the Ultrasoft Survey of Cordova et al. (1992). We present a revised version of the latter sample, based on the post publication discovery of a software error in the Einstein Rev-1b processing. We find that the mean spectral index of the AGNs between 0.1 and 0.6 keV is softer, and the distribution of indices wider, than previous estimates based on analyses of the X-ray spectra of optically selected AGNs. A subset of these AGNs exhibit flux variabiulity, some on timescales as short as 0.05 days. A correlation between radio and hard X-ray luminosity is confirmed, but the data do not support a correlation between the radio and soft X-ray luminosities, or between radio loudness and soft X-ray spectral slope. Evidence for physically distinct soft and hard X-ray components is found, along with the possibility of a bias in previous optically selected samples toward selection of AGNs with flatter X-ray spectra.

Thompson, R. J.↗

TADPLOT program, version 2.0: User's guide

The TADPLOT Program, Version 2.0 is described. The TADPLOT program is a software package coordinated by a single, easy-to-use interface, enabling the researcher to access several standard file formats, selectively collect specific subsets of data, and create full-featured publication and viewgraph quality plots. The user-interface was designed to be independent from any file format, yet provide capabilities to accommodate highly specialized data queries. Integrated with an applications software network, data can be assessed, collected, and viewed quickly and easily. Since the commands are data independent, subsequent modifications to the file format will be transparent, while additional file formats can be integrated with minimal impact on the user-interface. The graphical capabilities are independent of the method of data collection; thus, the data specification and subsequent plotting can be modified and upgraded as separate functional components. The graphics kernel selected adheres to the full functional specifications of the CORE standard. Both interface and postprocessing capabilities are fully integrated into TADPLOT.

Hammond, Dana P.↗

Plateau to River Model Predictive Simulations for All Ensemble Realizations to Support Modeling Work in Fiscal Year 2025

The purpose of this environmental calculation file (ECF) is to document predictions of flow and hydraulic head on the Central Plateau of the Hanford Site using the Plateau-to-River (P2R) Model (CP-57037, Model Package Report for the Plateau-to-River Model: Version 9.1). This calculation documents the simulation of the groundwater for the parent model domain of the P2R Model as a basis for use in other applications of the P2R Model. This application is unique from the standpoint that it will simulate all ensemble member models of the P2R Model whereas other applications may only utilize specific ensemble members. These simulations provide results that can be used in the process of selecting an appropriate subset of ensemble members for other applications.

54 ENVIRONMENTAL SCIENCES↗

GOES-16 Data for LASSO-CACTI Overview Paper

GOES-16 L1b satellite radiances have been obtained for the LASSO-CACTI case dates. Specifically, the period in the ARM subset is for select days in the period October 26, 2018 through March 15, 2019. These files were downloaded from Amazon Web Services using the GOES-2-Go library, https://blaylockbk.github.io/goes2go/_build/html/.

{"GOES-16 band 13",radiance,LASSO-CACTI}↗

Water quality parameter measurement using spectral signatures

Regression analysis is applied to the problem of measuring water quality parameters from remote sensing spectral signature data. The equations necessary to perform regression analysis are presented and methods of testing the strength and reliability of a regression are described. An efficient algorithm for selecting an optimal subset of the independent variables available for a regression is also presented.

White, P. E.↗

Rapid processing of multispectral scanner data using linear techniques.

Tests have been made to compare linear and quadratic techniques for processing multispectral scanner data. The tests have been limited to a few selected sets of agricultural data. Two aspects of processing were studied. The first, the selection of a subset of channels to be used in the decision function, was found to be faster by a factor of 50 when a linear method was used. Second, in recognition processing, our linear decision rule produced a lower error rate and utilized a larger number of channels for equal processing times. Nevertheless, when the criterion is lowest possible error rate, irregardless of processing time, the quadratic rule is preferable.

Crane, R. B.↗

Quality and use of ERTS radiometric information in geologic applications

Some techniques are described for making full use of the data contained in an ERTS MSS image. Only about one-fourth of the data in a single band can be displayed at one time on a black and white image; therefore, when all four bands are considered, only about 7% of the available data can be used by the interpreter. Selecting the proper subset of information for the photointerpreter is therefore a necessity. Ratio methods exclude the brightness information from the display. A field study in one area using a portable spectrometer has shown only fair correlation with ERTS radiometry after one normalization procedure. Plots of brightness of test areas with sun angle show discrepancies. Plots of ratios show discrepancies of lesser magnitude, although the error limits are large.

Goetz, A. F. H.↗

An algorithm for selecting a radiation transport subgrid for ablation and radiation coupled hypersonic viscous shock-layer problems

An algorithm is described for selecting a grid subset for calculating radiative transport. The subset spacing is determined by using the variation in aerothermal properties across the full grid of the shock layer. Results show that a radiation grid subset of 15 to 20 points can be used for viscous-shock layer calculations where approximately 50 grid points are required to define the aerothermal profiles. Results are presented for various planetary entry conditions with and without mass injection to demonstrate both the validity and utility of the algorithm.

Bolz, C. W., Jr.↗

An algorithm for a single machine scheduling problem with sequence dependent setup times and scheduling windows

An enumeration algorithm is presented for solving a scheduling problem similar to the single machine job shop problem with sequence dependent setup times. The scheduling problem differs from the job shop problem in two ways. First, its objective is to select an optimum subset of the available tasks to be performed during a fixed period of time. Secondly, each task scheduled is constrained to occur within its particular scheduling window. The algorithm is currently being used to develop typical observational timelines for a telescope that will be operated in earth orbit. Computational times associated with timeline development are presented.

Moore, J. E.↗

The reduction, verification and interpretation of Magsat magnetic data over Canada

The primary concern of this investigation is to detect and study variations in the magnetic field originating in the solid Earth, as measured by Magsat. Most of this field originates in the core, but an important part of the field is of lithospheric origin. Magnetic anomalies of lithospheric origin are weak at Magsat altitudes (20 to 30 nT at most), and they can easily be masked by much larger effects caused by field aligned and other currents at high latitudes. Most of Canada lies under the influence of ionospheric currents in the auroral zone and polar cap. Therefore, before Magsat data had become available, but after the October 30, 1979 launch, criteria were developed for selecting times when subsets of potentially usable Magsat data could be expected. Subsequently, as Magsat data became available, these critieria were applied.

Coles, R. L.↗