Search NASA⌕ Search

SEARCH · Search NASA

Results for “Discrimination Learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Optimal invariant sets for atomistic machine learning

The representation of atomic configurations for machine learning models has led to numerous sets of descriptors. However, many descriptor sets are incomplete and/or functionally dependent. Incomplete sets cannot faithfully represent atomic environments. Yet complete constructions often suffer from a high degree of functional dependence, where some descriptors are functions of others. These redundant descriptors do not improve discrimination between atomic environments. We employ pattern recognition techniques to remove dependent descriptors to produce the smallest possible set that satisfies completeness. We apply this in two ways: First, we refine an existing description, the atomic cluster expansion. Second, we augment an incomplete construction, yielding a new message-passing neural network architecture that can recognize up to 5-body patterns. This architecture shows strong accuracy on state-of-the-art benchmarks while retaining low computational cost. Our results demonstrate the utility of this strategy to optimize descriptor sets across a range of descriptors and application datasets.

97 MATHEMATICS AND COMPUTING↗

Remote Sensing of Lineage Functional Types for Modeling and Monitoring Biodiversity

Hyperspectral remote sensing has the potential to continuously scale plant function and plant diversity information from landscape to global extents. Numerous studies have indicated that VSWIR (400-2500 nm) reflectance properties of vegetation capture evolutionarily conserved biochemical, structural, and other functional attributes of plant species. Spectral properties conserved in plants provide the opportunity to both 1) aggregate species into lineages with improved classification accuracy and 2) link those lineages directly to plant traits. Full realization of this goal will enable parameterization of Land Surface Models (LSMs) with remotely sensed information, e.g., canopy nitrogen, and better representations of biodiversity and functional diversity in biogeographic studies. In this study, we use hyperspectral AVIRIS data from the 2013 HyspIRI campaign over the Southern Sierra Nevada, California flight box to investigate the potential for incorporating evolutionary thinking into landcover classification. We link the airborne hyperspectral data with vegetation plot data from roughly 1372 surveys and a phylogeny representing 1361 species. We aggregate species into lineages ranging from species level groups down to similar number of Plant Functional Types as often used in LSMs. We assessed the ability of Random Forest and Partial Least Squares Discriminant Analysis to discriminate across these different phylogenetic scales and determine the optimal number of lineages to classify. Although there are some temporal and spatial differences in our training data, our best approaches achieved moderate classification accuracy (Kappa > 0.65). Given an optimal number of lineages, we explored approaches to improve classifications including machine learning and unmixing approaches. This work suggests that lineage-based methods may be a promising way to leverage the huge amounts of data that will come from high resolution and high return interval hyperspectral data planned for the Surface Biology and Geology mission with sparsely sampled existing ground-based ecological data.

Hyperspectral↗

Learning for Autonomous Navigation

Robotic ground vehicles for outdoor applications have achieved some remarkable successes, notably in autonomous highway following (Dickmanns, 1987), planetary exploration (1), and off-road navigation on Earth (1). Nevertheless, major challenges remain to enable reliable, high-speed, autonomous navigation in a wide variety of complex, off-road terrain. 3-D perception of terrain geometry with imaging range sensors is the mainstay of off-road driving systems. However, the stopping distance at high speed exceeds the effective lookahead distance of existing range sensors. Prospects for extending the range of 3-D sensors is strongly limited by sensor physics, eye safety of lasers, and related issues. Range sensor limitations also allow vehicles to enter large cul-de-sacs even at low speed, leading to long detours. Moreover, sensing only terrain geometry fails to reveal mechanical properties of terrain that are critical to assessing its traversability, such as potential for slippage, sinkage, and the degree of compliance of potential obstacles. Rovers in the Mars Exploration Rover (MER) mission have got stuck in sand dunes and experienced significant downhill slippage in the vicinity of large rock hazards. Earth-based off-road robots today have very limited ability to discriminate traversable vegetation from non-traversable vegetation or rough ground. It is impossible today to preprogram a system with knowledge of these properties for all types of terrain and weather conditions that might be encountered.

mixture of Gaussians↗

Detection of Chlorophyll and Leaf Area Index Dynamics from Sub-weekly Hyperspectral Imagery

Temporally rich hyperspectral time-series can provide unique time critical information on within-field variations in vegetation health and distribution needed by farmers to effectively optimize crop production. In this study, a dense time series of images were acquired from the Earth Observing-1 (EO-1) Hyperion sensor over an intensive farming area in the center of Saudi Arabia. After correction for atmospheric effects, optimal links between carefully selected explanatory hyperspectral vegetation indices and target vegetation characteristics were established using a machine learning approach. A dataset of in-situ measured leaf chlorophyll (Chll) and leaf area index (LAI), collected during five intensive field campaigns over a variety of crop types, were used to train the rule-based predictive models. The ability of the narrow-band hyperspectral reflectance information to robustly assess and discriminate dynamics in foliar biochemistry and biomass through empirical relationships were investigated. This also involved evaluations of the generalization and reproducibility of the predictions beyond the conditions of the training dataset. The very high temporal resolution of the satellite retrievals constituted a specifically intriguing feature that facilitated detection of total canopy Chl and LAI dynamics down to sub-weekly intervals. The study advocates the benefits associated with the availability of optimum spectral and temporal resolution spaceborne observations for agricultural management purposes.

Houborg, Rasmus↗

Advances in Medical Analytics Solutions for Autonomous Medical Operations on Long-Duration Missions

A review will be presented on the progress made under STMDGame Changing Development Program Funding towards the development of a Medical Decision Support System for augmenting crew capabilities during long-duration missions, such as Mars Transit. To create an MDSS, initial work requires acquiring images and developing models that analyze and assess the features in such medical biosensor images that support medical assessment of pathologies. For FY17, the project has focused on ultrasound images towards cardiac pathologies: namely, evaluation and assessment of pericardial effusion identification and discrimination from related pneumothorax and even bladder-induced infections that cause inflammation around the heart. This identification is substantially changed due to uncertainty due to conditions of fluid behavior under space-microgravity. This talk will present and discuss the work-to-date in this Project, recognizing conditions under which various machine learning technologies, deep-learning via convolutional neural nets, and statistical learning methods for feature identification and classification can be employed and conditioned to graphical format in preparation for attachment to an inference engine that eventually creates decision support recommendations to remote crew in a triage setting.

Medical Decision Support Systems↗

Nearest-Neighbor Machine Learning Feature Selection for Interpretation of Microbial Molecular Signatures from Isotope Ratio Mass Spectrometry Data

Mass spectrometry (MS) promises to be a powerful tool for potential biosignature detection during astrobiological missions on ocean worlds in our solar system. Accurate and generalizable machine learning methods could enhance science return on investment by predicting seawater chemistry and classifying isotopic biosignatures, either as a signature consistent with microbial life (biotic) or as a novelty (unclassified/unique). However, machine learning models are likely to be complex and involve interactions between MS features, making biosignatures difficult to interpret. Feature selection methods provide biological and chemical context that help interpret the mechanisms of machine learning models, but these methods also need the ability to detect complex interactions. Previously, we developed a machine learning feature selection algorithm called nearest-neighbor projected distance regression (NPDR) that has the ability to identify important model features that involve complex interactions and automatically reduce correlation and the dimensionality in a high-dimensional variable space. The standard distance metrics used in NPDR – Manhattan and Euclidean – assume the multivariate data are isotropic, which is often violated in real data due to differences in the covariance between variables. Thus, we extend NPDR to include a random forest distance, and other anisotropic distance metrics, for computing nearest neighbors. We also augment the isotope-ratio MS data with time-series features from the raw MS signal to improve biotic classification. We test NPDR on our novel experimental ocean world seawater analog MS data. We measure isotope fractionations of volatile CO 2 that could be measured in exospheres or plumes. Samples include baseline abiotic conditions using a range of possible seawater chemistry consistent with Europa and Enceladus, and biotic samples that include microbes in these seawaters. We use penalized NPDR with random forest proximity to identify interpretable microbial molecular signatures. We compare features with random forest importance, and we train a classifier that discriminates between biotic and abiotic samples with high accuracy. These ML-trained ocean-world analog MS data could be used to assist in identifying biosignatures during future missions.

geochemistry↗

Heuristic Sonification Methods for Electromagnetic Signals Report

Sonification algorithms convert information to audible representations. The 2025 Seed Money project “Exploring Electromagnetic Signals through Sonification” seeks to create machine-learning based methods to convert electromagnetic signals to sound for human interpretation. However, as a precursor to ML-based methods, some heuristic methods have been investigated in preparation for the seed project. This report covers some example methods, applied to the “Flaming Moes” dataset of unintended radiative emissions (URE), with some simple metrics to study the device discrimination properties of the sonification as well as the “pleasantness” of the methods.

42 ENGINEERING↗

Velocity Extraction Using Complete Time-Domain Waveform Data and Audio Machine Learning

We developed a new machine learning-based tool for extracting information from interferometry measurements: MIDWAZE (Modular Interferometry Direct Waveform AnalyZEr). This paper showcases MIDWAZE’s ability to extract an object’s velocity information from Photonic Doppler Velocimetry (PDV) data at near-human accuracy with little to no human intervention. MIDWAZE can extract velocities roughly 350 times as fast as a human analyst "rushing" to complete their extractions, with similar extraction accuracy. MIDWAZE’s most outstanding feature is that it operates directly in waveform/temporal space, freeing analysis from certain limitations imposed by traditional spectrogram-based approaches and opening the way to "phase aware" PDV analysis. MIDWAZE also has limited ability to discriminate between different solid objects, which we develop as a first step towards automated discrimination of different kinds of objects such as ejecta clouds.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Sensitivity of a critical tracking task to alcohol impairment

A first order critical tracking task is evaluated for its potential to discriminate between sober and intoxicated performances. Mean differences between predrink and postdrink performances as a function of BAC are analyzed. Quantification of the results shows that intoxicated failure rates of 50% for blood alcohol concentrations (BACs) at or above 0.1%, and 75% for BACs at or above 0.14%, can be attained with no sober failure rates. A high initial rate of learning is observed, perhaps due to the very nature of the task whereby the operator is always pushed to his limit, and the scores approach a stable asymptote after approximately 50 trials. Finally, the implementation of the task as an ignition interlock system in the automobile environment is discussed. It is pointed out that lower critical performance limits are anticipated for the mechanized automotive units because of the introduction of larger hardware and neuromuscular lags. Whether such degradation in performance would reduce the effectiveness of the device or not will be determined in a continuing program involving a broader based sample of the driving population and performance correlations with both BACs and driving proficiency.

Tennant, J. A.↗

A flight expert system for on-board fault monitoring and diagnosis

An architecture for a flight expert system (FLES) to assist pilots in monitoring, diagnosing, and recovering from inflight faults is described. A prototype was implemented and an attempt was made to automate the knowledge acquisition process by employing a learning by being told methodology. The scope of acquired knowledge ranges from domain knowledge, including the information about objects and their relationships, to the procedural knowledge associated with the functionality of the mechanisms. AKAS (automatic knowledge acquisition system) is the constructed prototype for demonstration proof of concept, in which the expert directly interfaces with the knowledge acquisition system to ultimately construct the knowledge base for the particular application. The expert talks directly to the system using a natural language restricted only by the extent of the definitions in an analyzer dictionary, i.e., the interface understands a subset of concepts related to a given domain. In this case, the domain is the electrical system of the Boeing 737. Efforts were made to define and employ heuristics as well as algorithmic rules to conceptualize data produced by normal and faulty jet engine behavior examples. These rules were employed in developing the machine learning system (MLS). The input to MLS is examples which contain data of normal and faulty engine behavior and which are obtained from an engine simulation program. MLS first transforms the data into discrete selectors. Partial descriptions formed by those selectors are then generalized or specialized to generate concept descriptions about faults. The concepts are represented in the form of characteristic and discriminant descriptions, which are stored in the knowledge base and are employed to diagnose faults. MLS was successfully tested on jet engine examples.

Ali, Moonis↗

Advancing Wildfire Monitoring with TEMPO and ML tools: Hourly Smoke and Fire‑Front Mapping and Near‑Surface NO₂ Predictions

Wildfires impose substantial impacts on communities and regions downwind of wildfire smoke. We present a TEMPO‑enabled workflow that generates value‑added Level‑3 smoke‑plume masks and fire‑front maps for large wildfires, such as 2024 Park Fire, using the self‑supervised deep learning system SIT‑FUSE, along with near‑surface NO₂ predictions produced by a foundation model (Microsoft Aurora). We conclude by outlining a roadmap for expanding these capabilities to additional Western U.S. wildfire events and for delivering actionable tools to stakeholders. This open-source, reproducible workflow provides a scalable framework for cross-agency wildfire monitoring to overcome traditional limitations in smoke-cloud discrimination and air-quality forecasting by incorporating TEMPO data and beyond.

Xiaohua Pan↗

On the practical usefulness of the Hardware Efficient Ansatz

Variational Quantum Algorithms (VQAs) and Quantum Machine Learning (QML) models train a parametrized quantum circuit to solve a given learning task. The success of these algorithms greatly hinges on appropriately choosing an ansatz for the quantum circuit. Perhaps one of the most famous ansatzes is the one-dimensional layered Hardware Efficient Ansatz (HEA), which seeks to minimize the effect of hardware noise by using native gates and connectives. The use of this HEA has generated a certain ambivalence arising from the fact that while it suffers from barren plateaus at long depths, it can also avoid them at shallow ones. In this work, we attempt to determine whether one should, or should not, use a HEA. We rigorously identify scenarios where shallow HEAs should likely be avoided (e.g., VQA or QML tasks with data satisfying a volume law of entanglement). More importantly, we identify a Goldilocks scenario where shallow HEAs could achieve a quantum speedup: QML tasks with data satisfying an area law of entanglement. We provide examples for such scenario (such as Gaussian diagonal ensemble random Hamiltonian discrimination), and we show that in these cases a shallow HEA is always trainable and that there exists an anti-concentration of loss function values. Our work highlights the crucial role that input states play in the trainability of a parametrized quantum circuit, a phenomenon that is verified in our numerics.

97 MATHEMATICS AND COMPUTING↗

Variational autoencoders for at-source data reduction and anomaly detection in high energy particle detectors

Detectors in next-generation high-energy physics experiments face several daunting requirements, such as high data rates, damaging radiation exposure, and stringent constraints on power, space, and latency. To address these challenges, machine learning in readout electronics can be leveraged for smart detector designs, enabling intelligent inference and data reduction at-source. Variational autoencoders (VAEs) offer a variety of benefits for front-end readout; an on-sensor encoder can perform efficient lossy data compression while simultaneously providing a latent space representation that can be used for anomaly detection. Results are presented from low-latency and resource-efficient VAEs for front-end data processing in a futuristic silicon pixel detector. Encoder-based data compression is found to preserve good performance of off-detector analysis while significantly reducing the off-detector data rate as compared to a similarly sized data filtering approach. Furthermore, the latent space information is found to be a useful discriminator in the context of real-time sensor defect monitoring. Together, these results highlight the multifaceted utility of autoencoder-based front-end readout schemes and motivate their consideration in future detector designs.

47 OTHER INSTRUMENTATION↗

Dimensionality Reduction Through Classifier Ensembles

In data mining, one often needs to analyze datasets with a very large number of attributes. Performing machine learning directly on such data sets is often impractical because of extensive run times, excessive complexity of the fitted model (often leading to overfitting), and the well-known "curse of dimensionality." In practice, to avoid such problems, feature selection and/or extraction are often used to reduce data dimensionality prior to the learning step. However, existing feature selection/extraction algorithms either evaluate features by their effectiveness across the entire data set or simply disregard class information altogether (e.g., principal component analysis). Furthermore, feature extraction algorithms such as principal components analysis create new features that are often meaningless to human users. In this article, we present input decimation, a method that provides "feature subsets" that are selected for their ability to discriminate among the classes. These features are subsequently used in ensembles of classifiers, yielding results superior to single classifiers, ensembles that use the full set of features, and ensembles based on principal component analysis on both real and synthetic datasets.

Oza, Nikunj C.↗

In-Situ XRD/XRF to Support Life Detection on Mars

X-ray diffraction / X-ray fluorescence (XRD/XRF) analysis provides the most comprehensive mineralogical / compositional characterization of rocks and soils of any flight-capable technique. XRD data provide quantitative mineralogy (including abundance of X-ray amorphous materials) and crystal chemistry (structure, elemental composition and valence state), and XRF data provide complimentary major, minor, and some trace element abundances. Both types of data are important in evaluating habitability (environment of formation) and biosignature preservation/degradation (post-depositional diagenetic change). Whether or not a relict biosignature is detected, the mineral assemblage and its geochemistry can be used to determine the habitability of an ancient environment (e.g., salinity, pH, temperature), and to identify potential sources of energy for life (e.g., elements in different redox states). In this respect, a null result (a habitable environment lacking evidence of life) can play an important role in constraining the parameters of the search. Conversely, diagenetic alteration (taphonomic change) resulting from post-depositional variations in temperature, pressure or fluid chemistry can preserve evidence of biogenicity, erase such evidence completely or indeed can provide for post-depositional habitable conditions in the subsurface. XRD / XRF data are critical to these determinations. The CheMin instrument on the Mars Science Laboratory (MSL) Curiosity rover is the first XRD instrument flown in space. CheMin operates in transmission geometry with a Co X-ray source to minimize fluorescence from iron. Diffracted photons are collected with an energy-sensitive charge-coupled device (CCD). The position of the diffracted photons provides structural information for minerals, whereas the energy of sample-generated X-ray fluorescence photons provides elemental information, though these XRF data are qualitative. Mineralogical data from the CheMin XRD identified the three circumstances above: habitable depositional environments (e.g., Yellowknife Bay), habitable subsurface/diagenetic environments (e.g., throughout the Murray formation), and diagenetic conditions that may destroy evidence of habitability (e.g., oxidative and acidic environments at Vera Rubin ridge). Technological advances in X-ray technology and lessons learned from the operation of CheMin on Mars have resulted in a next-generation XRD/XRF, called CheMinX. Replacement of CheMin’s CCD with an array of hybrid pixel detectors and improvements in focusing optics dramatically decrease analysis time (15 minutes vs. 22 hours for MSL-CheMin) and result in a better angular resolution (0.18 vs. 0.30 °2θ for MSL-CheMin). This increased resolution improves mineral detection, including discrimination between types of pyroxenes, which is not possible with MSL-CheMin data. The hybrid pixel detectors do not require cooling like the MSL-CheMin CCD, therefore reducing the power needed to operate CheMinX. CheMinX has a silicon-drift detector (SDD) to measure fluoresced photons, enabling the quantification of major, minor, and some trace elements via XRF. XRD/XRF data are collected simultaneously in CheMinX, obviating the need for multiple compositional instruments. The CheMinX design also improves upon MSL-CheMin’s sample handling. Instead of sample cells on wheel, which are often not reusable and add complexity in commanding the instrument, CheMinX has single-use cells in a cartridge/dispenser configuration. Because of these improvements, CheMinX is an ideal instrument for Discovery-class life-detection missions, including Mars Life Explorer that was recommended for development in the Planetary Science and Astrobiology Decadal Survey 2023-2032.

E B Rampe↗

Mapping Rare Earths and Toxics in E-Waste via Hyperspectral Imaging and Machine Learning

Electronic waste (e-waste) presents a mounting challenge to environmental sustainability due to its complex composition, which includes high-value rare earth elements, hazardous organic compounds, and non-recyclable plastics. Accurate and scalable material classification is essential for enabling efficient resource recovery and safe recycling practices. This study introduces a confidence-aware classification pipeline that combines mid-infrared hyperspectral imaging (HSI), spectral angle mapping (SAM), and iterative machine learning to perform pixel-level material identification across e-waste devices. A curated spectral library encompassing artificial materials (e.g., plastic iron oxide, galvanized metals), minerals (e.g., allanite, hematite), and organic compounds (e.g., benzanthracene, toluene) was used to generate pseudo-labels, each assigned a confidence score based on SAM-derived spectral similarity. High-confidence samples from seven consumer electronics—digital cameras, keyboards, laptop fans, modems, motherboards, TV remotes, and speakers—were iteratively expanded and classified using models such as Support Vector Machine (SVM), Random Forest, Gradient Boosting Classifier, Partial Least Squares Discriminant Analysis (PLSDA) and Logistic Regression. The best-performing classifiers achieved macro F1 scores approaching 1.0. Results revealed widespread plastic content (dominated by plastic iron oxide), the presence of rare earth-bearing minerals like cerium-containing allanite, and pervasive detection of hazardous organics such as benzanthracene. Principal Component Analysis (PCA) visualizations and confusion matrices confirmed high separability and robust classification performance. This methodology enables precise, non-destructive, and scalable classification of heterogeneous e-waste streams. It supports automated, hazard-aware sorting in recycling workflows, facilitating selective recovery of critical materials and compliance with circular economy goals. The confidence-aware framework provides a foundation for real-time deployment in industrial settings, offering significant implications for smart e-recycling infrastructure and policy-driven material stewardship.

Circular economy↗

Machine Learning Algorithms for Aerosol and Cloud Detection Using CATS on the ISS

Clouds and aerosols are one of the largest uncertainties in understanding and forecasting the Earth’s changing climate system. The type and height of aerosols are important factors in determining the top-of-atmosphere (TOA) radiation budget, either direct reflection of solar radiation back to space and/or absorption of solar radiation. In addition to their impact on the Earth’s climate system, aerosols near the surface from wildfires, man-made pollution events, and dust storms are hazardous to human health. The phase and height of clouds also play a critical role in determining the role of clouds in the Earth’s climate system. Cirrus clouds in the upper troposphere can induce a significant daytime TOA warming effect, while liquid water clouds near the surface cause a large corresponding cooling effect. Lidar measurements provide accurate vertically resolved information about clouds and aerosols, including complex multi-layer scenes where passive sensors are challenged and at night, when passive sensors are unable to measure cloud and aerosol properties. The Cloud-Aerosol Transport System (CATS) is a lidar instrument that operated for 33 months on the International Space Station (ISS) at the 1064 nm wavelength to measure attenuated total backscatter and depolarization ratio. These fundamental measurements are used to derive “vertical feature mask” cloud and aerosol products, including layer top/base heights, layer geometrical thickness, aerosol type, and cloud phase. While space-based lidar systems like CATS provide cloud and aerosol vertical distributions that improve our understanding of the climate system, averaging of the daytime data from these sensors is required, at the expense of spatial resolution, to improve the daytime signal-to noise (SNR) and thus atmospheric layer detection. This presentation shows results from machine learning (ML) techniques that, when applied to CATS data: 1. improve the 1064 nm SNR 2. enable detection of atmospheric features during daytime with a horizontal resolution of 350 m or 5 km (compared to the 60 km required for standard CATS data products) 3. increase the number of atmospheric layers detected in the CATS data. A Convolutional Neural Network (CNN) trained using CATS standard data products also demonstrated the potential for improved cloud-aerosol discrimination, cloud phase, and aerosol typing compared to the operational CATS algorithms for cloud edges and complex near-surface scenes during daytime. The ML tools described in this paper can facilitate the development of smaller, low-cost lidar systems in the future and enable real-time accessibility of lidar data products from future lidar systems for monitoring and forecasting of hazardous events.

John Yorks↗

Spatial learning and memory is preserved in rats after early development in a microgravity environment

This study evaluated the cognitive mapping abilities of rats that spent part of their early development in a microgravity environment. Litters of male and female Sprague-Dawley rat pups were launched into space aboard the National Aeronautics and Space Administration space shuttle Columbia on postnatal day 8 or 14 and remained in space for 16 days. These animals were designated as FLT groups. Two age-matched control groups remained on Earth: those in standard vivarium housing (VIV) and those in housing identical to that aboard the shuttle (AGC). On return to Earth, animals were tested in three different tasks that measure spatial learning ability, the Morris water maze (MWM), and a modified version of the radial arm maze (RAM). Animals were also tested in an open field apparatus to measure general activity and exploratory activity. Performance and search strategies were evaluated in each of these tasks using an automated tracking system. Despite the dramatic differences in early experience, there were remarkably few differences between the FLT groups and their Earth-bound controls in these tasks. FLT animals learned the MWM and RAM as quickly as did controls. Evaluation of search patterns suggested subtle differences in patterns of exploration and in the strategies used to solve the tasks during the first few days of testing, but these differences normalized rapidly. Together, these data suggest that development in an environment without gravity has minimal long-term impact on spatial learning and memory abilities. Any differences due to development in microgravity are quickly reversed after return to earth normal gravity.

NASA Discipline Neuroscience↗