Search NASA⌕ Search

SEARCH · Search NASA

Results for “data mining”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 379 records · Page 21

The 1999 (Mw 7.1) Hector Mine, California, Earthquake: Near-Field Postseismic Deformation from ERS Interferometry

Interferometric synthetic aperture radar (InSAR) data over the area of the Hector Mine earthquake (Mw 7.1, 16 October 1999) reveal postseismic deformation of several centimeters over a spatial scale of 0.5 to 50 km. We analyzed seven SAR acquisitions to form interferograms over four time periods after the event. The main deformations seen in the line-of-sight (LOS) displacement maps are a region of subsidence (60 mm LOS increase) on the northern end of the fault, a region of uplift (45 mm LOS decrease) located to the northeast of the primary fault bend, and a linear trough running along the main rupture having a depth of up to 15 mm and a width of about 2 km. We correlate these features with a double left-bending, rightlateral, strike-slip fault that exhibits contraction on the restraining side and extension along the releasing side of the fault bends. The temporal variations in the near-fault postseismic deformation are consistent with a characteristic time scale of 135 + 42 or - 25 days, which is similar to the relaxation times following the 1992 Landers earthquake. High gradients in the LOS displacements occur on the fault trace, consistent with afterslip on the earthquake rupture. We derive an afterslip model by inverting the LOS data from both the ascending and descending orbits. Our model indicates that much of the afterslip occurs at depths of less than 3 to 4 km.

Jacobs, Allison↗

Transfer of terrestrial technology for lunar mining

The functions, operational procedures, and major items of equipment that comprise the terrestrial mining process are characterized. These data are used to synthesize a similar activity on the lunar surface. Functions, operations, and types of equipment that can be suitably transferred to lunar operation are identified. Shortfalls, enhancements, and technology development needs are described. The lunar mining process and what is required to adapt terrestrial equipment are highlighted. It is concluded that translation of terrestrial mining equipment and operational processes to perform similar functions on the lunar surface is practical. Adequate attention must be given to the harsh environment and logistical constraints of the lunar setting. By using earth-based equipment as a forcing function, near- and long-term benefits are derived (i.e., improved terrestrial mining in the near term vis-a-vis commercial production of helium-3 in the long term.

Hall, Robert A.↗

Efficient Implementation of an Optimal Interpolator for Large Spatial Data Sets

Interpolating scattered data points is a problem of wide ranging interest. A number of approaches for interpolation have been proposed both from theoretical domains such as computational geometry and in applications' fields such as geostatistics. Our motivation arises from geological and mining applications. In many instances data can be costly to compute and are available only at nonuniformly scattered positions. Because of the high cost of collecting measurements, high accuracy is required in the interpolants. One of the most popular interpolation methods in this field is called ordinary kriging. It is popular because it is a best linear unbiased estimator. The price for its statistical optimality is that the estimator is computationally very expensive. This is because the value of each interpolant is given by the solution of a large dense linear system. In practice, kriging problems have been solved approximately by restricting the domain to a small local neighborhood of points that lie near the query point. Determining the proper size for this neighborhood is a solved by ad hoc methods, and it has been shown that this approach leads to undesirable discontinuities in the interpolant. Recently a more principled approach to approximating kriging has been proposed based on a technique called covariance tapering. This process achieves its efficiency by replacing the large dense kriging system with a much sparser linear system. This technique has been applied to a restriction of our problem, called simple kriging, which is not unbiased for general data sets. In this paper we generalize these results by showing how to apply covariance tapering to the more general problem of ordinary kriging. Through experimentation we demonstrate the space and time efficiency and accuracy of approximating ordinary kriging through the use of covariance tapering combined with iterative methods for solving large sparse systems. We demonstrate our approach on large data sizes arising both from synthetic sources and from real applications.

Memarsadeghi, Nargess↗

Human factors model concerning the man-machine interface of mining crewstations

The U.S. Bureau of Mines is developing a computer model to analyze the human factors aspect of mining machine operator compartments. The model will be used as a research tool and as a design aid. It will have the capability to perform the following: simulated anthropometric or reach assessment, visibility analysis, illumination analysis, structural analysis of the protective canopy, operator fatigue analysis, and computation of an ingress-egress rating. The model will make extensive use of graphics to simplify data input and output. Two dimensional orthographic projections of the machine and its operator compartment are digitized and the data rebuilt into a three dimensional representation of the mining machine. Anthropometric data from either an individual or any size population may be used. The model is intended for use by equipment manufacturers and mining companies during initial design work on new machines. In addition to its use in machine design, the model should prove helpful as an accident investigation tool and for determining the effects of machine modifications made in the field on the critical areas of visibility and control reach ability.

Rider, James P.↗

Surface Mining and Reclamation Effects on Flood Response of Watersheds in the Central Appalachian Plateau Region

Surface mining of coal and subsequent reclamation represent the dominant land use change in the central Appalachian Plateau (CAP) region of the United States. Hydrologic impacts of surface mining have been studied at the plot scale, but effects at broader scales have not been explored adequately. Broad-scale classification of reclaimed sites is difficult because standing vegetation makes them nearly indistinguishable from alternate land uses. We used a land cover data set that accurately maps surface mines for a 187-km2 watershed within the CAP. These land cover data, as well as plot-level data from within the watershed, are used with HSPF (Hydrologic Simulation Program-Fortran) to estimate changes in flood response as a function of increased mining. Results show that the rate at which flood magnitude increases due to increased mining is linear, with greater rates observed for less frequent return intervals. These findings indicate that mine reclamation leaves the landscape in a condition more similar to urban areas rather than does simple deforestation, and call into question the effectiveness of reclamation in terms of returning mined areas to the hydrological state that existed before mining.

Ferrari, J. R.↗

Remote sensing of coal mine pollution in the upper Potomac River basin

A survey of remote sensing data pertinent to locating and monitoring sources of pollution resulting from surface and shaft mining operations was conducted in order to determine the various methods by which ERTS and aircraft remote sensing data can be used as a replacement for, or a supplement to traditional methods of monitoring coal mine pollution of the upper Potomac Basin. The gathering and analysis of representative samples of the raw and processed data obtained during the survey are described, along with plans to demonstrate and optimize the data collection processes.

Source record↗

GL4U: Using Space Biology Omics Data to Provide Bioinformatics Training for Students and Educators

NASA’s GeneLab project provides researchers open access to space-relevant multi-omics data via the Open Science Data Repository (OSDR) that can be mined to understand the effects of spaceflight on biological systems. To maximize the number of scientists who understand and utilize GeneLab data and data processing pipelines, GeneLab created GeneLab for Colleges and Universities (GL4U). GL4U provides space biology-relevant training in bioinformatics to the next generation of scientists through direct (training students) and indirect (training educators) approaches. The GL4U pilot programs were conducted in June 2021 (direct training) and 2022 (indirect training). During the pilots, students and educators at Historically Black Colleges and Universities (HBCUs) and Minority Serving Institutions (MSIs) participated in a week-long (direct training) or two-week-long (indirect training) bootcamp consisting of space biology-specific lectures and hands-on instruction using Jupyter Notebooks to analyze space biology RNA sequencing data from OSDR. During the educator pilot, participants received materials, training, and the necessary compute resources to enable them to run the bootcamp at their home institutions, thereby extending the reach of this initiative. In July 2023, GeneLab is partnering with JPL to expand GL4U to include amplicon sequencing (Amp-Seq) analysis training. During the GL4U Amp-Seq bootcamp, student and educator participants will receive training on how to analyze and interpret Amp-Seq data using the NASA GeneLab data processing pipeline. All bootcamp material, including instructions for requesting compute resources, will be made publicly available on GitHub for educators to teach the GL4U content in subsequent semesters. GL4U provides undergraduate students from underrepresented groups the opportunity to learn about NASA and Space Biology, and to enhance their career prospects by gaining hands-on experience analyzing omics data, a skillset that is highly applicable and marketable in the life sciences. We present results from pre- and post-training surveys completed by all participants of the Amp-Seq bootcamp.

Amanda M Saravia-Butler↗

Solutions Network Formulation Report. Landsat Data Continuity Mission Simulated Data Products for Bureau of Land Management and Environmental Protection Agency Abandoned Mine Lands Decision Support

Presently, the BLM (Bureau of Land Management) has identified a multitude of abandoned mine sites in primarily Western states for cleanup. These sites are prioritized and appropriate cleanup has been called in to reclaim the sites. The task is great in needing considerable amounts of agency resources. For instance, in Colorado alone there exists an estimated 23,000 abandoned mines. The problem is not limited to Colorado or to the United States. Cooperation for reclamation is sought at local, state, and federal agency level to aid in identification, inventory, and cleanup efforts. Dangers posed by abandoned mines are recognized widely and will tend to increase with time because some of these areas are increasingly used for recreation and, in some cases, have been or are in the process of development. In some cases, mines are often vandalized once they are closed. The perpetrators leave them open, so others can then access the mines without realizing the danger posed. Abandoned mine workings often fill with water or oxygen-deficient air and dangerous gases following mining. If the workings are accidentally entered into, water or bad air can prove fatal to those underground. Moreover, mine residue drainage negatively impacts the local watershed ecology. Some of the major hazards that might be monitored by higher-resolution satellites include acid mine drainage, clogged streams, impoundments, slides, piles, embankments, hazardous equipment or facilities, surface burning, smoke from underground fires, and mine openings.

Estep, Leland↗

GL4U: Training the next generation of bioinformaticians, one omics datatype at a time

Spaceflight modifies gene expression in every organism examined to date, including humans. Understanding how these gene expression changes affect physiology is crucial for the development of countermeasures to enable long-duration manned missions. NASA’s GeneLab project provides researchers open access to multi-omics data, including genetic and gene expression data, from spaceflight experiments that can be mined to understand the effects of spaceflight on biological systems. To ensure new knowledge generation through data re-use, it is important to maximize the number of scientists who utilize GeneLab data. Training students on the GeneLab platform is the best way to create long-term adopters of this NASA database and its tools. Turning students into future instructors and advocates will also accelerate the dissemination of these data and tools to the broader scientific community. Therefore, in collaboration with the GeneLab Educational Working Group (EWG), GeneLab has created GeneLab for Colleges and Universities (GL4U). GL4U provides space biology-relevant training in bioinformatics to the next generation of scientists through direct and indirect approaches. The GeneLab team plans to host two annual data processing bootcamps, one for college-level students (direct) and one for college educators (indirect – training of trainers), in which participants learn to analyze GeneLab’s space-relevant omics data. During the bootcamp, educators will receive materials and training to enable them to run the bootcamp at their home institutions or alternatively to adapt the content to implement within existing courses, thereby extending the reach of this initiative. The GL4U direct training pilot program was conducted in June 2021 in collaboration with USRA and San Jose State University (SJSU). During the pilot, SJSU students participated in a week-long bootcamp consisting of space biology-specific lectures and hands-on instruction using Jupyter Notebooks to analyze RNA sequence data. This pilot demonstrates the capacity of GL4U for training young scientists and encouraging data re-use.

Jonathan Matthew Galazka↗

Efficient Implementation of an Optimal Interpolator for Large Spatial Data Sets

Scattered data interpolation is a problem of interest in numerous areas such as electronic imaging, smooth surface modeling, and computational geometry. Our motivation arises from applications in geology and mining, which often involve large scattered data sets and a demand for high accuracy. The method of choice is ordinary kriging. This is because it is a best unbiased estimator. Unfortunately, this interpolant is computationally very expensive to compute exactly. For n scattered data points, computing the value of a single interpolant involves solving a dense linear system of size roughly n x n. This is infeasible for large n. In practice, kriging is solved approximately by local approaches that are based on considering only a relatively small'number of points that lie close to the query point. There are many problems with this local approach, however. The first is that determining the proper neighborhood size is tricky, and is usually solved by ad hoc methods such as selecting a fixed number of nearest neighbors or all the points lying within a fixed radius. Such fixed neighborhood sizes may not work well for all query points, depending on local density of the point distribution. Local methods also suffer from the problem that the resulting interpolant is not continuous. Meyer showed that while kriging produces smooth continues surfaces, it has zero order continuity along its borders. Thus, at interface boundaries where the neighborhood changes, the interpolant behaves discontinuously. Therefore, it is important to consider and solve the global system for each interpolant. However, solving such large dense systems for each query point is impractical. Recently a more principled approach to approximating kriging has been proposed based on a technique called covariance tapering. The problems arise from the fact that the covariance functions that are used in kriging have global support. Our implementations combine, utilize, and enhance a number of different approaches that have been introduced in literature for solving large linear systems for interpolation of scattered data points. For very large systems, exact methods such as Gaussian elimination are impractical since they require 0(n(exp 3)) time and 0(n(exp 2)) storage. As Billings et al. suggested, we use an iterative approach. In particular, we use the SYMMLQ method, for solving the large but sparse ordinary kriging systems that result from tapering. The main technical issue that need to be overcome in our algorithmic solution is that the points' covariance matrix for kriging should be symmetric positive definite. The goal of tapering is to obtain a sparse approximate representation of the covariance matrix while maintaining its positive definiteness. Furrer et al. used tapering to obtain a sparse linear system of the form Ax = b, where A is the tapered symmetric positive definite covariance matrix. Thus, Cholesky factorization could be used to solve their linear systems. They implemented an efficient sparse Cholesky decomposition method. They also showed if these tapers are used for a limited class of covariance models, the solution of the system converges to the solution of the original system. Matrix A in the ordinary kriging system, while symmetric, is not positive definite. Thus, their approach is not applicable to the ordinary kriging system. Therefore, we use tapering only to obtain a sparse linear system. Then, we use SYMMLQ to solve the ordinary kriging system. We show that solving large kriging systems becomes practical via tapering and iterative methods, and results in lower estimation errors compared to traditional local approaches, and significant memory savings compared to the original global system. We also developed a more efficient variant of the sparse SYMMLQ method for large ordinary kriging systems. This approach adaptively finds the correct local neighborhood for each query point in the interpolation process.

Memarsadeghi, Nargess↗

Application of EREP imagery to fracture-related mine safety hazards and environmental problems in mining

The author has identified the following significant results. Numerous fracture traces were detected on both the color transparencies and black and white spectral bands. Fracture traces of value to mining hazards analysis were noted on the EREP imagery which could not be detected on either the ERTS-1 or high altitude aircraft color infrared photography. Several areas of mine subsidence occurring in the Busseron Creek area near Sullivan, Indiana were successfully identified using color photography. Skylab photography affords an increase over comparable scale ERTS-1 imagery in level of information obtained in mined lands inventory and reclamation analysis. A review of EREP color photography permitted the identification of a substantial number of non-fuel mines within the Southern Indiana test area. A new mine was detected on the EREP photography without prior data. EREP has definite value for estimating areal changes in active mines and for detecting new non-fuel mines. Gob piles and slurry ponds of several acres could be detected on the S-190B color photography when observed in association with large scale mining operations. Apparent degradation of water quality resulting from acid mine drainage and/or siltation was noted in several ponds or small lakes and appear to be related to intensive mining activity near Sullivan, Indiana.

Wier, C. E.↗

An unsupervised classification approach for analysis of Landsat data to monitor land reclamation in Belmont county, Ohio

Two unsupervised classification procedures for analyzing Landsat data used to monitor land reclamation in a surface mining area in east central Ohio are compared for agreement with data collected from the corresponding locations on the ground. One procedure is based on a traditional unsupervised-clustering/maximum-likelihood algorithm sequence that assumes spectral groupings in the Landsat data in n-dimensional space; the other is based on a nontraditional unsupervised-clustering/canonical-transformation/clustering algorithm sequence that not only assumes spectral groupings in n-dimensional space but also includes an additional feature-extraction technique. It is found that the nontraditional procedure provides an appreciable improvement in spectral groupings and apparently increases the level of accuracy in the classification of land cover categories.

Brumfield, J. O.↗

Mapping alteration in the Goldfield Mining District, Nevada, with the Airborne Visible/Infrared Imaging Spectrometer (AVIRIS)

Analysis of the data acquired by the Airborne Visible/IR Imaging Spectrometer over the Goldfield Mining District, Nevada, demonstrates the unique capabilities of high resolution imaging spectrometers for alteration mapping and identification of mineralogy. The study focuses on the 2-2.45 micron window, as alteration minerals of interest have their distinctive spectral reflectance features in this region. An attempt was made to map the spatial distribution of the individual minerals and produce a mixture maps showing their relative abundance on a pixel-by-pixel basis, focussing on selected subareas. Software and SNR performances limited success in detailed mapping.

Carrere, Veronique↗

Ground-truthing AVIRIS mineral mapping at Cuprite, Nevada

Mineral abundance maps of 18 minerals were made of the Cuprite Mining District using 1990 AVIRIS data and the Multiple Spectral Feature Mapping Algorithm (MSFMA) as discussed in Clark et al. This technique uses least-squares fitting between a scaled laboratory reference spectrum and ground calibrated AVIRIS data for each pixel. Multiple spectral features can be fitted for each mineral and an unlimited number of minerals can be mapped simultaneously. Quality of fit and depth from continuum numbers for each mineral are calculated for each pixel and the results displayed as a multicolor image.

Swayze, Gregg↗

NASA Tech Briefs, March 2012

The topics include: 1) Spectral Profiler Probe for In Situ Snow Grain Size and Composition Stratigraphy; 2) Portable Fourier Transform Spectroscopy for Analysis of Surface Contamination and Quality Control; 3) In Situ Geochemical Analysis and Age Dating of Rocks Using Laser Ablation-Miniature Mass Spectrometer; 4) Physics Mining of Multi-Source Data Sets; 5) Photogrammetry Tool for Forensic Analysis; 6) Connect Global Positioning System RF Module; 7) Simple Cell Balance Circuit; 8) Miniature EVA Software Defined Radio; 9) Remotely Accessible Testbed for Software Defined Radio Development; 10) System-of-Systems Technology-Portfolio-Analysis Tool; 11) VESGEN Software for Mapping and Quantification of Vascular Regulators; 12) Constructing a Database From Multiple 2D Images for Camera Pose Estimation and Robot Localization; 13) Adaption of G-TAG Software for Validating Touch and Go Asteroid Sample Return Design Methodology; 14) 3D Visualization for Phoenix Mars Lander Science Operations; 15) RxGen General Optical Model Prescription Generator; 16) Carbon Nanotube Bonding Strength Enhancement Using Metal Wicking Process; 17) Multi-Layer Far-Infrared Component Technology; 18) Germanium Lift-Off Masks for Thin Metal Film Patterning; 19) Sealing Materials for Use in Vacuum at High Temperatures; 20) Radiation Shielding System Using a Composite of Carbon Nanotubes Loaded With Electropolymers; 21) Nano Sponges for Drug Delivery and Medicinal Applications; 22) Molecular Technique to Understand Deep Microbial Diversity; 23) Methods and Compositions Based on Culturing Microorganisms in Low Sedimental Fluid Shear Conditions; 24) Secure Peer-to-Peer Networks for Scientific Information Sharing; 25) Multiplexer/Demultiplexer Loading Tool (MDMLT); 26) High-Rate Data-Capture for an Airborne Lidar System; 27) Wavefront Sensing Analysis of Grazing Incidence Optical Systems; 28) Foam-on-Tile Damage Model; 29) Instrument Package Manipulation Through the Generation and Use of an Attenuated-Fluent Gas Fold; 30) Multicolor Detectors for Ultrasensitive Long-Wave Imaging Cameras; 31) Lunar Reconnaissance Orbiter (LRO) Command and Data Handling Flight Electronics Subsystem; and 32) Electro-Optic Segment-Segment Sensors for Radio and Optical Telescopes.

Source record↗

Nondimensional Representations for Occulter Design and Performance Evaluation

An occulter is a spacecraft with a precisely-shaped optical edges which ies in formation with a telescope, blocking light from a star while leaving light from nearby planets una ected. Using linear optimization, occulters can be designed for use with telescopes over a wide range of telescope aperture sizes, science bands, and starlight suppression levels. It can be shown that this optimization depends primarily on a small number of independent nondimensional parameters, which correspond to Fresnel numbers and physical scales and enter the optimization only as constraints. We show how these can be used to span the parameter space of possible optimized occulters; this data set can then be mined to determine occulter sizes for various mission scenarios and sets of engineering constraints.

Earth like planets↗

K-Means Cluster Study for Radiofrequency Propagation Characterization

The objective of this study is to design a simple method for mining radio frequency (RF) propagation data. The study explored the characteristics of a large dataset of propagation experiments conducted over the span of years and using several ground stations around the world. Furthermore, this study developed simple predictive models that can be used for link characterization and overall propagation behavior description, without the need for physical measurements on-site. It is understood that such statistical learning has several drawbacks in terms of accuracy and precision. K-means clustering was used to characterize the data set in a way never explored before in an attempt to create useful tools that reduce cost, time and risk. K-means clustering was used to characterize the data set. Cosine distance was used as a method to determine the optimal number for clustering each feature. Dependence and independence analysis was performed to explore intra and inter-sensitivity between the presented features, with respect to each other and time. Several predicative models were generated and evaluated with respect to a test set to assess a measure of prediction accuracy and precision. A simple method for data analysis was developed and tested as the basis for further studies and future refinement to produce optimal performing models.

Cognitive↗

Integrated Analysis of Multiple User Metrics - A “Sequel”; and Introducing the Google Analytic

For decades, the Goddard Earth Sciences Data and Information Services Center (GES DISC) has archived and distributed enormous volumes of NASA Earth science data (accompanied with many developed tools and services) to various research/applications communities and the general public. Being “immersed” in the Big Data era, we have inevitably faced the challenges of our continually increasing archived data in both volume and variety, as well as enhanced user needs and demands. In recent years, we have actively analyzed different types of user metrics, such as operational distribution metrics (recording numbers of distinct users and downloaded data files, size of distributed data volume): user publication metrics (mining info from our Giovanni users’ publications): and Bugzilla metrics (collecting info from user questions or feedback from user assistance tickets). Such metrics have helped us achieve a better understanding of user needs, demands, characteristics, and behaviors, which has then helped us improve our user services. Now we will present a “Sequel” of integrated analysis of multiple metrics at the GES DISC by introducing and adding one new kind of metrics acquired via utilizing our recently implemented Google Analytic 360 suite. Several “newer” reports, e.g., “What web site features and links are the most popular (and least)?” and “What are the top 25 dataset Keyword searches?” retrieved from this new metrics set will be presented, along with the aforementioned “traditional” metrics results.

Shie, Chung-Lin↗