Search NASA⌕ Search

SEARCH · Search NASA

Results for “Regression analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 541 records · Page 30

Numerical simulation of soil brightness temperatures at wavelength of 21 cm

A simulation model is applied to reproduce some observed brightness temperatures at a wavelength of 21 cm. The simulated results calculated with two different soil textures are compared directly with observations measured over fields in Arizona and South Dakota. It is found that good agreement is possible by properly adjusting the surface roughness parameter. Correlation analysis and linear regression of the brightness temperatures versus soil moistures are also carried out.

Mo, T.↗

Ordinary chondrites - Multivariate statistical analysis of trace element contents

The contents of mobile trace elements (Co, Au, Sb, Ga, Se, Rb, Cs, Te, Bi, Ag, In, Tl, Zn, and Cd) in Antarctic and non-Antarctic populations of H4-6 and L4-6 chondrites, were compared using standard multivariate discriminant functions borrowed from linear discriminant analysis and logistic regression. A nonstandard randomization-simulation method was developed, making it possible to carry out probability assignments on a distribution-free basis. Compositional differences were found both between the Antarctic and non-Antarctic H4-6 chondrite populations and between two L4-6 chondrite populations. It is shown that, for various types of meteorites (in particular, for the H4-6 chondrites), the Antarctic/non-Antarctic compositional difference is due to preterrestrial differences in the genesis of their parent materials.

Lipschutz, Michael E.↗

Study of the crust and mantle using magnetic surveys by Magsat and other satellites

This article summarizes some of the methodology and results for studies of the earth's mantle and crust. Mantle conductivity studies can be made either by studying signals impressed on the earth from outside, e.g., the ionosphere or magnetosphere, or by studing signals originating in the core and transmitted through the mantle. Crustal field studies begin with a careful selection of the data and subsequent removal of core and external fields by some sort of filtering. Average maps from different local times sometimes differ, presumably due to the remaining presence of fields of external origin. Several techniques for further filtering are discussed. Where large-area aeromagnetic maps are available, crustal maps derived from satellite data can be compared with upward continued data. In general, the comparisons show agreement, with some differences, particularly in and near the auroral belts. The satellite data are further reduced by various methods of inverse and forward modeling, sometimes including reduction to the pole. These techniques are generally unstable at the equator. Common methods of stabilizing the inversions include principle components analysis and ridge regression. Because of the presence of the core field, the entire crustal contribution from the field is not known. Also, there is a basic nonuniqueness to the inverse solutions. Nevertheless, magnetizations that are interpretable can be derived.

Langel, Robert A.↗

Estimating structural attributes of Douglas-fir/western hemlock forest stands from Landsat and SPOT imagery

Relationships between spectral and texture variables derived from SPOT HRV 10 m panchromatic and Landsat TM 30 m multispectral data and 16 forest stand structural attributes is evaluated to determine the utility of satellite data for analysis of hemlock forests west of the Cascade Mountains crest in Oregon and Washington, USA. Texture of the HRV data was found to be strongly related to many of the stand attributes evaluated, whereas TM texture was weakly related to all attributes. Data analysis based on regression models indicates that both TM and HRV imagery should yield equally accurate estimates of forest age class and stand structure. It is concluded that the satellite data are a valuable source for estimation of the standard deviation of tree sizes, mean size and density of trees in the upper canopy layers, a structural complexity index, and stand age.

Cohen, Warren B.↗

Multivariate statistical analysis: Principles and applications to coorbital streams of meteorite falls

Multivariate statistical analysis techniques (linear discriminant analysis and logistic regression) can provide powerful discrimination tools which are generally unfamiliar to the planetary science community. Fall parameters were used to identify a group of 17 H chondrites (Cluster 1) that were part of a coorbital stream which intersected Earth's orbit in May, from 1855 - 1895, and can be distinguished from all other H chondrite falls. Using multivariate statistical techniques, it was demonstrated that a totally different criterion, labile trace element contents - hence thermal histories - or 13 Cluster 1 meteorites are distinguishable from those of 45 non-Cluster 1 H chondrites. Here, we focus upon the principles of multivariate statistical techniques and illustrate their application using non-meteoritic and meteoritic examples.

Wolf, S. F.↗

Chemical studies of H chondrites. 4: New data and comparison of Antarctic suites

We report data for the trace elements Au, Co, Sb, Ga, Rb, Ag, Se, Cs, Te, Zn, Cd, Bi, Ti, and In (ordered by putative volatility during nebular condensation and accretion) determined by neutron activation analysis in 13 H5 chondrites from Victoria Land and 20 H4-6 chondrites from Queen Maud Land, Antarctica. These and earlier results provide Antarctic sample suites of 34 chondrites from Victoria Land and 25 from Queen Maud Land. Treatment of data for the most volatile 10 elements (Rb to In) in these studies by multivariate statistical techniques more robust, as well as more conservative, than conventional linear discriminant analysis and logistic regression demonstrates that compositions differ at marginally significant levels. This difference cannot be explained by trivial (terrestrial) causes and becomes more significant, despite the smaller size of the database, when comparisons are limited to data from a single analyst and when all upper limits are eliminated from consideration. The Victoria Land and Queen Maud Land suites have different mean terrestrial ages (approximately 300 kyr and approximately 100 kyr, respectively) and age distributions, suggesting that a time-dependent variation of chondritic sources with different thermal histories is responsible. As a result, these two Antarctic suites are, on average, chemically distinguishable from each other. Since H chondrites serve as a paradigm for other meteorite classes, these results indicate that the near-Earth populations of planetary materials varied with time on the 10(exp 5)-year timescale.

Wolf, Stephen F.↗

Chemical studies of H chondrites. 6: Antarctic/non-Antarctic compositional differences revisited

We report data for the trace elements Au, Co, Sb, Ga, Rb, Ag, Se, Cs, Te, Zn, Cd, Bi, T1, and In (ordered by putative volatility during nebular condensation and accretion) determined by radiochemical neutron activation analysis of 14 additional H5 and H6 chondrite falls. Data for the 10 most volatile elements (Rb to In) treated by the multivariate techniques of linear discriminant analysis and logistic regression in these and 44 other falls are compared with those of 59 H4-6 chondrites from Antarctica. Various populations are tested by the multivariate techniques, using the previously developed method of randomization-simulation to assess significance levels. An earlier conclusion, based on fewer examples, that H4-6 chondrite falls are compositionally distinguishable from the Antarctic suite is verified by the additional data. This distinctiveness is highly significant because of the presence of samples from Victoria Land in the Antarctic population, which differ compositionally from falls beyond any reasonable doubt. However, it cannot be proven unequivocally that falls and Antarctic samples from Queen Maud Land are compositionally distinguishable. Trivial causes (e.g., analyst bias, weathering) cannot explain the Victoria Land (Antarctic)/non-Antarctic compositional difference for paradigmatic H4-6 chondrites. This seems to reflect a time-dependent variation of near-Earth meteoroid source regions differing in average thermal history.

Wolf, Stephen F.↗

Efficient Parametric Uncertainty Analysis of an Earth Entry Vehicle Concept Using Least Angle Regression

The objective of this work was to outline and apply an efficient and accurate parametric un-certainty propagation approach to the analysis of convective heating on an Earth entry vehicle concept. The described approach was based on Least Angle Regression used to solve a sparse and underdetermined linear system in the point-collocation non-intrusive polynomial chaos surrogate method. This approach involved an iterative process to computing the non-zero terms of the underlying polynomial chaos model using only enough samples to converge uncertainty interval predictions and Sobol index values based global nonlinear sensitivity estimates. The Earth entry vehicle was analyzed at three points along a representative trajectory for a Mars return mission. 329 sources of uncertainty were identified in the computational fluid dynamics model used to predict the forebody convective heating. These included uncertainty in flow field chemical rates, collision integrals, heats of formation, surface finite rate char model reaction rates, wall roughness height, and the turbulent Schmidt number. Results from this study showed that convective heating uncertainty as high as 50% of the nominal was predicted with only about 50 evaluations of the computational model. This was far fewer than would be required for a sampling-based approach or a full basis polynomial chaos model, which would have required over 50,000 samples. Additionally, results showed that over 90% of the total convective heating uncertainty was due to uncertainty in the N2catalytic rate on the surface, while the remainder of the uncertainty was attributed to the turbulent Schmidt number and the wall roughness uncertainties.

Thomas K West IV↗

An investigation of the detection of tornadic thunderstorms by observing storm top features using geosynchronous satellite imagery

The number of tornado outbreak cases studied in detail was increased from the original 8. Detailed ground and aerial studies were carried out of two outbreak cases of considerable importance. It was demonstrated that multiple regression was able to predict the tornadic potential of a given thunderstorm cell by its cirrus anvil plume characteristics. It was also shown that the plume outflow intensity and the deviation of the plume alignment from storm relative winds at anvil altitude could account for the variance in tornadic potential for a given cell ranging from 0.37 to 0.82 for linear to values near 0.9 for quadratic regression. Several predictors were used in various discriminant analysis models and in censored regression models to obtain forecasts of whether a cell is tornadic and how strong tornadic it could be potentially. The experiments were performed with the synoptic scale vertical shear in the horizontal wind and with synoptic scale surface vorticity in the proximity of the cell.

Anderson, Charles E.↗

Analysis of Multivariate Experimental Data Using A Simplified Regression Model Search Algorithm

A new regression model search algorithm was developed that may be applied to both general multivariate experimental data sets and wind tunnel strain-gage balance calibration data. The algorithm is a simplified version of a more complex algorithm that was originally developed for the NASA Ames Balance Calibration Laboratory. The new algorithm performs regression model term reduction to prevent overfitting of data. It has the advantage that it needs only about one tenth of the original algorithm's CPU time for the completion of a regression model search. In addition, extensive testing showed that the prediction accuracy of math models obtained from the simplified algorithm is similar to the prediction accuracy of math models obtained from the original algorithm. The simplified algorithm, however, cannot guarantee that search constraints related to a set of statistical quality requirements are always satisfied in the optimized regression model. Therefore, the simplified algorithm is not intended to replace the original algorithm. Instead, it may be used to generate an alternate optimized regression model of experimental data whenever the application of the original search algorithm fails or requires too much CPU time. Data from a machine calibration of NASA's MK40 force balance is used to illustrate the application of the new search algorithm.

Ulbrich, Norbert M.↗

Subsonic Aircraft With Regression and Neural-Network Approximators Designed

At the NASA Glenn Research Center, NASA Langley Research Center's Flight Optimization System (FLOPS) and the design optimization testbed COMETBOARDS with regression and neural-network-analysis approximators have been coupled to obtain a preliminary aircraft design methodology. For a subsonic aircraft, the optimal design, that is the airframe-engine combination, is obtained by the simulation. The aircraft is powered by two high-bypass-ratio engines with a nominal thrust of about 35,000 lbf. It is to carry 150 passengers at a cruise speed of Mach 0.8 over a range of 3000 n mi and to operate on a 6000-ft runway. The aircraft design utilized a neural network and a regression-approximations-based analysis tool, along with a multioptimizer cascade algorithm that uses sequential linear programming, sequential quadratic programming, the method of feasible directions, and then sequential quadratic programming again. Optimal aircraft weight versus the number of design iterations is shown. The central processing unit (CPU) time to solution is given. It is shown that the regression-method-based analyzer exhibited a smoother convergence pattern than the FLOPS code. The optimum weight obtained by the approximation technique and the FLOPS code differed by 1.3 percent. Prediction by the approximation technique exhibited no error for the aircraft wing area and turbine entry temperature, whereas it was within 2 percent for most other parameters. Cascade strategy was required by FLOPS as well as the approximators. The regression method had a tendency to hug the data points, whereas the neural network exhibited a propensity to follow a mean path. The performance of the neural network and regression methods was considered adequate. It was at about the same level for small, standard, and large models with redundancy ratios (defined as the number of input-output pairs to the number of unknown coefficients) of 14, 28, and 57, respectively. In an SGI octane workstation (Silicon Graphics, Inc., Mountainview, CA), the regression training required a fraction of a CPU second, whereas neural network training was between 1 and 9 min, as given. For a single analysis cycle, the 3-sec CPU time required by the FLOPS code was reduced to milliseconds by the approximators. For design calculations, the time with the FLOPS code was 34 min. It was reduced to 2 sec with the regression method and to 4 min by the neural network technique. The performance of the regression and neural network methods was found to be satisfactory for the analysis and design optimization of the subsonic aircraft.

Patnaik, Surya N.↗

Tool for Forecasting Cool-Season Peak Winds Across Kennedy Space Center and Cape Canaveral Air Force Station

The expected peak wind speed for the day is an important element in the daily morning forecast for ground and space launch operations at Kennedy Space Center (KSC) and Cape Canaveral Air Force Station (CCAFS). The 45th Weather Squadron (45 WS) must issue forecast advisories for KSC/CCAFS when they expect peak gusts for >= 25, >= 35, and >= 50 kt thresholds at any level from the surface to 300 ft. In Phase I of this task, the 45 WS tasked the Applied Meteorology Unit (AMU) to develop a cool-season (October - April) tool to help forecast the non-convective peak wind from the surface to 300 ft at KSC/CCAFS. During the warm season, these wind speeds are rarely exceeded except during convective winds or under the influence of tropical cyclones, for which other techniques are already in use. The tool used single and multiple linear regression equations to predict the peak wind from the morning sounding. The forecaster manually entered several observed sounding parameters into a Microsoft Excel graphical user interface (GUI), and then the tool displayed the forecast peak wind speed, average wind speed at the time of the peak wind, the timing of the peak wind and the probability the peak wind will meet or exceed 35, 50 and 60 kt. The 45 WS customers later dropped the requirement for >= 60 kt wind warnings. During Phase II of this task, the AMU expanded the period of record (POR) by six years to increase the number of observations used to create the forecast equations. A large number of possible predictors were evaluated from archived soundings, including inversion depth and strength, low-level wind shear, mixing height, temperature lapse rate and winds from the surface to 3000 ft. Each day in the POR was stratified in a number of ways, such as by low-level wind direction, synoptic weather pattern, precipitation and Bulk Richardson number. The most accurate Phase II equations were then selected for an independent verification. The Phase I and II forecast methods were compared using an independent verification data set. The two methods were compared to climatology, wind warnings and advisories issued by the 45 WS, and North American Mesoscale (NAM) model (MesoNAM) forecast winds. The performance of the Phase I and II methods were similar with respect to mean absolute error. Since the Phase I data were not stratified by precipitation, this method's peak wind forecasts had a large negative bias on days with precipitation and a small positive bias on days with no precipitation. Overall, the climatology methods performed the worst while the MesoNAM performed the best. Since the MesoNAM winds were the most accurate in the comparison, the final version of the tool was based on the MesoNAM winds. The probability the peak wind will meet or exceed the warning thresholds were based on the one standard deviation error bars from the linear regression. For example, the linear regression might forecast the most likely peak speed to be 35 kt and the error bars used to calculate that the probability of >= 25 kt = 76%, the probability of >= 35 kt = 50%, and the probability of >= 50 kt = 19%. The authors have not seen this application of linear regression error bars in any other meteorological applications. Although probability forecast tools should usually be developed with logistic regression, this technique could be easily generalized to any linear regression forecast tool to estimate the probability of exceeding any desired threshold . This could be useful for previously developed linear regression forecast tools or new forecast applications where statistical analysis software to perform logistic regression is not available. The tool was delivered in two formats - a Microsoft Excel GUI and a Tool Command Language/Tool Kit (Tcl/Tk) GUI in the Meteorological Interactive Data Display System (MIDDS). The Microsoft Excel GUI reads a MesoNAM text file containing hourly forecasts from 0 to 84 hours, from one model run (00 or 12 UTC). The GUI then displays e peak wind speed, average wind speed, and the probability the peak wind will meet or exceed the 25-, 35- and 50-kt thresholds. The user can display the Day-1 through Day-3 peak wind forecasts, and separate forecasts are made for precipitation and non-precipitation days. The MIDDS GUI uses data from the NAM and Global Forecast System (GFS), instead of the MesoNAM. It can display Day-1 and Day-2 forecasts using NAM data, and Day-1 through Day-5 forecasts using GFS data. The timing of the peak wind is not displayed, since the independent verification showed that none of the forecast methods performed significantly better than climatology. The forecaster should use the climatological timing of the peak wind (2248 UTC) as a first guess and then adjust it based on the movement of weather features.

Barrett, Joe H., III↗

Algorithm For Solution Of Subset-Regression Problems

Reliable and flexible algorithm for solution of subset-regression problem performs QR decomposition with new column-pivoting strategy, enables selection of subset directly from originally defined regression parameters. This feature, in combination with number of extensions, makes algorithm very flexible for use in analysis of subset-regression problems in which parameters have physical meanings. Also extended to enable joint processing of columns contaminated by noise with those free of noise, without using scaling techniques.

Verhaegen, Michel↗

Machine Learning–Augmented Laser-Induced Breakdown Spectroscopy for Spectral Discrimination of Iron Oxalates

Enhanced characterization and phase identification of post-PUREX Pu Oxalates (PuOXA) are pivotal for nonproliferation and pre-detonation nuclear forensics. Despite significant advances in the characterization of PuO 2 samples, little is known about the impact of both the chemical structure and oxidation states of PuOXA (i.e., Pu(III) and Pu(IV)) have on optical emission signatures. Here, we demonstrate the analytical capabilities of laser-induced breakdown spectroscopy (LIBS) applied to Fe(II) and Fe(III) oxalate samples as surrogates for PuOXA, highlighting the discriminating features in the LIBS emission spectra arising from differences in the oxidation states within mixed FeOXA samples. We report the enhancement of spectral feature selection using Principal Component Analysis (PCA), which enables the analytical superiority of machine learning algorithms such as Linear Discriminant Analysis (LDA), Quadratic Discriminant Analysis (QDA), Partial Least Squares Regression (PLSR), Support Vector Regression (SVR), and Random Forest Regression (RFR) over conventional univariate techniques for phase discrimination and chemometric analysis. Cluster analysis revealed how both matrix effects and laser ablation influence cluster separability by introducing spectral artifacts that misdirect the maximization of variance. PCA-selected emission lines were used in the regression models, demonstrating that both univariate and multivariate linear regression models (i.e., PLSR and SVR) can achieve acceptable performance, with machine learning models outperforming conventional calibration regressions. Furthermore, the application of non-linearly activated PCA-selected emission lines illustrates how simplifying the data while retaining captured variance enables the use of less complex and more computationally efficient models. Furthermore, this is particularly evident in the underperformance of RFR, which suffers from increased computational costs and overfitting owing to its high complexity.

Oxalates↗

Resolution of a Reflector Shroud Fatigue Failure

Two cracks were observed on a reflector shroud for a space program after previously being subjected to the protoflight test campaign and several regression tests. After extensive analysis and investigations, the failure mechanism was identified to be fatigue as a result of the numerous vibration tests imposed on the unit. Two feasible corrective actions were proposed: first, a notched vibration profile which possesses sufficient margin from the anticipated acoustic and launch loads, while maintaining adequate fatigue life through launch and on-orbit operations, and second, a re-design of the shroud to strengthen the fatigue-susceptible areas. In this paper, we present the inspections, testing, and analysis performed to establish that the cracks were a result of fatigue failure. We discuss the conservative fatigue analysis methodology used in the development of both corrective action options. Finally, we review the lessons learned and the actions incorporated into the rework, subsequent regression testing, and the test plans to minimize the risk of recurrence in future units.

response limiting↗

Bayesian Symbolic Regression: Addressing Challenges in Estimating Fractional Bayes Factors and Application to Fatigue Crack Growth Modeling

This research pioneers advancements in computational mechanics by integrating Bayesian-based uncertainty quantification into symbolic regression, specifically focusing on the critical task of accurately estimating the fractional Bayes factor for selecting arbitrary equations. In our exploration, we rigorously study two prominent methods—sequential Monte Carlo and the Laplace approximation—employed for computing the fractional Bayes factor. Our findings underscore the limitations of the Laplace approximation, revealing its diminished accuracy in nonlinear and multimodal scenarios. Specifically, the Laplace approximation is shown to underpredict fractional Bayes factor on a wide set of equations associated with a symbolic regression benchmark. This comparative analysis sheds light on the nuanced performance of these techniques, guiding researchers toward more informed choices in uncertainty quantification within symbolic regression. Furthermore, we showcase the practical utility of these enhanced symbolic regression tools through their application to a real-world problem in fatigue crack growth modeling, emphasizing their efficacy in capturing the complexities of mechanical systems.

Geoffrey Bomarito↗

mvBayesR

SAND2025-11559O The mvBayesR tool performs multivariate Bayesian analysis on generic data. It includes tools for regression modeling, diagnosis, basis decomposition, sensitivity analysis, and visualization. The tool compiles state-of-the-art methodology into one easy-to-use package. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Tucker, James [Sandia National Lab. (SNL-CA), Live↗

mvBayesPy

SAND2025-11476O The mvBayesPy tool is a Python package that performs multivariate Bayesian analysis on generic data. It includes tools for regression modeling, diagnosis, basis decomposition, sensitivity analysis and visualization. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Tucker, James [Sandia National Lab. (SNL-CA), Live↗