Search NASA⌕ Search

SEARCH · Search NASA

Results for “data sets”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 271 records · Page 15

Bias Correction and Statistical Downscaling of Future Solar Irradiance Projections Using the NSRDB

Assessing renewable energy resources under future climate scenarios has been highlighted to understand potential impacts of future climate change in renewable generation on the power sector. Climate model projection has been recognized by the renewable energy community as a useful data set to analyze the impacts of future climate change on renewable resources. However, future climate projections generated from general circulation models (GCMs) contain inherent biases that need to be corrected for accurate analysis of future projections of climate variables. In addition, the coarse spatiotemporal resolution of GCMs needs to be improved for regional climate studies. In this work, we develop statistical methods to downscale future projections of global horizontal irradiance (GHI) in a computationally efficient way. Our approach builds statistical downscaling models that correct bias of climate projection of GHI and downscale the future GHI projection from daily-scale to hourly-scale. The National Solar Radiation Database (NSRDB) is used to calibrate the statistical models and validate the downscaled GHI projections across the contiguous United State (CONUS). Preliminary results show that the statistical approach efficiently downscales climate projections of GHI with a nBIAS of 3%, nMAE of 34 % and nRMSE of 46% calculated against NSRDB for CONUS. This study describes the implemented methodology and initial results as well as future research to create high-resolution climate data sets for solar energy applications.

analytical models↗

Thermal images collected during large-scale 3D printing with Ingersoll MasterPrint

This data set contains images produced to test the performance of anomaly and fault detection methods in the context of additive manufacturing. Two print jobs were executed using the Ingersoll MasterPrint with the Model 30 Strangpresse extruder. The material used was Techmer compounded polylactic acid (PLA) with wood flour as a filler (80/20 PLA/WF by weight). The first print job consists of the 3D printing of a hexagonal cylinder with a two-bead wall. This print was sliced at a gantry velocity of 3000 millimeter per minute, and an extruder screw speed of 68.14 rotations per minute. This are considered the normal operating conditions. A second hexagonal cylinder was printed using a reduced extruder screw speed, 15% lower than under normal operating conditions. The images were collected with a Teledyne FLIR Lepton 3.5 infra-red camera, a small form factor radiometric long-wave infrared camera with a spectral range of 8 µm to 14 µm. Sensor resolution was 160x120 pixels, with a pixel size of 12 µm, a temperature range of -10 - 450°C, and an accuracy of +/- 10°C in its low gain configuration. Images in this data set were collected with cameras oriented at the printer nozzle. The nozzle camera setup consisted of two cameras located 12.5cm from the nozzle center. These were mounted directly to the print head, so that the camera positions relative to the print direction would remain constant as it rotated around its C-axis to follow the print path. One camera was placed ahead of the nozzle to capture the previous layer immediately before being covered by the new layer of material, while the second camera was placed behind the nozzle and captured the freshly extruded bead. Images are collecting during each phase of the print: (a) idle (i.e., no material deposited), (b) extrusion (i.e., to prime the extruder), and (c) printing (deposition of material to manufacture the hexagonal cylinder).

36 MATERIALS SCIENCE↗

JOINT APPOINTEE: Evolution of ferroelectric properties in SmxBi1-xFeO3 via automated Piezoresponse Force Microscopy across combinatorial spread libraries

Combinatorial spread libraries offer a innovative approach to explore the evolution of material properties over broad concentration, temperature, and growth parameter spaces. However, traditional limitation of this approach is the requirement for the read-out of functional properties across the library. Here we develop automated Piezoresponse Force Microscopy (PFM) for the exploration of combinatorial spread libraries and demonstrate its application in the SmxBi1-xFeO3 system with the ferroelectric-antiferroelectric morphotropic phase boundary. This approach relies on the synergy of the quantitative nature of PFM and the implementation of automated experiments that allow PFM-based sampling over macroscopic samples. The concentration dependence of pertinent ferroelectric parameters has been determined and used to develop the mathematical framework based on Ginzburg-Landau theory describing the evolution of these properties across the concentration space. We pose that a combination of automated scanning probe microscope and combinatorial spread library approach will emerge as an efficient research paradigm to close the characterization gap in the high-throughput materials discovery. We make the data sets open to the community and hope that this will stimulate other efforts to interpret and understand the physics of these systems.

Automated Microscopy, Combinatorial Library, Ferro↗

Improving Bond Dissociations of Reactive Machine Learning Potentials through Physics-Constrained Data Augmentation

In the field of computational chemistry, predicting bond dissociation energies (BDEs) presents well-known challenges, particularly due to the multireference character of reactive systems. Many chemical reactions involve configurations where single-reference methods fall short, as the electronic structure can significantly change during bond breaking. As generating training data for partially broken bonds is a challenging task, even state-of-the-art reactive machine learning interatomic potentials (MLIPs) often fail to predict reliable BDEs and smooth dissociation curves. By contrast, simple and inexpensive physics-based models, such as the well-established Morse potential, do not suffer from any such limitations. This work leverages the Morse potential to improve reactive MLIPs by augmenting the training data set with inexpensive Morse data along the dissociation pathways. Further, this physics-constrained data augmentation (PCDA) approach results in MLIPs with smooth bond dissociation curves as well as near coupled-cluster level BDEs, all without requiring any expensive multireference quantum mechanical calculations. A case study for methane combustion demonstrates how the PCDA approach can improve an existing reactive MLIP, namely, ANI-1xnr. In conclusion, not only are the BDEs and bond dissociation curves for all radicals and molecules significantly improved compared to ANI-1xnr but the PCDA-trained MLIP retains the reliability of ANI-1xnr when performing reactive molecular dynamics simulations.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Modeling the Enceladus dust plume based on in situ measurements performed with the Cassini Cosmic Dust Analyzer

We analyzed data recorded by the Cosmic Dust Analyzer on board the Cassini spacecraft during Enceladus dust plume traversals. Our focus was on profiles of relative abundances of grains of different compositional types derived from mass spectra recorded with the Dust Analyzer subsystem during the Cassini flybys E5 and E17. The E5 profile, corresponding to a steep and fast traversal of the plume, has already been analyzed. In this paper, we included a second profile from the E17 flyby involving a nearly horizontal traversal of the south polar terrain at a significantly lower velocity. Additionally, we incorporated dust detection rates from the High Rate Detector subsystem during flybys E7 and E21. We derived grain size ranges in the different observational data sets and used these data to constrain parameters for a new dust plume model. This model was constructed using a mathematical description of dust ejection implemented in the software package DUDI. Further constraints included published velocities of gas ejection, positions of gas and dust jets, and the mass production rate of the plume. Our model employs two different types of sources: diffuse sources of dust ejected with a lower velocity and jets with a faster and more colimated emission. From our model, we derived dust mass production rates for different compositional grain types, amounting to at least 28 kg s –1 . Previously, salt-rich dust was believed to dominate the plume mass based on E5 data alone. The E17 profile shows a dominance of organic-enriched grains over the south polar terrain, a region not well constrained by E5 data. By including both E5 and E17 profiles, we find the salt-rich dust contribution to be at most 1% by mass. This revision also results from an improved understanding of grain masses of various compositional types that implies smaller sizes for salt-rich grains. Our new model can predict grain numbers and masses for future mission detectors during plume traversals.

79 ASTRONOMY AND ASTROPHYSICS↗

DEPRECATED AI-Batt-OS (Autonomous Identification of Battery Life Models - Open Source) [SWR 21-17]

DEPRECATED. This repository was archived by the owner on Jun 30, 2026. It is now read-only. Open source implementation of some of the methods utilized by AI-Batt, a battery lifetime modeling and analysis toolkit provided by the National Laboratory of the Rockies (NLR). This software demonstrates the use of bi-level optimization and symbolic regression techniques to semi-autonomously identify algebraic models predicting the capacity fade of lithium-ion batteries during calendar aging. Modeling the degradation of batteries is a complex task, due to the difficulty in separating the time-dependent and time-independent factors impacting cell level degradation, across multiple data series with different numbers of measurements and/or data quality. Bi-level optimization enables model parameters to be optimized to either the entire data set or to individual data series, allowing statistical disambiguation of global behaviors (data series independent) and local behaviors (data series dependent). Symbolic regression is used to automatically search for optimal low-dimesional models predicting the variation of locally optimized parameters versus time-independent experimental variables from millions of possible models, resulting in a more accurate and repeatable model identification process than is possible by a manual search. The provided tools also implement cross-validation and bootstrap resampling schemes, empowering statistical model comparison/selection and quantification of model uncertainties. An example script replicates the results from the manuscript "Challenging Practices of Algebraic Battery Life Models through Statistical Validation and Model Identification via Machine-Learning", submitted to ECS. All code is written in MATLAB. Requires the Statistics and Machine Learning Toolbox. Contact Dr. Paul Gasper at Paul.Gasper@nlr.gov for any questions.

Gasper, Paul [National Renewable Energy Lab. (NREL↗

Predicting Li-Ion Battery Capacity Fade Using Early-Life Data and a Hybrid Data-Driven Gaussian Process-Bayesian Regression Approach

Accurately predicting Li-ion battery capacity trajectories using early-life data can dramatically improve battery-life understandings and be used to rapidly evaluate design/cost/performance trade-offs when developing new battery materials. Accurate early-life predictions enable researchers to quickly iterate over cell designs and material precursor properties without consistently cycling cells to failure. To this end, we present a toolbox that uses a combined Gaussian Process and Bayesian regression approach that capitalizes on signals other than just capacity (e.g., dQ/dV, voltage drops) to rapidly predict capacity-fade trajectories. The prediction tool uses Bayesian regression to fit functional forms, e.g., power law, sigmoids, etc., to predict capacity-fade dynamics. By fitting functional forms, the capacity fade can be interrogated at any point in the future, allowing for early cell-failure prediction. Additionally, Bayesian regression allows for accurate uncertainty estimates that account for cell-to-cell variability (aleatoric uncertainty) and the lack of observation data (epistemic uncertainty). By only using early cycle data to predict the capacity fade trajectory, uncertainty bounds at end-of-life can be extremely large. The large uncertainty bounds are further exacerbated because there is no systematic way to define the prior distribution of the functional forms' parameters. We improve our the predicted trajectory confidence interval of our predicted trajectory using two methods. First, we shows that a small amount of held-out cycling data is sufficientuse some train cells, that have been cycled to failure to derive information regarding the appropriate prior distributions for the functional forms' parameters of the functional form, effectively leading to data-driven priors.. We propose constructing the data-driven priors by first running a Bayesian regression starting with uninformed priors to generate intermediate cell-specific posterior parameter distributions. These posterior distributions are combined using a Ggaussian mixture model for each parameter to create the data-driven priors. These mixture models serve as the data-driven prior distributions for the parameters for. Second, we derive multiple features, e.g., C_dchg 0.5 DoD 0.5, log (|mean(dQ/dV_(w_3-w_0 ) (V)|), etc., from the train cellsheld-out cycling data, identify which the features are that best predicting capacity at early/mid-life cycles, and then create Ggaussian process regression models that are used for predicting capacity at early/mid-life cycles for the test cells (see blue dots with error bars in Fig 1b). Finally, these predicted data-points are used in addition to the actual early cycle data capacity fade to construct the Bayesian regression trajectory for the test cell s. Notably. We note that these two methods are complementary and can be combined with each other. We evaluate the performance of our proposed method on an testing open-source dataset from Iowa State University and Iowa Lakes Community College (ISU-ILCC). This dataset comprises of 251 nickel-manganese-cobalt/graphite Lithium-ion cells that are cycled under 63 different conditions. We compute the mean average percentage error (MAPE) and negative log predictive density (NLPD) to quantify the efficacy of our method. Our initial findings suggest that, when only few observations are available, for test cells, when using only Bayesian regression with uninformed priors, a power law functional provides the most accurate predictions. with very few data points. However, asHowever, a the number of data points increases, a twin sigmoidal function becomes more accurate as the number of observations further increases. We also find that using as little as 10% of the data set towards generating data-driven priors can lead to significant improvement in prediction accuracy when using early cycle data. Lastly, we found that augmenting early-cycle data with Gaussian process-predicted capacity data for Bayesian regression greatly improves the prediction accuracy. We will present a comprehensive comparison of our methods to other methods available in the literature and apply this method to additional battery datasets.

42 ENGINEERING↗

Inter-Kingdom Viral Interactions

Please cite as : Josué A. Rodríguez-Ramos, Amy E. Zimmerman, Ruonan Wu, Sheryl Bell, Trinidad Alfaro, Kirsten Hofmockel, William C. Nelson. 2025. Inter-Kingdom Viral Interactions. [Data Set] PNNL DataHub. This data is published under a CC0 license. The authors encourage data reuse and request attribution by referencing the above citations for the data package and associated manuscript. Deciphering viral ecology in soils is challenging due to their high physiochemical and community complexity. To enhance detection of sub-communities of DNA and RNA viruses, we applied fractionation approaches to soils collected across a moisture gradient from a grassland field experiment. Analyses included metagenomics and metatranscriptomics of size-fractionated extracellular viruses (i.e., DNA and RNA viromes), metagenomics of bacteria/archaea- or eukaryote-enriched samples, and whole soil metatranscriptomes with rRNA-depletion or polyadenylation enrichment. While RNA virome and whole soil RNA methods captured similar viral diversity, RNA viromes identified longer, higher-quality genomes. Further, we showed that significantly more DNA viruses were active in higher moisture than lower moisture samples, whereas responses by overall diversity vary by genome type (DNA versus RNA genomes). Finally, we demonstrate the power of fractionation approaches for identifying distinct viral communities that infect unique hosts, which has significant implications for ecological investigations, particularly related to interkingdom interactions.

59 BASIC BIOLOGICAL SCIENCES↗

BSEC VPRM 10m Hourly Biogenic Fluxes in Baltimore (2021)

Model outputs from the Vegetation Photosynthesis and Respiration Model (VPRM: version from Horne et al. in prep). Model remote sensing inputs come from Sential 2-derived EVI and LSWI. Model meteorological inputs for two-meter air temperature and shortwave incoming come from the BSEC WRF 2021 Control Run (Foust, W. 2023). Plant functional Types (PFTs) are spatially classified using the Chesapeake Bay Program 2018 land use land cover product. The final biogenic flux (µmol CO2 m^-2 s^-1) outputs of NEE, RESP, and GEE are a weighted average based on the portion of PFTs within the cell. Individual PFT outputs are saved inside PFT directories (e.g., Crops, Grass, etc.) inside the specific month directory. Model outputs are denoted as a negative flux into the land system (i.e., photosynthesis) and a positive flux as a net release into the overlying atmosphere. Respiration (RESP) fluxes are positive and combine heterotrophic (only soil) and autotrophic sources. Gross ecosystem exchange (GEE) is a negative flux driven by only photosynthetic activity from vegetation, and the Net ecosystem exchange (NEE) is the sum of the two (i.e., NEE=RESP+GEE). Data Characteristics Spatial Resolution: 10m Temporal Resolution: Hourly File Format: VPRM_ _BSEC. .tif (Hour is in UTC) For more information on the model results, please email Jason Horne (jph6488@psu.edu). References: Foust, W. (2023). BSEC WRF 2021 Control Run Output (v0.1.0) [Data set]. MSD-LIVE Data Repository. https://data.msdlive.org/records/m0e6m-vvq17

Baltimore↗

SAXS Assistant: Automated SAXS analysis for structural discovery in biologics and polymeric nanoparticles

Small-angle x-ray scattering (SAXS) is a powerful technique for assessing macromolecular structure. High-throughput SAXS is limited by the time-consuming and, at times, subjective nature of SAXS data interpretation. Here, we present SAXS Assistant, a Python-based script that streamlines SAXS data analysis to extract features for machine learning (ML) and key structural parameters, including the Guinier radius of gyration (R g ), pair distance distribution function (PDDF)-derived R g , maximum particle dimension (D max ), and Kratky plots. The script builds upon BioXTAS RAW and validates reliability via Guinier/PDDF R g agreement, an important indicator of well-measured data sets. For assistance in D max estimation, a multilayer perceptron regressor was trained with 1940 data files from the Small Angle Scattering Biological Data Bank. The model achieved a test set performance R 2 = 0.90 and mean absolute error = 11.7 Å. Training exclusively with experimental data translates analyses from researchers, including experts in the field, to the ML model, which helps assess D max estimations from PDDF. Gaussian mixture model clustering was implemented to classify profiles into structural classes based on entries in the Small Angle Scattering Biological Data Bank. Users may therefore assess the similarity between experimental samples and known biomolecular shapes within the mapped repository entries. This probabilistic clustering aids in quantifying information from Kratky and generating shape-descriptive features. SAXS Assistant accelerates SAXS data analysis through enforced quality control, ML-ready outputs, and flags for low-confidence results. In addition to providing the ability to analyze large data sets at high throughput, this tool is versatile and may serve researchers in both biological and synthetic polymer research fields.

36 MATERIALS SCIENCE↗

Contrasting Time-Frequency Representations for Unknown Waveform Detection

Identifying unseen electromagnetic waveforms is critical for many applications, like interference management, electronic warfare and spectrum management. Traditionally this is done using statistical methods for anomaly detection, which has evolved to deep learning models for identifying the unseen data, formally termed as open set recognition. Some prior methods use a generative model to emulate open set data, which face challenges in generating synthetic samples for open set while simultaneously selecting an optimal discriminator for accurate classification. To alleviate this issue, we propose a discriminative model that effectively combines time and frequency domain features of communication signals for accurate predictions. We further introduce a cosine similarity loss that makes the domain specific features unique to enhance the prediction rate. Additionally, our model avoids generic feature vectors by extracting class-specific features during training, resulting in improved class representation. The experiment results show that this combined feature approach with cosine loss outperforms single-domain models and improves accuracy by 10% over models without cosine loss.

99 - GENERAL AND MISCELLANEOUS↗

Informing Robust Functional Relationship Benchmarks: An Evaluation of the Temperature Sensitivity of Ecosystem Respiration Across the Arctic-Boreal Region

During land model development, simulated carbon dynamics are often benchmarked against observational data sets to evaluate model performance. Functional relationship benchmarks are the relationship between a driving variable (e.g., temperature) and a response variable (e.g., ecosystem respiration) and are a promising tool for assessing model performance by evaluating modeled sensitivities to changing environmental conditions. However, observed functional relationships can be influenced by choices made during data collection and throughout the benchmarking process, impacting the inferred skill of land models. To avoid misrepresenting a model's true performance, it is necessary to systematically evaluate best practices when constructing functional relationship benchmarks. We developed a set of guidelines for constructing functional relationship benchmarks, considering the choice of data set, number of daily observations, temporal extent, and temporal resolution across Alaska and Canada over a 20-year period from 2001 to 2020. The temperature sensitivity of ecosystem respiration from observations, evaluated through an apparent Q 10 , is highly variable both spatially and as a result of the data processing approach applied in the benchmark formation. When benchmarking 13 models from the Warming Permafrost Model Intercomparison Project (WrPMIP), the range in inferred model skill is substantially impacted by the choices applied in constructing functional relationship benchmarks. The inferred performance of a given model is most sensitive to the number of daily observations and temporal extent, followed by choice of benchmark data set and temporal averaging. Results from this analysis can guide the development of consistent and robust functional relationships for future model evaluation studies.

Poe, Jeralyn [Northern Arizona University, Flagsta↗

Comparability of Liquid Chromatography Tandem Mass Spectrometry Analysis of Dissolved Organic Matter across Laboratories

Non-targeted liquid chromatography tandem highresolution mass spectrometry (LC−MS/MS) is increasingly applied for the structure-resolved chemical analysis of dissolved organic matter (DOM). With new developments in MS instrumentation and analysis software, the approach has gained substantial momentum over the past decade. However, achieving high-quality analytical data that is reproducible and comparable across laboratories can be a bottleneck in non-targeted metabolomics and organic matter chemical analysis, especially for data reuse in repository-scale analyses. Understanding the capabilities as well as challenges of comparing LC−MS/MS data from different laboratories is necessary for inferring global trends from public data sets. To illuminate instrumentation factors that drive differences and variability, we used a standardized data analysis pipeline, including classical (CMN) and featurebased molecular networking (FBMN), to analyze data from a ring trial by 24 laboratories on identical sample sets of algal and DOM extracts that were mixed in predefined concentrations and spiked with standards. Our results showed that data sets from similar mass spectrometer types with unified instrument parameters were qualitatively comparable, resolving the same general trends and shared mass spectral features. Interlaboratory comparability was best for high-intensity features, while low-intensity features showed greater detection variability. Our analysis also highlights challenges when comparing data from instruments with different acquisition rates or operating with less standardized methods. Lastly, we provide recommendations for data integration, public data sharing, standardization, and best practices for standardized LC−MS/MS data acquisition, which will be critical for long-term time series and intercomparability of DOM chemical analyses.

DOM↗

Graphene SHDMC Data

The data used to produce all of the figures and tables in the manuscript titled "Highly Accurate Many-Body Theory Reaches 2D Materials" can be found here. This data set includes: -SHDMC results for graphene -selected CI with and without re-normalized second-order perturbation (rPT2) theory corrections for graphene -data demonstrating that SHDMC displays an exponential rate of convergence -data used for sCI + rPT2 complete basis set extrapolation -data used to extrapolate SHDMC energies to the infinite basis limit -data used to demonstrate compactness of SHDMC wavefunction

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Physical, socio-psychological, and behavioural determinants of household energy consumption in the UK

Determining which attitudes and behaviours predict household energy consumption can help accelerate the low-carbon energy transition. Conventional approaches in this domain are limited, often relying on survey methods that produce data on individuals’ motivations and self-reported activities without pairing these with actual energy consumption records, which are particularly hard to collect for large, nationally representative samples. This challenge precludes the development of empirical evidence on which attitudes and behaviours influence patterns of energy consumption, thus limiting the extent to which these can inform energy interventions or conservation programs. This study demonstrates a novel methodology for estimating energy consumption in the absence of actual energy records by using a large, publicly available data set of energy consumption in the UK. We develop a predictive model using the Smart Energy Research Laboratory (SERL) data portal (with records from nearly 13,000 UK households) and then use this model to predict energy consumption (both electric and gas) for a sample of 1,000 UK householders for which we separately collect over 200 variables relating to climate change attitudes and practices. Our approach uses a set of over 50 independent variables that are shared between the data sets, allowing us to train a model on the SERL data and use it to analyse the relationship between energy consumption and the opinions, motivations, and daily practices of survey respondents. Results show that electricity consumption is influenced by a broader range of factors compared to gas. Household energy use is best explained by physical dwelling characteristics, socio-demographic variables, and certain behavioural and attitudinal measures. Notably, pro-environmental attitudes, frugality, and conscientiousness correlate with lower energy use, while income and consumerism are linked to higher consumption. We discuss how these findings can inform efforts to decarbonise home energy use in the UK.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Temperature and Water Levels Collectively Regulate Methane Emissions From Subtropical Freshwater Wetlands

Abstract Wetlands are the largest and most climate‐sensitive natural sources of methane. Accurately estimating wetland methane emissions involves reconciling inversion (“top‐down”) and process‐based (“bottom‐up”) models within the global methane budget. However, estimates from these two model types are inherently interdependent and often reveal substantial discrepancies. To enhance the reliability of both approaches, we need a comprehensive understanding of wetland methane emissions and an independent high‐resolution long‐term flux data set. Here, we employed a data‐driven random forest approach to identify key variables influencing methane emissions from subtropical freshwater wetlands in the Southeastern United States. The model‐estimated monthly mean methane fluxes fit well with measured methane fluxes ( R 2 = 0.67) at four representative FLUXNET‐CH4 wetland sites across the region. Variable importance analysis highlighted the sensitivity of subtropical freshwater wetland methane emissions to variations in both temperature and water levels. High temperatures facilitate methanogenesis by enhancing microbial activities, while elevated water levels maintain anaerobic conditions necessary for methane production. Notably, the response of methane emissions to water level fluctuations is contingent on temperature conditions, and vice versa. Moreover, we constructed the first high‐spatial‐resolution (∼1 km × 1 km) and long‐term (1982–2010) gridded regional wetland methane flux product for the Southeastern United States, estimating annual methane emissions from subtropical freshwater wetlands in the region at 4.93 ± 0.11 Tg CH 4 yr −1 for 1982–2010. This new benchmark product holds promise for validating and parameterizing uncertain wetland methane emission processes in bottom‐up models and provides improved prior information for top‐down models.

He, Keqi [Earth and Climate Sciences Nicholas Scho↗

cjohnson-LANL/GRL_Kilauea

The python routines are outlined in detail to perform the methods and results in the manuscript under review in the journal Geophysical Research Letters titled “Seismic features predict ground motions during repeating caldera collapse sequence” with LA-UR-23-33345. All routines are written in open source python and were applied to publicly available data sets. The codes formats the data into the appropriate structure required to train a boosted tree regression model. Other codes produce figure results.

Johnson, Christopher W↗