Search NASASearch

SEARCH · Search NASA

Results for “Data augmentation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

OLCF’s Advanced Computing Ecosystem (ACE): FY25 Update for Ongoing Efforts

The advent of widespread use of artificial intelligence (AI) and machine learning (ML) models in science, coupled with fast data production rates of scientific instruments strain the traditional batch-oriented high-performance computing (HPC) environment. As scientific exploration continues to require more data and faster processing and analysis, new emerging technologies and capabilities to enable cross-facility and time-sensitive workflows are required for seamless integration of HPC and experimental facilities. The Advanced Computing Ecosystem (ACE) is a strategic initiative within the Oak Ridge Leadership Computing Facility (OLCF) established in 2024 to support the development of cutting-edge technologies to advance computational research and infrastructure at OLCF and across the Department of Energy (DOE). Several DOE initiatives are spearheading the evolution of the scientific landscape by blurring facility boundaries and connecting the user facilities to advance scientific capabilities and ensure energy dominance. The DOE Integrated Research Infrastructure (IRI) program is one example that is laying a foundation to support complex cross-facility workflows. The IRI program aims to integrate diverse computational resources, data infrastructures, and scientific instruments to facilitate collaboration and accelerate scientific discovery. The Interconnected Science Ecosystem (INTERSECT) initiative at Oak Ridge National Laboratory (ORNL) is another example that aims to revolutionize scientific research through AI-driven, interconnected autonomous laboratories and research facilities. Finally, the American Science Cloud (AmSC), recently announced in the “One Big Beautiful Bill”, aims to leverage prior infrastructure efforts of the IRI and automation and AI efforts of INTERSECT (and others) to build a federated, AI-augmented AmSC platform to unify the DOE’s computing, experimental, and data resources to catalyze scientific innovation.

97 MATHEMATICS AND COMPUTING

A Physics-Based Digital Twin for Wave Elevation and Seabed Moment Estimation of Offshore Monopiles: Preprint

In this work, we present a proof of concept of a physics-based digital twin for a monopile structure (with overhead inertia) subjected to wave loading. The digital twin is formulated using reduced-order models derived from first principles and combined with a Kalman filter for state estimation. The proposed framework estimates the monopile top motion, the wave elevation, and the section forces and moments along the pile using primarily acceleration measurements at the monopile top. Key innovations include the use of a hydrodynamic shape function to represent distributed wave loading in a compact and computationally efficient manner, and the introduction of a shaping filter to augment the state-space with wave kinematics. Synthetic measurement data are generated using OpenFAST and used as a reference to assess the performance of the digital twin. Results demonstrate that the wave elevation can be accurately reconstructed without direct sea-state measurements as long as the wave regime is inertia-dominated. Under the ideal tested conditions, the total hydrodynamic force and sea-bed bending moment are estimated with relative errors on the order of 1% and correlation coefficients exceeding 96%. Future work will evaluate the estimator's performance under operational uncertainties and more complex loading conditions.

17 WIND ENERGY

Enhanced HLW glass property-composition models - phase 3

During the present phase (Phase 3) of work to enhance and expand the HLW glass property-composition models, test data for 137 glasses were collected and incorporated into the combined WTP/ORP database. The new data include those collected on glasses from two statistically designed matrices to augment the moderate alumina region (the HLW16-MA matrix covering the region of 9 wt% < Al2O3 < 12 wt% with 16 glasses) and the ultra-high alumina region (the HLW16-UHA matrix covering the region of 26 wt% < Al2O3 < 29 wt% with 29 glasses). The addition of ORP glasses over the different phases of testing has significantly expanded the combined WTP/ORP database. The compiled data include glass compositions, PCT (B, Li and Na) releases, spinel T1%, electrical conductivity, viscosity, TCLP-Cd releases, and nepheline formation upon CCC. With the exception of the CCC spinel formation data – the CCC data will be used to support the further development of nepheline models for Hanford HLW glasses.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W

A Comparison of Pre‐Construction and Operational Wake Loss Estimates for Land‐Based Wind Plants

The overall bias between pre‐construction energy yield assessment (EYA) estimates of wind plant energy production and the achieved operational production is improving in the wind industry, but uncertainty remains high for individual wind plants. Wake effects within wind plants are one of the largest sources of energy loss considered in the EYA process, and previous work shows wake loss estimates to be a major source of disagreement among wind energy consultants who perform EYAs. To better understand the accuracy of wake loss predictions, we compare overall operational wake loss estimates based on supervisory control and data acquisition data to pre‐construction estimates provided by six wind energy consultants for five land‐based wind plants in North America. By augmenting existing approaches for quantifying operational wake losses, we estimate wake losses during the period of record for which operational data are available as well as the expected long‐term wake losses, based on historical reanalysis weather data, to which the EYA estimates are compared. To account for power variations at different turbine locations caused by terrain‐induced wind resource heterogeneity, we correct the operational wake loss estimates using predicted freestream wind speed variations from the Wind Systems Engineering Reynolds‐averaged Navier–Stokes (RANS) tool. We identify long‐term corrected operational wake losses between 1.9% and 6.4% for the five plants, with a mean loss of 4%. For the project deemed most acceptable for operational wake loss assessment, which is located in the simplest terrain and isolated from neighboring plants, the mean EYA wake loss estimate is within 0.7 percentage points of the operational value of 6.4%. For most of the remaining plants, results suggest that wake losses are generally overpredicted by 2.6–6.3 percentage points. However, operational wake losses may be underestimated for many of these projects because of spatial wind resource variations not captured by the RANS model, external wake effects that are unaccounted for in the estimation process, and wind plant blockage effects. To better understand factors that contribute to the observed wake losses, we investigate operational wake losses as a function of wind direction and wind speed. As expected, wake losses are generally concentrated near wind directions that are aligned with rows of closely spaced turbines and at below‐rated wind speeds; however, for some projects, the energy produced by the wind plant exceeds the estimated potential energy of the plant without wake interactions for certain wind directions and wind speeds, suggesting inaccurate assumptions in the wake loss estimation method for those plants. Lastly, we compare predicted and operational wake losses for individual wind turbines, finding that even when overall wake losses are predicted accurately, large uncertainty exists at the turbine level.

17 WIND ENERGY

Dynamics of argon metastables in Ar–CH 4 radio frequency capacitively-coupled plasma: real-time monitoring with neural network-augmented broadband optical emission spectroscopy

In moderate-pressure radio frequency (RF) capacitively coupled plasmas generated in argon–methane mixtures, the density of argon metastable atoms (Ar 1 s 5 ) exhibits a non-monotonic dependence on methane (CH 4 ) concentration. Laser-induced fluorescence (LIF) was used to measure and compare local Ar 1 s 5 densities in Ar and Ar–CH 4 plasmas at 2.6 Torr and RF powers of 17–117 W. The addition of 1% CH 4 increases the metastable density, and 2% CH 4 triggers a strong depletion by an order of magnitude, compared to 1% CH 4 case. This non-monotonic behavior demonstrates the sensitivity of metastable populations to small gas admixtures, which is critical for processes where metastables drive precursor dissociation. For real-time monitoring of metastable population, broadband optical emission spectroscopy (OES) is augmented with a feedforward neural network (NN) to predict Ar s1 5 densities from spectral features. When trained on LIF data, the NN replicates the absolute densities and the dynamic trends of Ar 1 s 5 density variation. The NN-augmented broadband OES approach can be used as a simple and cost-effective tool for tracking Ar metastables in Ar-rich plasmas, facilitating industrial-scale optimization.

Yatom, Shurik [Princeton Plasma Physics Laboratory

Use of AI for Interpreting Technical Specifications for Power Uprates in Nuclear Power Plants

Powerpoint presentation. Background information provided on power plant uprates. Discussion of the current and proposed approaches to power plant uprates. Explanation of what data is used to draft a LAR. Methods such as retrieval augmented generation (RAG) and fine-tuning are discussed. Use case analysis is performed. Different failure types are examined. Conclusions are drawn from the analysis. Future work is proposed.

97 - MATHEMATICS AND COMPUTING

High-Performance Semiempirical Excited-State Molecular Dynamics Powered by Graphics Processing Units

Here, this Letter introduces excited-state molecular dynamics in PYSEQM, a GPU-accelerated semiempirical quantum chemistry engine implemented in PyTorch. The new module enables Born–Oppenheimer molecular dynamics (BOMD) using configuration-interaction singles and random phase approximation for excited states, allowing long trajectories and large statistical ensembles to be simulated efficiently on a single GPU. We also implement an extended Lagrangian excited-state BOMD (XL-ESMD) scheme that propagates auxiliary electronic variables, enabling relaxed ground and excited-state convergence thresholds without compromising energy conservation. The excited-state BOMD implementation scales smoothly from small chromophores to a nearly 900-atom dendrimer (taking 6.5 s per MD step). PYSEQM also supports batched execution, allowing many geometries or trajectories to be evaluated in a single GPU launch, substantially increasing throughput and making ensemble-based protocols routine. As a demonstration, we compute absorption, emission, and infrared spectra from trajectories propagated on the ground and first excited states. The XL-ESMD scheme yields identical spectra at significantly lower computational cost, establishing the role of extended Lagrangian based dynamics for efficient excited-state BOMD simulations. Beyond raw performance, PYSEQM’s PyTorch foundation provides automatic differentiation for forces, efficient GPU batching, and seamless interfacing with machine learning models. These capabilities position PYSEQM as a practical platform for machine learning-augmented excited-state dynamics and lay the foundation for future data-driven nonadiabatic excited-state dynamics modeling of ultrafast spectroscopic probes.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

RAG for FLAG: AI Assistance for a Physics Code

Artificial intelligence (AI) has quickly become an important tool in scientific research, where significant efforts are underway to develop tools that will expedite the research process. One area of particular impact is scientific software, which can be particularly complex, and therefore time consuming to learn and use effectively. AI assistants are increasingly helping to streamline the process by performing tasks such as interactively answering user questions or suggesting solutions. Los Alamos National Laboratory (LANL) develops several advanced scientific codes, such as FLAG, which can be used to run multiphysics simulations. With this study, our goal was to develop an AI assistant for FLAG that could help make the process of understanding the software and running physics simulations more efficient. To develop an AI assistant for FLAG, we used a method called retrieval-augmented generation (RAG), which is a technique that uses information from relevant data sources to enhance the accuracy of large language models (LLMs). We used the FLAG user manual and other FLAG documentation as the knowledge base for the RAG system. When a user provides a query, RAG retrieves relevant sections from the knowledge base in response, then uses those excerpts to generate grounded and contextually rich answers. We found that our AI assistant was able to provide context aware answers and source references to user queries. To evaluate performance, we developed a set of 40 benchmark questions and compared the accuracy of the responses to those of two standard LLMs without retrieval. Our AI assistant significantly outperformed the standard LLMs at answering FLAG-related questions, with an 82.5% accuracy rate, compared to 47.5% for both of the standard LLMs. This has the potential to make the process of learning and using FLAG much easier, especially for new users. Ultimately, it supports LANL’s broader mission by empowering scientists and engineers to focus more on discovery and analysis rather than on navigating complex software systems.

97 MATHEMATICS AND COMPUTING

The Importance of Being Adaptable: An Exploration of the Power and Limitations of Domain Adaptation for Simulation-Based Inference with Galaxy Clusters

The application of deep machine learning methods in astronomy has exploded in the last decade, with new models showing remarkably improved performance on benchmark tasks. Not nearly enough attention is given to understanding the models' robustness, especially when the test data are systematically different from the training data, or "out of domain." Domain shift poses a significant challenge for simulation-based inference, where models are trained on simulated data but applied to real observational data. In this paper, we explore domain shift and test domain adaptation methods for a specific scientific case: simulation-based inference for estimating galaxy cluster masses from X-ray profiles. We build datasets to mimic simulation-based inference: a training set from the Magneticum simulation, a scatter-augmented training set to capture uncertainties in scaling relations, and a test set derived from the IllustrisTNG simulation. We demonstrate that the Test Set is out of domain in subtle ways that would be difficult to detect without careful analysis. We apply three deep learning methods: a standard neural network (NN), a neural network trained on the scatter-augmented input catalogs, and a Deep Reconstruction-Regression Network (DRRN), a semi-supervised deep model engineered to address domain shift. Although the NN improves results by 17% in the Training Data, it performs 40% worse on the out-of-domain Test Set. Surprisingly, the Scatter-Augmented Neural Network (SANN) performs similarly. While the DRRN is successful in mapping the training and Test Data onto the same latent space, it consistently underperforms compared to a straightforward Yx scaling relation. These results serve as a warning that simulation-based inference must be handled with extreme care, as subtle differences between training simulations and observational data can lead to unforeseen biases creeping into the results.

Ntampaka, Michelle [Baltimore, Space Telescope Sci

Reducing AI RAG Hallucination by Optimizing Routing Techniques

Large Language Models (LLMs), such as ChatGPT, tend to “hallucinate”, meaning they confidently generate false information. Retrieval Augmented Generation (RAG) attempts to diminish hallucination by providing context to the LLM from data stores (indexes) containing relevant information. The LLM uses this context to formulate its response. RAG systems can still suffer from hallucination because of bad embeddings or ineffective routing. For example, a router will often return context from an irrelevant index, resulting in a hallucinated answer. In this study, we aim to minimize the frequency of routing hallucinations by optimizing Index Summary Routing.

97 MATHEMATICS AND COMPUTING

Measuring the Energy Consumption and Efficiency of Deep Neural Networks: An Empirical Analysis and Design Recommendations

Addressing the "Red-AI" trend of rising energy consumption by large-scale neural networks, this study investigates the measured energy consumption of training various fully connected neural network architectures. We introduce the BUTTER-E dataset, an augmentation to the BUTTER Empirical Deep Learning dataset, containing energy consumption and performance data from 41,129 individual experimental runs spanning 30,582 distinct configurations: 13 datasets, 20 sizes (trainable parameters), 8 "shapes", and 14 depths on both CPUs and GPUs using node-level watt-meters. This dataset reveals the complex relationship between dataset size, network structure, and energy use. Our analysis uncovers a surprising, hardware-mediated non-linear relationship between energy efficiency and network design, challenging the assumption that reducing the number of parameters or FLOPs is the best way to achieve greater energy efficiency. We propose a straightforward and effective energy model that accounts for network size, computing, and memory hierarchy. Highlighting the need for cache-considerate algorithm development, we suggest a codesign approach to energy efficient network, algorithm, and hardware design. This work contributes to the fields of sustainable computing and Green AI, offering practical guidance for creating more energy-efficient neural networks and promoting sustainable AI.

97 MATHEMATICS AND COMPUTING

Geophysical Observations of the 2023 September 24 OSIRIS-REx Sample Return Capsule Reentry

Sample return capsules (SRCs) entering Earth's atmosphere at hypervelocity from interplanetary space are a valuable resource for studying meteor phenomena. The 2023 September 24 arrival of the Origins, Spectral Interpretation, Resource Identification, and Security-Regolith Explorer SRC provided an unprecedented chance for geophysical observations of a well-characterized source with known parameters, including timing and trajectory. A collaborative effort involving researchers from 16 institutions executed a carefully planned geophysical observational campaign at strategically chosen locations, deploying over 400 ground-based sensors encompassing infrasound, seismic, distributed acoustic sensing, and Global Positioning System technologies. Additionally, balloons equipped with infrasound sensors were launched to capture signals at higher altitudes. This campaign (the largest of its kind so far) yielded a wealth of invaluable data anticipated to fuel scientific inquiry for years to come. The success of the observational campaign is evidenced by the near-universal detection of signals across instruments, both proximal and distal. This paper presents a comprehensive overview of the collective scientific effort, field deployment, and preliminary findings. The early findings have the potential to inform future space missions and terrestrial campaigns, contributing to our understanding of meteoroid interactions with planetary atmospheres. Furthermore, the data set collected during this campaign will improve entry and propagation models and augment the study of atmospheric dynamics and shock phenomena generated by meteoroids and similar sources.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND

2012 California Household Travel Survey Supplement

# 2012 California Household Travel Survey Supplement The 2012 California Household Travel Survey Supplement focused on gathering specific travel information from residents for the development of next-generation, activity-based models. Called the "Augment Survey," it supplemented the [2010–2012 California Household Travel Survey](https://www.nrel.gov/transportation/secure-transportation-data/tsdc-california-travel-survey). ## Data Collection Agency The Southern California Association of Governments (SCAG) hired Abt-SRBI, Inc. to conduct the survey. ## Methodology Travel data were collected from households via in-vehicle (625 vehicles) and wearable (244 participants) global positioning system (GPS) devices. ## Drive Cycle Processing and Filtering NREL has developed a GPS data filtration routine to filter erroneous data points in individual drive cycles sourced from GPS devices mounted in vehicles. Second-by-second drive cycle data collected from GPS-instrumented vehicles during this survey have passed through NREL's drive cycle processing and filtering routines. ## Survey Records Study records include 473 households. ## Transportation Data The SCAG data set contains data from 473 households that participated in one or more areas of study. Of these, 141 completed the wearable GPS portion of the study and 332 completed the vehicle GPS portion. There was no overlap between households participating in the two study areas (wearable and vehicle GPS). For details on available travel survey data and variable definitions, see the [data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/caltrans_scag_data_dictionary.pdf?sfvrsn=6ec36d7a_1). NREL-generated drive cycle data are also available for this survey. For details on available data and variable definitions, see the [drive cycle data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/drive_cycles_data_dictionary.pdf?sfvrsn=7de7e888_1). Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

2012 California Household Travel Survey Supplement

# 2012 California Household Travel Survey Supplement The 2012 California Household Travel Survey Supplement focused on gathering specific travel information from residents for the development of next-generation, activity-based models. Called the "Augment Survey," it supplemented the [2010–2012 California Household Travel Survey](https://www.nrel.gov/transportation/secure-transportation-data/tsdc-california-travel-survey). ## Data Collection Agency The Southern California Association of Governments (SCAG) hired Abt-SRBI, Inc. to conduct the survey. ## Methodology Travel data were collected from households via in-vehicle (625 vehicles) and wearable (244 participants) global positioning system (GPS) devices. ## Drive Cycle Processing and Filtering NREL has developed a GPS data filtration routine to filter erroneous data points in individual drive cycles sourced from GPS devices mounted in vehicles. Second-by-second drive cycle data collected from GPS-instrumented vehicles during this survey have passed through NREL's drive cycle processing and filtering routines. ## Survey Records Study records include 473 households. ## Transportation Data The SCAG data set contains data from 473 households that participated in one or more areas of study. Of these, 141 completed the wearable GPS portion of the study and 332 completed the vehicle GPS portion. There was no overlap between households participating in the two study areas (wearable and vehicle GPS). For details on available travel survey data and variable definitions, see the [data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/caltrans_scag_data_dictionary.pdf?sfvrsn=6ec36d7a_1). NREL-generated drive cycle data are also available for this survey. For details on available data and variable definitions, see the [drive cycle data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/drive_cycles_data_dictionary.pdf?sfvrsn=7de7e888_1). Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

2012 California Household Travel Survey Supplement

# 2012 California Household Travel Survey Supplement The 2012 California Household Travel Survey Supplement focused on gathering specific travel information from residents for the development of next-generation, activity-based models. Called the "Augment Survey," it supplemented the [2010–2012 California Household Travel Survey](https://www.nrel.gov/transportation/secure-transportation-data/tsdc-california-travel-survey). ## Data Collection Agency The Southern California Association of Governments (SCAG) hired Abt-SRBI, Inc. to conduct the survey. ## Methodology Travel data were collected from households via in-vehicle (625 vehicles) and wearable (244 participants) global positioning system (GPS) devices. ## Drive Cycle Processing and Filtering NREL has developed a GPS data filtration routine to filter erroneous data points in individual drive cycles sourced from GPS devices mounted in vehicles. Second-by-second drive cycle data collected from GPS-instrumented vehicles during this survey have passed through NREL's drive cycle processing and filtering routines. ## Survey Records Study records include 473 households. ## Transportation Data The SCAG data set contains data from 473 households that participated in one or more areas of study. Of these, 141 completed the wearable GPS portion of the study and 332 completed the vehicle GPS portion. There was no overlap between households participating in the two study areas (wearable and vehicle GPS). For details on available travel survey data and variable definitions, see the [data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/caltrans_scag_data_dictionary.pdf?sfvrsn=6ec36d7a_1). NREL-generated drive cycle data are also available for this survey. For details on available data and variable definitions, see the [drive cycle data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/drive_cycles_data_dictionary.pdf?sfvrsn=7de7e888_1). Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

2012 California Household Travel Survey Supplement

# 2012 California Household Travel Survey Supplement The 2012 California Household Travel Survey Supplement focused on gathering specific travel information from residents for the development of next-generation, activity-based models. Called the "Augment Survey," it supplemented the [2010–2012 California Household Travel Survey](https://www.nrel.gov/transportation/secure-transportation-data/tsdc-california-travel-survey). ## Data Collection Agency The Southern California Association of Governments (SCAG) hired Abt-SRBI, Inc. to conduct the survey. ## Methodology Travel data were collected from households via in-vehicle (625 vehicles) and wearable (244 participants) global positioning system (GPS) devices. ## Drive Cycle Processing and Filtering NREL has developed a GPS data filtration routine to filter erroneous data points in individual drive cycles sourced from GPS devices mounted in vehicles. Second-by-second drive cycle data collected from GPS-instrumented vehicles during this survey have passed through NREL's drive cycle processing and filtering routines. ## Survey Records Study records include 473 households. ## Transportation Data The SCAG data set contains data from 473 households that participated in one or more areas of study. Of these, 141 completed the wearable GPS portion of the study and 332 completed the vehicle GPS portion. There was no overlap between households participating in the two study areas (wearable and vehicle GPS). For details on available travel survey data and variable definitions, see the [data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/caltrans_scag_data_dictionary.pdf?sfvrsn=6ec36d7a_1). NREL-generated drive cycle data are also available for this survey. For details on available data and variable definitions, see the [drive cycle data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/drive_cycles_data_dictionary.pdf?sfvrsn=7de7e888_1). Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

2012 California Household Travel Survey Supplement

# 2012 California Household Travel Survey Supplement The 2012 California Household Travel Survey Supplement focused on gathering specific travel information from residents for the development of next-generation, activity-based models. Called the "Augment Survey," it supplemented the [2010–2012 California Household Travel Survey](https://www.nrel.gov/transportation/secure-transportation-data/tsdc-california-travel-survey). ## Data Collection Agency The Southern California Association of Governments (SCAG) hired Abt-SRBI, Inc. to conduct the survey. ## Methodology Travel data were collected from households via in-vehicle (625 vehicles) and wearable (244 participants) global positioning system (GPS) devices. ## Drive Cycle Processing and Filtering NREL has developed a GPS data filtration routine to filter erroneous data points in individual drive cycles sourced from GPS devices mounted in vehicles. Second-by-second drive cycle data collected from GPS-instrumented vehicles during this survey have passed through NREL's drive cycle processing and filtering routines. ## Survey Records Study records include 473 households. ## Transportation Data The SCAG data set contains data from 473 households that participated in one or more areas of study. Of these, 141 completed the wearable GPS portion of the study and 332 completed the vehicle GPS portion. There was no overlap between households participating in the two study areas (wearable and vehicle GPS). For details on available travel survey data and variable definitions, see the [data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/caltrans_scag_data_dictionary.pdf?sfvrsn=6ec36d7a_1). NREL-generated drive cycle data are also available for this survey. For details on available data and variable definitions, see the [drive cycle data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/drive_cycles_data_dictionary.pdf?sfvrsn=7de7e888_1). Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

2012 California Household Travel Survey Supplement

# 2012 California Household Travel Survey Supplement The 2012 California Household Travel Survey Supplement focused on gathering specific travel information from residents for the development of next-generation, activity-based models. Called the "Augment Survey," it supplemented the [2010–2012 California Household Travel Survey](https://www.nrel.gov/transportation/secure-transportation-data/tsdc-california-travel-survey). ## Data Collection Agency The Southern California Association of Governments (SCAG) hired Abt-SRBI, Inc. to conduct the survey. ## Methodology Travel data were collected from households via in-vehicle (625 vehicles) and wearable (244 participants) global positioning system (GPS) devices. ## Drive Cycle Processing and Filtering NREL has developed a GPS data filtration routine to filter erroneous data points in individual drive cycles sourced from GPS devices mounted in vehicles. Second-by-second drive cycle data collected from GPS-instrumented vehicles during this survey have passed through NREL's drive cycle processing and filtering routines. ## Survey Records Study records include 473 households. ## Transportation Data The SCAG data set contains data from 473 households that participated in one or more areas of study. Of these, 141 completed the wearable GPS portion of the study and 332 completed the vehicle GPS portion. There was no overlap between households participating in the two study areas (wearable and vehicle GPS). For details on available travel survey data and variable definitions, see the [data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/caltrans_scag_data_dictionary.pdf?sfvrsn=6ec36d7a_1). NREL-generated drive cycle data are also available for this survey. For details on available data and variable definitions, see the [drive cycle data dictionary](https://www.nrel.gov/media/docs/libraries/tsdc/drive_cycles_data_dictionary.pdf?sfvrsn=7de7e888_1). Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI