Search NASA⌕ Search

SEARCH · Search NASA

Results for “building extraction”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Boundary-Aware Adversarial Learning Domain Adaption and Active Learning for Cross-Sensor Building Extraction

The use of convolutional neural networks (CNNs) for building extraction from remote sensing images has been widely studied and many public datasets have been made available for accelerating development of these CNN models. Yet adapting pretrained models at scale in real-world scenarios remains a challenging task. The main barrier is that certain new labels are still needed to compensate for domain shifting between the labeled data and new images that potentially cover new geographic locations or that are from a different sensor. In this article, we propose to add informatively labeled samples from a new image pool under the paradigm of active learning. To select the most useful samples based on model uncertainty, we first tackle the problem of uncalibrated uncertainty estimation due to distribution shifting by adapting feature extractors with boundary-based adversarial learning. Calibrated uncertainty is used as the query criterion in the active learning process, where the most uncertain samples are selected for annotation and included for model retraining. The proposed workflow was tested with three data pairs in which each workflow represents a scenario often encountered in real-world applications, including adapting pretrained models to new images collected with different sensors or to new geographic areas where appearances and types of buildings are very different. Compared to several baselines, including random sampling, temperature scaling (a well-known uncertainty calibration technique), different query strategies, and active domain adaptation methods, the proposed workflow shows that strategically querying a smaller set of samples for labeling achieves comparable or better building extraction performance. The proposed method reduces the number of labeled samples required to achieve sufficient model accuracy, thus significantly reducing hundreds of person-hours for labeled data creation. In addition, we include a few considerations when deploying this workflow in a GPU cluster that can be easily adapted to achieve operational building extraction model retraining.

97 MATHEMATICS AND COMPUTING↗

Improving Building Footprint Extraction Using NAIP and 3DEP Lidar Derived Features with Deep Learning

Accurate building footprint extraction is critical for applications ranging from population estimation to disaster management. Although optical imagery provides detailed spectral information, it often struggles with shadows, occlusions, and background clutter in dense urban environments. Lidar data, by contrast, offer precise elevation and structural attributes but face challenges such as variable point density and noise. This study integrates multispectral imagery from the U.S. Department of Agriculture (USDA) National Agriculture Imagery Program (NAIP) with lidar-derived feature height and intensity from the U.S. Geological Survey (USGS) 3D Elevation Program (3DEP) to improve footprint extraction using a U-Net–based deep learning model. A six-band input stack (RGB, near-infrared, height, intensity) was developed, normalized, and tiled for training and evaluation against Microsoft Global Building Footprints (GBF). Results from the Houston, TX test site show that the six-band model achieved a precision of 0.86, recall of 0.88, F1 score of 0.87, and Intersection-over-Union (IoU) of 0.76, consistently outperforming four-band baselines by reducing false positives while maintaining sensitivity. Predictions on withheld Houston tiles confirmed strong within-region generalization, yielded a precision of 0.78, recall of 0.81, F1 score of 0.79, and IoU of 0.66. Qualitative analysis further revealed limitations stemming from both training label quality and vegetation–building confusion. These findings demonstrate the complementary value of integrating spectral and structural information for robust building footprint extraction and how domain adaptation strategies can be used to enhance cross-regional transferability.

Liu, Jung Kuan [United States Geological Survey (U↗

Efficient Extraction Of Building Elevation Attributes For Flood Risk Management Using Airborne LiDAR Data

In this paper, we address the need for extracting two key building elevation attributes—Lowest Adjacent Grade (LAG) and Highest Adjacent Grade (HAG)—which are crucial for effective flood risk management. Conventional methods, involving onsite surveying or the use of optical imagery-derived building footprints combined with Digital Elevation Models (DEMs), often face misalignment and time discrepancy issues due to varied remote sensing sources. We introduce a new, scalable method that exclusively relies on airborne LiDAR data to overcome these challenges. Our approach employs an object-based ground filtering technique, and the results were evaluated using two different DEMs and building footprint sets. The findings demonstrate that our single-source method, utilizing only airborne LiDAR data, significantly improves the accuracy of LAG and HAG calculations compared to traditional methods that use hand-digitized building footprints. The proposed approach offers a solution for comprehensive flood risk management endeavors.

Song, Hunsoo↗

ORBITaL-Net: A labeled training library for large-scale building feature extraction

Over the course of several years, nearly 1.5 million building outlines have been created from approximately 128,000 training tiles covering roughly 7,000 km 2 of very high-resolution multispectral overhead imagery, primarily dated between 2010 and 2020. This dataset, dubbed the Oak Ridge Building Image and TrAining Label Net (ORBITaL-Net), is designed for machine learning applications and is global in scope, with samples drawn from 72 countries across North America, South America, Africa, Europe, and Asia. ORBITaL-Net captures a great diversity in geographic setting, structural characteristics, land use (urban and rural), terrain, and imagery conditions. While the labeled building outlines are themselves valuable, the dataset’s true strength lies in the pairing of these labels with corresponding reference imagery, which is being released for open source use. Similar to SpaceNet and Replicable AI For Microplanning (ramp), this building outline dataset will allow the larger computer vision community from academia, government, and industry the opportunity to develop robust, scalable, and generalizable geospatial machine learning techniques. Unlike SpaceNet and ramp, which offer high resolution labels and imagery primarily for large urban cities, ORBITaL-Net is not focused on training samples from heavily populated areas but instead aims to capture the innate variability of conditions present in both the physical environment and imagery collections.

Geography↗

Building morphologies of the USA structures database; a gauntlet feature set

In recent years there has been a proliferation of methods and data to extract building footprints from satellite imagery. However there has been very little effort to provide additional insight about these buildings beyond their spatial location and shape. Features derived from their geometries can be used to better characterize these buildings which are critical for further research and development. In this work a set of 65 unique features for every building for more than 131 million buildings covering the US has been developed. This rich feature dataset will enable researchers, policymakers and various agencies to derive additional building characteristics like height, occupancy type, and help to gain valuable and new insights of the built environment.

Environmental sciences↗

Gauntlet

Gauntlet (Geographic Augmentation of Extracted Building Features Tool) generates 65 measures of a building’s morphology. These morphology features can be used for various classification tasks and modeling the built environment.

Hauser, Taylor [Oak Ridge National Laboratory (ORN↗

Leveraging Open-Source Satellite-Derived Building Footprints for Height Inference

At a global scale, cities are growing and characterizing the built environment is essential for deeper understanding of human population patterns, urban development, energy usage, climate change impacts, among others. Buildings are a key component of the built environment and significant progress has been made in recent years to scale building footprint extractions from satellite datum and other remotely sensed products. Billions of building footprints have recently been released by companies such as Microsoft and Google at a global scale. However, research has shown that depending on the methods leveraged to produce a footprint dataset, discrepancies can arise in both the number and shape of footprints produced. Therefore, each footprint dataset should be examined and used on a case-by-case study. In this work, we find through two experiments on Oak Ridge National Laboratory and Microsoft footprints within the same geographic extent that our approach of inferring height from footprint morphology features is source agnostic. Regardless of the differences associated with the methods used to produce a building footprint dataset, our approach of inferring height was able to overcome these discrepancies between the products and generalize, as evidenced by 98% of our results being within 3m of the ground-truthed height. This signifies that our approach can be applied to the billions of open-source footprints which are freely available to infer height, a key building metric. This work impacts the broader domain of urban science in which building height is a key, and limiting factor.

Stipek, Clinton [ORNL] (ORCID:0000000280501096)↗

Forensic Analysis of SOHO Router Binaries

Small Office/Home Office (SOHO) routers are used by millions of consumers across the United States, and are commensurately vulnerable. Forensic analysis of SOHO router firmware helps to understand and mitigate those vulnerabilities. This poster focused particularly on analysis of BusyBox executables, a software suite that provides several Unix utilities in a single file. Three main tools were used to analyze the binaries. BinWalk was used to extract the files, but also to build entropy graphs, extract Linux kernel images, and identify CPU architectures; WiiBin processed the binaries to find endianness, architecture, the percent compressed/encrypted, and compiler data; and @DisCo, a machine learning tool used to determine function similarity in disassembled binaries, analyzed similarities and determined versions of extracted BusyBox files from each router. These tools found that venders from all five routers utilized the same version of the BusyBox software across different firmware updates, demonstrating the importance of constant firmware scrutiny to protect against security vulnerabilities.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Performance evaluation of automated data-driven feature extraction and selection methods for practical and scalable building energy consumption prediction models

Here, this study quantifies the impact of automated feature engineering methods (feature extraction and selection) on the quality and accuracy of machine learning models that predict building energy consumption. The case study compares model performance for three main scenarios: baseline (no feature extraction and selection), feature extraction only, and feature extraction combined with feature selection (filter and/or wrapper methods) for fully trained machine learning models for 200 metered/sub-metered energy measurements across 118 real buildings. For consistency, the same machine learning model architecture (a black box deep learning neural network with probabilistic forecast output) was used for all scenarios. Based on results, all feature engineering methods provided noticeable prediction accuracy improvements (e.g., 29%-68% median prediction improvement) compared to baseline scenarios. However, in this application, feature selection methods provide little practical value due to their limited performance gains and high computational cost. Smarter algorithm development supported by better computational environments will be needed before feature selection methods can reliably and efficiently improve predictive model performance.

97 MATHEMATICS AND COMPUTING↗

Extraction and Analysis of Time Series Data from Building Automation Systems Using Large Language Models

Semantic schemas like Haystack 4, Brick and ASHRAE standard 223 enable the structured, standardized, and machine-readable representation of building data, facilitating interoperability, data integration, and advanced analytics. However, extracting information from these models requires specialized expertise in SPARQL and other programming languages, skills that are not commonly found among building professionals. Recent advancements in Large Language Models (LLMs), such as ChatGPT, enable the construction of queries using natural language, making it easier for individuals to interact with these systems in a manner that resembles everyday speech. However, these methods have not yet been tested on building semantic ontologies. This paper introduces a novel workflow and tool for enabling users to ask questions about a specific building's data, using natural language and receive answers automatically generated by GPT-4o. Our approach integrates semantic ontologies with advanced LLM capabilities to automate three critical steps: (1) generating SPARQL queries to retrieve time series references from ontological models, (2) extracting the corresponding time series data from the Building Automation System, and (3) performing computations and visualizations tailored to the user's query. The proposed method simplifies access to BAS data, allowing both domain experts and non-specialists to conduct sophisticated analyses without needing extensive technical knowledge of semantic web technologies. By demonstrating this pipeline, we facilitate more accessible and scalable data-driven decision-making in building operations and management.

Mulayim, Ozan Baris↗

Extraction and Analysis of Time Series Data from Building Automation Systems Using Large Language Models

Semantic schemas like Haystack 4, Brick and ASHRAE standard 223 enable the structured, standardized, and machine-readable representation of building data, facilitating interoperability, data integration, and advanced analytics. However, extracting information from these models requires specialized expertise in SPARQL and other programming languages, skills that are not commonly found among building professionals. Recent advancements in Large Language Models (LLMs), such as ChatGPT, enable the construction of queries using natural language, making it easier for individuals to interact with these systems in a manner that resembles everyday speech. However, these methods have not yet been tested on building semantic ontologies. This paper introduces a novel workflow and tool for enabling users to ask questions about a specific building's data, using natural language and receive answers automatically generated by GPT-4o. Our approach integrates semantic ontologies with advanced LLM capabilities to automate three critical steps: (1) generating SPARQL queries to retrieve time series references from ontological models, (2) extracting the corresponding time series data from the Building Automation System, and (3) performing computations and visualizations tailored to the user's query. The proposed method simplifies access to BAS data, allowing both domain experts and non-specialists to conduct sophisticated analyses without needing extensive technical knowledge of semantic web technologies. By demonstrating this pipeline, we facilitate more accessible and scalable data-driven decision-making in building operations and management.

Mulayim, Ozan Baris↗

Decoding Ethiopian Abodes: Towards Classifying Buildings by Occupancy Type Using Footprint Morphology

Building occupancy classification plays a crucial role in urban planning, disaster management, and population modeling. Traditional methods often require extensive field surveys or detailed datasets, which can be time-consuming, expensive, and may yield incomplete or erroneous data. In this paper, we present a novel approach for classifying buildings as residential or non-residential using only building footprint data. By extracting geometric shape derivatives that characterize building morphology, we developed a high-accuracy classification model employing a combination of unsupervised and supervised learning methods. We utilized open-source data from Open Street Map, aggregating it to create binary labels for buildings based on their respective human use type. Our approach demonstrates the potential for scalability without the need for additional data sources other than building footprints and labels, offering a more efficient solution for building occupancy classification.

Adams, Daniel↗

Establishing a versatile toolkit of flux enhanced strains and cell extracts for pathway prototyping

Building and optimizing biosynthetic pathways in engineered cells holds promise to address societal needs in energy, materials, and medicine, but it is often time-consuming. Cell-free synthetic biology has emerged as a powerful tool to accelerate design-build-test-learn cycles for pathway engineering with increased tolerance to toxic compounds. However, most cell-free pathway prototyping to date has been performed in extracts from wildtype cells which often do not have sufficient flux towards the pathways of interest, which can be enhanced by engineering. Here, in this study, to address this gap, we create a set of engineered Escherichia coli and Saccharomyces cerevisiae strains rewired via CRISPR-dCas9 to achieve high-flux toward key metabolic precursors; namely, acetyl-CoA, shikimate, triose-phosphate, oxaloacetate, α-ketoglutarate, and glucose-6-phosphate. Cell-free extracts generated from these strains are used for targeted enzyme screening in vitro. As model systems, we assess in vivo and in vitro production of triacetic acid lactone from acetyl-CoA and muconic acid from the shikimate pathway. The need for these platforms is exemplified by the fact that muconic acid cannot be detected in wildtype extracts provided with the same biosynthetic enzymes. We also perform metabolomic comparison to understand biochemical differences between the cellular and cell-free muconic acid synthesis systems (E. coli and S. cerevisiae cells and cell extracts with and without metabolic rewiring). While any given pathway has different interfaces with metabolism, we anticipate that this set of pre-optimized, flux enhanced cell extracts will enable prototyping efforts for new biosynthetic pathways and the discovery of biochemical functions of enzymes.

59 BASIC BIOLOGICAL SCIENCES↗

Greenhouse Gas Emissions and Decarbonization Potential of Global Fired Clay Brick Production

Fired clay bricks (FCBs) are a dominant building material globally due to their low cost and simplicity of production, especially in low- and middle-income countries. With a projected rising housing demand, commensurate growth in brick demand is anticipated, the production of which could result in significant greenhouse gas (GHG) emissions. Robust models are needed to estimate brick demand and emissions to systematically address decarbonization pathways. Few sources report production values; hence, we present two novel proxy models: (i) a consumption prediction model, relying on country-specific clay extraction data, dynamic building stock modeling, and average material intensity use allowing for projections to 2050; and (ii) a GHG emissions model, using literature-based data and production technology-specific inputs. Based on these models, the current global FCB consumption is estimated as 2.18 Gt annually, resulting in approximately 500 million tCO2e (1% of current global GHG emissions). If unaddressed, this fraction could increase to 3.5-5% in 2050 considering a moderate SSP 2-4.5 climate change mitigation scenario. Consequently, we explored three potential decarbonization pathways: (i) improving energy efficiency; (ii) shifting production to best practices; and (iii) replacing half of FCB demand with hollow concrete blocks, resulting in 27%, 49%, and 51% reduction in GHG emissions, respectively.

Olsson, Josefine A↗

Improving and Automating Building Model Data Exchange

There are many instances throughout a project’s lifecycle where there arises a need for quick and accurate risk assessment of building designs. For example, an unexpected design change during construction may necessitate structural engineers to perform a seismic risk assessment on analytical models of the updated building design using high fidelity structural analysis software, such as ANSYS or Abaqus. However, the efficiency of such workflows often depends upon the interoperability of architectural design software and structural analysis software. When the quality of this interoperability is lacking or even non-existent, the efficiency of virtual engineering workflows is hampered, which increases project costs. A McGraw Hill industry survey of professional users of Building Information Modeling (BIM) technologies found that there is high demand for BIM interoperability for structural analysis, but that the value/difficulty ratio is currently too low for practical use. There have been efforts by the academic community to facilitate model data exchange between the architectural design and structural analysis domains, but such solutions have not been widely adopted by industry, face technical challenges, and oftentimes are limited in applicability for users of various BIM software. Therefore, INL is developing capabilities to improve, automate, and generalize model data exchange between architectural BIM software (e.g., Revit) and structural analysis software (e.g., SAP2000, ANSYS). The goal is to help expedite and automate as much of the pre-processing step for creating analytical models in finite element analysis software as reasonably as possible. Such a "BIM-to-FEA" conversion tool should provide direct benefit to end-users through accuracy, automation, quick turn-around, and wide applicability. To generalize the application of this BIM-to-FEA conversion tool and increase its useability among the many different commercial BIM software currently used by industry, the program is being developed with the concept of openBIM. OpenBIM is the application of non-proprietary, open data standards that allow for BIM model data exchange in a format that is accessible, retainable, and useable for all users. The most widely used open, non-proprietary data exchange format for BIM is the Industry Foundation Classes (IFC) schema. IFC is developed by buildingSMART international and is ISO certified (ISO 16739-1:2018). The BIM-to-FEA conversion tool is being developed for compatibility with typical commercial building designs of steel framed structures. The tool is currently capable of importing architectural BIM data of framed building structures, recognizing and extracting the aspects of the model that are required for structural analysis, adjusting the connectivity of frame members, and finally exporting to an analytical model stored in the IFC format. The exported IFC analytical model can then be imported into various openBIM compliant software, such as SAP2000. Such capabilities have already been tested on commercial software, as shown above, and continue to be improved. Work is underway to test the conversion on various commercial BIM software, develop a user-friendly interface, incorporate the program into the broader DeepLynx data warehouse project being developed by INL, and to eventually open-source the tool for the benefit of the community. Future development of the tool envisions the ability for efficient iterative risk assessment of generative building designs, all within a workflow utilizing open-source tools. One such open-source tool will be MOOSE, an advanced finite element analysis tool developed at INL. The conversion tool will also branch out from typical commercial building designs and will aim to incorporate nuclear construction. The aim will be to convert both structural and non-structural components of nuclear facilities, such as curved concrete containment structures and piping systems, respectively.

97 MATHEMATICS AND COMPUTING↗

Toward Lithium Recovery Using Modular (and Membraneless) Phase Separation and Extraction (MPSE) Technology with Ionic Liquid (IL) Solvents: Effect of Coatings

In this work, a single-channel slope-plate Modular and Membrane-less Phase Separation and Extraction (MPSE) device was fabricated using 3D printing. The slope-plate was designed with an insertable glass slide to enable surface modification with coatings. Self-assembled monolayers (SAMs) and perfluoropolyether (Zdol) coatings were applied to tailor surface wettability and enhance phase separation. Using a model biphasic system of water and hexadecane, both coatings significantly improved the separation efficiency compared to the uncoated device by promoting selective wetting. Building on these results, lithium extraction from a simulated saline solution was investigated using 1-ethyl-3-methylimidazolium bis(trifluoromethylsulfonyl)imide ([EMIM][NTf₂]) as the ionic liquid (IL) extractant. The IL effectively extracted Li⁺ from the aqueous phase in bulk extraction, demonstrating its potential for lithium recovery. However, the low interfacial tension between the IL and aqueous phases posed challenges for phase separation. The application of SAM and Zdol coatings effectively mitigated this issue. Overall, integrating tailored surface chemistry with the slope-plate MPSE design shows great promise as an efficient and scalable platform for studying liquid–liquid separation and optimizing ionic liquid–based extraction processes for lithium recovery from saline sources.

3D printing↗

Microstructure Validation of Graph Theory Model-Derived Cooling Rates in the Wire Arc Additive Manufacturing of ER70S-6 Steel

Wire arc additive manufacturing (WAAM) enables high-rate fabrication of large metallic components, but spatial variations in thermal history can lead to microstructural heterogeneity that requires efficient process models to evaluate. This study evaluates whether cooling rates extracted from a graph theory model (GTM)-based thermal simulation are consistent with the microstructural evolution observed in an ER70S-6 WAAM wall. Thermal histories from the model were analyzed at selected build heights, and cooling rates were extracted from the final thermal excursion through the austenite phase field. Microstructures at corresponding locations were characterized using electron backscatter diffraction (EBSD) to quantify grain size distributions, and pearlite interlamellar spacing was used as an additional indicator of cooling behavior. The modeled cooling rates were highest near the substrate and generally decreased with build height, consistent with the observed reduction in the fine grain fraction and the progressive shift in the grain size distribution as build height increased. Pearlite spacing trends also supported the modeled cooling rate variation. These results indicate that GTM-derived thermal histories can be post-processed into metallurgically meaningful cooling rate estimates for WAAM steel builds and linked to dataset specific empirical grain size distribution relationships for process–thermal history–microstructure assessment.

36 MATERIALS SCIENCE↗