Search NASA⌕ Search

Engineering topics

Lebakula, Viswadeep

Publications and source records attributed to Lebakula, Viswadeep.

LandScan Global 30 Arcsecond Annual Global Gridded Population Datasets from 2000 to 2022

Abstract Oak Ridge National Laboratory (ORNL) annually develops the LandScan Global (LSG) dataset, a 30 arcsecond global gridded population dataset representing global ambient human population distribution. This multivariable dasymetric model disaggregates census counts within administrative boundaries using ancillary data. Each country’s distribution reflects cultural and socioeconomic patterns; manual validations yield a unique global dataset for assessing populations at risk. For over two decades, LSG has been a standard for estimating populations at risk, aiding U.S. federal government, academia and humanitarian organizations. During disasters such as the 2004 Indian Ocean tsunami and the 2010 Haiti earthquake and geopolitical crises such as the Syrian civil war and the 2022 Russian invasion of Ukraine, LSG supported scientific and operational communities in emergency response and recovery. In 2022, LSG datasets from 2000 onward were made publicly available through ORNL’s LandScan Portal. This data descriptor details our methodology and the application of geospatial science and machine learning to geographic and demographic data, highlighting uses in urban resiliency, emergency management, disaster response, and human health and security.

Science & Technology - Other Topics↗

A structural equation modeling approach to leveraging the power of extant sentiment analysis tools

Machine-derived sentiment analysis has become a pervasive and useful tool to address a wide array of issues in natural language processing. Leading technology companies such as Google now provide sentiment analysis tools (SATs) as readily accessible online products. Academic researchers develop and make available SATs to support the research enterprise. One of the major challenges with SATs is the inconsistencies in results among the various SATs. Consequently, the selection of a SAT for a specific purpose may significantly impact the application. This study addresses the foregoing problem by utilizing structural equation modeling to merge the outputs of SATs to develop a combined sentiment metric without the need for a labeled training dataset. This method is applicable to a wide range of text-based problems, is data-driven, and replicable. It was tested using three publicly available datasets and compared against seven different SATs. The results indicate that as a continous measure, the proposed method outperformed other SATs in the movie reviews and SemEval datasets, and achieved a tie for first place with IBM Watson on the Sentiment 140 dataset. Also, compared to the published major alternatives, the arithmetic mean solution, this approach performed better across these three datasets.

97 MATHEMATICS AND COMPUTING↗

LandScan Global 2023: Silver Edition

For a quarter of a century, the LandScan Global (LSG) project has annually released a global, high-resolution gridded population dataset representing the ambient or unwarned population at a 30 arcsecond resolution. LSG supports a range of applications such as emergency management, disaster response, and human health and security for understanding populations at risk. The 2023 release of LSG, the LandScan Silver Edition, represents a major methodological leap forward while also leveraging previous knowledge—the previous year was the baseline for the current annual update carrying forward valuable knowledge of the built environment for the past quarter century—to train the machine learning models. Compared with annual releases over the past 24years, multiple advancements were made to different aspects of the methodology to achieve reproducibility, transparency, and consistent global propagation of solutions to modeling or population distribution issues identified during the review process. These novel changes include incorporation of the latest available geospatial inputs across the globe, machine learning models instead of manual modifications, population feature importance analysis, open-source solutions vs. proprietary software, generation of multiple global versions, analytic validations, and human-in-the-loop revisions to produce the final version. Additionally, algorithms—such as anomaly detection—were introduced to quickly identify areas of focus to develop a new and robust systematic review. Significant changes in modeled population distributions were observed between the 2022 and 2023 releases, largely attributable to improvements in data and methods and discussed thoroughly within this report. In summation, the LandScan Silver Edition leverages the best of the past quarter century of LSG legacy knowledge and continues a tradition of applying cutting-edge enhancements to serve as a new benchmark for accurate, actionable gridded population data

Lebakula, Viswadeep↗

Synchronization and Rebound Effects in Residential Loads

Increasing fuel prices and capacity investment deferral place an increasing demand for peak reduction from distribution level systems. Residential and commercial devices, such as HVAC systems and water heaters, are increasingly involved in load control programs, and their use may generate synchronization and rebound effects, such as artificial peaks caused by device optimization. While there have been concerns over device synchronization, few studies quantify the extent of this effect with numerical values. In this study, we attempt to investigate whether control efforts result in device synchronization or rebound effects. We focus on three clustering methods – Ward’s clustering, Euclidean K-means, and Density-based spatial clustering of applications with noise – to evaluate the extent of synchronization of a fleet of water heaters and HVAC systems in Atlanta, Georgia. Our findings show that synchronization and rebound effects are present in the neighborhood’s water heaters, but none were found in the HVAC systems. Further, high usage water heaters are more susceptible to synchronization and rebound effects.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Explaining Health Risk Behaviors in the U.S. with Social Deprivation at Local and Regional Levels

Health risk behaviors are precursors to many chronic health outcomes, and hence, they pose a challenge to public health. Social deprivation undoubtedly creates circumstances that limit access to healthy habits. Moreover, broad regional effects (weather patterns, political ideology, social norms), and local characteristics (cultural notions and barriers, urban places) also influence lifestyle choices and must be accounted for to truly understand the impact of social deprivation on risky behaviors. This research fills the knowledge gap in epidemiological modeling of health risk behaviors by leveraging machine learning to find associations between social deprivation and health risk behaviors, when adjusted by regional and local effects. Four health risk behaviors, namely, binge drinking, smoking, lack of sleep, and lack of physical activity from the CDC PLACES project are considered in a single framework to understand and compare the interplay between local/regional characteristics and seven measures of social deprivation. Our results indicate that local and/or regional factors rise to the top for three out of four risk behaviors (binge drinking, smoking and lack of sleep) out-competing social deprivation measures. Un-entangling the geographical effects reveals that poverty, educational attainment and non-employment are the three deprivation measures most significantly associated with all four health risk factors. The research thus indicates that public health policies to promote healthy lifestyle behaviors must seek to remedy social deprivation, but using socially and culturally sensitive interventions.

Gokhale, Swapna↗

Autonomous Anomaly Detection for MPC Forecasts of HVAC Systems in Residential Communities

The use of residential heating, ventilation, and air conditioning (HVAC) to shift peak demand or provide ancillary services is a potential solution in the presence of older grids and distributed renewables. However, to ensure the efficient use of devices, utilities need to accurately forecast the load and adopt error correction schemes when necessary. While significant theoretical research exists in the area of predictive control of HVAC, little experimental evidence exists. The lack of experimental data in turn causes researchers to be unprepared for unsystematic errors which emerge due to the higher complexity of the data generating process. This study offers an anomaly detection methodology that uses unsupervised machine learning algorithms to detect and isolate these errors with different forecast error ranges. The results of anomaly detection procedure can then be used for error correction and would eventually help develop better predictive controllers. The methodology is tested using real world data from a smart neighborhood that currently operates in Atlanta. GA.

Lebakula, Viswadeep↗

Structural Differences between Morning and Evening Peak in Optimized Water Heaters

Peak reduction is an important concern that can help reduce the growing stress on distribution grids and allow to defer investments in new capacity. Water heaters represent a convenient way of reducing peak, depending on controllability of devices. But while controlling water heaters does allow to shift peak, it also results in rebound effects, which require additional understanding before water heater fleets can be used on a large scale. We attempt to investigate the nature of peak behaviors of water heaters and demonstrate that water heaters are not homogenous in their behavior. Depending on the overall intensity of the use of water, part of the population has higher rebound effect, while part of the population has little or no rebound effect. Even though we do not have sufficient data to statistically evaluate our findings, we use a sample of 42 water heaters in a connected neighborhood to provide an early attempt at discovering and reporting this diversity.

Tsybina, Eve↗

Empirically Categorizing the Built Environment in Relation to Height

Buildings are a core component of the urban environment and affect human populations, energy usage, city development, city planning, and urban heat islands. Buildings span an enormous range of sizes, from a 2m tall shelter to the Burj Khalifa; and at the same time there are widely recognized categories of similar buildings, with homes, office buildings, or skyscrapers as some examples. Currently, there is no consistent method to quantitatively determine how a building should be categorized by its height, or how many categories there should be within the built environment. Additionally, these categories vary spatially, leading to multiple definitions at local scales of what it means to be a tall, medium, or short building. Here, we find across 17.59 million buildings in the United States, Germany, and Japan, that applying a K-nearest neighbor approach to quantitatively bin the built environment outperforms the current state-of-the-art, subjective domain knowledge. This was evidenced as our method of leveraging a K-nearest neighbor improved upon the existing approach of using domain knowledge by 10% with respect to precision, recall, F1-score and accuracy. Our results showcase the finding that it is possible to generate a global and consistent approach to categorizing the built environment in relation to height. This is significant in that there is now a quantitative way to categorize the built environment based on building height at a global scale, allowing researchers a consistent platform for comparison and collaboration across various applications.

Stipek, Clinton↗