EconPapers    
Economics at your fingertips  
 

Evaluating missing data handling methods for developing building energy benchmarking models

Kyungjae Lee, Hyunwoo Lim, Jeongyun Hwang and Doyeon Lee

Energy, 2024, vol. 308, issue C

Abstract: This study explored methods for handling missing data in the development of machine learning-based energy benchmarking models, assessing their training time, performance, and variance. Unlike the common assumption of missing completely at random, this study adopted a missing at random (MAR) perspective, which is more appropriate for building data. We compared the inherent missing data handling method of extreme gradient boosting (XGBoost) with the Median, k-nearest neighbors (KNN), and classification and regression trees (CART) methods, alongside Shapley additive explanation (SHAP) method for model interpretability. The findings indicate that, despite its computational demands, the CART method most accurately mirrors the original data distribution, thereby enhancing model performance and stability. The KNN method is effective, while the XGBoost method is viable under computational time constraints. This work highlights the importance of reliable test data for performing accurate evaluations of imputation methods. These results offer guidelines for the selection of imputation methods in model development, contributing to the improved accuracy of energy benchmarking models. The MAR-based approach for missing data analysis holds promise for future research on building energy data, providing crucial insights for accurate energy benchmark model performance assessments.

Keywords: Building energy benchmarking model; Building energy performance; Building energy data; Missing value imputation; Machine learning (search for similar items in EconPapers)
Date: 2024
References: View references in EconPapers View complete reference list from CitEc
Citations:

Downloads: (external link)
http://www.sciencedirect.com/science/article/pii/S0360544224027531
Full text for ScienceDirect subscribers only

Related works:
This item may be available elsewhere in EconPapers: Search for items with the same title.

Export reference: BibTeX RIS (EndNote, ProCite, RefMan) HTML/Text

Persistent link: https://EconPapers.repec.org/RePEc:eee:energy:v:308:y:2024:i:c:s0360544224027531

DOI: 10.1016/j.energy.2024.132979

Access Statistics for this article

Energy is currently edited by Henrik Lund and Mark J. Kaiser

More articles in Energy from Elsevier
Bibliographic data for series maintained by Catherine Liu ().

 
Page updated 2025-03-19
Handle: RePEc:eee:energy:v:308:y:2024:i:c:s0360544224027531