CRISP-DM-Based Data-Driven Approach for Building Energy Prediction Utilizing Indoor and Environmental Factors
Notice bibliographique
Résumé
The significant energy consumption associated with the built environment demands comprehensive energy prediction modelling. Leveraging their ability to capture intricate patterns without extensive domain knowledge, supervised data-driven approaches present a marked advantage in adaptability over traditional physical-based building energy models. This study employs various machine learning models to predict energy consumption for an office building in Berkeley, California. To enhance the accuracy of these predictions, different feature selection techniques, including principal component analysis (PCA), decision tree regression (DTR), and Pearson correlation analysis, were adopted to identify key attributes of energy consumption and address collinearity. The analyses yielded nine influential attributes: heating, ventilation, and air conditioning (HVAC) system operating parameters, indoor and outdoor environmental parameters, and occupancy. To overcome missing occupancy data in the datasets, we investigated the possibility of occupancy-based Wi-Fi prediction using different machine learning algorithms. The results of the occupancy prediction modelling indicate that Wi-Fi can be used with acceptable accuracy in predicting occupancy count, which can be leveraged to analyze occupant comfort and enhance the accuracy of building energy models. Six machine learning models were tested for energy prediction using two different datasets: one before and one after occupancy prediction. Using a 10-fold cross-validation with an 8:2 training-to-testing ratio, the Random Forest algorithm emerged superior, exhibiting the highest R2 value of 0.92 and the lowest RMSE of 3.78 when occupancy data were included. Additionally, an error propagation analysis was conducted to assess the impact of the occupancy-based Wi-Fi prediction model’s error on the energy prediction model. The results indicated that Wi-Fi-based occupancy prediction can improve the data inputs for building energy models, leading to more accurate energy consumption predictions. The findings underscore the potential of integrating the developed energy prediction models with fault detection systems, model predictive controllers, and energy load shape analysis, ultimately enhancing energy management practices.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction distillée sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.
Scores Codex et Gemma par catégorie
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,000 | 0,000 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,000 |
| Méta-épidémiologie (sens large) | 0,000 | 0,000 |
| Bibliométrie | 0,000 | 0,000 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,000 | 0,000 |
| Science ouverte | 0,000 | 0,000 |
| Intégrité de la recherche | 0,000 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,000 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule tête enseignante, pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».