Transfer Learning on Transformers for Building Energy Consumption Forecasting -- A Comparative Study

Spencer, Robert; Ranathunga, Surangika; Boulic, Mikael; van Heerden, Andries; Susnjak, Teo

Computer Science > Machine Learning

arXiv:2410.14107 (cs)

[Submitted on 18 Oct 2024 (v1), last revised 21 Nov 2024 (this version, v3)]

Title:Transfer Learning on Transformers for Building Energy Consumption Forecasting -- A Comparative Study

Authors:Robert Spencer, Surangika Ranathunga, Mikael Boulic, Andries van Heerden, Teo Susnjak

View PDF

Abstract:This study investigates the application of Transfer Learning (TL) on Transformer architectures to enhance building energy consumption forecasting. Transformers are a relatively new deep learning architecture, which has served as the foundation for groundbreaking technologies such as ChatGPT. While TL has been studied in the past, prior studies considered either one data-centric TL strategy or used older deep learning models such as Recurrent Neural Networks or Convolutional Neural Networks. Here, we carry out an extensive empirical study on six different data-centric TL strategies and analyse their performance under varying feature spaces. In addition to the vanilla Transformer architecture, we also experiment with Informer and PatchTST, specifically designed for time series forecasting. We use 16 datasets from the Building Data Genome Project 2 to create building energy consumption forecasting models. Experimental results reveal that while TL is generally beneficial, especially when the target domain has no data, careful selection of the exact TL strategy should be made to gain the maximum benefit. This decision largely depends on the feature space properties such as the recorded weather features. We also note that PatchTST outperforms the other two Transformer variants (vanilla Transformer and Informer). Our findings advance the building energy consumption forecasting using advanced approaches like TL and Transformer architectures.

Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2410.14107 [cs.LG]
	(or arXiv:2410.14107v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2410.14107

Submission history

From: Surangika Ranathunga [view email]
[v1] Fri, 18 Oct 2024 01:26:04 UTC (1,531 KB)
[v2] Tue, 19 Nov 2024 22:19:12 UTC (1,281 KB)
[v3] Thu, 21 Nov 2024 05:19:42 UTC (1,281 KB)

Computer Science > Machine Learning

Title:Transfer Learning on Transformers for Building Energy Consumption Forecasting -- A Comparative Study

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Transfer Learning on Transformers for Building Energy Consumption Forecasting -- A Comparative Study

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators