Federated Learning for Big Data: A Survey on Opportunities, Applications, and Future Directions

Gadekallu, Thippa Reddy; Pham, Quoc-Viet; Huynh-The, Thien; Feng, Hailin; Fang, Kai; Pandya, Sharnil; Liyanage, Madhusanka; Wang, Wei; Nguyen, Thanh Thi

Computer Science > Machine Learning

arXiv:2110.04160 (cs)

[Submitted on 8 Oct 2021 (v1), last revised 7 Jul 2025 (this version, v3)]

Title:Federated Learning for Big Data: A Survey on Opportunities, Applications, and Future Directions

Authors:Thippa Reddy Gadekallu, Quoc-Viet Pham, Thien Huynh-The, Hailin Feng, Kai Fang, Sharnil Pandya, Madhusanka Liyanage, Wei Wang, Thanh Thi Nguyen

View PDF

Abstract:In the recent years, generation of data have escalated to extensive dimensions and big data has emerged as a propelling force in the development of various machine learning advances and internet-of-things (IoT) devices. In this regard, the analytical and learning tools that transport data from several sources to a central cloud for its processing, training, and storage enable realization of the potential of big data. Nevertheless, since the data may contain sensitive information like banking account information, government information, and personal information, these traditional techniques often raise serious privacy concerns. To overcome such challenges, Federated Learning (FL) emerges as a sub-field of machine learning that focuses on scenarios where several entities (commonly termed as clients) work together to train a model while maintaining the decentralisation of their data. Although enormous efforts have been channelized for such studies, there still exists a gap in the literature wherein an extensive review of FL in the realm of big data services remains unexplored. The present paper thus emphasizes on the use of FL in handling big data and related services which encompasses comprehensive review of the potential of FL in big data acquisition, storage, big data analytics and further privacy preservation. Subsequently, the potential of FL in big data applications, such as smart city, smart healthcare, smart transportation, smart grid, and social media are also explored. The paper also highlights various projects pertaining to FL-big data and discusses the associated challenges related to such implementations. This acts as a direction of further research encouraging the development of plausible solutions.

Comments:	Submitted for peer review in a journal
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2110.04160 [cs.LG]
	(or arXiv:2110.04160v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2110.04160

Submission history

From: Gadekallu Thippa Reddy [view email]
[v1] Fri, 8 Oct 2021 14:36:43 UTC (1,375 KB)
[v2] Sun, 17 Oct 2021 15:55:33 UTC (1,809 KB)
[v3] Mon, 7 Jul 2025 15:45:16 UTC (2,824 KB)

Computer Science > Machine Learning

Title:Federated Learning for Big Data: A Survey on Opportunities, Applications, and Future Directions

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Federated Learning for Big Data: A Survey on Opportunities, Applications, and Future Directions

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators