Leveraging Unlabeled Data in Federated Learning: A Review
- 1. Marmara Univ, Fac Engn, Dept Comp Engn, TR-34854 Istanbul, Turkiye
- 2. Ozyegin Univ, Dept Elect & Elect Engn, TR-34794 Istanbul, Turkiye
Description
In centralized machine learning, both the data and model to be trained reside on a single server, which may cause problems regarding data privacy as sensitive or personal data need to be transferred from clients to the server. Federated learning has been proposed to provide a solution to this problem by allowing the training of a model without the data leaving the clients. This training takes place between a coordinating server and the clients by continuously exchanging the model parameters instead of exchanging data. In real-life applications, the data on some of the clients or the server may be partially labeled or completely unlabeled, which poses a severe challenge to federated learning. In this paper, we present a survey of recently proposed methods that leverage unlabeled data in a federated learning setting to improve model performance. We also present a novel taxonomy of the methods that leverage unlabeled data based on whether the unlabeled data is assigned a pseudo-label during the process or not. We summarize the datasets, main data modalities, and application areas of federated learning with unlabeled data methods in the literature and highlight future research directions. We believe that this survey will be a useful guide for researchers planning to work on federated learning with partially labeled data.
Files
bib-9de84773-172d-41ea-9582-3f2d0affcb49.txt
Files
(139 Bytes)
| Name | Size | Download all |
|---|---|---|
|
md5:3d93cd77d39e33ab187f0c14facdd154
|
139 Bytes | Preview Download |