MEGA Hub

FedCC: Towards Addressing Label Distribution Skews in Distillation-Based Federated Learning

Authors

Do you know Wenxuan Ye?You can claim authorship or link another user.Do you know Onur Ayan?You can claim authorship or link another user.Do you know Xueli An?You can claim authorship or link another user.Do you know Georg Carle?You can claim authorship or link another user.

Abstract

Federated Learning (FL) enables distributed clients to collaboratively train models without sharing raw data, making it promising for leveraging massive devices in communication networks. In distillation-based FL, each client applies its local model on an unlabeled public dataset, and shares only prediction results with the server. While heterogeneous local data introduces label distribution skew, thus biasing client models toward majority classes and leading to potentially inaccurate predictions. The lack of ground-truth labels in the public dataset hampers the server's ability to calibrate predictions, which ultimately degrades overall performance. To address this, we propose FedCC, a simple and effective algorithm for mitigating client misclassification. Instead of being forced to classify and risking error propagation, clients are allowed to tag ambiguous samples as 'unknown'. This additional class, together with calibrated pseudo-labels on the public data, balances confidence in majority classes against uncertainty in under-represented ones. Extensive experiments demonstrate that FedCC significantly outperforms existing methods, especially under severe label skew. In the extreme scenario where each client holds samples from only one of ten classes, FedCC achieves 67.3% accuracy, while baselines collapse to near-random results.

Community

00

Publication notes

Author note
Accepted to Globecom 2026