Score: 0

Federated Learning With L0 Constraint Via Probabilistic Gates For Sparsity

Published: December 28, 2025 | arXiv ID: 2512.23071v1

By: Krishna Harsha Kovelakuntla Huthasana, Alireza Olama, Andreas Lundell

Potential Business Impact:

Makes AI learn better with less data.

Business Areas:

Machine Learning Artificial Intelligence, Data and Analytics, Software

Federated Learning (FL) is a distributed machine learning setting that requires multiple clients to collaborate on training a model while maintaining data privacy. The unaddressed inherent sparsity in data and models often results in overly dense models and poor generalizability under data and client participation heterogeneity. We propose FL with an L0 constraint on the density of non-zero parameters, achieved through a reparameterization using probabilistic gates and their continuous relaxation: originally proposed for sparsity in centralized machine learning. We show that the objective for L0 constrained stochastic minimization naturally arises from an entropy maximization problem of the stochastic gates and propose an algorithm based on federated stochastic gradient descent for distributed learning. We demonstrate that the target density (rho) of parameters can be achieved in FL, under data and client participation heterogeneity, with minimal loss in statistical performance for linear and non-linear models: Linear regression (LR), Logistic regression (LG), Softmax multi-class classification (MC), Multi-label classification with logistic units (MLC), Convolution Neural Network (CNN) for multi-class classification (MC). We compare the results with a magnitude pruning-based thresholding algorithm for sparsity in FL. Experiments on synthetic data with target density down to rho = 0.05 and publicly available RCV1, MNIST, and EMNIST datasets with target density down to rho = 0.005 demonstrate that our approach is communication-efficient and consistently better in statistical performance.

Lightweight Federated Learning in Mobile Edge Computing with Statistical and Device Heterogeneity Awareness

Systems and Control

Makes phones learn together without sharing private data.

29 Oct 2025 1

90%

Communication-Efficient Zero-Order and First-Order Federated Learning Methods over Wireless Networks

Machine Learning (CS)

Makes phones learn together without sharing secrets.

11 Aug 2025 0

90%

Minimizing Layerwise Activation Norm Improves Generalization in Federated Learning

Machine Learning (CS)

Makes AI learn better without sharing private data.

9 Dec 2025 3

View PDF Login to Bookmark

Country of Origin

🇫🇮 Finland

Page Count

19 pages

Federated Learning With L0 Constraint Via Probabilistic Gates For Sparsity

Makes AI learn better with less data.

Technical Abstract

Lightweight Federated Learning in Mobile Edge Computing with Statistical and Device Heterogeneity Awareness

Communication-Efficient Zero-Order and First-Order Federated Learning Methods over Wireless Networks

Minimizing Layerwise Activation Norm Improves Generalization in Federated Learning