The Benefits of Balance: From Information Projections to Variance Reduction

doi:10.48550/arXiv.2408.15065

The Benefits of Balance: From Information Projections to Variance Reduction

Data balancing across multiple modalities/sources appears in various forms in several foundation models (e.g., CLIP and DINO) achieving universal representation learning. We show that this iterative algorithm, usually used to avoid representation collapse, enjoys an unsuspected benefit: reducing the variance of estimators that are functionals of the empirical distribution over these sources. We provide non-asymptotic bounds quantifying this variance reduction effect and relate them to the eigendecays of appropriately defined Markov operators. We explain how various forms of data balancing in contrastive multimodal learning and self-supervised clustering can be interpreted as instances of this variance reduction scheme.

Publication:

arXiv e-prints

Pub Date:

August 2024

DOI:

10.48550/arXiv.2408.15065

arXiv:

arXiv:2408.15065

Bibcode:

2024arXiv240815065L

Keywords:

Statistics - Machine Learning;
Computer Science - Machine Learning;
Mathematics - Statistics Theory

NASA/ADS

The Benefits of Balance: From Information Projections to Variance Reduction

Abstract