An aggregate learning approach for interpretable semi-supervised population prediction and disaggregation using ancillary data

doi:10.48550/arXiv.1907.00270

An aggregate learning approach for interpretable semi-supervised population prediction and disaggregation using ancillary data

Census data provide detailed information about population characteristics at a coarse resolution. Nevertheless, fine-grained, high-resolution mappings of population counts are increasingly needed to characterize population dynamics and to assess the consequences of climate shocks, natural disasters, investments in infrastructure, development policies, etc. Dissagregating these census is a complex machine learning, and multiple solutions have been proposed in past research. We propose in this paper to view the problem in the context of the aggregate learning paradigm, where the output value for all training points is not known, but where it is only known for aggregates of the points (i.e. in this context, for regions of pixels where a census is available). We demonstrate with a very simple and interpretable model that this method is on par, and even outperforms on some metrics, the state-of-the-art, despite its simplicity.

Publication:

arXiv e-prints

Pub Date:

June 2019

DOI:

10.48550/arXiv.1907.00270

arXiv:

arXiv:1907.00270

Bibcode:

2019arXiv190700270D

Keywords:

Computer Science - Machine Learning;
Statistics - Machine Learning

E-Print:

Accepted at ECML-PKDD 2019 Data on Zenodo: https://zenodo.org/record/3260713

NASA/ADS

An aggregate learning approach for interpretable semi-supervised population prediction and disaggregation using ancillary data

Abstract