Dimension estimation using random connection models

Open Access
Authors
Publication date 11-2017
Journal Journal of Machine Learning Research
Article number 138
Volume | Issue number 18
Number of pages 35
Organisations
  • Faculty of Science (FNWI) - Korteweg-de Vries Institute for Mathematics (KdVI)
Abstract
Information about intrinsic dimension is crucial to perform dimensionality reduction, compress information, design efficient algorithms, and do statistical adaptation. In this paper we propose an estimator for the intrinsic dimension of a data set. The estimator is based on binary neighbourhood information about the observations in the form of two adjacency matrices, and does not require any explicit distance information. The underlying graph is modelled according to a subset of a specific random connection model, sometimes referred to as the Poisson blob model. Computationally the estimator scales like n log n, and we specify its asymptotic distribution and rate of convergence. A simulation study on both real and simulated data shows that our approach compares favourably with some competing methods from the literature, including approaches that rely on distance information.
Document type Article
Language English
Published at https://www.jmlr.org/papers/volume18/16-232/16-232.pdf
Other links https://www.scopus.com/pages/publications/85040725812
Downloads
Permalink to this page
Back