
arXiv: 2210.14298
Archetypal analysis is an unsupervised machine learning method that summarizes data using a convex polytope. In its original formulation, for fixed k, the method finds a convex polytope with k vertices, called archetype points, such that the polytope is contained in the convex hull of the data and the mean squared Euclidean distance between the data and the polytope is minimal. In the present work, we consider an alternative formulation of archetypal analysis based on the Wasserstein metric, which we call Wasserstein archetypal analysis (WAA). In one dimension, there exists a unique solution of WAA and, in two dimensions, we prove existence of a solution, as long as the data distribution is absolutely continuous with respect to Lebesgue measure. We discuss obstacles to extending our result to higher dimensions and general data distributions. We then introduce an appropriate regularization of the problem, via a Renyi entropy, which allows us to obtain existence of solutions of the regularized problem for general data distributions, in arbitrary dimensions. We prove a consistency result for the regularized problem, ensuring that if the data are iid samples from a probability measure, then as the number of samples is increased, a subsequence of the archetype points converges to the archetype points for the limiting data distribution, almost surely. Finally, we develop and implement a gradient-based computational approach for the two-dimensional problem, based on the semi-discrete formulation of the Wasserstein metric. Our analysis is supported by detailed computational experiments.
FOS: Computer and information sciences, Numerical optimization and variational techniques, Computer Science - Machine Learning, Optimal transportation, 62H12, 62H30, 65K10, 49Q22, Wasserstein metric, Estimation in multivariate analysis, Probability (math.PR), Machine Learning (stat.ML), unsupervised learning, Machine Learning (cs.LG), Density estimation, optimal transport, Statistics - Machine Learning, Optimization and Control (math.OC), FOS: Mathematics, archetypal analysis, multivariate data summarization, Mathematics - Optimization and Control, Mathematics - Probability
FOS: Computer and information sciences, Numerical optimization and variational techniques, Computer Science - Machine Learning, Optimal transportation, 62H12, 62H30, 65K10, 49Q22, Wasserstein metric, Estimation in multivariate analysis, Probability (math.PR), Machine Learning (stat.ML), unsupervised learning, Machine Learning (cs.LG), Density estimation, optimal transport, Statistics - Machine Learning, Optimization and Control (math.OC), FOS: Mathematics, archetypal analysis, multivariate data summarization, Mathematics - Optimization and Control, Mathematics - Probability
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 0 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Average | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Average |
