References & Citations
Statistics > Methodology
Title: Wasserstein Regression
(Submitted on 17 Jun 2020 (this version), latest version 6 Jul 2021 (v2))
Abstract: The analysis of samples of random objects that do not lie in a vector space has found increasing attention in statistics in recent years. An important class of such object data is univariate probability measures defined on the real line. Adopting the Wasserstein metric, we develop a class of regression models for such data, where random distributions serve as predictors and the responses are either also distributions or scalars. To define this regression model, we utilize the geometry of tangent bundles of the metric space of random measures with the Wasserstein metric. The proposed distribution-to-distribution regression model provides an extension of multivariate linear regression for Euclidean data and function-to-function regression for Hilbert space valued data in functional data analysis. In simulations, it performs better than an alternative approach where one first transforms the distributions to functions in a Hilbert space and then applies traditional functional regression. We derive asymptotic rates of convergence for the estimator of the regression coefficient function and for predicted distributions and also study an extension to autoregressive models for distribution-valued time series. The proposed methods are illustrated with data on human mortality and distributions of house prices.
Submission history
From: Yaqing Chen [view email][v1] Wed, 17 Jun 2020 05:12:59 GMT (449kb,D)
[v2] Tue, 6 Jul 2021 01:24:00 GMT (236kb,D)
Link back to: arXiv, form interface, contact.