
arXiv: 1806.09048
We show how the problem of estimating conditional Kendall's tau can be rewritten as a classification task. Conditional Kendall's tau is a conditional dependence parameter that is a characteristic of a given pair of random variables. The goal is to predict whether the pair is concordant (value of $1$) or discordant (value of $-1$) conditionally on some covariates. We prove the consistency and the asymptotic normality of a family of penalized approximate maximum likelihood estimators, including the equivalent of the logit and probit regressions in our framework. Then, we detail specific algorithms adapting usual machine learning techniques, including nearest neighbors, decision trees, random forests and neural networks, to the setting of the estimation of conditional Kendall's tau. Finite sample properties of these estimators and their sensitivities to each component of the data-generating process are assessed in a simulation study. Finally, we apply all these estimators to a dataset of European stock indices.
30 pages
FOS: Computer and information sciences, Measures of association (correlation, canonical correlation, etc.), classification task, Mathematics - Statistics Theory, Machine Learning (stat.ML), Statistics Theory (math.ST), stock indices, Statistics - Computation, Methodology (stat.ME), machine learning, Statistics - Machine Learning, FOS: Mathematics, conditional Kendall's tau, Computational methods for problems pertaining to statistics, Nonparametric estimation, conditional dependence measure, Statistics - Methodology, Computation (stat.CO)
FOS: Computer and information sciences, Measures of association (correlation, canonical correlation, etc.), classification task, Mathematics - Statistics Theory, Machine Learning (stat.ML), Statistics Theory (math.ST), stock indices, Statistics - Computation, Methodology (stat.ME), machine learning, Statistics - Machine Learning, FOS: Mathematics, conditional Kendall's tau, Computational methods for problems pertaining to statistics, Nonparametric estimation, conditional dependence measure, Statistics - Methodology, Computation (stat.CO)
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 8 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Top 10% | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Top 10% |
