High-Dimensional Multi-Task Averaging and Application to Kernel Mean Embedding - Inria - Institut national de recherche en sciences et technologies du numérique Accéder directement au contenu
Communication Dans Un Congrès Année : 2021

High-Dimensional Multi-Task Averaging and Application to Kernel Mean Embedding

Résumé

We propose an improved estimator for the multi-task averaging problem, whose goal is the joint estimation of the means of multiple distributions using separate, independent data sets. The naive approach is to take the empirical mean of each data set individually, whereas the proposed method exploits similarities between tasks, without any related information being known in advance. First, for each data set, similar or neighboring means are determined from the data by multiple testing. Then each naive estimator is shrunk towards the local average of its neighbors. We prove theoretically that this approach provides a reduction in mean squared error. This improvement can be significant when the dimension of the input space is large, demonstrating a "blessing of dimensionality" phenomenon. An application of this approach is the estimation of multiple kernel mean embeddings, which plays an important role in many modern applications. The theoretical results are verified on artificial and real world data.
Fichier principal
Vignette du fichier
arxiv_paper2.pdf (335.78 Ko) Télécharger le fichier
Origine : Fichiers produits par l'(les) auteur(s)
Loading...

Dates et versions

hal-03002342 , version 1 (12-11-2020)

Identifiants

Citer

Hannah Marienwald, Jean-Baptiste Fermanian, Gilles Blanchard. High-Dimensional Multi-Task Averaging and Application to Kernel Mean Embedding. AISTATS 2021 - 24th International Conference on Artificial Intelligence and Statistics, Apr 2021, Virtual, United States. pp.1963-1971. ⟨hal-03002342⟩
86 Consultations
64 Téléchargements

Altmetric

Partager

Gmail Facebook X LinkedIn More