Linear regression through PAC-Bayesian truncation

Jean-Yves Audibert; Olivier Catoni

Preprints, Working Papers, ... Year : 2011

Linear regression through PAC-Bayesian truncation

(1, 2) , (3, 4)

1
2
3
4

Jean-Yves Audibert

Function : Author
PersonId : 931557

imagine [Marne-la-Vallée]

Statistical Machine Learning and Parsimony

Olivier Catoni

Function : Author
PersonId : 858015

Département de Mathématiques et Applications - ENS Paris

Computational Learning, Aggregation, Supervised Statistical, Inference, and Classification

Abstract

We consider the problem of predicting as well as the best linear combination of d given functions in least squares regression under L^\infty constraints on the linear combination. When the input distribution is known, there already exists an algorithm having an expected excess risk of order d/n, where n is the size of the training data. Without this strong assumption, standard results often contain a multiplicative log(n) factor, complex constants involving the conditioning of the Gram matrix of the covariates, kurtosis coefficients or some geometric quantity characterizing the relation between L^2 and L^\infty-balls and require some additional assumptions like exponential moments of the output. This work provides a PAC-Bayesian shrinkage procedure with a simple excess risk bound of order d/n holding in expectation and in deviations, under various assumptions. The common surprising factor of these results is their simplicity and the absence of exponential moment condition on the output distribution while achieving exponential deviations. The risk bounds are obtained through a PAC-Bayesian analysis on truncated differences of losses. We also show that these results can be generalized to other strongly convex loss functions.

Keywords

Linear regression Generalization error Shrinkage PAC-Bayesian theorems Risk bounds Robust statistics Resistant estimators Gibbs posterior distributions Randomized estimators Statistical learning theory

Domains

Statistics [math.ST] Statistics Theory [stat.TH]

Fichier principal

dovern2long.pdf (291.04 Ko)

Origin : Files produced by the author(s)

Jean-Yves Audibert : Connect in order to contact the contributor

https://hal.science/hal-00522536

Submitted on : Sunday, September 11, 2011-5:26:49 PM

Last modification on : Saturday, April 27, 2024-3:14:18 AM

Long-term archiving on: Monday, December 12, 2011-2:21:46 AM

Dates and versions

hal-00522536 , version 1 (30-09-2010)

hal-00522536 , version 2 (11-09-2011)

Identifiers

HAL Id : hal-00522536 , version 2
ARXIV : 1010.0072

Cite

Jean-Yves Audibert, Olivier Catoni. Linear regression through PAC-Bayesian truncation. 2011. ⟨hal-00522536v2⟩

Export

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

ENS-PARIS ENPC CNRS INRIA INSMI PARISTECH LIGM IMAGINE INRIA2 PSL MATH_ENS_PARIS UNIV-EIFFEL JSE2024

594 View

309 Download

Linear regression through PAC-Bayesian truncation

Abstract

Keywords

Domains

Dates and versions

Identifiers

Cite

Export

Collections

Altmetric

Share