Sandwich boosting for accurate estimation in partially linear models for grouped data

Elliot H Young,Rajen D Shah

doi:10.1093/jrsssb/qkae032

Abstract

Abstract We study partially linear models in settings where observations are arranged in independent groups but may exhibit within-group dependence. Existing approaches estimate linear model parameters through weighted least squares, with optimal weights (given by the inverse covariance of the response, conditional on the covariates) typically estimated by maximizing a (restricted) likelihood from random effects modelling or by using generalized estimating equations. We introduce a new ‘sandwich loss’ whose population minimizer coincides with the weights of these approaches when the parametric forms for the conditional covariance are well-specified, but can yield arbitrarily large improvements in linear parameter estimation accuracy when they are not. Under relatively mild conditions, our estimated coefficients are asymptotically Gaussian and enjoy minimal variance among estimators with weights restricted to a given class of functions, when user-chosen regression methods are used to estimate nuisance functions. We further expand the class of functional forms for the weights that may be fitted beyond parametric models by leveraging the flexibility of modern machine learning methods within a new gradient boosting scheme for minimizing the sandwich loss. We demonstrate the effectiveness of both the sandwich loss and what we call ‘sandwich boosting’ in a variety of settings with simulated and real-world data.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Sandwich boosting for accurate estimation in partially linear models for grouped data

Abstract

Talk to us

Similar Papers

More From: Journal of the Royal Statistical Society Series B: Statistical Methodology

Lead the way for us

Journal: Journal of the Royal Statistical Society Series B: Statistical Methodology	Publication Date: May 9, 2024
License type: CC BY 4.0

Similar Papers

COVARIANCES FOR FIXED INTERVAL SMOOTHED KALMAN FILTER PARAMETER ESTIMATES
Stephen Haslett
Australian Journal of Statistics | VOL. 32
Stephen HaslettStephen Haslett
01 Jun 1990
Australian Journal of Statistics | VOL. 32

Error-Free and Best-Fit Extensions of Partially Defined Boolean Functions
Endre Boros ... Toshihide Ibaraki
Information and Computation | VOL. 140
Endre Boros, et. al.Endre Boros ... Toshihide Ibaraki
01 Feb 1998
Information and Computation | VOL. 140

Best linear unbiased estimation for varying probability with and without replacement sampling
Stephen Haslett
Special Matrices | VOL. 7
Stephen HaslettStephen Haslett
01 Jan 2019
Special Matrices | VOL. 7

Application of Linear Model Fitting in Image Edge Fast Detection
Peng Wang ... Zhao Wei
-
Peng Wang, et. al.Peng Wang ... Zhao Wei
01 Jan 2008
01 Jan 2008

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Sandwich boosting for accurate estimation in partially linear models for grouped data

Abstract

Talk to us

Similar Papers

More From: Journal of the Royal Statistical Society Series B: Statistical Methodology