Back to Search Start Over

Group linear algorithm with sparse principal decomposition: a variable selection and clustering method for generalized linear models.

Authors :
Laria, Juan C.
Aguilera-Morillo, M. Carmen
Lillo, Rosa E.
Source :
Statistical Papers; Feb2023, Vol. 64 Issue 1, p227-253, 27p
Publication Year :
2023

Abstract

This paper introduces the Group Linear Algorithm with Sparse Principal decomposition, an algorithm for supervised variable selection and clustering. Our approach extends the Sparse Group Lasso regularization to calculate clusters as part of the model fit. Therefore, unlike Sparse Group Lasso, our idea does not require prior specification of clusters between variables. To determine the clusters, we solve a particular case of sparse Singular Value Decomposition, with a regularization term that follows naturally from the Group Lasso penalty. Moreover, this paper proposes a unified implementation to deal with, but not limited to, linear regression, logistic regression, and proportional hazards models with right-censoring. Our methodology is evaluated using both biological and simulated data, and details of the implementation in R and hyperparameter search are discussed. [ABSTRACT FROM AUTHOR]

Details

Language :
English
ISSN :
09325026
Volume :
64
Issue :
1
Database :
Complementary Index
Journal :
Statistical Papers
Publication Type :
Academic Journal
Accession number :
161416278
Full Text :
https://doi.org/10.1007/s00362-022-01313-z