Back to Search
Start Over
Group linear algorithm with sparse principal decomposition: a variable selection and clustering method for generalized linear models.
- Source :
- Statistical Papers; Feb2023, Vol. 64 Issue 1, p227-253, 27p
- Publication Year :
- 2023
-
Abstract
- This paper introduces the Group Linear Algorithm with Sparse Principal decomposition, an algorithm for supervised variable selection and clustering. Our approach extends the Sparse Group Lasso regularization to calculate clusters as part of the model fit. Therefore, unlike Sparse Group Lasso, our idea does not require prior specification of clusters between variables. To determine the clusters, we solve a particular case of sparse Singular Value Decomposition, with a regularization term that follows naturally from the Group Lasso penalty. Moreover, this paper proposes a unified implementation to deal with, but not limited to, linear regression, logistic regression, and proportional hazards models with right-censoring. Our methodology is evaluated using both biological and simulated data, and details of the implementation in R and hyperparameter search are discussed. [ABSTRACT FROM AUTHOR]
Details
- Language :
- English
- ISSN :
- 09325026
- Volume :
- 64
- Issue :
- 1
- Database :
- Complementary Index
- Journal :
- Statistical Papers
- Publication Type :
- Academic Journal
- Accession number :
- 161416278
- Full Text :
- https://doi.org/10.1007/s00362-022-01313-z