Nonsparse learning with latent variables
From MaRDI portal
Abstract: As a popular tool for producing meaningful and interpretable models, large-scale sparse learning works efficiently when the underlying structures are indeed or close to sparse. However, naively applying the existing regularization methods can result in misleading outcomes due to model misspecification. In particular, the direct sparsity assumption on coefficient vectors has been questioned in real applications. Therefore, we consider nonsparse learning with the conditional sparsity structure that the coefficient vector becomes sparse after taking out the impacts of certain unobservable latent variables. A new methodology of nonsparse learning with latent variables (NSL) is proposed to simultaneously recover the significant observable predictors and latent factors as well as their effects. We explore a common latent family incorporating population principal components and derive the convergence rates of both sample principal components and their score vectors that hold for a wide class of distributions. With the properly estimated latent variables, properties including model selection consistency and oracle inequalities under various prediction and estimation losses are established for the proposed methodology. Our new methodology and results are evidenced by simulation and real data examples.
Recommendations
Cites work
- A general framework for consistency of principal component analysis
- A Selective Overview of Variable Selection in High Dimensional Feature Space (Invited Review Article)
- Adaptive estimation in structured factor models with applications to overlapping clustering
- Asymptotic Equivalence of Regularization Methods in Thresholded Parameter Space
- Asymptotics of empirical eigenstructure for high dimensional spiked covariance
- Asymptotics of sample eigenstructure for a large dimensional spiked covariance model
- Confidence Intervals and Hypothesis Testing for High-Dimensional Regression
- Confidence intervals for low dimensional parameters in high dimensional linear models
- Controlling the false discovery rate via knockoffs
- Factor models and variable selection in high-dimensional regression analysis
- Fisher lecture: Dimension reduction in regression
- High dimensional thresholded regression and shrinkage effect
- High-dimensional macroeconomic forecasting and variable selection via penalized regression
- scientific article; zbMATH DE number 5957408 (Why is no real title available?)
- scientific article; zbMATH DE number 3673370 (Why is no real title available?)
- scientific article; zbMATH DE number 845714 (Why is no real title available?)
- Impacts of high dimensionality in finite samples
- Large covariance estimation by thresholding principal orthogonal complements. With discussion and authors' reply
- Latent variable graphical model selection via convex optimization
- Linear hypothesis testing in dense high-dimensional linear models
- Maximum Likelihood Estimation of Misspecified Models
- Model selection principles in misspecified models
- Nonconcave Penalized Likelihood With NP-Dimensionality
- On asymptotically optimal confidence regions and tests for high-dimensional models
- On consistency and sparsity for principal components analysis in high dimensions
- On model selection from a finite family of possibly misspecified time series models
- On principal components and regression: a statistical explanation of a natural phenomenon
- On the distribution of the largest eigenvalue in principal components analysis
- Panning for Gold: ‘Model-X’ Knockoffs for High Dimensional Controlled Variable Selection
- PCA consistency in high dimension, low sample size context
- Penalized high-dimensional empirical likelihood
- Prediction by Supervised Principal Components
- Principal component analysis of compositional data
- Program evaluation and causal inference with high-dimensional data
- RANK: Large-Scale Inference With Graphical Nonlinear Knockoffs
- Regression analysis of additive hazards model with latent variables
- Regularization and Variable Selection Via the Elastic Net
- Regularization methods for high-dimensional instrumental variables regression with an application to genetical genomics
- Scaled sparse linear regression
- Simultaneous analysis of Lasso and Dantzig selector
- Square-root lasso: pivotal recovery of sparse signals via conic programming
- Statistical optimization in high dimensions
- Statistics for high-dimensional data. Methods, theory and applications.
- The Dantzig selector: statistical estimation when \(p\) is much larger than \(n\). (With discussions and rejoinder).
- Uniformly valid post-regularization confidence regions for many functional parameters in z-estimation framework
- Variable inclusion and shrinkage algorithms
- Variable selection for sparse Dirichlet-multinomial regression with an application to microbiome data analysis
- Variable selection in regression with compositional covariates
- Variable Selection via Nonconcave Penalized Likelihood and its Oracle Properties
- Variance estimation using refitted cross-validation in ultrahigh dimensional regression
Cited in
(12)- A generalized information criterion for high-dimensional PCA rank selection
- Parallel integrative learning for large-scale multi-response regression with incomplete outcomes
- L0-Regularized Learning for High-Dimensional Additive Hazards Regression
- Subgroup analysis method for accelerated failure time model
- High-Dimensional Interaction Detection With False Sign Rate Control
- A new model for counterfactual analysis for functional data
- Randomized tensor decomposition and optimization in the Tucker and tensor train formats
- Model selection for generalized linear models with weak factors
- Hard-thresholding regularization method for high-dimensional heterogeneous models
- SOFARI: High-Dimensional Manifold-Based Inference
- Adaptive reweighting for joint estimation of sparse and non-sparse components
- Joint estimation of sparse and dense components through structured iteration
This page was built for publication: Nonsparse learning with latent variables
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q4994162)