Individualized Multidirectional Variable Selection
From MaRDI portal
Abstract: In this paper we propose a heterogeneous modeling framework which achieves individual-wise feature selection and individualized covariates' effects subgrouping simultaneously. In contrast to conventional model selection approaches, the new approach constructs a separation penalty with multi-directional shrinkages, which facilitates individualized modeling to distinguish strong signals from noisy ones and selects different relevant variables for different individuals. Meanwhile, the proposed model identifies subgroups among which individuals share similar covariates' effects, and thus improves individualized estimation efficiency and feature selection accuracy. Moreover, the proposed model also incorporates within-individual correlation for longitudinal data to gain extra efficiency. We provide a general theoretical foundation under a double-divergence modeling framework where the number of individuals and the number of individual-wise measurements can both diverge, which enables inference on both an individual level and a population level. In particular, we establish strong oracle property for the individualized estimator to ensure its optimal large sample property under various conditions. An efficient ADMM algorithm is developed for computational scalability. Simulation studies and applications to post-trauma mental disorder analysis with genetic variation and an HIV longitudinal treatment study are illustrated to compare the new approach to existing methods.
Recommendations
- Heterogeneous quantile regression for longitudinal data with subgroup structures
- Group selection in the Cox model with a diverging number of covariates
- A group bridge approach for variable selection
- Subgroup identification and variable selection for treatment decision making
- A random-effect model approach for group variable selection
Cites work
- Asymptotic results with generalized estimating equations for longitudinal data
- Asymptotics for generalized estimating equations with large cluster sizes
- Cluster analysis of longitudinal profiles with subgroups
- Cluster analysis: unsupervised learning via supervised learning with a non-convex penalty
- Coordinate descent algorithms for nonconvex penalized regression, with applications to biological feature selection
- Distributed optimization and statistical learning via the alternating direction method of multipliers
- Estimating the dimension of a model
- Estimating the number of clusters in a data set via the gap statistic
- Finding the Number of Clusters in a Dataset
- Fused Lasso approach in regression coefficients clustering -- learning parameter heterogeneity in data integration
- Grouping pursuit through a regularization solution surface
- scientific article; zbMATH DE number 845714 (Why is no real title available?)
- Longitudinal clustering for heterogeneous binary data
- Longitudinal data analysis using generalized linear models
- Model Selection and Estimation in Regression with Grouped Variables
- Nearly unbiased variable selection under minimax concave penalty
- Pairwise Variable Selection for High-Dimensional Model-Based Clustering
- Penalized generalized estimating equations for high-dimensional longitudinal data analysis
- Penalized model-based clustering with application to variable selection
- Properties and refinements of the fused Lasso
- Regularization and Variable Selection Via the Elastic Net
- Shrinkage tuning parameter selection with a diverging number of parameters
- Simultaneous Regression Shrinkage, Variable Selection, and Supervised Clustering of Predictors with OSCAR
- Simultaneous supervised clustering and feature selection over a graph
- Sparsity and Smoothness Via the Fused Lasso
- Testing for Qualitative Interactions between Treatment Effects and Patient Subsets
- The Adaptive Lasso and Its Oracle Properties
- Tuning parameter selectors for the smoothly clipped absolute deviation method
- Variable Selection for Model-Based Clustering
- Variable Selection via Nonconcave Penalized Likelihood and its Oracle Properties
Cited in
(32)- Pursuing Sources of Heterogeneity in Modeling Clustered Population
- Grouped Generalized Estimating Equations for Longitudinal Data Analysis
- Gene-environment interaction analysis under the Cox model
- Community Detection in General Hypergraph Via Graph Embedding
- Query-Augmented Active Metric Learning
- Center-Augmented ℓ2-Type Regularization for Subgroup Learning
- Heterogeneous Mediation Analysis on Epigenomic PTSD and Traumatic Stress in a Predominantly African American Cohort
- Structure learning via unstructured kernel-based M-estimation
- Integrated subgroup identification from multi-source data
- Heterogeneous quantile regression for longitudinal data with subgroup structures
- Crowdsourcing Utilizing Subgroup Structure of Latent Factor Modeling
- Subgroup analysis with concave pairwise fusion penalty for ordinal response
- Robust integrative analysis via quantile regression with homogeneity and sparsity
- De-confounding Causal Inference Using Latent Multiple-Mediator Pathways
- High-dimensional variable selection accounting for heterogeneity in regression coefficients across multiple data sources
- Fused mean structure learning in data integration with dependence
- Health care provider clustering using fusion penalty in quasi-likelihood
- Local Clustering for Functional Data
- Heterogeneous Functional Regression for Subgroup Analysis
- Integrating quantile regression for multi-source subgroup identification
- Reinforcement learning for individual optimal policy from heterogeneous data
- High-dimensional subgroup regression analysis
- Subgroup learning for multiple mixed-type outcomes with block-structured covariates
- A forward k-means algorithm for regression clustering
- Simultaneous Coefficient Clustering and Sparsity for Multivariate Mixed Models
- A novel communication-efficient heterogeneous federated positive and unlabeled learning method for credit scoring
- Integrative subgroup analysis for high-dimensional mixed-type multi-response data
- Variable selection in modelling clustered data via within-cluster resampling
- Learning social relationships: a network embedding-based approach for community detection via cosine-similarity
- Dynamic Decision Making With Individualized Variable Selection
- Regression coefficients clustering for longitudinal data in the presence of heteroscedasticity
- Subgroup identification and membership prediction
This page was built for publication: Individualized Multidirectional Variable Selection
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6040687)