Sure Independence Screening for Ultrahigh Dimensional Feature Space

Summary Variable selection plays an important role in high dimensional statistical modelling which nowadays appears in many areas and is key to various scientific discoveries. For problems of large scale or dimensionality p, accuracy of estimation and computational cost are two top concerns. Recently, Candes and Tao have proposed the Dantzig selector using L1-regularization and showed that it achieves the ideal risk up to a logarithmic factor log(p). Their innovative procedure and remarkable result are challenged when the dimensionality is ultrahigh as the factor log(p) can be large and their uniform uncertainty principle can fail. Motivated by these concerns, we introduce the concept of sure screening and propose a sure screening method that is based on correlation learning, called sure independence screening, to reduce dimensionality from high to a moderate scale that is below the sample size. In a fairly general asymptotic framework, correlation learning is shown to have the sure screening property for even exponentially growing dimensionality. As a methodological extension, iterative sure independence screening is also proposed to enhance its finite sample performance. With dimension reduced accurately from high to below sample size, variable selection can be improved on both speed and accuracy, and can then be accomplished by a well-developed method such as smoothly clipped absolute deviation, the Dantzig selector, lasso or adaptive lasso. The connections between these penalized least squares methods are also elucidated.

Variable Selection viaNonconcave Penalized…Variable Selection via Nonconcave Penalized Likelihood and its Oracle PropertiesVariable Selection forCox's proportional…Variable Selection for Cox's proportional Hazards Model and Frailty ModelNonconcave penalizedlikelihood with a…Nonconcave penalized likelihood with a diverging number of parametersThe Adaptive Lasso andIts Oracle PropertiesThe Adaptive Lasso and Its Oracle PropertiesStatistical Challengeswith High…Statistical Challenges with High Dimensionality: Feature Selection in Knowledge DiscoveryHigh-dimensionalclassification using…High-dimensional classification using features annealed independence rulesShrinkage TuningParameter Selection wit…Shrinkage Tuning Parameter Selection with a Diverging number of ParametersAdaptive Lasso forsparse high-dimensional…Adaptive Lasso for sparse high-dimensional regression modelsThe sparsity and bias ofthe Lasso selection in…The sparsity and bias of the Lasso selection in high-dimensional linear regressionSimultaneous analysis ofLasso and Dantzig…Simultaneous analysis of Lasso and Dantzig selectorUsing GeneralizedCorrelation to Effect…Using Generalized Correlation to Effect Variable Selection in Very High Dimensional ProblemsOn the adaptiveelastic-net with a…On the adaptive elastic-net with a diverging number of parametersPenalized CompositeQuasi-Likelihood for…Penalized Composite Quasi-Likelihood for Ultrahigh Dimensional Variable SelectionOn tuning parameterselection of lasso-type…On tuning parameter selection of lasso-type methods - a monte carlo studyUltrahigh-dimensionalvariable selection…Ultrahigh-dimensional variable selection method for whole-genome gene-gene interaction analysisBayesian Methods forHigh Dimensional Linear…Bayesian Methods for High Dimensional Linear ModelsOn asymptoticallyoptimal confidence…On asymptotically optimal confidence regions and tests for high-dimensional modelsTwo tales of variableselection for high…Two tales of variable selection for high dimensional regression: Screening and model buildingConditional SureIndependence ScreeningConditional Sure Independence ScreeningClassification ofclinical outcomes using…Classification of clinical outcomes using high-throughput informatics: Part 1 - nonparametric method reviewsSparse estimation basedon square root nonconve…Sparse estimation based on square root nonconvex optimization in high-dimensional dataSparse MinimumDiscrepancy Approach to…Sparse Minimum Discrepancy Approach to Sufficient Dimension Reduction with Simultaneous Variable Selection in Ultrahigh DimensionModel-Free ForwardScreening Via Cumulativ…Model-Free Forward Screening Via Cumulative DivergenceComment: FeatureScreening and Variable…Comment: Feature Screening and Variable Selection via Iterative Ridge RegressionSure IndependenceScreening for Ultrahigh…Sure Independence Screening for Ultrahigh Dimensional Feature SpaceEarlier referencesFocus paperCiting papersOlderNewer

Click a node to pin it, click the empty canvas to go back to this paper, or hover to preview. Open a node’s page from its title.