High-Dimensional Feature Selection by Feature-Wise Kernelized Lasso

The goal of supervised feature selection is to find a subset of input features that are responsible for predicting output values. The least absolute shrinkage and selection operator (Lasso) allows computationally efficient feature selection based on linear dependency between input features and output values. In this letter, we consider a feature-wise kernelized Lasso for capturing nonlinear input-output dependency. We first show that with particular choices of kernel functions, nonredundant features with strong statistical dependence on output values can be found in terms of kernel-based independence measures such as the Hilbert-Schmidt independence criterion. We then show that the globally optimal solution can be efficiently computed; this makes the approach scalable to high-dimensional problems. The effectiveness of the proposed method is demonstrated through feature selection experiments for classification and regression with thousands of features.

Regression Shrinkage andSelection Via the LassoRegression Shrinkage and Selection Via the LassoFeature Selection Basedon Mutual Information…Feature Selection Based on Mutual Information: Criteria of Max-Dependency, Max-Relevance, and Min-RedundancyRegularization andVariable Selection Via…Regularization and Variable Selection Via the Elastic NetAn Interior-Point Methodfor Large-Scale ell…An Interior-Point Method for Large-Scale ell _1-Regularized Least SquaresSparse Additive ModelsSparse Additive ModelsQuadratic ProgrammingFeature SelectionQuadratic Programming Feature SelectionFromTransformation-Based…From Transformation-Based Dimensionality Reduction to Feature SelectionStability SelectionStability SelectionSuper-Linear Convergenceof Dual Augmented…Super-Linear Convergence of Dual Augmented Lagrangian Algorithm for Sparsity Regularized EstimationAlgorithms for LearningKernels Based on…Algorithms for Learning Kernels Based on Centered AlignmentGlobal SensitivityAnalysis with Dependenc…Global Sensitivity Analysis with Dependence MeasuresNeural Decoding withKernel-Based Metric…Neural Decoding with Kernel-Based Metric LearningGlobal SensitivityAnalysis with Dependenc…Global Sensitivity Analysis with Dependence MeasuresNeural Decoding withKernel-Based Metric…Neural Decoding with Kernel-Based Metric LearningKernel learning andoptimization with…Kernel learning and optimization with Hilbert-Schmidt independence criterionLocally weighted kernelpartial least squares…Locally weighted kernel partial least squares regression based on sparse nonlinear features for virtual sensing of nonlinear time-varying processesTwo-Stage Fuzzy MultipleKernel Learning Based o…Two-Stage Fuzzy Multiple Kernel Learning Based on Hilbert-Schmidt Independence CriterionHigh-dimensionalsupervised feature…High-dimensional supervised feature selection via optimized kernel mutual informationBlock HSIC Lasso:model-free biomarker…Block HSIC Lasso: model-free biomarker detection for ultra-high dimensional dataA Kernel DiscriminantInformation Approach to…A Kernel Discriminant Information Approach to Non-linear Feature SelectionA novel Grangercausality method based…A novel Granger causality method based on HSIC-Lasso for revealing nonlinear relationship between multivariate time seriesImproved metabolomicdata-based prediction o…Improved metabolomic data-based prediction of depressive symptoms using nonlinear machine learning with feature selectionGraphLIME: LocalInterpretable Model…GraphLIME: Local Interpretable Model Explanations for Graph Neural NetworksDependency maximizationforward feature…Dependency maximization forward feature selection algorithms based on normalized cross-covariance operator and its approximated form for high-dimensional dataHigh-Dimensional FeatureSelection by…High-Dimensional Feature Selection by Feature-Wise Kernelized Lasso過去の参考文献中心の論文この論文を引用する論文古い新しい

ノードをクリックするとフォーカスを固定、空白をクリックすると本論文に戻ります。ホバーで一時的にプレビューできます。各ノードのページはタイトルから開けます。