Sufficient Dimension Reduction via Squared-Loss Mutual Information Estimation

The goal of sufficient dimension reduction in supervised learning is to find the low-dimensional subspace of input features that contains all of the information about the output values that the input features possess. In this letter, we propose a novel sufficient dimension-reduction method using a squared-loss variant of mutual information as a dependency measure. We apply a density-ratio estimator for approximating squared-loss mutual information that is formulated as a minimum contrast estimator on parametric or nonparametric models. Since cross-validation is available for choosing an appropriate model, our method does not require any prespecified structure on the underlying distributions. We elucidate the asymptotic bias of our estimator on parametric models and the asymptotic convergence rate on nonparametric models. The convergence analysis utilizes the uniform tail-bound of a U-process, and the convergence rate is characterized by the bracketing entropy of the model. We then develop a natural gradient algorithm on the Grassmann manifold for sufficient subspace search. The analytic formula of our estimator allows us to compute the gradient efficiently. Numerical experiments show that the proposed method compares favorably with existing dimension-reduction approaches on artificial and benchmark data sets.

A survey of kernels forstructured dataA survey of kernels for structured dataDimensionality Reductionfor Supervised Learning…Dimensionality Reduction for Supervised Learning with Reproducing Kernel Hilbert SpacesSufficient DimensionReduction via Inverse…Sufficient Dimension Reduction via Inverse RegressionEstimating divergencefunctionals and the…Estimating divergence functionals and the likelihood ratio by penalized convex risk minimizationApproximating MutualInformation by Maximum…Approximating Mutual Information by Maximum Likelihood Density Ratio EstimationMutual informationestimation reveals…Mutual information estimation reveals global associations between stimuli and biological processesKernel dimensionreduction in regressionKernel dimension reduction in regressionA Least-squares Approachto Direct Importance…A Least-squares Approach to Direct Importance EstimationEstimating DivergenceFunctionals and the…Estimating Divergence Functionals and the Likelihood Ratio by Convex Risk MinimizationFeature Selection viaL1-Penalized…Feature Selection via L1-Penalized Squared-Loss Mutual InformationDirect DivergenceApproximation between…Direct Divergence Approximation between Probability Distributions and Its Applications in Machine LearningConditional DensityEstimation with…Conditional Density Estimation with Dimensionality Reduction via Squared-Loss Conditional Entropy MinimizationDirect density-ratioestimation with…Direct density-ratio estimation with dimensionality reduction via least-squares hetero-distributional subspace searchStatistical outlierdetection using direct…Statistical outlier detection using direct density ratio estimationRelative Density-RatioEstimation for Robust…Relative Density-Ratio Estimation for Robust Distribution ComparisonCanonical dependencyanalysis based on…Canonical dependency analysis based on squared-loss mutual informationFeature Selection viaL1-Penalized…Feature Selection via L1-Penalized Squared-Loss Mutual InformationDensity-DifferenceEstimationDensity-Difference EstimationDirect Approximation ofDivergences Between…Direct Approximation of Divergences Between Probability DistributionsDirect DivergenceApproximation between…Direct Divergence Approximation between Probability Distributions and Its Applications in Machine LearningDivergence estimationfor machine learning an…Divergence estimation for machine learning and signal processingConditional DensityEstimation with…Conditional Density Estimation with Dimensionality Reduction via Squared-Loss Conditional Entropy MinimizationCanonical kerneldimension reductionCanonical kernel dimension reductionDirect Estimation of theDerivative of Quadratic…Direct Estimation of the Derivative of Quadratic Mutual Information with Application in Supervised Dimension ReductionSufficient DimensionReduction via…Sufficient Dimension Reduction via Squared-Loss Mutual Information Estimation過去の参考文献中心の論文この論文を引用する論文古い新しい

ノードをクリックするとフォーカスを固定、空白をクリックすると本論文に戻ります。ホバーで一時的にプレビューできます。各ノードのページはタイトルから開けます。