Minimax-Optimal Rates For Sparse Additive Models Over Kernel Classes Via Convex Programming

Sparse additive models are families of $d$-variate functions that have the additive decomposition $f^* = \sum_{j \in S} f^*_j$, where $S$ is an unknown subset of cardinality $s \ll d$. In this paper, we consider the case where each univariate component function $f^*_j$ lies in a reproducing kernel Hilbert space (RKHS), and analyze a method for estimating the unknown function $f^*$ based on kernels combined with $\ell_1$-type convex regularization. Working within a high-dimensional framework that allows both the dimension $d$ and sparsity $s$ to increase with $n$, we derive convergence rates (upper bounds) in the $L^2(\mathbb{P})$ and $L^2(\mathbb{P}_n)$ norms over the class $\MyBigClass$ of sparse additive models with each univariate function $f^*_j$ in the unit ball of a univariate RKHS with bounded kernel function. We complement our upper bounds by deriving minimax lower bounds on the $L^2(\mathbb{P})$ error, thereby showing the optimality of our method. Thus, we obtain optimal minimax rates for many interesting classes of sparse additive models, including polynomials, splines, and Sobolev classes. We also show that if, in contrast to our univariate conditions, the multivariate function class is assumed to be globally bounded, then much faster estimation rates are possible for any sparsity $s = Ω(\sqrt{n})$, showing that global boundedness is a significant restriction in the high-dimensional setting.

Empirical Processes inM-EstimationEmpirical Processes in M-EstimationGeometric Parameters ofKernel MachinesGeometric Parameters of Kernel MachinesComponent selection andsmoothing in…Component selection and smoothing in multivariate nonparametric regressionSpAM: Sparse AdditiveModelsSpAM: Sparse Additive ModelsNonnegative GarroteComponent Selection in…Nonnegative Garrote Component Selection in Functional ANOVA modelsSparse Recovery in LargeEnsembles of Kernel…Sparse Recovery in Large Ensembles of Kernel Machines On-Line Learning and BanditsConsistency of the GroupLasso and Multiple…Consistency of the Group Lasso and Multiple Kernel LearningHigh-dimensionaladditive modelingHigh-dimensional additive modelingA unified framework forhigh-dimensional…A unified framework for high-dimensional analysis of M-estimators with decomposable regularizersMinimax rates ofestimation for…Minimax rates of estimation for high-dimensional linear regression over ell _q-ballsA unified framework forhigh-dimensional…A unified framework for high-dimensional analysis of M-estimators with decomposable regularizersFast Learning Rate ofMultiple Kernel…Fast Learning Rate of Multiple Kernel Learning: Trade-Off between Sparsity and SmoothnessA unified framework forhigh-dimensional…A unified framework for high-dimensional analysis of M-estimators with decomposable regularizersEarly stopping fornon-parametric…Early stopping for non-parametric regression: An optimal data-dependent stopping ruleA unified framework forhigh-dimensional…A unified framework for high-dimensional analysis of M-estimators with decomposable regularizersFast Learning Rate ofMultiple Kernel…Fast Learning Rate of Multiple Kernel Learning: Trade-Off between Sparsity and SmoothnessEarly stopping andnon-parametric…Early stopping and non-parametric regression: an optimal data-dependent stopping ruleMinimax Optimal Rates ofEstimation in High…Minimax Optimal Rates of Estimation in High Dimensional Additive Models: Universal Phase TransitionLearning rates for therisk of kernel-based…Learning rates for the risk of kernel-based quantile regression estimators in additive modelsMinimax-optimalnonparametric regressio…Minimax-optimal nonparametric regression in high dimensionsDivide and ConquerKernel Ridge Regression…Divide and Conquer Kernel Ridge Regression: A Distributed Algorithm with Minimax Optimal RatesMinimax optimal rates ofestimation in high…Minimax optimal rates of estimation in high dimensional additive modelsRandomized sketches forkernels: Fast and…Randomized sketches for kernels: Fast and optimal nonparametric regressionSparse Modal AdditiveModelSparse Modal Additive ModelMinimax-Optimal RatesFor Sparse Additive…Minimax-Optimal Rates For Sparse Additive Models Over Kernel Classes Via Convex Programming過去の参考文献中心の論文この論文を引用する論文古い新しい

ノードをクリックするとフォーカスを固定、空白をクリックすると本論文に戻ります。ホバーで一時的にプレビューできます。各ノードのページはタイトルから開けます。