Statistical inverse learning and \ell^1-regularization
Abstract
We study the recovery of sparse functions from finite, noisy, and indirect observations in the framework of statistical inverse learning. The unknown is modeled as an element of , and observations are generated through a possibly nonlinear forward operator , where is a vector-valued reproducing kernel Hilbert space. We propose an -regularized empirical risk minimizer and develop a theoretical analysis of its statistical properties. Under mild assumptions, we establish almost-sure consistency and derive non-asymptotic high-probability convergence rates in both the prediction and reconstruction norms. The rates depend on the source smoothness parameter , characterized by a variational source condition, and the effective dimension exponent , describing the polynomial spectral decay of the covariance operator. We further prove matching minimax lower bounds, showing that the obtained convergence rates are optimal. To relate the theory to practical sparsity models, we consider finitely smoothing operators of the form , where is a synthesis operator, and show that approximation-space assumptions imply the required variational source conditions. In particular, we prove that membership in the approximation space is equivalent to polynomial decay of the best -term approximation error. Finally, we verify the assumptions for two representative inverse problems: reaction coefficient identification in elliptic PDEs and sparse computed tomography. For filtered Radon transforms, we derive explicit effective-dimension asymptotics, yielding concrete convergence rates for standard image models and sparsifying systems.
Explore similar work
Separating Oblivious and Adaptive Models of Variable Selection
for each'') and \emph{adaptive} (for all'') models of sparse recovery. We show that under an oblivious model, the optimal error is attainable in near-linear time with samples, whereas in an adaptive model, samples are necessary for any algorithm to achieve this bound. This establishes a surprising contrast with the standard setting, where samples suffice even for adaptive sparse recovery. We conclude with a preliminary examination of a \emph{partially-adaptive} model, where we show nontrivial variable selection guarantees are possible with measurements.