• Home
  • Search
  • Adaptive model selection using empirical complexities
  • Cite Icon60
  • https://doi.org/10.1214/aos/1017939241Copy DOI Icon

Adaptive model selection using empirical complexities

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Given $n$ independent replicates of a jointly distributed pair $(X, Y) \in \mathscr{R}^d \times \mathscr{R}$, we wish to select from a fixed sequence of model classes $\mathscr{F}_1, \mathscr{F}_2,\dots$ a deterministic prediction rule $f: \mathscr{R}^d \to \mathscr{R}$ whose risk is small.We investigate the possibility of empirically assessing the complexity of each model class, that is, the actual difficulty of the estimation problem within each class. The estimated complexities are in turn used to define an adaptive model selection procedure, which is based on complexity penalized empirical risk. The available data are divided into two parts. The first is used to form an empirical cover of each model class, and the second is used to select a candidate rule from each cover based on empirical risk. The covering radii are determined empirically to optimize a tight upper bound on the estimation error. An estimate is chosen from the list of candidates in order to minimize the sum of class complexity and empirical risk. A distinguishing feature of the approach is that the complexity of each model class is assessed empirically, based on the size of its empirical cover. Finite sample performance bounds are established for the estimates, and these bounds are applied to several nonparametric estimation problems. The estimates are shown to achieve a favorable trade-off between approximation and estimation error and to perform as well as if the distribution-dependent complexities of the model classes were known beforehand. In addition, it is shown that the estimate can be consistent, and even possess near optimal rates of convergence, when each model class has an infinite VC or pseudo dimension. For regression estimation with squared loss we modify our estimate to achieve a faster rate of convergence.

Similar Papers
  • Research Article
  • Citations30

FAST RATES FOR ESTIMATION ERROR AND ORACLE INEQUALITIES FOR MODEL SELECTION

  • Jan 15, 2008
  • Econometric Theory
  • Peter L Bartlett
  • Research Article
  • Citations1

AMSPM: Adaptive Model Selection and Partition Mechanism for Edge Intelligence-driven 5G Smart City with Dynamic Computing Resources

  • Mar 16, 2024
  • ACM Transactions on Sensor Networks
  • Xin Niu +3
  • Research Article
  • Citations9

Optimal convergence and superconvergence of semi-Lagrangian discontinuous Galerkin methods for linear convection equations in one space dimension

  • Mar 10, 2020
  • Mathematics of Computation
  • Yang Yang +2
  • Conference Article
  • Citations11

Too much TV is bad: Dense reconstruction from sparse laser with non-convex regularisation

  • May 01, 2015
  • Pedro Pinies +2
  • Research Article
  • Citations23

A prediction algorithm for time series based on adaptive model selection

  • Dec 05, 2007
  • Expert Systems With Applications
  • Jiangjiao Duan +4
  • Research Article
  • Citations7

Model selection: A Lagrange optimization approach

  • Mar 09, 2009
  • Journal of Statistical Planning and Inference
  • Yongli Zhang
  • Research Article
  • Citations204

Adaptive Model Selection

  • Mar 01, 2002
  • Journal of the American Statistical Association
  • Xiaotong Shen +1
  • Research Article
  • Citations2

On rate-optimal nonparametric wavelet regression with long memory moving average errors

  • Jun 06, 2013
  • Statistical Inference for Stochastic Processes
  • Linyuan Li +1
  • Book Chapter
  • Citations27

Model Selection by Linear Programming

  • Jan 01, 2014
  • Joseph Wang +3
  • Research Article
  • Citations10

Technical note—Constructing confidence intervals for nested simulation

  • Aug 22, 2022
  • Naval Research Logistics (NRL)
  • Hong‐Fa Cheng +2
  • Research Article
  • Citations3

Conservative SPDEs as fluctuating mean field limits of stochastic gradient descent

  • Jan 30, 2025
  • Probability Theory and Related Fields
  • Benjamin Gess +2
  • Research Article
  • Citations166

Variance Function Estimation in Regression: The Effect of Estimating the Mean

  • Sep 01, 1989
  • Journal of the Royal Statistical Society Series B: Statistical Methodology
  • Peter Hall +1
  • Research Article
  • Citations14

Distress Classification of Road Structures via Adaptive Bayesian Network Model Selection

  • May 19, 2017
  • Journal of Computing in Civil Engineering
  • K Maeda +3
  • Research Article
  • Citations4

Remeshing and Refining with Moving Finite Elements. Application to Nonlinear Wave Problems

  • Jan 01, 2006
  • Cmes-computer Modeling in Engineering & Sciences
  • Abigail Wacher +1
  • Research Article
  • Citations10

Contention Grading and Adaptive Model Selection for Machine Vision in Embedded Systems

  • Sep 30, 2022
  • ACM Transactions on Embedded Computing Systems
  • Basar Kutukcu +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.