Estimating a regression function in exponential families by model selection
At a glance
- Citations
- 0
- References
- 26
- Comments
- 0
Öz
Let (W1,Y1),…,(Wn,Yn) be n pairs of independent random variables. We assume that, for each i∈{1,…,n}, the conditional distribution of Yi given Wi belongs to a one-parameter exponential family with parameter γ⋆(Wi)∈R. The statistical goal is to estimate these conditional distributions. We consider a model selection procedure which works based on a general assumption that each of the model is VC-subgraph. We establish a non-asymptotic risk bound for the resulting estimator with respect to a Hellinger-type distance. By leveraging this result, we extend several findings previously explored in Gaussian regression to the regression in exponential families. Specifically, we address the curse of dimensionality by imposing structural assumptions, such as general additive and multiple index structures, on γ⋆. We also study model selection for ReLU neural networks, and provide a concrete example of how ReLU neural networks can achieve a significantly faster convergence rate than traditional models. When γ⋆ is close to a composition of several Hölder functions, we show that under a suitable parametrization of the exponential family, our estimator achieves the same rate of convergence as in the Gaussian case. Combining with a lower bound, the rate is minimax optimal up to a logarithmic term. Finally, we apply the model selection procedure to address adaptation and variable selection problems in exponential families.
Publication details
- DOI
- 10.3150/23-bej1649
- OpenAlex
- W4221147925
- Document type
- article
- Language
- EN
- Source
- Bernoulli
- Last metadata update
Comments
Oturum Açın to join the discussion.