Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

Finite Neural Networks as Mixtures of Gaussian Processes: From Provable Error Bounds to Prior Selection

Дата публикации: 17-08-2026 20:26:00


Infinitely wide or deep neural networks (NNs) with independent and identically distributed (i.i.d.) parameters have been shown to be equivalent to Gaussian processes. Because of the favorable properties of Gaussian processes, this equivalence is commonly employed to analyze neural networks and has led to various breakthroughs over the years. However, neural networks and Gaussian processes are equivalent only in the limit; in the finite case there are currently no methods available to approximate a trained neural network with a Gaussian model with bounds on the approximation error. In this work, we present an algorithmic framework to approximate a neural network of finite width and depth, and with not necessarily i.i.d. parameters, with a mixture of Gaussian processes with bounds on the approximation error. In particular, we consider the Wasserstein distance to quantify the closeness between probabilistic models and, by relying on tools from optimal transport and Gaussian processes, we iteratively approximate the output distribution of each layer of the neural network as a mixture of Gaussian processes. Crucially, for any NN and $\epsilon >0$ our approach is able to return a mixture of Gaussian processes that is $\epsilon$-close to the NN at a finite set of input points. Furthermore, we rely on the differentiability of the resulting error bound to show how our approach can be employed to tune the parameters of a NN to mimic the functional behavior of a given Gaussian process, e.g., for prior selection in the context of Bayesian inference. We empirically investigate the effectiveness of our results on both regression and classification problems with various neural network architectures. Our experiments highlight how our results can represent an important step towards understanding neural network predictions and formally quantifying their uncertainty.

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1 A Mean-Field Analysis of Neural Stochastic Gradient Descent-Ascent for Functional Minimax Optimization 09.8217-08-2026
2 Neural Network Parameter-optimization of Gaussian Pre-marginalized Directed Acyclic Graphs 05.717-08-2026
3 Statistical Learning Theory for Neural Operators 010.2117-08-2026
4 High-Dimensional Analysis of Gradient Flow for Extensive-Width Quadratic Neural Networks 08.717-08-2026
5 End-to-End Deep Learning for Predicting Metric Space-Valued Outputs 010.6617-08-2026
6 The surrogate Gibbs-posterior of a corrected stochastic MALA: Towards uncertainty quantification for neural networks 08.3217-08-2026
7 Generative Bayesian Inference with GANs 06.6217-08-2026
8 A Functional-Space Mean-Field Theory of Partially-Trained Three-Layer Neural Networks 010.9717-08-2026
9 Nonparametric Estimation of a Factorizable Density using Diffusion Models 08.5917-08-2026
10 Learning Bayesian Network Classifiers to Minimize Class Variable Parameters 05.8617-08-2026

Классификация: Пресс-релизы. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 4.23. Источник: jmlr.org.