Expectation propagation (EP) is a powerful approximate inference algorithm. However, a critical barrier in applying EP is that the moment matching in message updates can be intractable. Handcrafting approximations is usually tricky, and lacks generalizability. Importance sampling is very expensive. While Laplace propag…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Corrected moment-based methods improve inference in topic model regression.
Many inference problems involving questions of optimality ask for the maximum or the minimum of a finite set of unknown quantities. This technical report derives the first two posterior moments of the maximum of two correlated Gaussian variables and the first two posterior moments of the two generating variables (corre…
Bayesian learning is often hampered by large computational expense. As a powerful generalization of popular belief propagation, expectation propagation (EP) efficiently approximates the exact Bayesian computation. Nevertheless, EP can be sensitive to outliers and suffer from divergence for difficult cases. To address t…
A new filter reduces density fitting to a linear solve, improving performance on nonlinear systems.
New methods for uncertainty in neural networks with leaky ReLU activations.
To deepen our understanding of graph neural networks, we investigate the representation power of Graph Convolutional Networks (GCN) through the looking glass of graph moments, a key property of graph topology encoding path of various lengths. We find that GCNs are rather restrictive in learning graph moments. Without c…
A new framework for lightweight BNNs learns heteroscedastic uncertainties efficiently.
Paper presents a method to accurately quantify neural network uncertainty without sampling.
New EP variants improve inference stability and efficiency.
Proposes a new neural network architecture inspired by biology to improve learning and information flow.
This paper makes two contributions to Bayesian machine learning algorithms. Firstly, we propose stochastic natural gradient expectation propagation (SNEP), a novel alternative to expectation propagation (EP), a popular variational inference algorithm. SNEP is a black box variational algorithm, in that it does not requi…
Stochastic regularisation is an important weapon in the arsenal of a deep learning practitioner. However, despite recent theoretical advances, our understanding of how noise influences signal propagation in deep neural networks remains limited. By extending recent work based on mean field theory, we develop a new frame…
Unified framework for efficient Gaussian process inference.
Many neural networks use the tanh activation function, however when given a probability distribution as input, the problem of computing the output distribution in neural networks with tanh activation has not yet been addressed. One important example is the initialization of the echo state network in reservoir computing…
Novel methods for splitting Gaussian mixtures improve uncertainty propagation in nonlinear systems.
Uniform-in-time analysis for Stein Variational Gradient Descent across various metrics.
Exponential family distributions are highly useful in machine learning since their calculation can be performed efficiently through natural parameters. The exponential family has recently been extended to the t-exponential family, which contains Student-t distributions as family members and thus allows us to handle noi…
We address the problem of estimating statistics of hidden units in a neural network using a method of analytic moment propagation. These statistics are useful for approximate whitening of the inputs in front of saturating non-linearities such as a sigmoid function. This is important for initialization of training and f…
Wide neural networks learn features under P, identifying weights and decomposing support.
QP improves Gaussian process inference by minimizing Wasserstein distance.
Conditional DGP learns effective kernels from low-fidelity data.
Develops a Gaussian-based message-passing algorithm for noisy matrix completion.
Graph-based Bayesian SSL uses graph theory to propagate labels from a few to many unlabeled features.
Paper improves spectral learning of HMMs to avoid local optima and improve robustness.
A fast method for neural networks that provides uncertainty measures.
Proposes a method to adapt DNNs to drift in data distribution.
A new neural network initialization method is proposed for faster and more accurate training.
Generative model captures complex dependence in financial data.
Expectation Propagation (EP) provides a framework for approximate inference. When the model under consideration is over a latent Gaussian field, with the approximation being Gaussian, we show how these approximations can systematically be corrected. A perturbative expansion is made of the exact but intractable correcti…
New initialization schemes preserve fractional moments of weights in deep networks, improving training and test performance.
Optimal stock price prediction model using recurrent neural networks with RMSprop optimizer.
Favour speeds up variance estimation in BNNs, making them practical for performance-critical tasks.
A new method learns posterior and predictive distributions together, reducing computational cost.
SPIDER uses deep neural networks for streaming tensor factorization.
We study analytically and numerically Minsky instability as a combination of top-down, bottom-up and peer-to-peer positive feedback loops. The peer-to-peer interactions are represented by the links of a network formed by the connections between firms, contagion leading to avalanches and percolation phase transitions pr…
We address the problem of learning the parameters in graphical models when inference is intractable. A common strategy in this case is to replace the partition function with its Bethe approximation. We show that there exists a regime of empirical marginals where such Bethe learning will fail. By failure we mean that th…
The aim of this paper is to provide new theoretical and computational understanding on two loss regularizations employed in deep learning, known as local entropy and heat regularization. For both regularized losses we introduce variational characterizations that naturally suggest a two-step scheme for their optimizatio…
A network of independently trained Gaussian processes (StackedGP) is introduced to obtain predictions of quantities of interest with quantified uncertainties. The main applications of the StackedGP framework are to integrate different datasets through model composition, enhance predictions of quantities of interest thr…
Two wave fronts and that originated at some points of the manifold are said to be causally related if one of them passed through the origin of the other before the other appeared. We define the causality relation invariant to be the algebraic number of times the earlier born front pass…
Generalized belief propagation converges to optimal solutions on graphs with motifs.
Belief propagation recovers backpropagation results.
A susceptibility propagation that is constructed by combining a belief propagation and a linear response method is used for approximate computation for Markov random fields. Herein, we formulate a new, improved susceptibility propagation by using the concept of a diagonal matching method that is based on mean-field app…
FAB-COST improves cold-start recommendation accuracy with less data.
A new method calculates fractional moments using the moment-generating function.
Study compares weak and homotopy moment maps in multisymplectic geometry.
Variational inference is a powerful concept that underlies many iterative approximation algorithms; expectation propagation, mean-field methods and belief propagations were all central themes at the school that can be perceived from this unifying framework. The lectures of Manfred Opper introduce the archetypal example…
For a GJR-GARCH specification with a generic innovation distribution we derive analytic expressions for the first four conditional moments of the forward and aggregated returns and variances. Moment for the most commonly used GARCH models are stated as special cases. We also the limits of these moments as the time hori…