Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

0.3%0.5%0.8%0.1% · Feb 202319922001200920172026
6 results for over-parameterised

Wide neural networks with asymmetrical node scaling converge globally and learn features.

problem Global convergence and feature learning in over-parameterised shallow networks.
method Gradient-based optimisation of wide, shallow neural networks with asymmetrical node scaling.
result Gradient flow and gradient descent converge to a global minimum and learn features, unlike in the NTK parameterisation.

Wide stochastic networks show Gaussian behavior and improve training with PAC-Bayesian methods.

problem Analyzing and training over-parameterised neural networks with large width.
method Establishing Gaussian behavior for a stochastic architecture, applying PAC-Bayesian training.
result PAC-Bayesian training on large but finite-width networks outperforms standard methods.

Study on SGD dynamics in neural networks, revealing generalisation patterns.

problem Understanding generalisation in over-parameterised neural networks.
method Analysis of SGD dynamics in a teacher-student setup using differential equations.
result Network size affects generalisation error differently depending on training layers and activation functions.

VINNAS uses variational inference to avoid mode collapse in neural architecture search.

problem Mode collapse in gradient-based NAS methods, leading to suboptimal architectures.
method Differentiable variational inference with variational dropout and automatic relevance determination.
result State-of-the-art accuracy with up to twice fewer non-zero parameters.

This research explores the dynamics of linearised neural nets, revealing distinct learning phases and layer growth rates.

problem Understanding the fundamental mechanics of neural nets and their learning dynamics.
method Derivation of properties of learning dynamics in general multi-layer linear neural nets, including orthogonal networks.
result Linear multi-layer neural nets exhibit distinct phases of learning with different layer growth rates, and nonlinearity affects these dynamics.