Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

74148221295 · Jun 202019922001200920172026
48 results for loss minimisation

Loss minimisation fails to capture epistemic uncertainty in second-order predictors.

problem Capturing epistemic uncertainty in machine learning models.
method Analysis of a second-order learner approach using loss minimisation.
result Loss minimisation does not faithfully represent epistemic uncertainty in second-order predictors.

In few-shot learning, typically, the loss function which is applied at test time is the one we are ultimately interested in minimising, such as the mean-squared-error loss for a regression problem. However, given that we have few samples at test time, we argue that the loss function that we are interested in minimising…

2019-11-29abs ↗pdf ↗

A neural flow method minimizes Willmore energy for 2-surfaces in 3D space.

problem Minimizing Willmore energy for closed oriented 2-surfaces in 3D space.
method Introducing neural Willmore flow to model and minimize the Willmore energy using neural architectures.
result The neural flow reproduces expected round sphere and Clifford torus for genus 0 and 1 surfaces, respectively, and finds minimal Willmore surfaces for genus 2.

Denoising autoencoders (DAEs) are powerful deep learning models used for feature extraction, data generation and network pre-training. DAEs consist of an encoder and decoder which may be trained simultaneously to minimise a loss (function) between an input and the reconstruction of a corrupted version of the input. The…

2017-08-28abs ↗pdf ↗

Playlist recommendation involves producing a set of songs that a user might enjoy. We investigate this problem in three cold-start scenarios: (i) cold playlists, where we recommend songs to form new personalised playlists for an existing user; (ii) cold users, where we recommend songs to form new playlists for a new us…

2019-01-18abs ↗pdf ↗

The paper optimizes forecasting for risk-adjusted decisions under trading frictions.

problem Optimizing forecasting accuracy for investment decisions in the presence of transaction costs.
method Develops a utility-weighted calibration criterion to minimize decision loss net of costs.
result Utility-weighted calibration reduces decision loss by over 30% and improves Sharpe ratio.

New findings show second-order scoring rules can't accurately represent epistemic uncertainty.

problem Lack of epistemic uncertainty representation in second-order learners.
method Generalised second-order scoring rules introduced to prove theoretical limitations.
result No loss function incentivizes second-order learners to accurately represent epistemic uncertainty.

The paper studies properties of Sliced Wasserstein energy for discrete measures.

problem Optimizing discrete probability measures using Sliced Wasserstein loss.
method Investigates the regularity and optimisation properties of the Sliced Wasserstein energy and its Monte-Carlo approximation.
result Convergence results on the critical points of Monte-Carlo approximations to the Sliced Wasserstein energy.

Study robust linear regression with outliers, providing exact asymptotics for ERM performance.

problem Robust linear regression in high-dimension with outliers.
method Analyzes 2\ell_2, 1\ell_1, and Huber losses, providing asymptotic performance metrics.
result Optimally-regularised ERM is asymptotically consistent with simple calibration, but Huber loss requires norm calibration.

EVILL uses randomised perturbations to improve exploration in bandit problems.

problem Improving exploration in structured stochastic bandit problems.
method Solves for the minimiser of a linearly perturbed regularised negative log-likelihood function.
result EVILL matches the performance of Thompson-sampling-style methods in theory and practice.

In this paper we revisit the weighted likelihood bootstrap, a method that generates samples from an approximate Bayesian posterior of a parametric model. We show that the same method can be derived, without approximation, under a Bayesian nonparametric model with the parameter of interest defined as minimising an expec…

2017-09-22abs ↗pdf ↗

New algorithm reduces regret in stochastic bandit convex optimization.

problem Optimizing decisions in uncertain environments with convex losses.
method Introduces a second-order method for zeroth-order stochastic convex bandits.
result Regret bound of (1+r/d)[d1.5n+d3]polylog(n,d,r)(1 + r/d)[d^{1.5} \sqrt{n} + d^3] polylog(n, d, r).

The study of a machine learning problem is in many ways is difficult to separate from the study of the loss function being used. One avenue of inquiry has been to look at these loss functions in terms of their properties as scoring rules via the proper-composite representation, in which predictions are mapped to probab…

2019-02-19abs ↗pdf ↗

We study the problem of finding strain-minimising stream surfaces in a divergence-free vector field. These surfaces are generated by motions of seed curves that propagate through the field in a strain minimising manner, i.e., they move without stretching or shrinking, preserving the length of their arbitrary arc. In ge…

2014-11-05abs ↗pdf ↗

Let ΩR3Ω\subset \mathbb{R}^3 be a Lipschitz domain, and consider a harmonic map v:ΩS2v: Ω\rightarrow \mathbb{S}^2 with boundary data vΩ=φv|\partialΩ= \varphi which minimises the Dirichlet energy. For p2p\geq 2, we show that any energy minimiser uu whose boundary map ψψ has a small W1,pW^{1,p}-distance to φ\varphi is close t…

2018-10-24abs ↗pdf ↗

Wasserstein GANs fail to approximate Wasserstein distance, leading to their success.

problem Approximating Wasserstein distance in deep generative models.
method Analysis of differences between theoretical setup and training reality.
result Wasserstein GANs' success is due to their failure to approximate Wasserstein distance.

Supervised learning requires the specification of a loss function to minimise. While the theory of admissible losses from both a computational and statistical perspective is well-developed, these offer a panoply of different choices. In practice, this choice is typically made in an \emph{ad hoc} manner. In hopes of mak…

2020-02-10abs ↗pdf ↗

Bayesian adversaries can outsmart traditional adversarial attacks.

problem Bayesian adversaries can manipulate machine learning models through small perturbations.
method Developed a continuous-time particle system (Abram) to approximate the gradient flow of the Bayesian adversarial robustness problem.
result Abram approximates the minimizer of the Bayesian adversarial robustness problem under certain assumptions.

Prevalidated ridge regression simplifies logistic regression for high-dimensional data.

problem Efficient probabilistic classification in high-dimensional data with logistic regression.
method Developed a prevalidated ridge regression model that matches logistic regression's performance but is more computationally efficient.
result Prevalidated ridge regression achieves similar classification error and log-loss to logistic regression for high-dimensional data.

Study optimizes Bitcoin futures hedging to reduce liquidation risk.

problem Optimizing hedging strategies to minimize liquidation risk in Bitcoin futures.
method Derived a semi-closed form optimal hedging strategy considering spot and futures extreme returns, loss aversion, leverage, and collateral management.
result Optimal strategy reduces both hedged portfolio variance and liquidation probability.

Study characterizes hulls and capacities on Riemannian manifolds, proving isoperimetric inequalities.

problem Characterizing hulls and capacities on Riemannian manifolds.
method Investigates strictly outward minimising hulls and uses p-capacities to recover their areas.
result Sharp isoperimetric inequality on complete noncompact manifolds with nonnegative Ricci curvature.

Mirror flow optimizes separable data problems, converging to a maximum margin classifier.

problem Optimizing classification problems with separable data using mirror flow.
method Examine mirror flow on linearly separable classification problems, focusing on the horizon function of the mirror potential.
result Mirror flow converges to a maximum margin classifier for separable data under certain conditions.

The study analyzes multi-class teacher-student perceptron performance and generalization errors.

problem Analyzing multi-class classification with the teacher-student perceptron.
method Deriving asymptotic expressions for Bayes-optimal and empirical risk minimization (ERM) generalization errors.
result Regularised cross-entropy minimization yields close-to-optimal accuracy for multi-class classification.

In this paper, we develop a new aligned vertex convolutional network model to learn multi-scale local-level vertex features for graph classification. Our idea is to transform the graphs of arbitrary sizes into fixed-sized aligned vertex grid structures, and define a new vertex convolution operation by adopting a set of…

2019-02-26abs ↗pdf ↗

Characterizes harmonic morphisms preserving minimal submanifolds and finds novel area-minimising hypercones.

problem Understanding harmonic morphisms and their relationship to minimal submanifolds.
method Characterization of harmonic morphisms as weakly horizontally conformal maps preserving minimal submanifold equations, derivation of reduction properties for other co-dimensions, application to find novel area-minimising hypercones.
result Novel family of degree 4 area-minimising hypercones in R^m, m≥32.

We study the Calabi functional on a ruled surface over a genus two curve. For polarisations which do not admit an extremal metric we describe the behaviour of a minimising sequence splitting the manifold into pieces. We also show that the Calabi flow starting from a metric with suitable symmetry gives such a minimising…

2007-03-19abs ↗pdf ↗

Proves strict inequality for minimizers of Willmore energy under isoperimetric constraints.

problem Minimizing the Willmore energy under isoperimetric constraints.
method Connected sum approach, building on previous work by Keller-Mondino-Rivière.
result Existence of minimizers for the isoperimetric constrained Willmore problem in every genus.

Current approaches in approximate inference for Bayesian neural networks minimise the Kullback-Leibler divergence to approximate the true posterior over the weights. However, this approximation is without knowledge of the final application, and therefore cannot guarantee optimal predictions for a given task. To make mo…

2018-05-10abs ↗pdf ↗

The Duffing oscillator's parameters are identified online using variational message passing.

problem Estimating parameters of a nonlinear Duffing oscillator in real-time.
method Variational message passing on a factor graph of the Duffing oscillator's generative model.
result The online inference procedure performs as well as offline methods.

We consider a one-period Kyle (1985) framework where the insider can be subject to a penalty if she trades. We establish existence and uniqueness of equilibrium for virtually any penalty function when noise is uniform. In equilibrium, the demand of the insider and the price functions are in general non-linear and remai…

2018-09-20abs ↗pdf ↗

The paper analyzes greedy algorithms for MMD minimization, showing their efficiency and approximation error.

problem Minimizing Maximum Mean Discrepancy (MMD) for probability measure quantization.
method Iterative algorithms including kernel herding, greedy MMD minimization, and Sequential Bayesian Quadrature (SBQ).
result The greedy algorithms have a lower approximation error than SBQ, but are significantly faster.