InfoNCE objective is equivalent to ELBO in RPM, linking to self-supervised learning.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Automatic continuous speech recognition (CSR) is sufficiently mature that a variety of real world applications are now possible including large vocabulary transcription and interactive spoken dialogues. This paper reviews the evolution of the statistical modelling techniques which underlie current-day systems, specific…
We introduce a new approach to learning in hierarchical latent-variable generative models called the "distributed distributional code Helmholtz machine", which emphasises flexibility and accuracy in the inferential process. In common with the original Helmholtz machine and later variational autoencoder algorithms (but …
Probabilistic programming has emerged as a powerful paradigm in statistics, applied science, and machine learning: by decoupling modelling from inference, it promises to allow modellers to directly reason about the processes generating data. However, the performance of inference algorithms can be dramatically affected …
We show how to specify preferred parameterisations on a homogeneous curve in an arbitrary homogeneous space. We apply these results to limit the natural parameters on distinguished curves in parabolic geometries.
Wide neural networks with asymmetrical node scaling converge globally and learn features.
Extends hyperparameter transfer across model sizes and modules, improving training speed.
Just as an explicit parameterisation of system dynamics by state, i.e., a choice of coordinates, can impede the identification of general structure, so it is too with an explicit parameterisation of system dynamics by control. However, such explicit and fixed parameterisation by control is commonplace in control theory…
The Gaussian process state space model (GPSSM) is a non-linear dynamical system, where unknown transition and/or measurement mappings are described by GPs. Most research in GPSSMs has focussed on the state estimation problem, i.e., computing a posterior of the latent state given the model. However, the key challenge in…
In this article we propose a generalisation of the recent work of Gatheral and Jacquier on explicit arbitrage-free parameterisations of implied volatility surfaces. We also discuss extensively the notion of arbitrage freeness and Roger Lee's moment formula using the recent analysis by Roper. We further exhibit an arbit…
New approach improves linear-time attention for language models.
We introduce Clique Matrices as an alternative representation of undirected graphs, being a generalisation of the incidence matrix representation. Here we use clique matrices to decompose a graph into a set of possibly overlapping clusters, de ned as well-connected subsets of vertices. The decomposition is based on a s…
New statistical model improves protein alignment accuracy.
The Industrial Internet of Things drastically increases connectivity of devices in industrial applications. In addition to the benefits in efficiency, scalability and ease of use, this creates novel attack surfaces. Historically, industrial networks and protocols do not contain means of security, such as authentication…
Gaussian process model learns Hamiltonian systems from noisy data.
Gating is a key technique used for integrating information from multiple sources by long short-term memory (LSTM) models and has recently also been applied to other models such as the highway network. Although gating is powerful, it is rather expensive in terms of both computation and storage as each gating unit uses a…
Wide stochastic networks show Gaussian behavior and improve training with PAC-Bayesian methods.
Improves financial instrument pricing using neural networks.
Parameterised actions in reinforcement learning are composed of discrete actions with continuous action-parameters. This provides a framework for solving complex domains that require combining high-level actions with flexible control. The recent P-DQN algorithm extends deep Q-networks to learn over such action spaces. …
Study constant mean curvature tori in R^3 using spectral data and Whitham deformations.
PRZI traders adapt their quote-prices based on a strategy parameter s, affecting market dynamics.
Study on projective structures on a hyperbolic 3-orbifold using tetrahedra.
Automatically learns flexible symmetry constraints in neural networks using gradients.
Study solves sub-Laplacian equivalence on a specific Heisenberg group.
We study here the large-time behaviour of all continuous affine stochastic volatility models (in the sense of Keller-Ressel) and deduce a closed-form formula for the large-maturity implied volatility smile. Based on refinements of the Gartner-Ellis theorem on the real line, our proof reveals pathological behaviours of …
New kernel interprets 3D anisotropic data with rotations and improved predictions.
EmbraceNet fusion model for multi-sensor activity recognition.
With the recent renaissance of deep convolution neural networks, encouraging breakthroughs have been achieved on the supervised recognition tasks, where each class has sufficient training data and fully annotated training data. However, to scale the recognition to a large number of classes with few or now training samp…
A new method for discrete data normalizing flows using latent transformations.
In this work, we explore the dependencies between speaker recognition and emotion recognition. We first show that knowledge learned for speaker recognition can be reused for emotion recognition through transfer learning. Then, we show the effect of emotion on speaker recognition. For emotion recognition, we show that u…
In this paper we investigate whether electroencephalography (EEG) features can be used to improve the performance of continuous visual speech recognition systems. We implemented a connectionist temporal classification (CTC) based end-to-end automatic speech recognition (ASR) model for performing recognition. Our result…
Neural network framework for language recognition considers sequence information and improves accuracy.
Meta-gradient RL learns to learn from experience.
This paper presents a novel method for structural data recognition using a large number of graph models. In general, prevalent methods for structural data recognition have two shortcomings: 1) Only a single model is used to capture structural variation. 2) Naive recognition methods are used, such as the nearest neighbo…
AV-CPL uses continuous pseudo-labels for AVSR combining labeled and unlabeled data.
Using the parameterisation of the deformation space of GHMC anti-de Sitter structures on by the cotangent bundle of the Teichmüller space of , we study how some geometric quantities, such as the Lorentzian Hausdorff dimension of the limit set, the width of the convex core and the Hölder exponen…
A new approach to unsupervised learning using recognition-parametrised models.
Long Short-Term Memory (LSTM) is a recurrent neural network (RNN) architecture that has been designed to address the vanishing and exploding gradient problems of conventional RNNs. Unlike feedforward neural networks, RNNs have cyclic connections making them powerful for modeling sequences. They have been successfully u…
In this paper we demonstrate end-to-end continuous speech recognition (CSR) using electroencephalography (EEG) signals with no speech signal as input. An attention model based automatic speech recognition (ASR) and connectionist temporal classification (CTC) based ASR systems were implemented for performing recognition…
Bayesian approach optimizes quantum circuits for noisy hardware.
A comparative study of the application of Gaussian Mixture Model (GMM) and Radial Basis Function (RBF) in biometric recognition of voice has been carried out and presented. The application of machine learning techniques to biometric authentication and recognition problems has gained a widespread acceptance. In this res…
Todays interactive devices such as smart-phone assistants and smart speakers often deal with short-duration speech segments. As a result, speaker recognition systems integrated into such devices will be much better suited with models capable of performing the recognition task with short-duration utterances. In this pap…
Today's proliferation of powerful facial recognition systems poses a real threat to personal privacy. As Clearview.ai demonstrated, anyone can canvas the Internet for data and train highly accurate facial recognition models of individuals without their knowledge. We need tools to protect ourselves from potential misuse…
VoiceFilter-Lite separates speech from background in real-time for on-device speech recognition.
TransFall uses transfer learning to improve activity recognition from mobile sensors.
This paper presents a comparison of a traditional hybrid speech recognition system (kaldi using WFST and TDNN with lattice-free MMI) and a lexicon-free end-to-end (TensorFlow implementation of multi-layer LSTM with CTC training) models for German syllable recognition on the Verbmobil corpus. The results show that expli…
Directly estimates CQC, improving interpretability and accuracy.
We study projective structures on a surface having poles of prescribed orders. We obtain a monodromy map from a complex manifold parameterising such structures to the stack of framed local systems on the associated marked bordered surface. We prove that the image of this map is contained in…