Two machine learning models detect anomalies in ER claims, saving up to 40% in improper payments.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Experience replay (ER) is a fundamental component of off-policy deep reinforcement learning (RL). ER recalls experiences from past iterations to compute gradient estimates for the current policy, increasing data-efficiency. However, the accuracy of such updates may deteriorate when the policy diverges from past behavio…
ReF-ER algorithm improved performance in multi-agent reinforcement learning.
Model for hedging price and quantity risks in electricity markets.
Soft Actor-Critic (SAC) is an off-policy actor-critic deep reinforcement learning (DRL) algorithm based on maximum entropy reinforcement learning. By combining off-policy updates with an actor-critic formulation, SAC achieves state-of-the-art performance on a range of continuous-action benchmark tasks, outperforming pr…
Entity resolution (ER) presents unique challenges for evaluation methodology. While crowdsourcing platforms acquire ground truth, sound approaches to sampling must drive labelling efforts. In ER, extreme class imbalance between matching and non-matching records can lead to enormous labelling requirements when seeking s…
Study assesses environmental management accounting practices in Bangladesh.
ER-GNN uses experience replay to prevent GNNs from forgetting previous tasks.
Deep learning predicts road GHG emissions with speed, density, and past ERs.
Entity resolution (ER; also known as record linkage or de-duplication) is the process of merging noisy databases, often in the absence of unique identifiers. A major advancement in ER methodology has been the application of Bayesian generative models, which provide a natural framework for inferring latent entities with…
We come up with infinite-dimensional prequantum line bundles and moment map interpretations of three different sets of equations - the generalised Monge-Amp`ere equation, the almost Hitchin system, and the Calabi-Yang-Mills equations. These are all perturbations of already existing equations. Our construction for the g…
Extends XCS with Experience Replay for improved sample efficiency in single-step tasks.
Usually considered as a classification problem, entity resolution (ER) can be very challenging on real data due to the prevalence of dirty values. The state-of-the-art solutions for ER were built on a variety of learning models (most notably deep neural networks), which require lots of accurately labeled training data.…
Study compares SPG and PPO for racing games, finding SPG more stable with weighted actions.
Study on median algebra structures on Euclidean spaces and manifolds with local CAT(0) cubulation.
Entity resolution (ER) is the task of identifying records belonging to the same entity (e.g. individual, group) across one or multiple databases. Ironically, it has multiple names: deduplication and record linkage, among others. In this paper we survey metrics used to evaluate ER results in order to iteratively improve…
Proposes simplified SHAP for faster black-box model explanations.
Test assesses if a linear classifier is random or significant.
New algorithm achieves almost exact graph matching in almost quadratic time.
A user-friendly interface constructs effective background knowledge from ER diagrams.
Identifying latent structure in large data matrices is essential for exploring biological processes. Here, we consider recovering gene co-expression networks from gene expression data, where each network encodes relationships between genes that are locally co-regulated by shared biological mechanisms. To do this, we de…
LiDER refreshes past experiences in RL by dreaming about them.
Solves complex Monge-Ampère equations on Kähler manifolds.
Entity resolution (ER) is one of the fundamental problems in data integration, where machine learning (ML) based classifiers often provide the state-of-the-art results. Considerable human effort goes into feature engineering and training data creation. In this paper, we investigate a new problem: Given a dataset D_T fo…
ERDMD discovers sparse, nonuniformly timed DMD models from chaotic attractors.
We prove existence and uniqueness results for patterns of circles with prescribed intersection angles in constant curvature surfaces. Our method is based on two new functionals--one for the Euclidean and one for the hyperbolic case. We show how Colin de Verdi`ere's, Br"agger's and Rivin's functionals can be derived fro…
A simple technique improves continual learning by 50% on image datasets.
While a wide range of interpretable generative procedures for graphs exist, matching observed graph topologies with such procedures and choices for its parameters remains an open problem. Devising generative models that closely reproduce real-world graphs requires domain knowledge and time-consuming simulation. While e…
Conventional prior for Variational Auto-Encoder (VAE) is a Gaussian distribution. Recent works demonstrated that choice of prior distribution affects learning capacity of VAE models. We propose a general technique (embedding-reparameterization procedure, or ER) for introducing arbitrary manifold-valued variables in VAE…
On 2 November 2009, the Financial Bubble Experiment was launched within the Financial Crisis Observatory (FCO) at ETH Zurich (\url{http://www.er.ethz.ch/fco/}). In that initial report, we diagnosed and announced three bubbles on three different assets. In this latest release of 23 December 2009 in this ongoing experime…
The proof of Theorem 7.12 of "Uniqueness of smooth cohomology theories" by the authors of this note is not correct. The said theorem identifies the flat part of a differential extension of a generalized cohomology theory E with ER/Z (there called "smooth extension"). In this note, we give a correct proof. Moreover, we …
An automatic machine learning (AutoML) task is to select the best algorithm and its hyper-parameters simultaneously. Previously, the hyper-parameters of all algorithms are joint as a single search space, which is not only huge but also redundant, because many dimensions of hyper-parameters are irrelevant with the selec…
Radiomics aims to extract and analyze large numbers of quantitative features from medical images and is highly promising in staging, diagnosing, and predicting outcomes of cancer treatments. Nevertheless, several challenges need to be addressed to construct an optimal radiomics predictive model. First, the predictive p…
We improve a graph generation model to accurately recover Barabási-Albert graph parameters.
Graph neural networks denote a group of neural network models introduced for the representation learning tasks on graph data specifically. Graph neural networks have been demonstrated to be effective for capturing network structure information, and the learned representations can achieve the state-of-the-art performanc…
Expectile regression is a nice tool for investigating conditional distributions beyond the conditional mean. It is well-known that expectiles can be described with the help of the asymmetric least square loss function, and this link makes it possible to estimate expectiles in a non-parametric framework by a support vec…
A Z-structure on a group G, defined by M. Bestvina, is a pair (\hat{X}, Z) of spaces such that \hat{X} is a compact ER, Z is a Z-set in \hat{X}, G acts properly and cocompactly on X=\hat{X}\Z, and the collection of translates of any compact set in X forms a null sequence in \hat{X}. It is natural to ask whether a given…
DIAL learns embeddings to maximize recall and accuracy for entity resolution.
ERTS uses Thompson sampling for Gaussian entropic risk bandits, achieving regret bounds.
In this work we deal with coverings and actions of Lie group- groupoids being a sort of the structured Lie groupoids. Firstly, we define an action of a Lie group-groupoid on some Lie group and the smooth coverings of Lie group-groupoids. Later, we show the equivalence of the category of smooth actions of Lie group-grou…
We present a unifying framework for designing and analysing distributional reinforcement learning (DRL) algorithms in terms of recursively estimating statistics of the return distribution. Our key insight is that DRL algorithms can be decomposed as the combination of some statistical estimator and a method for imputing…
We study asymptotically harmonic manifolds of negative curvature, without any cocompactness or homogeneity assumption. We show that asymptotic harmonicity provides a lot of information on the asymptotic geometry of these spaces: in particular, we determine the volume entropy, the spectrum and the relative densities of …
Unified framework for image restoration using equivariant denoisers.
This paper proposes a novel policy for a group of agents to, individually as well as collectively, solve a multi armed bandit (MAB) problem. The policy relies solely on the information that an agent has obtained through sampling of the options on its own and through communication with neighbors. The option selection po…
The almost complex Lie algebroids over smooth manifolds are introduced in the paper. In the first part we give some examples and we obtain a Newlander-Nirenberg type theorem on almost complex Lie algebroids. Next the almost Hermitian Lie algebroids and some related structures on the associated complex Lie algebroid are…
Generative adversarial networks (GANs) are pow- erful generative models based on providing feed- back to a generative network via a discriminator network. However, the discriminator usually as- sesses individual samples. This prevents the dis- criminator from accessing global distributional statistics of generated samp…
Growing interest in automatic speaker verification (ASV)systems has lead to significant quality improvement of spoofing attackson them. Many research works confirm that despite the low equal er-ror rate (EER) ASV systems are still vulnerable to spoofing attacks. Inthis work we overview different acoustic feature spaces…
This work tackles continual learning with semi-supervised data, showing that even with minimal labeled data, performance can match full-supervised methods.