We examine the question of when and how parametric models are most useful in reinforcement learning. In particular, we look at commonalities and differences between parametric models and experience replay. Replay-based learning algorithms share important traits with model-based approaches, including the ability to plan…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
We present a new replay-based method of continual classification learning that we term "conditional replay" which generates samples and labels together by sampling from a distribution conditioned on the class. We compare conditional replay to another replay-based continual learning paradigm (which we term "marginal rep…
D-CBRS manages memory for continual learning by accounting for intra-class diversity.
CLOPS improves deep learning for continuous physiological data.
Paper proposes a new method to select memory data for online class-incremental learning.
Develops a framework for continual learning in anomaly detection.
This project combines recent advances in experience replay techniques, namely, Combined Experience Replay (CER), Prioritized Experience Replay (PER), and Hindsight Experience Replay (HER). We show the results of combinations of these techniques with DDPG and DQN methods. CER always adds the most recent experience to th…
Modern deep reinforcement learning methods have departed from the incremental learning required for eligibility traces, rendering the implementation of the -return difficult in this context. In particular, off-policy methods that utilize experience replay remain problematic because their random sampling of minibatch…
La-MAML improves fast online continual learning with a look-ahead approach.
CAM-GAN improves GANs for continual learning with efficient feature map transformations.
This research tackles imbalanced continual learning with a new sampling strategy.
Neural network tackles continual learning with neuromodulation and local error signals.
The paper improves reinforcement learning stability and efficiency with a new theoretical framework.
ER-GNN uses experience replay to prevent GNNs from forgetting previous tasks.
SynthER uses generative models to augment limited RL experience.
Geometric approach combines asset returns and investor views for better portfolio optimization.
Paper proposes an alternative method to price American options using HJM approach.
We study inference and learning based on a sparse coding model with `spike-and-slab' prior. As in standard sparse coding, the model used assumes independent latent sources that linearly combine to generate data points. However, instead of using a standard sparse prior such as a Laplace distribution, we study the applic…
Quantum machine learning: Adiabatic quantum SVM outperforms classical methods.
We develop a semi-analytic approach to the valuation of auto-callable structures with accrual features subject to barrier conditions. Our approach is based on recent studies of multi-assed binaries, present in the literature. We extend these studies to the case of time-dependent parameters. We compare numerically the s…
Two ML approaches learn local volatility surfaces from option prices, with GP being arbitrage-free.
This paper critiques the Standardized Measurement Approach (SMA) for operational risk and recommends maintaining Advanced Measurement Approach (AMA).
Distributed Acoustic Sensing (DAS) using fiber optic cables is a promising new technology for pipeline monitoring and protection. In this work, we applied and compared two approaches for event detection using DAS: Classic machine learning approach and the approach based on image processing and deep learning. Although w…
A new Euclidean approach reveals the pentagram map's beauty.
Two approaches extend knowledge distillation to Gaussian Processes, showing relationships to existing methods.
We discuss the relative merits of optimistic and randomized approaches to exploration in reinforcement learning. Optimistic approaches presented in the literature apply an optimistic boost to the value estimate at each state-action pair and select actions that are greedy with respect to the resulting optimistic value f…
Classical approaches for approximate inference depend on cleverly designed variational distributions and bounds. Modern approaches employ amortized variational inference, which uses a neural network to approximate any posterior without leveraging the structures of the generative models. In this paper, we propose Amorti…
Online boosting method improves weak to strong learner.
Predicting potential credit default accounts in advance is challenging. Traditional statistical techniques typically cannot handle large amounts of data and the dynamic nature of fraud and humans. To tackle this problem, recent research has focused on artificial and computational intelligence based approaches. In this …
In this paper, we present a new wrapper feature selection approach based on Jensen-Shannon (JS) divergence, termed feature selection with maximum JS-divergence (FSMJ), for text categorization. Unlike most existing feature selection approaches, the proposed FSMJ approach is based on real-valued features which provide mo…
Two new methods for option pricing without or with a riskless asset.
An integrated and extendable approach for stress-testing loan portfolios
Bayesian symbolic regression automates model discovery from data.
New approach interprets Nyström for kernel machines with geometric insight.
Common Representation Learning (CRL), wherein different descriptions (or views) of the data are embedded in a common subspace, is receiving a lot of attention recently. Two popular paradigms here are Canonical Correlation Analysis (CCA) based approaches and Autoencoder (AE) based approaches. CCA based approaches learn …
Two approaches detect EV charging patterns at stations.
The floating body approach to affine surface area is adapted to a holomorphic context providing an alternate approach to Fefferman's invariant hypersurface measure.
We review classical approaches to the problem of isometrically embedding a Riemannian surface into Euclidean 3-space, including coordinate-based approaches exposited by Darboux and Eisenhart, as well as the moving-frames based approaches advocated by Cartan. In particular, the first approach involves reducing the probl…
Paper evaluates CNN-based facial landmark detection methods.
This paper provides a comprehensive benchmark and taxonomy for certifiably robust DNN defenses.
VB approach for dynamic network models improves efficiency and accuracy.
New approach for prudent risk evaluation using model aggregation.
Study proposes a new approach for deep hedging using artificial market simulations.
Some of recent developments, including recent results, ideas, techniques, and approaches, in the study of degenerate partial differential equations are surveyed and analyzed. Several examples of nonlinear degenerate, even mixed, partial differential equations, are presented, which arise naturally in some longstanding, …
We compare several approaches to learn an Optimal Map, represented as a neural network, between probability distributions. The approaches fall into two categories: ``Heuristics'' and approaches with a more sound mathematical justification, motivated by the dual of the Kantorovitch problem. Among the algorithms we consi…
This paper examines how optimization methods affect the reliability of detecting inputs outside a model's training distribution.
Deep Neural Networks have shown tremendous success in the area of object recognition, image classification and natural language processing. However, designing optimal Neural Network architectures that can learn and output arbitrary graphs is an ongoing research problem. The objective of this survey is to summarize and …
In this paper, we propose three approaches for the estimation of the Tucker decomposition of multi-way arrays (tensors) from partial observations. All approaches are formulated as convex minimization problems. Therefore, the minimum is guaranteed to be unique. The proposed approaches can automatically estimate the numb…