Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

68135203270 · Jun 202019922001200920172026
48 results for Adaptive epsilon decay

ALPE improves mid-price forecasting in HFT with real-time data.

problem Real-time mid-price forecasting in high-frequency trading.
method Adaptive Learning Policy Engine (ALPE) using RL and adaptive epsilon decay.
result ALPE outperforms other models in mid-price forecasting.

We obtain a lower asymptotic bound on the decay rate of the probability of a portfolio's underperformance against a benchmark over a large time horizon. It is assumed that the prices of the securities are governed by geometric Brownian motions with the coefficients depending on an economic factor, possibly nonlinearly.…

2016-02-05abs ↗pdf ↗

Bayesian approach improves ε\varepsilon-greedy exploration in RL.

problem Improving ε\varepsilon-greedy exploration in model-free RL.
method Introducing a Bayesian model update for ε\varepsilon based on BMC.
result Proposed ε\varepsilon- exttt{BMC} algorithm efficiently balances exploration and exploitation.

LPF provides formal guarantees for aggregating multi-evidence in probabilistic tasks.

problem Lack of formal guarantees for multi-evidence reasoning in AI.
method LPF uses variational autoencoders and Sum-Product Networks to aggregate evidence items.
result Proves multiple formal guarantees including calibration preservation and error decay.

We define a concordance invariant, epsilon(K), associated to the knot Floer complex of K, and give a formula for the Ozsváth-Szabó concordance invariant tau of K_{p,q}, the (p,q)-cable of a knot K, in terms of p, q, tau(K), and epsilon(K). We also describe the behavior of epsilon under cabling, allowing one to compute …

2012-02-07abs ↗pdf ↗

Epsilon-machines are minimal, unifilar presentations of stationary stochastic processes. They were originally defined in the history machine sense, as hidden Markov models whose states are the equivalence classes of infinite pasts with the same probability distribution over futures. In analyzing synchronization, though…

2011-11-18abs ↗pdf ↗

Paper calculates third coefficient in Kaehler-Einstein metric expansion.

problem Understanding Kaehler-Einstein metrics and their epsilon functions.
method Computes the third coefficient in the TYCZ-expansion of the epsilon function.
result Discovers the vanishing of the third coefficient's significance.

Ozsvath-Stipsicz-Szabo recently defined a one-parameter family, upsilon of K at t, of concordance invariants associated to the knot Floer complex. We compare their invariant to the {-1, 0, 1}-valued concordance invariant epsilon, which is also associated to the knot Floer complex. In particular, we give an example of a…

2014-07-30abs ↗pdf ↗

Regularization in the optimization of deep neural networks is often critical to avoid undesirable over-fitting leading to better generalization of model. One of the most popular regularization algorithms is to impose L-2 penalty on the model parameters resulting in the decay of parameters, called weight-decay, and the …

2019-07-21abs ↗pdf ↗

Adaptive weights improve physics-informed neural networks and deep operator networks.

problem Training physics-informed neural networks and deep operator networks can be challenging, leading to unsatisfactory accuracy and efficiency.
method Proposes a pointwise adaptive weighting method that balances the residual decay rate across different training points.
result Our proposed approach of balanced residual decay rates offers advantages including bounded weights, high prediction accuracy, fast convergence rate, low training uncertainty, low computational cost, and ease of hyperparameter tuning.

This paper investigates the effectiveness of decoupled weight decay at the start of training.

problem The traditional approach to weight decay is not effective throughout training.
method The authors investigate decoupled weight decay, applying it only at the start of training.
result Applying weight decay only at the start of training stabilizes network weights and improves performance.

Adaptive time decay functions improve financial product recommendation accuracy.

problem Inaccurate recommendations due to static historical data in finance.
method Time-dependent collaborative filtering with personalized decay functions.
result Significant improvements over state-of-the-art benchmarks in financial product recommendation.

We analyze how an observer synchronizes to the internal state of a finite-state information source, using the epsilon-machine causal representation. Here, we treat the case of exact synchronization, when it is possible for the observer to synchronize completely after a finite number of observations. The more difficult …

2010-08-25abs ↗pdf ↗

We develop an epsilon-controlled algebraic L-theory, extending our earlier work on epsilon-controlled algebraic K-theory. The controlled L-theory is very close to being a generalized homology theory; we study analogues of the homology exact sequence of a pair, excision properties, and the Mayer--Vietoris exact sequence…

2004-02-13abs ↗pdf ↗

Let M be a closed 3-manifold which can be triangulated with N simplices. We prove that any map from M to a genus 2 surface has Hopf invariant at most C^N. Let X be a closed oriented hyperbolic 3-manifold with injectivity radius less than epsilon at one point. If there is a degree non-zero map from M to X, then we prove…

2007-09-09abs ↗pdf ↗

Minimax optimization plays a key role in adversarial training of machine learning algorithms, such as learning generative models, domain adaptation, privacy preservation, and robust learning. In this paper, we demonstrate the failure of alternating gradient descent in minimax optimization problems due to the discontinu…

2018-05-29abs ↗pdf ↗

Proposes a new approach to regression learning that addresses overfitting and underfitting.

problem Regression learning issues, including overfitting and underfitting.
method Introduces epsilon-Confidence Approximately Correct (epsilon CoAC) framework using Kullback Leibler divergence.
result Demonstrates improved learnability and accuracy compared to cross-validation.

For any family of measurable sets in a probability space, we show that either (i) the family has infinite Vapnik-Chervonenkis (VC) dimension or (ii) for every epsilon > 0 there is a finite partition pi such the pi-boundary of each set has measure at most epsilon. Immediate corollaries include the fact that a family wit…

2010-10-21abs ↗pdf ↗

A near-identity nilpotent pseudogroup of order m >= 1 is a family f_1, ..., f_n: (-1,1) -> R of C^2 functions for which: |f_i - id|_{C^1} < epsilon for some small positive real number epsilon < 1/10^{m+1} and commutators of the functions f_i of order at least m equal the identity. We present a classification of near-id…

2004-04-07abs ↗pdf ↗

Proposes a new adversarial model to avoid accuracy vs. adversarial accuracy tradeoff.

problem Inherent tradeoff between accuracy and adversarial accuracy in existing adversarial robustness definitions.
method Introduces Voronoi-epsilon adversary that balances perturbation constraints.
result Voronoi-epsilon adversary avoids accuracy vs. adversarial accuracy tradeoff even with large εε.

ε-Consistent Mixup improves semi-supervised classification accuracy.

problem Improving semi-supervised classification accuracy with limited labeled data.
method Combines Mixup's linear interpolation with consistency regularization, using an adaptive tradeoff between the two.
result ε-Consistent Mixup yields the largest gains in low label-availability scenarios.

Compactness theorems for G2G_2-solitons established with scalar curvature and potential function constraints.

problem Establishing compactness theorems for G2G_2-solitons under specific conditions.
method Proved Gromov-Hausdorff convergence and derived epsilon-regularity estimates.
result Smooth convergence of G2G_2-solitons under uniform energy bounds at half the dimension.

The paper explores MAB strategies for very short horizons, introducing new methods and showing improved performance.

problem Short horizon multi-armed bandit problems in games.
method Regression oracles, forced exploration, UCBT strategy.
result Combination of epsilon-greedy or epsilon-decreasing with regression oracles outperforms other strategies.

In recent years there has been an increasing interest in learning Bayesian networks from data. One of the most effective methods for learning such networks is based on the minimum description length (MDL) principle. Previous work has shown that this learning procedure is asymptotically successful: with probability one,…

2013-02-13abs ↗pdf ↗

The regularization path of the Lasso can be shown to be piecewise linear, making it possible to "follow" and explicitly compute the entire path. We analyze in this paper this popular strategy, and prove that its worst case complexity is exponential in the number of variables. We then oppose this pessimistic result to a…

2012-05-01abs ↗pdf ↗

Estimates matrix trace optimization with statistical learning theory.

problem Optimizing trace of parameter-dependent matrices.
method Monte Carlo estimator with bounds derived from epsilon nets and generic chaining.
result Predicts small sampling amount for matrices with small off-diagonal mass.

Exploiting a relationship between closed geodesics on a generic closed hyperbolic surface S and a certain unipotent flow on the product space T_1(S) x T_1(S), we obtain a local asymptotic equidistribution result for long closed geodesics on S. Applications include asymptotic estimates for the number of pants immersions…

2005-05-23abs ↗pdf ↗

An SU(3)- or SU(1,2)-structure on a 6-dimensional manifold N^6 can be defined as a pair of a 2-form omega and a 3-form rho. We prove that any analytic SU(3)- or SU(1,2)-structure on N^6 with d omega^2 =0 can be extended to a parallel Spin(7)- or Spin_0(3,4)-structure Phi that is defined on the trivial disc bundle N^6\t…

2010-10-08abs ↗pdf ↗

New research shows fixed-budget best-arm identification cannot match static oracle performance.

problem Fixed-budget best-arm identification's performance limitations.
method Analysis of various adaptive and static algorithms for best-arm identification.
result For any algorithm, there exists at least one instance where the error decay rate is at most \((1 + \frac{\log(K)}{8})^{-1}\) times that of the static oracle.

We introduce a Bayesian approach to discovering patterns in structurally complex processes. The proposed method of Bayesian Structural Inference (BSI) relies on a set of candidate unifilar HMM (uHMM) topologies for inference of process structure from a data series. We employ a recently developed exact enumeration of to…

2013-09-05abs ↗pdf ↗

Deep neural networks are traditionally trained using human-designed stochastic optimization algorithms, such as SGD and Adam. Recently, the approach of learning to optimize network parameters has emerged as a promising research topic. However, these learned black-box optimizers sometimes do not fully utilize the experi…

2018-11-22abs ↗pdf ↗