Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,181 papers · 148 categories

Trend · papers per month

220440659879 · Jun 202019922001200920182026
48 results for on-line algorithm

Paper models foreign exchange markets and develops an on-line portfolio selection algorithm.

problem Modeling and predicting returns in foreign exchange markets.
method Matrix-valued time series model, trading matrices, and cross rate method.
result Proves the profitability and universality of the on-line portfolio selection algorithm.

In this paper, we establish a robustification of an on-line algorithm for modelling asset prices within a hidden Markov model (HMM). In this HMM framework, parameters of the model are guided by a Markov chain in discrete time, parameters of the asset returns are therefore able to switch between different regimes. The p…

2013-04-07abs ↗pdf ↗

Research compares ML and Time Series methods for generating trading signals.

problem Efficiency of on-line learning Algorithms in generating trading signals.
method Used technical indicators and ensemble of Random Forests, also Kalman Filter.
result Kalman Filter outperformed Random Forests in on-line learning predictions of stock prices.

On-line portfolio selection has attracted increasing interests in machine learning and AI communities recently. Empirical evidences show that stock's high and low prices are temporary and stock price relatives are likely to follow the mean reversion phenomenon. While the existing mean reversion strategies are shown to …

2012-06-18abs ↗pdf ↗

This paper re-examines the problem of parameter estimation in Bayesian networks with missing values and hidden variables from the perspective of recent work in on-line learning [Kivinen & Warmuth, 1994]. We provide a unified framework for parameter estimation that encompasses both on-line learning, where the model is c…

2013-02-06abs ↗pdf ↗

The problem of on-line off-policy evaluation (OPE) has been actively studied in the last decade due to its importance both as a stand-alone problem and as a module in a policy improvement scheme. However, most Temporal Difference (TD) based solutions ignore the discrepancy between the stationary distribution of the beh…

2017-02-23abs ↗pdf ↗

We consider an on-line system identification setting, in which new data become available at given time steps. In order to meet real-time estimation requirements, we propose a tailored Bayesian system identification procedure, in which the hyper-parameters are still updated through Marginal Likelihood maximization, but …

2016-01-17abs ↗pdf ↗

Most of machine learning deals with vector parameters. Ideally we would like to take higher order information into account and make use of matrix or even tensor parameters. However the resulting algorithms are usually inefficient. Here we address on-line learning with matrix parameters. It is often easy to obtain onlin…

2015-06-16abs ↗pdf ↗

In some applications and in order to address real world situations better, data may be more complex than simple vectors. In some examples, they can be known through their pairwise dissimilarities only. Several variants of the Self Organizing Map algorithm were introduced to generalize the original algorithm to this fra…

2012-12-27abs ↗pdf ↗

SparseMix clusters sparse high dimensional binary data efficiently.

problem Clustering sparse high dimensional binary data.
method SparseMix is a mixture model designed for sparse data, using an on-line Hartigan optimization algorithm.
result SparseMix builds partitions with higher compatibility with reference grouping than related methods.

Generalized cross validation (GCV) is one of the most important approaches used to estimate parameters in the context of inverse problems and regularization techniques. A notable example is the determination of the smoothness parameter in splines. When the data are generated by a state space model, like in the spline c…

2017-06-08abs ↗pdf ↗

It is well-known that the precision of data, hyperparameters, and internal representations employed in learning systems directly impacts its energy, throughput, and latency. The precision requirements for the training algorithm are also important for systems that learn on-the-fly. Prior work has shown that the data and…

2016-07-03abs ↗pdf ↗

Meta-learning consists in learning learning algorithms. We use a Long Short Term Memory (LSTM) based network to learn to compute on-line updates of the parameters of another neural network. These parameters are stored in the cell state of the LSTM. Our framework allows to compare learned algorithms to hand-made algorit…

2016-10-19abs ↗pdf ↗

Paper analyzes gradient descent with noisy data copies for linear regression, showing regularization and acceleration effects.

problem Improving generalization in machine learning through data augmentation with noise.
method Gradient descent with on-line noisy copies for linear regression analysis.
result Training with on-line noisy copies is equivalent to ridge regularization with a specific regularization parameter.

A principle for specialized decision-making divides complex problems into manageable parts.

problem Complex decision-making problems beyond individual capabilities.
method An on-line learning rule that learns a partitioning of the problem space for specialized linear policies.
result The approach solves problems that exceed individual decision-makers' capabilities.

Invariant Kähler metrics on line bundles are derived from the Calabi ansatz.

problem Finding invariant scalar-flat Kähler metrics on line bundles over generalized flag varieties.
method Proved using the Calabi ansatz and uniqueness in each Kähler class.
result Existence of a unique scalar-flat Kähler metric in each Kähler class.

Improved direction finding for closely-spaced sources using iterative ESPRIT.

problem Improving DOA estimation for closely-spaced, uncorrelated and correlated sources.
method Iterative ESPRIT algorithm that incorporates prior knowledge and reduces covariance matrix disturbance.
result Improves DOA estimation accuracy for closely-spaced sources.

New scheme detects anomalies in time series data quickly and accurately.

problem Detecting anomalies in time series data with multi-seasonality and unlabeled data.
method Prediction-driven, unsupervised anomaly detection scheme combining decomposition and inference.
result Our scheme outperforms existing algorithms in AUC metric while maintaining efficiency.

Researchers use statistical physics to model neural network learning dynamics.

problem Understanding the learning dynamics of ReLU neural networks.
method Developed a system of differential equations using statistical physics techniques.
result ReLU networks exhibit distinct learning behavior compared to sigmoidal networks.

We show that a suitable notion of Dirac-Jacobi structure on a generic line bundle LL, is provided by Dirac structures in the omni-Lie algebroid of LL. Dirac-Jacobi structures on line bundles generalize Wade's E1(M)\mathcal E^1 (M)-Dirac structures and unify generic (i.e.~non-necessarily coorientable) precontact distribu…

2015-02-18abs ↗pdf ↗

We revisit the problem of inferring the overall ranking among entities in the framework of Bradley-Terry-Luce (BTL) model, based on available empirical data on pairwise preferences. By a simple transformation, we can cast the problem as that of solving a noisy linear system, for which a ready algorithm is available in …

2016-05-09abs ↗pdf ↗

Automated feature extraction for bearing health monitoring.

problem Predicting mechanical faults in process industries to prevent shutdowns.
method Stacked autoencoder neural network and OSELM for automated feature extraction.
result 100% detection accuracy for bearing health states.

Kernel-based reinforcement learning (KBRL) stands out among reinforcement learning algorithms for its strong theoretical guarantees. By casting the learning problem as a local kernel approximation, KBRL provides a way of computing a decision policy which is statistically consistent and converges to a unique solution. U…

2014-07-21abs ↗pdf ↗

We solve the dynamics of the on-line minority game, with general types of decision noise, using generating functional techniques a la De Dominicis and the temporal regularization procedure of Bedeaux et al. The result is a macroscopic dynamical theory in the form of closed equations for correlation- and response functi…

2001-07-30abs ↗pdf ↗

Conformal prediction uses past experience to determine precise levels of confidence in new predictions. Given an error probability εε, together with a method that makes a prediction y^\hat{y} of a label yy, it produces a set of labels, typically containing y^\hat{y}, that also contains yy with probability 1ε1-ε. Con…

2007-06-21abs ↗pdf ↗

We study holomorphic locally homogeneous geometric structures modelled on line bundles over the projective line. We classify these structures on primary Hopf surfaces. We write out the developing map and holonomy morphism of each of these structures explicitly on each primary Hopf surface.

2009-10-02abs ↗pdf ↗

Paper optimizes traffic signal control for better traffic flow.

problem Optimizing traffic signal control to reduce congestion and improve safety.
method Value-based reinforcement learning with interpretable policy functions (polynomial functions).
result Deep Regulatable Hardmax Q-learning variant reduces vehicle delay by up to 19.4%.

We present a new analysis of the problem of learning with drifting distributions in the batch setting using the notion of discrepancy. We prove learning bounds based on the Rademacher complexity of the hypothesis set and the discrepancy of distributions both for a drifting PAC scenario and a tracking scenario. Our boun…

2012-05-19abs ↗pdf ↗

Twitter, a popular social network, presents great opportunities for on-line machine learning research. However, previous research has focused almost entirely on learning from passively collected data. We study the problem of learning to acquire followers through normative user behavior, as opposed to the mass following…

2015-04-16abs ↗pdf ↗

The paper improves safety in autonomous systems using adversarial learning.

problem Ensuring safety in real-time control systems of autonomous vehicles.
method The paper introduces a dual anomaly detection framework (CFAM and SFAM) using generative adversarial networks (GANs) and video prediction.
result Demonstrated effectiveness on both indoor and outdoor autonomous ground vehicles.

Co-adaptation is a special form of on-line learning where an algorithm A\mathcal{A} must assist an unknown algorithm B\mathcal{B} to perform some task. This is a general framework and has applications in recommendation systems, search, education, and much more. Today, the most common use of co-adaptive algorithms is …

2016-11-29abs ↗pdf ↗

QActor optimizes learning from noisy labeled data streams by querying experts for clean labels.

problem Learning from noisy labeled data in continuous streams with limited oracle queries.
method Combines quality models for filtering and oracle queries for true labels, dynamically adjusting query limits.
result QActor nearly matches optimal accuracy with up to 6% additional ground truth data from experts.