Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

4590135180 · Jun 202019922001200920172026
48 results for length averages

Non-parametric estimators improve quickest changepoint detection under irregular sequence lengths.

problem Limited and irregular sequence lengths hinder application of ARL and ADD in QCD.
method Analogies with survival analysis to model detection probabilities under truncation.
result KM-ARL and KM-ADD non-parametric estimators are asymptotically unbiased.

Paper finds optimal shapes for minimizing average lengths of billiard trajectories in specific polygons.

problem Finding optimal shapes to minimize the average length of billiard trajectories.
method Used techniques from Teichmüller theory.
result Optimal shapes minimize average lengths of billiard trajectories in specific polygons.

Since the pioneering work of Ghys, Langevin and Walczak among others, it has been known that several methods of dynamical systems theory can be adopted to study of foliations. Our aim in this paper is to investigate complexity of foliations, by generalising existence problem of time averages in dynamical systems theory…

2018-10-17abs ↗pdf ↗

SummerTime summarizes variable-length time series for machine learning applications.

problem Classical machine learning methods struggle with variable-length time series data.
method Summarizes time series into a fixed-length feature vector using Gaussian Mixture Models (GMM).
result Improves classification and regression performance in physical activity analysis.

Two-Tailed Averaging improves generalization by optimizing the number of leading iterates to ignore.

problem Improving generalization in stochastic optimization with limited resources and hyperparameters.
method An anytime adaptive algorithm that balances the number of leading iterates to ignore for better generalization.
result Approximates the optimal tail at all optimization steps, improving generalization without hyperparameters.

Distributed statistical learning problems arise commonly when dealing with large datasets. In this setup, datasets are partitioned over machines, which compute locally, and communicate short messages. Communication is often the bottleneck. In this paper, we study one-step and iterative weighted parameter averaging in s…

2018-09-30abs ↗pdf ↗

Average signature measures geodesics in Lie groups.

problem Understanding geometric properties of Lie groups through geodesic paths.
method Introducing average signature A(G)\mathbb A(G) and using it with trace operation to recover geometric properties.
result Average signature can recover geometric properties like dimension, diameter, volume, and scalar curvature.
Agents Play Mix-gamephysics.soc-ph

In mix-game which is an extension of minority game, there are two groups of agents; group1 plays the majority game, but the group2 plays the minority game. This paper studies the change of the average winnings of agents and volatilities vs. the change of mixture of agents in mix-game model. It finds that the correlatio…

2005-05-17abs ↗pdf ↗

The variance reduction class of algorithms including the representative ones, SVRG and SARAH, have well documented merits for empirical risk minimization problems. However, they require grid search to tune parameters (step size and the number of iterations per inner loop) for optimal performance. This work introduces `…

2019-08-25abs ↗pdf ↗

Length spectral rigidity is the question of under what circumstances the geometry of a surface can be determined, up to isotopy, by knowing only the lengths of its closed geodesics. It is known that this can be done for negatively curved Riemannian surfaces, as well as for negatively-curved cone surfaces. Steps are tak…

2012-07-26abs ↗pdf ↗

Randomized positional encodings boost transformer performance on longer sequences.

problem Transformers struggle with generalizing to sequences of arbitrary length.
method Introduced randomized positional encodings that simulate longer sequences and randomly select positions.
result Randomized positional encodings increase test accuracy by 12.0% on average for sequences of unseen length.

This paper improves forecasts for diverse time series by averaging similar ones.

problem Forecasting challenges in heterogeneous time series.
method Dynamic Time Warping to find similar time series, k-Nearest Neighbor averaging.
result Averaging improves forecasts of simple models.

We investigate the average-case complexity of decision problems for finitely generated groups, in particular the word and membership problems. Using our recent results on ``generic-case complexity'' we show that if a finitely generated group GG has the word problem solvable in subexponential time and has a subgroup of…

2002-06-25abs ↗pdf ↗

New bounds show transformers need longer training for length generalization.

problem Understanding when transformers can generalize to longer inputs.
method Analyzing different settings of transformers, providing quantitative bounds.
result Transformers need training data longer than previously thought for length generalization.

We present a technique for clustering categorical data by generating many dissimilarity matrices and averaging over them. We begin by demonstrating our technique on low dimensional categorical data and comparing it to several other techniques that have been proposed. Then we give conditions under which our method shoul…

2015-06-26abs ↗pdf ↗

It is shown that given any link-manifold, there is an algorithm to decide if the manifold contains an embedded, essential planar surface; if it does, the algorithm will construct one. If a slope on the boundary of the link-manifold is given, there is an algorithm to determine if the slope bounds an embedded punctured-d…

2006-08-28abs ↗pdf ↗

This is an up-to-date introduction to and overview of the Minimum Description Length (MDL) Principle, a theory of inductive inference that can be applied to general problems in statistics, machine learning and pattern recognition. While MDL was originally based on data compression ideas, this introduction can be read w…

2019-08-21abs ↗pdf ↗

Deep neural network approximates flow averages for rough walls in multiscale simulations.

problem Approximating flow averages in rough-wall Stokes flow simulations.
method Fourier neural operator for local averages, parameterized by local wall geometry.
result Stable and accurate HMM solution with reduced micro problem solving cost.

Model learns brevity by exposing to easy problems, improving efficiency without explicit length penalties.

problem Excessive verbosity in step-by-step reasoning models trained with RLVR.
method Retaining and up-weighting moderately easy problems as implicit length regularizers.
result Model generates solutions that are, on average, nearly twice as short without explicit length penalties.

R. Schwartz's inequality provides an upper bound for the Schwarzian derivative of a parameterization of a circle in the complex plane and on the potential of Hill's equation with coexisting periodic solutions. We prove a discrete version of this inequality and obtain a version of the planar Blaschke-Santalo inequality …

2010-06-07abs ↗pdf ↗

Study improves queue length estimation from connected vehicles by filtering parameters.

problem Large errors in estimated queue lengths at low market penetration rates.
method Used Kalman and Particle filters as multilevel real-time estimators.
result Filters reduce estimation errors and improve accuracy within 15 minutes.

This paper presents a tensor multiplication based smoothing algorithm that follows a two step denoising method. Unlike other traditional averaging approaches, our approach uses an element based normal voting tensor to compute smooth surfaces. By introducing a binary optimization on the proposed tensor together with a l…

2016-07-20abs ↗pdf ↗

We introduce a variant of Farber's topological complexity, defined for smooth compact orientable Riemannian manifolds, which takes into account only motion planners with the lowest possible "average length" of the output paths. We prove that it never differs from topological complexity by more than 11, thus showing th…

2016-07-04abs ↗pdf ↗

Fix a translation surface XX, and consider the measures on XX coming from averaging the uniform measures on all the saddle connections of length at most RR. Then as RR\to\infty, the weak limit of these measures exists and is equal to the Lebesgue measure on XX. We also show that any weak limit of a subsequence of …

2017-05-30abs ↗pdf ↗

The determinants of the velocity of money have been examined based on life-cycle hypothesis. The velocity of money can be expressed by reciprocal of the average value of holding time which is defined as interval between participating exchanges for one unit of money. This expression indicates that the velocity is govern…

2005-07-21abs ↗pdf ↗

An online framework optimizes efficiency in conformal prediction with a target miscoverage rate.

problem Achieving coverage and minimizing interval length in a sequential, online setting.
method Optimizes efficiency by directly optimizing the average length of intervals while maintaining coverage.
result Shows a gap between optimal performance for exchangeable and arbitrary sequences, and provides a matching algorithm for the Pareto-optimal settings.

We define a new class of knot energies (known as renormalization energies) and prove that a broad class of these energies are uniquely minimized by the round circle. Most of O'Hara's knot energies belong to this class. This proves two conjectures of O'Hara and of Freedman, He, and Wang. We also find energies not minimi…

2001-05-16abs ↗pdf ↗

BCI provides calibrated prediction intervals for time series forecasts.

problem Calibration of prediction intervals for time series forecasts.
method BCI wraps around any time series forecasting models and optimizes interval lengths using dynamic programming.
result BCI achieves long-term coverage under arbitrary distribution shifts and temporal dependence.

Recurrent Neural Networks (RNN) are a type of statistical model designed to handle sequential data. The model reads a sequence one symbol at a time. Each symbol is processed based on information collected from the previous symbols. With existing RNN architectures, each symbol is processed using only information from th…

2017-03-03abs ↗pdf ↗

A new control chart detects shifts in binary data streams quickly and reliably.

problem Early detection of small shifts in multiple binary data streams.
method Cumulative Standardized Binomial EWMA (CSB-EWMA) chart with exact variance derivation.
result Adaptive control limits ensure robust detection across different data distributions.

In this paper we will study the statistics of the unit geodesic flow normal to the boundary of a hyperbolic manifold with non-empty totally geodesic boundary. Viewing the time it takes this flow to hit the boundary as a random variable, we derive a formula for its moments in terms of the orthospectrum. The first moment…

2013-03-26abs ↗pdf ↗

We study the curve diffusion flow for closed curves immersed in the Minkowski plane M\mathcal{M}, which is equivalent to the Euclidean plane endowed with a closed, symmetric, convex curve called an indicatrix that scales the length of a vector in M\mathcal{M} depending on its length. The indiactrix $\partial\mathcal{…

2017-06-07abs ↗pdf ↗

Deep RL algorithm trades high-dimensional stock portfolios.

problem Trading high-dimensional stock portfolios with data gaps and non-unique history lengths.
method Deep Q-learning algorithm, sequentially setting up environments, rewarding based on asset returns and cash reservation.
result Algorithm outperforms all passive and active benchmarks by a large margin.