Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

90180270360 · Jun 202019922001200920172026
48 results for value lower-bound

We give three lower bounds for the Morse index of a constant mean curvature torus in Euclidean 3-space in terms of its spectral genus g. The first two lower bounds grow linearly in g and are stronger for smaller values of g, while the third grows quadratically in g but is weaker for smaller values of g.

2004-10-06abs ↗pdf ↗

This paper improves reinforcement learning policies in a scalable way.

problem Ensuring monotonic policy improvement in entropy-regularized RL.
method Derives an entropy-aware lower bound and proposes a novel RL algorithm.
result Demonstrates effectiveness in continuous-state tasks using a linear function approximator.

BCPO optimizes offline RL policies by converting uncertainty into conservative bounds.

problem Offline RL's fragility under distribution shifts and model errors.
method Bayesian approach with credible lower bounds and KL regularization.
result BCPO yields an uncertainty-calibrated policy that avoids exploiting model errors.

Study on identifying most preferred policy in bandits with vector-valued rewards.

problem Identifying the most preferred policy in bandits with vector-valued rewards.
method Derive a novel lower bound on sample complexity, design the Preference-based Track and Stop (PreTS) algorithm, and derive a new concentration inequality.
result The sample complexity of PreTS is asymptotically tight.

New theory of sensitivity for unbiased estimators using Wasserstein geometry.

problem Estimating the instability of estimators under small perturbations.
method Developed a new theory based on Wasserstein geometry, analogous to classical Cramér-Rao theory.
result Wasserstein-Cramér-Rao lower bound for sensitivity of unbiased estimators.

The study optimizes polynomial regression for learning under Gaussian distributions.

problem Agnostic learning of Boolean and real-valued functions under Gaussian distributions.
method LP duality and polynomial degree analysis for L1L^1-regression.
result Optimal SQ lower bounds for various function classes.

New research shows exponential lower bounds for planning in MDPs with linearly-realizable optimal action-value functions.

problem Determining the minimum number of queries needed for sound planners in MDPs with linear function approximation.
method Analyzing fixed-horizon and discounted MDPs with a generative model, showing lower bounds on the number of queries required.
result Sound planners need at least exponential number of queries in both fixed-horizon and discounted settings.

CQL learns conservative Q-functions to improve offline RL performance.

problem Leveraging large, static datasets in reinforcement learning without further interaction.
method Conservative Q-learning (CQL) which learns a conservative Q-function to lower-bound policy values.
result CQL substantially outperforms existing offline RL methods, often achieving 2-5 times higher final returns.

The paper sets lower bounds for a Kirby-Thompson invariant of 4-manifolds.

problem Determining the Kirby-Thompson invariant of specific 4-manifolds.
method Using trisections, the paper establishes lower bounds and calculates the invariant for specific examples.
result The paper calculates the Kirby-Thompson invariant of the spin of L(2,1)L(2,1) and shows the existence of 4-manifolds with arbitrarily large invariants.

We estimate risk measures in Markov cost processes with lower and upper bounds.

problem Estimating risk measures in infinite-horizon discounted costs within Markov processes.
method Truncation scheme and lower/upper bounds for CVaR and variance estimation.
result Upper and lower bounds for CVaR and variance estimation match up to logarithmic factors.

In his work on singularities, expanders and topology of maps, Gromov showed, using isoperimetric inequalities in graded algebras, that every real valued map on the nn-torus admits a fibre whose homological size is bounded below by some universal constant depending on nn. He obtained similar estimates for maps with va…

2017-03-07abs ↗pdf ↗

The study finds the maximum spectrum of 3D manifolds with lower scalar curvature.

problem Finding the maximum spectrum of 3D manifolds with lower scalar curvature.
method Establishing an analogous result to Cheng's theorem for 3D manifolds with scalar curvature lower bound.
result A splitting theorem for 3D manifolds with the maximal bottom spectrum.

Berry et al. (1997) initiated the development of the infinite arms bandit problem. They derived a regret lower bound of all allocation strategies for Bernoulli rewards with uniform priors, and proposed strategies based on success runs. Bonald and Proutière (2013) proposed a two-target algorithm that achieves the regret…

2018-05-30abs ↗pdf ↗

For a given knot, we study the minimal number of positive eigenvalues of the double branched cover over spanning surfaces for the knot. The value gives a lower bound for various genera, the dealternating number and the alternation number of knots, and we prove that Batson's bound for the non-orientable 4-genus gives an…

2017-09-17abs ↗pdf ↗

The paper introduces a method to learn and apply value envelopes for faster online reinforcement learning.

problem Accelerating online reinforcement learning using offline data with theoretical grounding.
method A two-stage framework: offline data for learning value bounds, online algorithms for applying them.
result Substantial regret reductions in empirical tests on tabular MDPs.

We introduce a new class of lower bounds on the log partition function of a Markov random field which makes use of a reversed Jensen's inequality. In particular, our method approximates the intractable distribution using a linear combination of spanning trees with negative weights. This technique is a lower-bound count…

2012-03-15abs ↗pdf ↗

The Barankin bound is generalized to the vector case in the mean square error sense. Necessary and sufficient conditions are obtained to achieve the lower bound. To obtain the result, a simple finite dimensional real vector valued generalization of the Riesz representation theorem for Hilbert spaces is given. The bound…

2017-06-30abs ↗pdf ↗

TensorPlan shows an exponential lower bound for planning in MDPs with linearly realizable value functions.

problem Finding an exponential lower bound for planning in MDPs with linearly realizable value functions.
method TensorPlan and a few action lower bound approach.
result An exponentially large lower bound is shown for planning in MDPs with linearly realizable value functions.

We consider the Max KK-Armed Bandit problem, where a learning agent is faced with several stochastic arms, each a source of i.i.d. rewards of unknown distribution. At each time step the agent chooses an arm, and observes the reward of the obtained sample. Each sample is considered here as a separate item with the rewa…

2015-12-23abs ↗pdf ↗

New method estimates optimal Q-values with better accuracy for specific problems.

problem Estimating optimal Q-values in reinforcement learning is difficult and varies by problem instance.
method Local minimax framework and variance-reduced Q-learning.
result Sharp lower bounds on estimation accuracy for Q-learning.

Let G, a subset of O(4), act isometrically on the 3-sphere. In this article we calculate a lower bound for the diameter of the quotient spaces S3/GS^3/G. We find it to be 1/2arccos(tan(3π10)3){1/2}\arccos(\frac{\tan(\frac{3 π}{10})}{\sqrt3}), which is exactly the value of the lower bound for diameters of the spherical space forms. In the p…

2007-02-23abs ↗pdf ↗

We study the stochastic multi-armed bandit problem when one knows the value μ()μ^{(\star)} of an optimal arm, as a well as a positive lower bound on the smallest positive gap ΔΔ. We propose a new randomized policy that attains a regret {\em uniformly bounded over time} in this setting. We also prove several lower bound…

2013-02-06abs ↗pdf ↗

S. Nelson, M. Orrison, V. Rivera {\cite{S}} modified Kauffman's construction of bracket. Their invariant ΦXβΦ^β_X takes value in a finite ring Z2[t]/(1+t+t3)Z_2[t]/(1+t+t^3). In this paper, the author generalizes this invariant. The new invariant takes value in a polynomial ring. Furthermore, for a tricolorable link diagram, the au…

2017-02-11abs ↗pdf ↗

The paper improves bounds on knot crossings and tabulates minimal diagrams.

problem Improving bounds on knot crossings and tabulating minimal diagrams.
method Analyzing triple-crossing and delta-crossing numbers, proving tangle existence, generating tables.
result Improved bounds on knot crossings and tabulated minimal diagrams for prime knots up to delta-crossing number 4.

The paper calculates the value of information in high-dimensional decision making.

problem Determining the value of acquiring new information in high-dimensional decision problems.
method Using tools from sub-Gaussian processes and generic chaining for asymptotic analysis.
result Asymptotic results on the expected value of information as dimensionality increases.

New methods for evaluating and optimizing policies in offline RL with unobserved confounders.

problem Evaluating and optimizing policies in the presence of unobserved confounders.
method Characterized settings and algorithms for consistent value estimates and lower bounds, with sample complexity guarantees.
result Proved local convergence guarantees for offline policy improvement.

Transductive learning considers a training set of mm labeled samples and a test set of uu unlabeled samples, with the goal of best labeling that particular test set. Conversely, inductive learning considers a training set of mm labeled samples drawn iid from P(X,Y)P(X,Y), with the goal of best labeling any future sample…

2016-02-09abs ↗pdf ↗

Study sharpens unlinking number bounds for special alternating links.

problem Determining the exact unlinking number for special alternating links.
method Analyzes links in the 3-sphere, focusing on special alternating links and their crossing changes.
result Sharp lower bounds for unlinking number realized by crossing changes in alternating diagrams.

The paper sets limits on the accuracy of macroeconomic forecasts based on statistical moments and trade volumes.

problem Uncertainty in predicting macroeconomic variables like prices and returns.
method Defines theoretical lower bounds of uncertainty and upper limits on forecast accuracy based on statistical moments and trade volumes.
result Accuracy of forecasts of probabilities of macroeconomic variables doesn't exceed Gaussian approximations.