Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

59118177236 · May 202619922001200920172026
48 results for Zero Duality Gap

This paper shows how to solve complex reinforcement learning problems with zero duality gap.

problem Complex reinforcement learning problems with conflicting objectives.
method Formulate as a constrained RL problem and solve using Primal-Dual methods.
result The problem has zero duality gap, making it convex and solvable exactly in the dual domain.

Screening rules allow to early discard irrelevant variables from the optimization in Lasso problems, or its derivatives, making solvers faster. In this paper, we propose new versions of the so-called safe rules\textit{safe rules} for the Lasso. Based on duality gap considerations, our new rules create safe test regions whose d…

2015-05-13abs ↗pdf ↗

New approach reduces simulator exploitation by learning robust models.

problem Simulator exploitation leading to reality gap in reinforcement learning.
method Formulated as a zero-sum minimax game between model player and policy player, providing theoretical guarantees and a convergent active data selection algorithm.
result Reduces prediction error in strategically important regions by 1.5-2.2 times and enables near-optimal real-world performance.

Study duality of zero mean curvature surfaces in Heisenberg group.

problem Understanding the duality of zero mean curvature surfaces in the Lorentzian Heisenberg group.
method Investigation of a transformation surface associated with zero mean curvature surfaces in the Heisenberg group under two metrics.
result Derivation of the Sym formula for the dual surface in both metric cases.

New approach reduces simulator exploitation by improving strategic robustness.

problem Simulator exploitation leading to reality gap between simulation and real-world performance.
method Formulated as a zero-sum minimax game, providing theoretical guarantees and a convergent active data selection algorithm.
result Proves convergence and reduces prediction error in strategically important regions by 1.5-2.2 times.

High dimensional regression benefits from sparsity promoting regularizations. Screening rules leverage the known sparsity of the solution by ignoring some variables in the optimization, hence speeding up solvers. When the procedure is proven not to discard features wrongly the rules are said to be \emph{safe}. In this …

2015-06-11abs ↗pdf ↗

The paper explores the information-theoretic nature of excess risk in machine learning.

problem Understanding the excess risk in machine learning models.
method Formulates the minimax excess risk as a zero-sum game and modifies it to allow swapping of the order of play.
result Proves that under certain conditions, the duality gap is zero, allowing for the application of Bayesian results to provide bounds on minimax excess risk.

This paper studies dynamic stochastic optimization problems parametrized by a random variable. Such problems arise in many applications in operations research and mathematical finance. We give sufficient conditions for the existence of solutions and the absence of a duality gap. Our proof uses extended dynamic programm…

2011-05-04abs ↗pdf ↗

Efficient reinforcement learning for simultaneous-move zero-sum games using optimistic value iteration.

problem Learning optimal strategies in simultaneous-move zero-sum Markov games with function approximation.
method Developed an optimistic variant of least-squares minimax value iteration algorithm for offline and online settings.
result Achieved an upper bound of ildeO(d3H3T) ilde O(\sqrt{d^3 H^3 T}) on duality gap and regret.

In this paper, we extend the T-duality isomorphism by Gualtieri and Cavalcanti, from invariant exact Courant algebroids, to exotic exact Courant algebroids such that the momentum and winding numbers are exchanged, filling in a gap in the literature.

2019-09-16abs ↗pdf ↗

We deconstruct the performance of GANs into three components: 1. Formulation: we propose a perturbation view of the population target of GANs. Building on this interpretation, we show that GANs can be viewed as a generalization of the robust statistics framework, and propose a novel GAN architecture, termed as Cascade …

2019-01-27abs ↗pdf ↗

We formulate a theory of pointed manifolds, accommodating both embeddings and Pontryagin-Thom collapse maps, so as to present a common generalization of Poincaré duality in topology and Koszul duality in En\mathcal{E}_n-algebra.

2014-09-09abs ↗pdf ↗

This paper tackles constrained statistical learning problems by proposing a new approach.

problem Statistical learning problems with constraints are challenging and scarce.
method Directly tackling the constrained problem using finite dimensional parameterizations, sample averages, and duality theory.
result We bound the empirical duality gap, showing the effectiveness of the constrained formulation.

We study the optimal transport between two probability measures on the real line, where the transport plans are laws of one-step martingales. A quasi-sure formulation of the dual problem is introduced and shown to yield a complete duality theory for general marginals and measurable reward (cost) functions: absence of a…

2015-07-02abs ↗pdf ↗

This paper proposes a general duality framework for the problem of minimizing a convex integral functional over a space of stochastic processes adapted to a given filtration. The framework unifies many well-known duality frameworks from operations research and mathematical finance. The unification allows the extension …

2010-06-21abs ↗pdf ↗

In a model free discrete time financial market, we prove the superhedging duality theorem, where trading is allowed with dynamic and semi-static strategies. We also show that the initial cost of the cheapest portfolio that dominates a contingent claim on every possible path ωΩω\in Ω, might be strictly greater than the …

2015-06-22abs ↗pdf ↗

Algorithm identifies correct hypothesis from alternatives in bandit problems.

problem Efficiently identifying the correct hypothesis from a finite set of alternatives in structured stochastic multi-armed bandits.
method Frank-Wolfe Self-Play (FWSP) reformulates the game as a saddle-point problem, using a differential-inclusion argument to prove convergence.
result Convergence of the game value for best-arm identification in linear bandits, with uniform global convergence to the optimal value.

Paper tackles fast convergence for non-convex strongly-concave min-max problems.

problem Non-convex strongly-concave min-max problems in deep learning.
method Proximal stage-based method with PL condition for faster convergence.
result Established fast convergence in primal objective gap and duality gap.

Using our earlier proposal for Ramond-Ramond fields in an H-flux on loop space, we extend the Hori isomorphism of Bouwknegt-Evslin-Mathai from invariant differential forms, to invariant exotic differential forms such that the momentum and winding numbers are exchanged, filling in a gap in the literature. We also extend…

2017-10-19abs ↗pdf ↗

We propose a randomized block-coordinate variant of the classic Frank-Wolfe algorithm for convex optimization with block-separable constraints. Despite its lower iteration cost, we show that it achieves a similar convergence rate in duality gap as the full Frank-Wolfe algorithm. We also show that, when applied to the d…

2012-07-19abs ↗pdf ↗

Optimal algorithm for two-player zero-sum games with linear parameterization.

problem Finding Nash Equilibrium in two-player zero-sum Markov games with linear transition.
method Nash-UCRL algorithm, Coarse Correlated Equilibrium, Optimism-in-Face-of-Uncertainty.
result Proves ildeO(dHT) ilde{O}(dH\sqrt{T}) regret bound, matching lower bound up to logarithmic factors.

Study stability of contingent claim solutions under probabilistic perturbations.

problem Stability of solutions to discrete-time contingent-claim problems under uncertainty.
method Use Rockafellian perturbations to analyze stability of solutions.
result Establishes convergence of dual problems and shadow prices.

This is the topological part of two papers on the cohomology of Kaehler groups. In this paper we show that if a linear duality group of dimension larger than 6 is the fundamental group of a compact Kaehler manifold then its second or its fourth Betti number is non-zero. As a corollary a cocompact p-adic lattice of rank…

2010-05-17abs ↗pdf ↗

New method finds rare dense clusters in asymmetric binary perceptrons, resolving algorithmic hardness.

problem Resolving algorithmic hardness in asymmetric binary perceptrons.
method Fully lifted random duality theory (fl RDT) and large deviation upgrade (sfl LD RDT).
result Local entropy breaks down for constraint densities in (0.77, 0.78) interval, matching current solver limits.

CInA method uses attention to improve causal inference.

problem Challenges in causal inference, especially in complex tasks.
method CInA method utilizes self-supervised causal learning with multiple unlabeled datasets and transformer-type architecture.
result CInA effectively generalizes to out-of-distribution datasets and various real-world datasets.

Optimal transport reformulates multiple quantile hedging problem.

problem Multiple quantile hedging problem in incomplete markets.
method Reformulated as Monge optimal transport problem, introduced Kantorovitch version, proved no duality gap.
result Multiple quantile hedging problem can be seen as semi-discrete optimal transport problem.

Study potential computational gaps in symmetric binary perceptrons using fl-RDT.

problem Potential statistical-computational gaps in symmetric binary perceptrons.
method Parametric utilization of fully lifted random duality theory (fl-RDT).
result Observation of a computational gap SCG=αcαaSCG=α_c-α_a in SBP.

We derive expressions for the predicitive information rate (PIR) for the class of autoregressive Gaussian processes AR(N), both in terms of the prediction coefficients and in terms of the power spectral density. The latter result suggests a duality between the PIR and the multi-information rate for processes with mutua…

2012-06-01abs ↗pdf ↗