We study the online saddle point problem, an online learning problem where at each iteration a pair of actions need to be chosen without knowledge of the current and future (convex-concave) payoff functions. The objective is to minimize the gap between the cumulative payoffs and the saddle point value of the aggregate …
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Cumulant expansion is used to derive accurate closed-form approximation for Monthly Sum Options in case of constant volatility model. Payoff of Monthly Sum Option is based on sum of caped (and probably floored) returns. It is noticed, that can be used as a small parameter in Edgeworth expansion. First …
Most decision theories, including expected utility theory, rank dependent utility theory and cumulative prospect theory, assume that investors are only interested in the distribution of returns and not in the states of the economy in which income is received. Optimal payoffs have their lowest outcomes when the economy …
Explicit robust hedging strategies for convex or concave payoffs under a continuous semimartingale model with uncertainty and small transaction costs are constructed. In an asymptotic sense, the upper and lower bounds of the cumulative volatility enable us to super-hedge convex and concave payoffs respectively. The ide…
This paper describes an agent-based model of interacting firms, in which interacting firm agents rationally invest capital and labor in order to maximize payoff. Both transactions and production are taken into account in this model. First, the performance of individual firms on a real transaction network was simulated.…
Algorithm learns to bid optimally in repeated first-price auctions with censored feedback.
Algorithm learns from changing zero-sum games with no regret.
This paper studies the bail-out optimal dividend problem with regime switching under the constraint that the cumulative dividend strategy is absolutely continuous. We confirm the optimality of the regime-modulated refraction-reflection strategy when the underlying risk model follows a general spectrally negative Markov…
Proper balance between exploitation and exploration is what makes good decisions, which achieve high rewards like payoff or evolutionary fitness. The Infomax principle postulates that maximization of information directs the function of diverse systems, from living systems to artificial neural networks. While specific a…
A reinforcement learning agent tries to maximize its cumulative payoff by interacting in an unknown environment. It is important for the agent to explore suboptimal actions as well as to pick actions with highest known rewards. Yet, in sensitive domains, collecting more data with exploration is not always possible, but…
We consider the performance of non-optimal hedging strategies in exponential Lévy models. Given that both the payoff of the contingent claim and the hedging strategy admit suitable integral representations, we use the Laplace transform approach of Hubalek et al. (2006) to derive semi-explicit formulas for the resulting…
Optimal exit strategies of CPT gamblers in unfair gambles
Bandit problem on graphs aims to recommend items with high expected ratings.
Method constructs CFMMs matching desired payoffs.
Optimal payoff choice constrained by Bregman-Wasserstein divergence.
Optimal portfolio yields a digital option payoff.
Study finds cheapest possible payoff under ambiguity, linking to maxmin expected utility.
We introduce signature payoffs, a family of path-dependent derivatives that are given in terms of the signature of the price path of the underlying asset. We show that these derivatives are dense in the space of continuous payoffs, a result that is exploited to quickly price arbitrary continuous payoffs. This approach …
The paper uncovers the impact of price and payoff autocorrelations in multi-period asset pricing models.
New method uses neural networks for better financial hedging.
Paper shows how to replicate payoffs without oracles in CFMMs.
We study a non-parametric multi-armed bandit problem with stochastic covariates, where a key complexity driver is the smoothness of payoff functions with respect to covariates. Previous studies have focused on deriving minimax-optimal algorithms in cases where it is a priori known how smooth the payoff functions are. I…
New method optimizes resource allocation for uncertain tasks.
Can one parallelize complex exploration exploitation tradeoffs? As an example, consider the problem of optimal high-throughput experimental design, where we wish to sequentially design batches of experiments in order to simultaneously learn a surrogate function mapping stimulus to response and identify the maximum of t…
Develops a new method for robust risk measurement by averaging nearby payoffs.
Agent optimizes perpetual contract liquidation with transaction costs and risk.
Multi-armed bandit problems are the most basic examples of sequential decision problems with an exploration-exploitation trade-off. This is the balance between staying with the option that gave highest payoffs in the past and exploring new options that might give higher payoffs in the future. Although the study of band…
New findings show pure strategy equilibria are more robust in a war of attrition game.
We study the use of the multilevel Monte Carlo technique in the context of the calculation of Greeks. The pathwise sensitivity analysis differentiates the path evolution and reduces the payoff's smoothness. This leads to new challenges: the inapplicability of pathwise sensitivities to non-Lipschitz payoffs often makes …
Study optimal stopping for diffusion processes using data-driven methods.
New algorithms for stochastic linear bandits with heavy-tailed payoffs achieve nearly optimal regret.
The game-theoretic risk management framework put forth in the precursor work "Towards a Theory of Games with Payoffs that are Probability-Distributions" (arXiv:1506.07368 [q-fin.EC]) is herein extended by algorithmic details on how to compute equilibria in games where the payoffs are probability distributions. Our appr…
The study analyzes how wartime controls influenced zaibatsu stock prices in Japan.
Study of zero-sum games with noisy observations and commitments.
Study bandit problem on smooth graph functions for recommender systems.
This paper studies robust payoff allocation in submodular games, especially against replication.
We study the problem of repeated play in a zero-sum game in which the payoff matrix may change, in a possibly adversarial fashion, on each round; we call these Online Matrix Games. Finding the Nash Equilibrium (NE) of a two player zero-sum game is core to many problems in statistics, optimization, and economics, and fo…
Study on optimal information acquisition in Kyle model with entropy cost.
We derive a formula for liquidity providers' payoff on DEXs, linking it to volatility.
In an online contract selection problem there is a seller which offers a set of contracts to sequentially arriving buyers whose types are drawn from an unknown distribution. If there exists a profitable contract for the buyer in the offered set, i.e., a contract with payoff higher than the payoff of not accepting any c…
The paper examines bounds for stop-loss payoffs using transformed random variables.
New decision-theoretic calibration error metric improves prediction reliability.
Quantum Monte Carlo speeds up option pricing for complex payoff functions.
New method uses DistRL to estimate entire payoff distribution for financial derivatives.
Motivated by the observation that overexposure to unwanted marketing activities leads to customer dissatisfaction, we consider a setting where a platform offers a sequence of messages to its users and is penalized when users abandon the platform due to marketing fatigue. We propose a novel sequential choice model to ca…
We consider a sequential learning problem with Gaussian payoffs and side information: after selecting an action , the learner receives information about the payoff of every action in the form of Gaussian observations whose mean is the same as the mean payoff, but the variance depends on the pair (and may…
Novel approach to financial derivatives pricing using rough path theory.
New algorithms improve on bandit feedback in matrix games with unknown payoff matrices.