New algorithm reduces switching costs in multinomial logit bandit problems.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
New polynomial invariants derived from birack and switch structures.
In this paper, we study optimal switching problems under ambiguity. To characterize the optimal switching under ambiguity in the finite horizon, we use multidimensional reflected backward stochastic differential equations (multidimensional RBSDEs) and show that a value function of the optimal switching under ambiguity …
This work extends identifiability analysis to sequential latent variable models, focusing on Switching Dynamical Systems.
The problem of optimal switching between nonlinear autonomous subsystems is investigated in this study where the objective is not only bringing the states to close to the desired point, but also adjusting the switching pattern, in the sense of penalizing switching occurrences and assigning different preferences to util…
Code-switching, the alternation of languages within a conversation or utterance, is a common communicative phenomenon that occurs in multilingual communities across the world. This survey reviews computational approaches for code-switched Speech and Natural Language Processing. We motivate why processing code-switched …
New algorithms improve sampling from complex distributions.
Squirrel switches between optimizers for better performance.
Study approximates financial market with discrete-time models.
This paper studies deep learning methodologies for portfolio optimization in the US equities market. We present a novel residual switching network that can automatically sense changes in market regimes and switch between momentum and reversal predictors accordingly. The residual switching network architecture combines …
Optimizes control of hybrid systems with multiple switching processes.
Study tackles balancing policy switching costs in offline RL.
New algorithm learns switching dynamics from multiple neural signals.
Paper tackles utility maximization with job-switching and retirement constraints.
New RL algorithm reduces policy switching cost to loglog(T) with similar regret.
Paper presents an efficient algorithm for linear MDP with low switching cost.
This paper studies the impact of limited switches on resource-constrained dynamic pricing with demand learning. We focus on the classical price-based blind network revenue management problem and extend our results to the bandits with knapsacks problem. In both settings, a decision maker faces stochastic and distributio…
Study strategic competition in commodity markets using impulse-switching controls.
One type of switch simplifies operations on lattice knots.
Optimal switching regret for all segmentations in online convex optimisation.
As a metric to measure the performance of an online method, dynamic regret with switching cost has drawn much attention for online decision making problems. Although the sublinear regret has been provided in many previous researches, we still have little knowledge about the relation between the dynamic regret and the s…
This paper tackles near-optimal adversarial RL with switching costs, providing algorithms and matching lower bounds.
This paper addresses parameter estimation for wave equations with Markovian switching.
Algorithm for bandits with switching costs achieves optimal regret bounds.
In this paper, we derive the family switching formula of -n two-sphere fiber bundle embedded in a smooth four-manifold fiber bundle. In the smooth category, it is a partial generalization of Fintushel-Stern's argument for four-manifolds. We also derive an algebraic analogue of the family switching formula, allowing the…
Regime switching volatility models provide a tractable method of modelling stochastic volatility. Currently the most popular method of regime switching calibration is the Hamilton filter. We propose using the Baum-Welch algorithm, an established technique from Engineering, to calibrate regime switching models instead. …
Audit fees change based on company and economic factors during auditor switching.
We study the problem of switching-constrained online convex optimization (OCO), where the player has a limited number of opportunities to change her action. While the discrete analog of this online learning task has been studied extensively, previous work in the continuous setting has neither established the minimax ra…
LaMBO optimizes modular systems with switching costs, achieving better results than existing methods.
New algorithm reduces switching costs in RL beyond linear MDPs.
Develops identifiability theory for multi-lag regime-switching models.
We study online learning when partial feedback information is provided following every action of the learning process, and the learner incurs switching costs for changing his actions. In this setting, the feedback information system can be represented by a graph, and previous works studied the expected regret of the le…
Paper derives analytical formulas for NLD-CEV moments with regime switching.
OMGD algorithm optimizes online convex optimization with switching costs and delayed gradients.
Pricing financial or real options with arbitrary payoffs in regime-switching models is an important problem in finance. Mathematically, it is to solve, under certain standard assumptions, a general form of optimal stopping problems in regime-switching models. In this article, we reduce an optimal stopping problem with …
In this paper, we consider the problem of pricing discretely-sampled variance swaps based on a hybrid model of stochastic volatility and stochastic interest rate with regime-switching. Our modelling framework extends the Heston stochastic volatility model by including the CIR stochastic interest rate and model paramete…
New algorithm reduces RL complexity with low switching costs.
The paper deals with regression problems, in which the nonsmooth target is assumed to switch between different operating modes. Specifically, piecewise smooth (PWS) regression considers target functions switching deterministically via a partition of the input space, while switching regression considers arbitrary switch…
Many complex dynamical phenomena can be effectively modeled by a system that switches among a set of conditionally linear dynamical modes. We consider two such models: the switching linear dynamical system (SLDS) and the switching vector autoregressive (VAR) process. Our Bayesian nonparametric approach utilizes a hiera…
Label switching is a phenomenon arising in mixture model posterior inference that prevents one from meaningfully assessing posterior statistics using standard Monte Carlo procedures. This issue arises due to invariance of the posterior under actions of a group; for example, permuting the ordering of mixture components …
This paper first describes a class of uncertain stochastic control systems with Markovian switching, and derives an Itô-Liu formula for Markov-modulated processes. And we characterize an optimal control law, which satisfies the generalized Hamilton-Jacobi-Bellman (HJB) equation with Markovian switching. Then, by using …
Optimal credit and consumption strategies in a switching market with default contagion.
The stochastic knapsack has been used as a model in wide ranging applications from dynamic resource allocation to admission control in telecommunication. In recent years, a variation of the model has become a basic tool in studying problems that arise in revenue management and dynamic/flexible pricing; and it is in thi…
We have developed a statistical technique to test the model assumption of binary regime switching extension of the geometric Brownian motion (GBM) model by proposing a new discriminating statistics. Given a time series data, we have identified an admissible class of the regime switching candidate models for the statist…
Latent force models (LFMs) are hybrid models combining mechanistic principles with non-parametric components. In this article, we shall show how LFMs can be equivalently formulated and solved using the state variable approach. We shall also show how the Gaussian process prior used in LFMs can be equivalently formulated…
Adaptive Bayesian Optimization for resource-constrained experiments with switching costs.
Study optimal portfolios in a non-Markovian regime-switching model with random time horizon.
We study the problem of dynamically trading futures in a regime-switching market. Modeling the underlying asset price as a Markov-modulated diffusion process, we present a utility maximization approach to determine the optimal futures trading strategy. This leads to the analysis of the associated system of Hamilton-Jac…