Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,181 papers · 148 categories

Trend · papers per month

95191286381 · Jun 202019922001200920182026
48 results for gradient regulation

Policy gradient methods converge globally and efficiently for linear quadratic regulators.

problem Policy gradient methods struggle with non-convex optimization in linear quadratic regulators.
method Model-free policy gradient methods are shown to globally converge and have polynomial complexity in sample and computational terms.
result Policy gradient methods globally converge to the optimal solution for linear quadratic regulators with polynomial efficiency.

Study shows a linear quadratic regulator's imitation learning converges globally.

problem Global convergence of imitation learning for linear quadratic regulators.
method Analyzed alternating gradient algorithm and established Q-linear rate of convergence.
result Established a unique saddle point for globally optimal policy and reward function.

Policy gradient methods converge for LQR problems with noisy state dynamics.

problem Finding optimal policies in noisy LQR problems over finite time horizons.
method Policy gradient methods with convergence guarantees for finite time and stochastic state dynamics.
result Global linear convergence for policy gradient methods in LQR problems with weak assumptions.

Policy gradient converges to globally optimal policy in nearly linear-quadratic systems.

problem Finding optimal policies in nonlinear control systems with partial information.
method Policy gradient algorithm designed for nearly linear-quadratic regulators with small Lipschitz nonlinear components.
result Policy gradient algorithm converges to globally optimal policy with linear rate.

URN neural network dynamically generates various neural structures during training.

problem Creating neural networks with flexible, dynamic structures during training.
method Introduced Unstructured Recursive Network (URN) and used gradient descent on a single loss function.
result Different neural structures can emerge from a single URN during training.

The Conant-Ashby theorem is verified for hypergraph observers, leading to unique learning rules.

problem Verifying conditions for hypergraph observers to maintain internal models.
method Formalizing persistent observers, applying the Conant-Ashby theorem, and using natural gradient descent.
result Natural gradient descent is the unique admissible learning rule for hypergraph observers.

EAGC boosts GCD by regulating gradient entanglement, improving known and novel category separability.

problem Gradient entanglement distorts supervised gradients and overlaps known and novel class representations.
method EAGC uses AGA and EEP to align and project gradients, reducing entanglement and overlap.
result EAGC consistently boosts GCD performance, setting new state-of-the-art results.

This paper studies how gradient descent in control systems can perform well on unseen data.

problem The extent of a learned controller's ability to extrapolate to unseen initial states.
method Theoretical study of policy gradient in Linear Quadratic Regulator (LQR) problems, focusing on the role of exploration.
result The performance of a learned controller on unseen initial states depends on the degree of exploration induced by the system.

IEBN normalizes noise by enhancing instance-specific information, improving deep learning performance.

problem Improving deep learning performance by regulating noise in batch normalization.
method Integrates self-attention mechanism to recalibrate channel information in BN.
result IEBN outperforms BN with improved generalization and stability.

Model proposes how regulators should oversee complex algorithms in high-stakes applications.

problem Regulating complex algorithms used in high-stakes applications like lending, testing, and hiring.
method Proposes a model where regulators are limited in learning about complex algorithms with misaligned preferences, and explores different regulatory approaches.
result Complex algorithms can improve welfare, but regulation should focus on the source of incentive misalignment for optimal results.

Regulated curves on Banach manifolds with continuous projections and regulated derivatives are studied.

problem Regulated curves on Banach manifolds with continuous projections and regulated derivatives.
method Building a Banach manifold structure on the set of such curves.
result Existence of a 'local addition' on such a manifold for any Banach manifold.

We show that any objective risk measurement algorithm mandated by central banks for regulated financial entities will result in more risk being taken on by those financial entities than would otherwise be the case. Furthermore, the risks taken on by the regulated financial entities are far more systemically concentrate…

2010-04-10abs ↗pdf ↗

Hyperboost uses gradient boosting for hyperparameter optimization, outperforming state-of-the-art methods.

problem Hyperparameter tuning for machine learning algorithms
method Gradient boosting surrogate model with quantile regression and distance metric
result Hyperboost outperforms state-of-the-art techniques in empirical tests

Optimal insurance investment under VaR regulation improves policyholders' utility.

problem Optimal investment for participating insurance contracts under VaR-regulation.
method Martingale approach for constrained non-concave optimization problems.
result VaR constraints lead to more prudent investment, improving policyholders' utility.

Study shows model-based methods require fewer samples than model-free methods for LQR tasks.

problem Comparing model-based and model-free methods in reinforcement learning for continuous control tasks.
method An asymptotic analysis of sample complexity for policy evaluation in LQR tasks.
result Model-based methods require asymptotically less samples than model-free methods for policy evaluation in LQR tasks.

Proposes a game-theoretic framework for ML trust regulation.

problem Lack of coordination between ML model builders and regulators.
method Formulates trustworthy ML as a multi-objective multi-agent optimization problem and introduces regulation games and ParetoPlay.
result Enables efficient enforcement of ML model specifications without discouraging participation.

The FCA improved insider trading regulation after 2012, reducing abnormal returns.

problem Regulation of insider trading before and after the UK Financial Services Act 2012.
method Event study methodology using abnormal returns analysis.
result Abnormal returns were reduced after the FCA took over from the FSA.

A deterministic trading strategy by a representative investor on a single market asset, which generates complex and realistic returns with its first four moments similar to the empirical values of European stock indices, is used to simulate the effects of financial regulation that either pricks bubbles, props up crashe…

2010-02-11abs ↗pdf ↗

An asset network systemic risk (ANWSER) model is presented to investigate the impact of how shadow banks are intermingled in a financial system on the severity of financial contagion. Particularly, the focus of this study is the impact of the following three representative topologies of an interbank loan network betwee…

2014-09-30abs ↗pdf ↗

Develops new methods for isospectral orbifolds and regulator quotients.

problem Isospectral orbifolds and regulator quotients in Vignéras constructions.
method New sufficient criteria for isospectrality and regulator quotients, linking torsion homology and Galois representations.
result Produces small exotic isospectral orbifolds and sufficient criteria for regulator quotients.

A neural network method tackles high-dimensional diffeomorphic mapping problems.

problem High-dimensional diffeomorphic mapping struggles with the curse of dimensionality.
method Combines variational principles with quasi-conformal theory for accurate, bijective mappings.
result Validated accuracy, robustness, and effectiveness in complex registration scenarios.

This study examines how ChiNext IPOs' initial returns are influenced by regulation regime changes.

problem Investors' behavior and pricing of ChiNext IPOs under different regulation regimes.
method Analysis of three time periods with two different regulation regimes and three sets of listing day trading restrictions.
result Regulation regime changes significantly impact ChiNext IPO pricing and overreaction.

This paper tackles robust control of LQR systems with multiplicative noise using policy gradient methods.

problem Robustness in reinforcement learning control of complex systems with multiplicative noise.
method Policy gradient algorithms with gradient domination property for non-convex cost functions.
result Global convergence of policy gradient algorithms to the globally optimum control policy.

We investigate a randomization procedure undertaken in real option games which can serve as a basic model of regulation in a duopoly model of preemptive investment. We recall the rigorous framework of [M. Grasselli, V. Leclère and M. Ludkovsky, Priority Option: the value of being a leader, International Journal of Theo…

2013-09-07abs ↗pdf ↗

Study optimal liquidation strategies in lit and dark pools with and without regulation.

problem Optimal liquidation strategies in dark and lit pools with execution uncertainty.
method Design optimal make-take fee policies, solve HJB-Fokker-Planck systems, use BSDEs.
result Explicit solutions for optimal strategies in both competitive and regulated markets.