New eco-systemic prudential policies aim to finance green companies, reducing systemic financial risk.
problem Insufficient financing for green companies despite available savings and monetary management.
method Reorient corporate accounting towards socio-environmental solvency, facilitating access with public guarantees.
result Green financing increases, reducing systemic financial risk and promoting less leveraged investments.
Smart Close-out Netting aims to automate close-out netting processes.
problem Inefficiencies in close-out netting processes for financial institutions.
method Standardisation and automation of legal and regulatory processes using a data-driven framework and controlled natural language.
result Standardisation and automation can improve close-out netting processes for prudentially regulated financial institutions.
Optimizes retirement income with MBGs and neural networks for longevity risk.
problem Maximizing lifetime withdrawals while managing longevity risk.
method Neural-network optimization under stochastic mortality.
result International diversification and longevity pooling improve retirement outcomes.
Inspired by the importance of diversity in biological system, we built an heterogeneous system that could achieve this goal. Our architecture could be summarized in two basic steps. First, we generate a diverse set of classification hypothesis using both Convolutional Neural Networks, currently the state-of-the-art tec…
In order to adapt to the liberalization of the financial sphere started in the Eighties, marked in particular by the end of the framing of credit, the disappearance of the various forms of protection of the State whose profited the banks, and the privatization of the near total of the establishments in Europe, the bank…
Financial contagion from liquidity shocks has being recently ascribed as a prominent driver of systemic risk in interbank lending markets. Building on standard compartment models used in epidemics, in this work we develop an EDB (Exposed-Distressed-Bankrupted) model for the dynamics of liquidity shocks reverberation be…
We propose a new model of the liquidity driven banking system focusing on overnight interbank loans. This significant branch of the interbank market is commonly neglected in the banking system modeling and systemic risk analysis. We construct a model where banks are allowed to use both the interbank and the securities …
We take a closer look at the life and legacy of Micheal Milken. We discuss why Michael Milken, also know as the Junk Bond King, was not just any other King or run-of-the-mill Junk Dealer, but "The Junk Dealer". We find parallels between the three parts to any magic act and what Micheal Milken did, showing that his acco…
Novel method for multiclass ROC curves using multidimensional Gini index.
problem Multiclass performance evaluation, especially for imbalanced datasets.
method Extends ROC curve methodology to multiclass settings using multidimensional Gini index.
result Validated through case studies in health care and finance.
Unified framework maps financial market dynamics using TE and KM, revealing directional information flow.
problem Challenges in traditional correlation analysis of financial markets, especially during crises.
method Combines Transfer Entropy (TE) and Kramers-Moyal (KM) expansion to analyze dynamic interactions among major indices.
result Increased directional information flow during crises, highlighting gold-dollar and oil-equity linkages.
Model assesses how supply chain disruptions affect financial stability.
problem Systemic risk in production networks and its financial implications.
method Data-driven econo-financial stress-testing framework combining supply chain and interbank networks.
result Increase of up to 28% in financial systemic risk due to production network contagion.
GPflux simplifies deep Gaussian processes for Python.
problem Challenges in implementing deep Gaussian processes.
method Python library for Bayesian deep learning with DGPs.
result Efficient, modular, and extensible library for DGPs.
For credit risk management purposes in general, and for allocation of regulatory capital by banks in particular (Basel II), numerical assessments of the credit-worthiness of borrowers are indispensable. These assessments are expressed in terms of probabilities of default (PD) that should incorporate a certain degree of…
The Basel II internal ratings-based (IRB) approach to capital adequacy for credit risk implements an asymptotic single risk factor (ASRF) model. Measurements from the ASRF model of the prevailing state of Australia's economy and the level of capitalisation of its banking sector find general agreement with macroeconomic…
Companies do not operate in a vacuum. As companies move towards an increasingly specialized production function and their reach is becoming truly global, their aptitude in managing and shaping their inter-organizational network is a determining factor in measuring their health. Current models of company financial healt…
The Basel II internal ratings-based (IRB) approach to capital adequacy for credit risk plays an important role in protecting the Australian banking sector against insolvency. We outline the mathematical foundations of regulatory capital for credit risk, and extend the model specification of the IRB approach to a more g…
Paper proposes real-time risk metrics for stablecoin protocols.
problem Lack of risk management frameworks for stablecoins.
method Developed two risk metrics: capitalization and liquidity.
result Demonstrated practical benefits of real-time on-chain data.
Measurement and management of credit concentration risk is critical for banks and relevant for micro-prudential requirements. While several methods exist for measuring credit concentration risk within institutions, the systemic effect of different institutions' exposures to the same counterparties has been less explore…
This paper reviews the economic and theoretical foundations of insolvency risk measurement and capital adequacy rules. The proposed new measure of insolvency risk is constructed by disentangling assets, debt and equity at the micro-prudential firm level. This new risk index is the Firm Insolvency Risk Index (FIRI) whic…
Study how firm liquidation regimes affect shareholder value and stability.
problem Balancing shareholder value and financial stability during firm liquidation.
method Modelled forced liquidation in reduced form, solved singular stochastic control problem.
result Combining distress regions below and above ruin threshold improves both shareholder value and firm survival.
SwiGAN generates drought scenarios for climate risk management.
problem Natural catastrophes and droughts increase insurance costs.
method Conditional GANs for generating spatio-temporal SWI maps.
result Simulates drought patterns up to 2050 for French regions.
Model predicts Mozambique bank failures, aiding risk management.
problem Lack of bankruptcy prediction model in Mozambique banking sector.
method Linear Discriminant Analysis method, using financial indicators.
result Model accurately predicted 84% of bank failures 1 year before Central Bank intervention.
The study provides a practical strategy for pricing and hedging equity-release mortgages guarantees.
problem Pricing and hedging the No-Negative-Equity-Guarantee in incomplete markets.
method Discrete-time model, Excess-of-Loss reinsurance, numerical illustrations.
result Superhedge cost decreases with more lives in the portfolio, making it more realistic.
IDA makes DFMM's asset tradeable, enhancing cross-chain finance efficiency.
problem Making DFMM's asset tradeable to improve cross-chain finance efficiency.
method Introducing IDA as a tradeable asset, leveraging DFMM's robust liquidity and dynamic AMM.
result IDA enhances cross-chain finance efficiency through tradeable asset and dynamic AMM.
The theory of multilayer networks is in its early stages, and its development provides vital methods for understanding complex systems. Multilayer networks, in their multiplex form, have been introduced within the last three years to analysing the structure of financial systems, and existing studies have modelled and e…
Adapts GRPO for off-policy RL, improving reward.
problem Improving training stability and efficiency in RL.
method Adapts GRPO to off-policy setting, uses clipped surrogate objectives.
result Off-policy GRPO outperforms on-policy GRPO in empirical tests.
Paper tackles efficient evaluation of natural stochastic policies in offline RL.
problem Efficiency issues in evaluating natural stochastic policies due to unknown evaluation policy.
method Derive efficiency bounds for tilting and modified treatment policies, propose nonparametric estimators.
result Proposed estimators attain efficiency bounds under lax conditions and enjoy partial double robustness.
New framework studies policy learning problems under data scarcity.
problem Learning improving policies when data is insufficient.
method Developed a mathematical framework for policy learning problems.
result Reduced policy learning problems to simpler ones in sample complexity.
New algorithms improve policy evaluation in reinforcement learning.
problem Off-policy stability and on-policy efficiency issues in policy evaluation.
method Introduced novel algorithms using oblique projection method.
result Demonstrated both off-policy stability and on-policy efficiency.
We study the problem of off-policy policy optimization in Markov decision processes, and develop a novel off-policy policy gradient method. Prior off-policy policy gradient approaches have generally ignored the mismatch between the distribution of states visited under the behavior policy used to collect data, and what …
Stabilizes policy optimization with off-policy data using divergence augmentation.
problem Premature convergence and instability in policy optimization with off-policy data.
method Incorporates Bregman divergence between behavior and current policies to ensure safe policy updates.
result Empirically shows better performance in data-scarce scenarios compared to other algorithms.
We consider the problem of off-policy evaluation in Markov decision processes. Off-policy evaluation is the task of evaluating the expected return of one policy with data generated by a different, behavior policy. Importance sampling is a technique for off-policy evaluation that re-weights off-policy returns to account…
New methods estimate policy value and gradients for deterministic policies from off-policy data.
problem Estimating policy value and gradients for deterministic policies from off-policy data.
method Proposed new doubly robust estimators based on kernelization approaches.
result Demonstrated a rate independent of horizon length for policy value and gradient estimation.
Study designs logging policies to minimize off-policy evaluation error.
problem Minimizing OPE error with logging policies for target policies.
method Characterizes reward-coverage tradeoff, proposes a unifying framework, derives optimal policies.
result Provides actionable guidance for firms choosing recommendation systems.
DSPI connects natural policy gradient to policy iteration, proving global convergence.
problem Optimizing policies in reinforcement learning.
method DSPI framework, combining smoothed policy iteration and natural policy gradient.
result DSPI achieves geometric convergence and optimal complexity for policy optimization.
POTEC tackles off-policy learning in large action spaces, improving effectiveness.
problem Existing OPL methods fail in large discrete action spaces due to bias or variance issues.
method Two-stage algorithm: cluster selection via policy-based approach, action selection via regression-based approach.
result POTEC provides substantial improvements in off-policy learning effectiveness, especially in large and structured action spaces.
Monotonic policy improvement and off-policy learning are two main desirable properties for reinforcement learning algorithms. In this paper, by lower bounding the performance difference of two policies, we show that the monotonic policy improvement is guaranteed from on- and off-policy mixture samples. An optimization …
Memory-efficient algorithm reduces variance in off-policy RL.
problem High variance in off-policy policy optimization.
method Memory-efficient, stochastically variance-reduced algorithm using off-policy samples.
result Empirically validated effectiveness of the proposed algorithm.
In this work, we consider the problem of estimating a behaviour policy for use in Off-Policy Policy Evaluation (OPE) when the true behaviour policy is unknown. Via a series of empirical studies, we demonstrate how accurate OPE is strongly dependent on the calibration of estimated behaviour policy models: how precisely …
Protects proprietary policies from imitation learning by training adversarial policy ensembles.
problem Protecting policies from external observers cloning them.
method Introduces a reinforcement learning framework that trains an ensemble of near-optimal policies, making demonstrations useless for external observers.
result Demonstrates the existence of 'non-clonable' ensembles and provides a solution to the optimization problem.
Extends OPE to evaluate policies using diverse logging data.
problem Evaluate policies using log data from different policies.
method Develops an OPE method for various logging policies.
result Method's predictions converge to true performance as sample size increases.
PS framework selects best policy from library for CSO problems.
problem Policy selection in CSO with heterogeneous performance across covariate space.
method PS framework constructs library of candidate policies and learns a meta-policy to select the best one.
result PS consistently outperforms best single policy in heterogeneous CSO problems.
Entropy regularization improves policy optimization in reinforcement learning.
problem Improving policy optimization in reinforcement learning.
method Entropy regularization is introduced to soften the greedy policy towards a more diverse softmax policy, leading to a continuously parameterized algorithm that interpolates between policy gradient and Q-learning.
result An intermediate algorithm can improve performance in reinforcement learning.
New method optimizes treatment policies to avoid winner's curse.
problem Winner's curse in treatment policy optimization.
method Inference-aware policy optimization.
result Optimizes for both estimated performance and downstream evaluation.
PBVFs generalize across policies using learned value functions.
problem RL algorithms forget information about old policies when updating value functions to track the learned policy.
method Introduce Parameter-Based Value Functions (PBVFs) that include policy parameters in their inputs, enabling them to generalize across different policies.
result PBVFs enable zero-shot learning of new policies that outperform any policy seen during training.
New method improves off-policy critic evaluation in reinforcement learning.
problem High variance and instability in off-policy policy evaluation.
method Doubly robust estimators applied to actor-critic algorithms.
result Doubly robust estimation significantly improves performance in continuous control tasks.
Policy gradient aims to maximize expected return using gradient ascent.
problem Finding a policy that maximizes expected return in a given class of policies.
method Gradient ascent applied to a differentiable model of the policy, estimating the gradient of expected return.
result Policy gradient methods require on-policy data for gradient estimation, limiting sample efficiency.
Optimizes Thompson sampling policies using policy gradient methods.
problem Improving Thompson sampling in bandit problems.
method Applies policy gradient algorithms to optimize Thompson sampling policies.
result Direct policy search on Thompson sampling improves performance.