Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,291 papers · 148 categories

Trend · papers per month

57114171228 · May 202619922001200920182026
48 results for feedback-loop control

Quantum reinforcement learning protocols implemented in superconducting circuits.

problem Improving quantum devices through learning processes.
method Implementation of quantum reinforcement learning protocols using superconducting circuits.
result Feasibility analysis of quantum reinforcement learning protocols in superconducting circuits.

CAFL breaks feedback loops in recommender systems using causal inference.

problem Feedback loops in recommender systems compromise recommendation quality and homogenize user behavior.
method Causal Adjustment for Feedback Loops (CAFL) algorithm that breaks feedback loops using causal inference.
result CAFL improves recommendation quality compared to prior correction methods.

Runaway feedback loops in predictive policing cause crime rate disparities.

problem Runaway feedback loops in predictive policing systems exacerbate crime rate disparities.
method Developed a mathematical model to explain and demonstrate interventions to prevent runaway feedback loops.
result Interventions can prevent runaway feedback loops and allow true crime rates to be learned.

The paper analyzes feedback loops in recommender systems causing echo chambers and filter bubbles.

problem Feedback loops in recommender systems leading to echo chambers and filter bubbles.
method Theoretical analysis of user dynamics and recommender system behavior.
result Solutions to slow down system degeneracy and understanding echo chambers and filter bubbles.

Theoretical model for iterative user discovery in recommender systems.

problem Iterative feedback loops in recommender systems and their biases.
method Theoretical framework to model system evolution and convergence properties.
result Theoretical bounds and convergence properties on user discovery and blind spots.

Study uses geometric algebra to analyze credit cycles, revealing dangerous feedback loops.

problem Understanding and predicting dangerous feedback loops in credit cycles.
method Represent economic states as multi-vectors in Clifford algebra, focusing on bivector elements for rotational coupling.
result Geometric relationship between unemployment and credit contraction shifts from simple correlation to dangerous rotational dynamics during crises.

The study aims to prevent unfair content presentation in recommender systems.

problem Over- and under-presentation of content leads to biased user preference estimates.
method Two models are considered: one that ignores systematic and limited exposure, and another that conditions on limited exposure.
result Ignoring systematic presentations overestimates promoted options and underestimates censored alternatives.

The paper optimizes exceptions in a statistical production system using machine learning.

problem Lack of curated and labeled training data for machine learning in data quality assurance.
method Explainable supervised machine learning to identify and prioritize exceptions.
result Improvement in the quality and efficiency of exceptions generated and authenticated by users.

DeepMPC uses neural networks to control complex fluid flows efficiently.

problem Controlling complex fluid flows in real-time is challenging due to high dimensionality and multi-scale dynamics.
method Deep learning, specifically recurrent neural networks (RNNs), embedded in model predictive control (MPC) framework.
result Significant improvements in control performance achieved through online updates to prediction accuracy.

We review the evidence that the erratic dynamics of markets is to a large extent of endogenous origin, i.e. determined by the trading activity itself and not due to the rational processing of exogenous news. In order to understand why and how prices move, the joint fluctuations of order flow and liquidity - and the way…

2010-09-15abs ↗pdf ↗

We propose a method for learning cyclic causal models from a combination of observational and interventional equilibrium data. Novel aspects of the proposed method are its ability to work with continuous data (without assuming linearity) and to deal with feedback loops. Within the context of biochemical reactions, we a…

2013-09-26abs ↗pdf ↗

New algorithm mitigates affinity bias in hiring feedback loops.

problem Mitigating affinity bias in hiring decisions to avoid unconscious favoritism.
method Introducing affinity bandits, a new bandit variant that accounts for evolving biased feedback.
result Elimination-style algorithm nearly matches the derived regret bound, outperforming classical algorithms.

Economic growth is unpredictable unless demand is quantified. We solve this problem by introducing the demand for unpaid spare time and a user quantity named human capacity. It organizes and amplifies spare time required for enjoying affluence like physical capital, the technical infrastructure for production, organize…

2012-06-12abs ↗pdf ↗

Causal methods for GRN inference from single-cell data often fail in real-world benchmarks.

problem Understanding when and why causal methods for GRN inference from single-cell data fail in real-world benchmarks.
method Introduced a controlled diagnostic framework to isolate and measure seven pathologies.
result Causal methods dominate in clean and structurally favorable regimes but fail in specific pathologies.

PolySwarm uses a swarm of LLMs to predict and arbitrage prediction markets.

problem Real-time prediction market trading and latency arbitrage inefficiencies.
method PolySwarm employs a swarm of 50 diverse LLMs, Bayesian combination, and risk-controlled execution.
result Swarm aggregation outperforms single-model baselines in prediction tasks.

Automated suggestions help train technicians diagnose incidents faster.

problem Manual and time-consuming incident diagnosis by train maintenance technicians.
method Developed and deployed a learning machine to suggest diagnostics to technicians.
result The model refines its accuracy through feedback from experts and uses feature engineering.

Feedback loops amplify dataset biases, affecting future model performance.

problem Feedback loops amplify biases in datasets, risking future model reliability.
method Formalized system where model interactions are recorded and reused, analyzed for bias amplification.
result Models that behave like samples from the training distribution are more stable and calibrated.

Model captures mini-flash crashes caused by liquidity, feedback loops, and market fragmentation.

problem Understanding and predicting mini-flash crashes in financial markets.
method Develops a mathematical model borrowing from optimal execution literature.
result Mini-flash crashes can occur even when participants are uncertain of their impact.

Partially performative prediction studies how predictive models influence future data.

problem Distribution shift in predictive models due to endogenous and exogenous factors.
method Generalizing performative prediction to capture both endogenous and exogenous sources of distribution shift.
result Developed online analogues of performative stability and optimality for partially performative environments.

New method prevents RLHF alignment collapse by accounting for policy's influence on reward model updates.

problem Iterative RLHF leads to alignment collapse where policies exploit RM's blind spots.
method Foresighted policy optimization (FPO) restores missing steering term via regularization.
result FPO prevents alignment collapse on LLM alignment pipelines using Llama-3.2-1B.

Study state-dependent Hawkes processes for limit order book modeling.

problem Modeling feedback loop between order flow and limit order book shape.
method Existence and uniqueness of state-dependent Hawkes processes, simulation, maximum likelihood estimation.
result Excitation effects in order flow are strongly state-dependent.

Algorithmic recommendation systems can homogenize user behavior, reducing utility.

problem Algorithmic feedback loops homogenize user behavior in recommendation systems.
method Simulations of recommendation systems using confounded data.
result Using confounded data decreases utility without increasing diversity.

We present a simple agent-based model of a financial system composed of leveraged investors such as banks that invest in stocks and manage their risk using a Value-at-Risk constraint, based on historical observations of asset prices. The Value-at-Risk constraint implies that when perceived risk is low, leverage is high…

2014-07-20abs ↗pdf ↗

We first review empirical evidence that asset prices have had episodes of large fluctuations and been inefficient for at least 200 years. We briefly review recent theoretical results as well as the neurological basis of trend following and finally argue that these asset price properties can be attributed to two fundame…

2016-05-02abs ↗pdf ↗

Edge language models show bias over time, especially on resource-constrained devices.

problem Bias in edge language models on resource-constrained devices.
method Comparative analysis of text-based bias across edge, cloud, and desktop environments; optimized Llama-2 model on Raspberry Pi 4; feedback loop mechanism to correct bias.
result Llama-2 on Raspberry Pi 4 shows 43.23% and 21.89% more bias over time compared to cloud and desktop models.

Model predicts insolvency risks in banks due to liquidity and credit risks.

problem Determining insolvency regions in banks due to non-linear interaction between liquidity and credit risks.
method Developed a continuous-time structural dynamic model integrating Basel III requirements into a stochastic optimal control framework. Used Hamilton-Jacobi-Bellman (HJB) equation to solve for insolvency boundary. Derived surrogate analytical approximation for real-time monitoring.
result Calibrated model reveals significant non-linear threshold effects and accelerates insolvency transition.

Proposes a model for clearing prices in financial markets due to margin calls.

problem Determining prices in financial markets following margin calls and short squeezes.
method Developed an explicit formulation for clearing prices after margin calls and short squeezes.
result Identified a threshold short interest ratio leading to discontinuity in clearing prices.

A predictor that is deployed in a live production system may perturb the features it uses to make predictions. Such a feedback loop can occur, for example, when a model that predicts a certain type of behavior ends up causing the behavior it predicts, thus creating a self-fulfilling prophecy. In this paper we analyze p…

2013-10-10abs ↗pdf ↗

Self-poisoning in adaptive OOD detectors is explained with a sharp threshold theory and certified calibration.

problem Self-poisoning in adaptive OOD detectors.
method Modeling bank impurity as a generalized Pólya urn, proving almost-sure convergence to a mean-field equilibrium.
result A certified admission gate removes the transition at every contamination rate, controlling false positives label-free.