Quantum reinforcement learning protocols implemented in superconducting circuits.
problem Improving quantum devices through learning processes.
method Implementation of quantum reinforcement learning protocols using superconducting circuits.
result Feasibility analysis of quantum reinforcement learning protocols in superconducting circuits.
CAFL breaks feedback loops in recommender systems using causal inference.
problem Feedback loops in recommender systems compromise recommendation quality and homogenize user behavior.
method Causal Adjustment for Feedback Loops (CAFL) algorithm that breaks feedback loops using causal inference.
result CAFL improves recommendation quality compared to prior correction methods.
DFNets uses feedback-looped filters for better graph CNN performance.
problem Improving CNN performance on graph structured data.
method DFNets incorporates feedback-looped spectral graph filters.
result DFNets outperforms state-of-the-art methods in document and entity classification tasks.
Bayesian model eliminates feedback loops in personalization systems.
problem Feedback loops in user choice systems based on limited exposure.
method Bayesian choice model based on Luce axioms, fair and efficient.
result Low regret in learning to present, accurate preference estimates with minimal interactions.
Runaway feedback loops in predictive policing cause crime rate disparities.
problem Runaway feedback loops in predictive policing systems exacerbate crime rate disparities.
method Developed a mathematical model to explain and demonstrate interventions to prevent runaway feedback loops.
result Interventions can prevent runaway feedback loops and allow true crime rates to be learned.
The paper analyzes feedback loops in recommender systems causing echo chambers and filter bubbles.
problem Feedback loops in recommender systems leading to echo chambers and filter bubbles.
method Theoretical analysis of user dynamics and recommender system behavior.
result Solutions to slow down system degeneracy and understanding echo chambers and filter bubbles.
Theoretical model for iterative user discovery in recommender systems.
problem Iterative feedback loops in recommender systems and their biases.
method Theoretical framework to model system evolution and convergence properties.
result Theoretical bounds and convergence properties on user discovery and blind spots.
Study uses geometric algebra to analyze credit cycles, revealing dangerous feedback loops.
problem Understanding and predicting dangerous feedback loops in credit cycles.
method Represent economic states as multi-vectors in Clifford algebra, focusing on bivector elements for rotational coupling.
result Geometric relationship between unemployment and credit contraction shifts from simple correlation to dangerous rotational dynamics during crises.
Retraining stabilizes model influence on data.
problem Performativity in predictive models leads to feedback loops.
method Developed the stable signal principle to address retraining dynamics.
result Repeated risk minimization converges geometrically to stable signal direction.
The study aims to prevent unfair content presentation in recommender systems.
problem Over- and under-presentation of content leads to biased user preference estimates.
method Two models are considered: one that ignores systematic and limited exposure, and another that conditions on limited exposure.
result Ignoring systematic presentations overestimates promoted options and underestimates censored alternatives.
The paper optimizes exceptions in a statistical production system using machine learning.
problem Lack of curated and labeled training data for machine learning in data quality assurance.
method Explainable supervised machine learning to identify and prioritize exceptions.
result Improvement in the quality and efficiency of exceptions generated and authenticated by users.
Online algorithms stabilize in feedback loops of performative prediction.
problem Feedback loops in algorithmic predictions influence data distributions.
method Martingale argument and randomization to avoid distributional assumptions.
result No-regret algorithms converge to performatively stable equilibria.
DeepMPC uses neural networks to control complex fluid flows efficiently.
problem Controlling complex fluid flows in real-time is challenging due to high dimensionality and multi-scale dynamics.
method Deep learning, specifically recurrent neural networks (RNNs), embedded in model predictive control (MPC) framework.
result Significant improvements in control performance achieved through online updates to prediction accuracy.
We review the evidence that the erratic dynamics of markets is to a large extent of endogenous origin, i.e. determined by the trading activity itself and not due to the rational processing of exogenous news. In order to understand why and how prices move, the joint fluctuations of order flow and liquidity - and the way…
Cycles in causal learning cause feedback loops under intervention.
problem Cyclic causal structures lead to feedback loops in causal inference.
method Theoretical observations about self-referential distributions and their factorizations.
result Cyclic causal dependence can exist even when observational data suggest independence.
Develops a framework for analyzing multi-agent and many-body systems with feedback loops.
problem Optimal order of multi-agent and general many-body systems
method Derive macroscopic properties and optimal degree of order
result Optimal degree of order balances productivity, stability, and adaptability
We propose a method for learning cyclic causal models from a combination of observational and interventional equilibrium data. Novel aspects of the proposed method are its ability to work with continuous data (without assuming linearity) and to deal with feedback loops. Within the context of biochemical reactions, we a…
New algorithm mitigates affinity bias in hiring feedback loops.
problem Mitigating affinity bias in hiring decisions to avoid unconscious favoritism.
method Introducing affinity bandits, a new bandit variant that accounts for evolving biased feedback.
result Elimination-style algorithm nearly matches the derived regret bound, outperforming classical algorithms.
Economic growth is unpredictable unless demand is quantified. We solve this problem by introducing the demand for unpaid spare time and a user quantity named human capacity. It organizes and amplifies spare time required for enjoying affluence like physical capital, the technical infrastructure for production, organize…
New method detects when models influence their own drift in real-time data streams.
problem Models can induce concept drift in real-time data streams.
method CheckerBoard Performative Drift Detection (CB-PDD)
result CB-PDD effectively detects performative drift in real-time data streams.
Financial models shape markets through performativity, creating self-fulfilling prophecies.
problem Lack of mathematical formulation for performativity in financial markets.
method Embedding the model in the market process, creating a closed feedback loop.
result Performative market makers can reverse engineer dominant strategies and arbitrage them.
Causal methods for GRN inference from single-cell data often fail in real-world benchmarks.
problem Understanding when and why causal methods for GRN inference from single-cell data fail in real-world benchmarks.
method Introduced a controlled diagnostic framework to isolate and measure seven pathologies.
result Causal methods dominate in clean and structurally favorable regimes but fail in specific pathologies.
PolySwarm uses a swarm of LLMs to predict and arbitrage prediction markets.
problem Real-time prediction market trading and latency arbitrage inefficiencies.
method PolySwarm employs a swarm of 50 diverse LLMs, Bayesian combination, and risk-controlled execution.
result Swarm aggregation outperforms single-model baselines in prediction tasks.
This communication is based on an original approach linking economical factors to technical and methodological ones. This work is applied to the decision process for mix production. This approach is relevant for costing driving systems. The main interesting point is that the quotation factors (linked to time indicators…
Automated suggestions help train technicians diagnose incidents faster.
problem Manual and time-consuming incident diagnosis by train maintenance technicians.
method Developed and deployed a learning machine to suggest diagnostics to technicians.
result The model refines its accuracy through feedback from experts and uses feature engineering.
Feedback loops amplify dataset biases, affecting future model performance.
problem Feedback loops amplify biases in datasets, risking future model reliability.
method Formalized system where model interactions are recorded and reused, analyzed for bias amplification.
result Models that behave like samples from the training distribution are more stable and calibrated.
New method tackles complex systems with hidden confounders and feedback loops.
problem Understanding complex systems with hidden confounders and feedback loops.
method Robust Causal Analysis of Linear Cyclic Systems with Hidden Confounders (LLC)
result LLC method can robustly analyze cyclic systems with hidden confounders.
We study analytically and numerically Minsky instability as a combination of top-down, bottom-up and peer-to-peer positive feedback loops. The peer-to-peer interactions are represented by the links of a network formed by the connections between firms, contagion leading to avalanches and percolation phase transitions pr…
Model captures mini-flash crashes caused by liquidity, feedback loops, and market fragmentation.
problem Understanding and predicting mini-flash crashes in financial markets.
method Develops a mathematical model borrowing from optimal execution literature.
result Mini-flash crashes can occur even when participants are uncertain of their impact.
Partially performative prediction studies how predictive models influence future data.
problem Distribution shift in predictive models due to endogenous and exogenous factors.
method Generalizing performative prediction to capture both endogenous and exogenous sources of distribution shift.
result Developed online analogues of performative stability and optimality for partially performative environments.
New method prevents RLHF alignment collapse by accounting for policy's influence on reward model updates.
problem Iterative RLHF leads to alignment collapse where policies exploit RM's blind spots.
method Foresighted policy optimization (FPO) restores missing steering term via regularization.
result FPO prevents alignment collapse on LLM alignment pipelines using Llama-3.2-1B.
Study state-dependent Hawkes processes for limit order book modeling.
problem Modeling feedback loop between order flow and limit order book shape.
method Existence and uniqueness of state-dependent Hawkes processes, simulation, maximum likelihood estimation.
result Excitation effects in order flow are strongly state-dependent.
Paper proves conformal prediction works for any data distribution.
problem Quantifying risk in AI systems with non-exchangeable data.
method Developed a method to extend conformal prediction to any data distribution.
result Valid conformal prediction guarantees for any data distribution.
Study counterfactuals in cyclic systems with shifts and scales.
problem Counterfactual inference in cyclic systems with shifts and scales.
method Shift-scale interventions in cyclic SCMs.
result Valid inference in cyclic systems with shifts and scales.
Algorithmic recommendation systems can homogenize user behavior, reducing utility.
problem Algorithmic feedback loops homogenize user behavior in recommendation systems.
method Simulations of recommendation systems using confounded data.
result Using confounded data decreases utility without increasing diversity.
We present a simple agent-based model of a financial system composed of leveraged investors such as banks that invest in stocks and manage their risk using a Value-at-Risk constraint, based on historical observations of asset prices. The Value-at-Risk constraint implies that when perceived risk is low, leverage is high…
We first review empirical evidence that asset prices have had episodes of large fluctuations and been inefficient for at least 200 years. We briefly review recent theoretical results as well as the neurological basis of trend following and finally argue that these asset price properties can be attributed to two fundame…
Predictions can shape outcomes, study helps predict these effects.
problem Understanding how predictions influence real-world outcomes.
method Causal identifiability analysis of prediction-covariate-outcome relationships.
result Standard supervised learning can identify transferable relationships from predictions.
Edge language models show bias over time, especially on resource-constrained devices.
problem Bias in edge language models on resource-constrained devices.
method Comparative analysis of text-based bias across edge, cloud, and desktop environments; optimized Llama-2 model on Raspberry Pi 4; feedback loop mechanism to correct bias.
result Llama-2 on Raspberry Pi 4 shows 43.23% and 21.89% more bias over time compared to cloud and desktop models.
Model predicts insolvency risks in banks due to liquidity and credit risks.
problem Determining insolvency regions in banks due to non-linear interaction between liquidity and credit risks.
method Developed a continuous-time structural dynamic model integrating Basel III requirements into a stochastic optimal control framework. Used Hamilton-Jacobi-Bellman (HJB) equation to solve for insolvency boundary. Derived surrogate analytical approximation for real-time monitoring.
result Calibrated model reveals significant non-linear threshold effects and accelerates insolvency transition.
Proposes a model for clearing prices in financial markets due to margin calls.
problem Determining prices in financial markets following margin calls and short squeezes.
method Developed an explicit formulation for clearing prices after margin calls and short squeezes.
result Identified a threshold short interest ratio leading to discontinuity in clearing prices.
A predictor that is deployed in a live production system may perturb the features it uses to make predictions. Such a feedback loop can occur, for example, when a model that predicts a certain type of behavior ends up causing the behavior it predicts, thus creating a self-fulfilling prophecy. In this paper we analyze p…
Study shows naive recommendation models are flawed due to user feedback.
problem Flaws in naive recommendation models due to user feedback.
method Proposed a model with heterogeneous user preferences and proved the inconsistency of naive estimators.
result Consistent estimators are efficient in the presence of myopic agents.
Self-poisoning in adaptive OOD detectors is explained with a sharp threshold theory and certified calibration.
problem Self-poisoning in adaptive OOD detectors.
method Modeling bank impurity as a generalized Pólya urn, proving almost-sure convergence to a mean-field equilibrium.
result A certified admission gate removes the transition at every contamination rate, controlling false positives label-free.
Improves search performance by transferring knowledge from recommender system.
problem Cold start and feedback loop problems in search retrieval.
method Zero-Shot Heterogeneous Transfer Learning framework.
result Significant improvements in relevance and user interactions over production system.
Generative models map simple samples to complex target samples.
problem Improving Monte-Carlo sampling techniques.
method Variational learning of dynamical maps between base and target measures.
result Improved sampling efficiency through feedback loops.
Model shows how relaxed leverage can lead to asset price bubbles.
problem Understanding how financial leverage affects asset prices and growth.
method Developed a macro-finance model with feedback loops between investment and land prices.
result Relaxed leverage can cause unbalanced growth and asset price bubbles.
Paper combines latent state space with CRF for improved autoregressive text generation.
problem Autoregressive models expose hidden state trajectory to biases.
method Combines latent state space model with CRF observation model.
result Improved performance on unconditional sentence generation compared to RNN and GAN baselines.