Heterogeneous SVO leads to diverse policies in sequential social dilemmas.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Research tackles alliance formation in many-player zero-sum games, showing reinforcement learning fails but a contract mechanism can help.
This paper considers stochastic bandits with side observations, a model that accounts for both the exploration/exploitation dilemma and relationships between arms. In this setting, after pulling an arm i, the decision maker also observes the rewards for some other actions related to i. We will see that this model is su…
Controversies around race and machine learning have sparked debate among computer scientists over how to design machine learning systems that guarantee fairness. These debates rarely engage with how racial identity is embedded in our social experience, making for sociological and psychological complexity. This complexi…
This survey outlines methods to ensure fairness in machine learning.
SLIM model tackles graph classification by resolving part-interaction dilemmas.
We propose a unified mechanism for achieving coordination and communication in Multi-Agent Reinforcement Learning (MARL), through rewarding agents for having causal influence over other agents' actions. Causal influence is assessed using counterfactual reasoning. At each timestep, an agent simulates alternate actions t…
Many real-world problems like Social Influence Maximization face the dilemma of choosing the best out of options at a given time instant. This setup can be modeled as a combinatorial bandit which chooses out of arms at each time, with an aim to achieve an efficient trade-off between exploration and expl…
NFTs raise concerns like scams, racism, and sexism; centralization vs decentralization debate.
The paper proposes a framework to recommend suitable conferences for authors.
Adding data can sometimes hurt model performance in multi-source healthcare tasks.
The Data Clustering (DC) problem is of central importance for the area of Machine Learning (ML), given its usefulness to represent data structural similarities from input spaces. Differently from Supervised Machine Learning (SML), which relies on the theoretical frameworks of the Statistical Learning Theory (SLT) and t…
Leo Breiman's Rashomon Effect and Occam Dilemma are re-evaluated in the context of modern machine learning.
Margin enlargement over training data has been an important strategy since perceptrons in machine learning for the purpose of boosting the robustness of classifiers toward a good generalization ability. Yet Breiman (1999) showed a dilemma that a uniform improvement on margin distribution does NOT necessarily reduces ge…
The study examines causal razors and their logical relations, highlighting a dilemma in causal discovery.
Whenever customers' choices (e.g. to buy or not a given good) depend on others choices (cases coined 'positive externalities' or 'bandwagon effect' in the economic literature), the demand may be multiply valued: for a same posted price, there is either a small number of buyers, or a large one -- in which case one says …
New method improves OOD detection without sacrificing generalization.
Bayesian active learning tackles nuisance parameters, leading to bias and dilemmas.
The explore{exploit dilemma is one of the central challenges in Reinforcement Learning (RL). Bayesian RL solves the dilemma by providing the agent with information in the form of a prior distribution over environments; however, full Bayesian planning is intractable. Planning with the mean MDP is a common myopic approxi…
We present PredRNN++, an improved recurrent network for video predictive learning. In pursuit of a greater spatiotemporal modeling capability, our approach increases the transition depth between adjacent states by leveraging a novel recurrent unit, which is named Causal LSTM for re-organizing the spatial and temporal m…
Model-based reinforcement learning (MBRL) is widely seen as having the potential to be significantly more sample efficient than model-free RL. However, research in model-based RL has not been very standardized. It is fairly common for authors to experiment with self-designed environments, and there are several separate…
Bayesian model averaging, model selection and its approximations such as BIC are generally statistically consistent, but sometimes achieve slower rates og convergence than other methods such as AIC and leave-one-out cross-validation. On the other hand, these other methods can br inconsistent. We identify the "catch-up …
Paper finds efficient OPE estimator for multiple logging policies with minimum variance.
Generative AI predicts Arctic sea ice dynamics over decades.
Users of a personalised recommendation system face a dilemma: recommendations can be improved by learning from data, but only if the other users are willing to share their private information. Good personalised predictions are vitally important in precision medicine, but genomic information on which the predictions are…
SAE improves VAE's latent space precision in high dimensions.
AANets balance stability and plasticity in CIL.
We consider a novel stochastic multi-armed bandit problem called {\em good arm identification} (GAI), where a good arm is defined as an arm with expected reward greater than or equal to a given threshold. GAI is a pure-exploration problem that a single agent repeats a process of outputting an arm as soon as it is ident…
A major challenge in cognitive science and AI has been to understand how autonomous agents might acquire and predict behavioral and mental states of other agents in the course of complex social interactions. How does such an agent model the goals, beliefs, and actions of other agents it interacts with? What are the com…
Proposes SWA for adversarial training to improve model robustness.
Investigates model selection challenges in heterogeneous treatment effect estimation.
PHASE dataset simulates complex social interactions in physical environments.
Improved neural model for social recommendation by integrating social and interest networks.
Social media enhances or diminishes scientific status, depending on usage.
We derive an optimal strategy for minimizing the expected loss in the two-period economy when a pivotal decision needs to be made during the first time period and cannot be subsequently reversed. Our interest in the problem has been motivated by the classical shopper's dilemma during the Black Friday promotion period, …
Social learning can make financial markets inefficient, but individual learning can fix this.
SINN combines social science and deep learning for predicting opinion dynamics.
Optimizes social interactions for profit, people, and planet using mathematical models.
New algorithms for batch decision-making with high-dimensional user data.
Model predicts increased social unrest during COVID-19 using social media data.
Enhances traditional MV model for socially responsible investors.
The rise in online social networking has brought about a revolution in social relations. However, its effects on offline interactions and its implications for collective well-being are still not clear and are under-investigated. We study the ecology of online and offline interaction in an evolutionary game framework wh…
We present the quantum model of Bertrand duopoly and study the entanglement behavior on the profit functions of the firms. Using the concept of optimal response of each firm to the price of the opponent, we found only one Nash equilibirum point for maximally entangled initial state. The very presence of quantum entangl…
Secure social recommendation framework using secret sharing.
Predicting event attendance using social influence from social networks.
Algorithm improves learning by integrating diverse agents' behaviors.
Events are happening in real-world and real-time, which can be planned and organized occasions involving multiple people and objects. Social media platforms publish a lot of text messages containing public events with comprehensive topics. However, mining social events is challenging due to the heterogeneous event elem…
SBO uses dual voting to build consensus in noisy feedback settings.