A decentralized deep RL controller improves hexapod locomotion learning.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
This paper proposes a decentralized reinforcement learning method for multi-agent resource allocation.
Flexible decentralized MARL framework for cooperative multi-agent learning.
A decentralized approach for agents to learn and optimize collectively.
DePAint solves MARL for agents with local constraints, privacy, and no central controller.
Study uses multi-agent reinforcement learning to control self-assembly with high-resolution external control.
A decentralized routing framework for lunar exploration robots.
A new method for learning policies in multiple environments.
Improved exploration in cooperative multi-agent reinforcement learning.
Despite the success of single-agent reinforcement learning, multi-agent reinforcement learning (MARL) remains challenging due to complex interactions between agents. Motivated by decentralized applications such as sensor networks, swarm robotics, and power grids, we study policy evaluation in MARL, where agents with jo…
Despite the increasing interest in multi-agent reinforcement learning (MARL) in multiple communities, understanding its theoretical foundation has long been recognized as a challenging problem. In this work, we address this problem by providing a finite-sample analysis for decentralized batch MARL with networked agents…
New MARL algorithms resolve the curse of multiagency with function approximation.
We investigate a classification problem using multiple mobile agents capable of collecting (partial) pose-dependent observations of an unknown environment. The objective is to classify an image over a finite time horizon. We propose a network architecture on how agents should form a local belief, take local actions, an…
Reinforcement Learning (RL) is a learning paradigm concerned with learning to control a system so as to maximize an objective over the long term. This approach to learning has received immense interest in recent times and success manifests itself in the form of human-level performance on games like \textit{Go}. While R…
Swarm systems constitute a challenging problem for reinforcement learning (RL) as the algorithm needs to learn decentralized control policies that can cope with limited local sensing and communication abilities of the agents. While it is often difficult to directly define the behavior of the agents, simple communicatio…
Proposes three decentralized multi-agent reinforcement learning algorithms to reduce network congestion.
We consider the problem of \emph{fully decentralized} multi-agent reinforcement learning (MARL), where the agents are located at the nodes of a time-varying communication network. Specifically, we assume that the reward functions of the agents might correspond to different tasks, and are only known to the corresponding…
A new method improves ridesharing efficiency using QMIX.
V-learning tackles multiagent reinforcement learning by reducing sample complexity.
Improves data efficiency in multi-agent control tasks using model-based reinforcement learning.
This paper proposes a gossip-based algorithm for distributed bilevel optimization over networks.
Enhances crypto-asset AMM with deep learning for better liquidity and efficiency.
Paper tackles low sample and communication complexities in decentralized bilevel optimization.
We consider the networked multi-agent reinforcement learning (MARL) problem in a fully decentralized setting, where agents learn to coordinate to achieve the joint success. This problem is widely encountered in many areas including traffic control, distributed control, and smart grids. We assume that the reward functio…
Smart grid uses deep learning to optimize household energy use.
Recent developments in deep reinforcement learning are concerned with creating decision-making agents which can perform well in various complex domains. A particular approach which has received increasing attention is multi-agent reinforcement learning, in which multiple agents learn concurrently to coordinate their ac…
New algorithm reduces learning regret in multi-agent systems with unknown dynamics.
Bayesian network approach for efficient cooperative MARL.
Paper analyzes convergence of decentralized algorithms with noise and bias.
Policy-gradient method controls multiple non-cohesive targets.
Mobile edge computing (MEC) emerges recently as a promising solution to relieve resource-limited mobile devices from computation-intensive tasks, which enables devices to offload workloads to nearby MEC servers and improve the quality of computation experience. Nevertheless, by considering a MEC system consisting of mu…
Multi-agent reinforcement learning (MARL) has long been a significant and everlasting research topic in both machine learning and control. With the recent development of (single-agent) deep RL, there is a resurgence of interests in developing new MARL algorithms, especially those that are backed by theoretical analysis…
We propose a method to model multi-agent behaviors with limited observation and mechanical constraints.
Motivated by the emerging use of multi-agent reinforcement learning (MARL) in engineering applications such as networked robotics, swarming drones, and sensor networks, we investigate the policy evaluation problem in a fully decentralized setting, using temporal-difference (TD) learning with linear function approximati…
Autogov uses RL to automate DeFi governance, improving security and profitability.
We explore value-based solutions for multi-agent reinforcement learning (MARL) tasks in the centralized training with decentralized execution (CTDE) regime popularized recently. However, VDN and QMIX are representative examples that use the idea of factorization of the joint action-value function into individual ones f…
MARLA uses deep reinforcement learning for multi-agent AHT, reducing Bayes risk.
In reinforcement learning, agents learn by performing actions and observing their outcomes. Sometimes, it is desirable for a human operator to \textit{interrupt} an agent in order to prevent dangerous situations from happening. Yet, as part of their learning process, agents may link these interruptions, that impact the…
A DRL-based strategy improves vehicle tracking accuracy while saving energy.
Deep reinforcement learning boosts throughput in RF-powered cognitive radio networks.
Deep reinforcement learning (DRL) is a booming area of artificial intelligence. Many practical applications of DRL naturally involve more than one collaborative learners, making it important to study DRL in a multi-agent context. Previous research showed that effective learning in complex multi-agent systems demands fo…
We propose a unified mechanism for achieving coordination and communication in Multi-Agent Reinforcement Learning (MARL), through rewarding agents for having causal influence over other agents' actions. Causal influence is assessed using counterfactual reasoning. At each timestep, an agent simulates alternate actions t…
Humans are capable of attributing latent mental contents such as beliefs or intentions to others. The social skill is critical in daily life for reasoning about the potential consequences of others' behaviors so as to plan ahead. It is known that humans use such reasoning ability recursively by considering what others …
MaxMax Q-Learning improves coordination in multi-agent reinforcement learning by refining action selection.
When multiple agents learn in a decentralized manner, the environment appears non-stationary from the perspective of an individual agent due to the exploration and learning of the other agents. Recently proposed deep multi-agent reinforcement learning methods have tried to mitigate this non-stationarity by attempting t…
New algorithm reduces complexity in multi-agent reinforcement learning.
QLAMMP optimizes fees on AMMs using Q-Learning.
Recently, deep reinforcement learning (RL) methods have been applied successfully to multi-agent scenarios. Typically, these methods rely on a concatenation of agent states to represent the information content required for decentralized decision making. However, concatenation scales poorly to swarm systems with a large…