State-only imitation learning improves dexterous manipulation learning from videos.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
We use reinforcement learning (RL) to learn dexterous in-hand manipulation policies which can perform vision-based object reorientation on a physical Shadow Dexterous Hand. The training is performed in a simulated environment in which we randomize many of the physical properties of the system like friction coefficients…
This work tackles real-world robotic reinforcement learning challenges.
SAVO actor improves reinforcement learning by avoiding local optima in complex Q-functions.
This work improves RL for complex robotic tasks by guiding exploration with task-specific goal distributions.
Paper tackles multi-object reinforcement learning, improving skill extrapolation.
We adapt the ideas underlying the success of Deep Q-Learning to the continuous action domain. We present an actor-critic, model-free algorithm based on the deterministic policy gradient that can operate over continuous action spaces. Using the same learning algorithm, network architecture and hyper-parameters, our algo…
AWAC combines offline and online data to accelerate RL learning.
The study of dexterous manipulation has provided important insights in humans sensorimotor control as well as inspiration for manipulation strategies in robotic hands. Previous work focused on experimental environment with restrictions. Here we describe a method using the deformation and color distribution of the finge…
KL-regularized RL from expert demos can lead to slow, unstable learning.
We propose a plan online and learn offline (POLO) framework for the setting where an agent, with an internal model, needs to continually act and learn in the world. Our work builds on the synergistic relationship between local model-based control, global value function learning, and exploration. We study how local traj…
Paper proposes a sequential statistical test for comparing imitation learning policies with near-optimal stopping.
ROBEL is an open-source platform of cost-effective robots designed for reinforcement learning in the real world. ROBEL introduces two robots, each aimed to accelerate reinforcement learning research in different task domains: D'Claw is a three-fingered hand robot that facilitates learning dexterous manipulation tasks, …
Non-invasive myoelectric prostheses require a long training time to obtain satisfactory control dexterity. These training times could possibly be reduced by leveraging over training efforts by previous subjects. So-called domain adaptation algorithms formalize this strategy and have indeed been shown to significantly r…
Model-free deep reinforcement learning (RL) algorithms have been successfully applied to a range of challenging sequential decision making and control tasks. However, these methods typically suffer from two major challenges: high sample complexity and brittleness to hyperparameters. Both of these challenges limit the a…
Paper teaches robots to play piano with touch and learning.
A game-theoretic approach simplifies MBRL design and improves sample efficiency.
Deep learning is an established framework for learning hierarchical data representations. While compute power is in abundance, one of the main challenges in applying this framework to robotic grasping has been obtaining the amount of data needed to learn these representations, and structuring the data to the task at ha…
AI learns market manipulation through simulation, suggesting regulation.
The paper analyzes how leverage affects manipulation in event-linked markets, offering new insights into regulation.
New model improves neural network robustness against input manipulations.
This paper studies poisoning attacks in episodic RL and discovers their effectiveness depends on reward bounds.
Manipulating video content is easier than ever. Due to the misuse potential of manipulated content, multiple detection techniques that analyze the pixel data from the videos have been proposed. However, clever manipulators should also carefully forge the metadata and auxiliary header information, which is harder to do …
In this work we propose a model that can manipulate individual visual attributes of objects in a real scene using examples of how respective attribute manipulations affect the output of a simulation. As an example, we train our model to manipulate the expression of a human face using nonphotorealistic 3D renders of a f…
Prediction markets can be manipulated by traders who can move contract settlements, harming price discovery.
DIGIT is a low-cost tactile sensor for in-hand manipulation.
Volunteer labor can temporarily yield lower benefits to charities than its costs. In such instances, organizations may wish to defer volunteer donations to a later date. Exploiting a discontinuity in blood donations' eligibility criteria, we show that deferring donors reduces their future volunteerism. In our setting, …
Manipulating data, such as weighting data examples or augmenting with new instances, has been increasingly used to improve model training. Previous work has studied various rule- or learning-based approaches designed for specific types of data manipulation. In this work, we propose a new method that supports learning d…
Paper presents a new port-Hamiltonian model for vehicle manipulators.
Optimal benchmark design varies based on costs in financial manipulation.
Study examines reasons for Nutek India's share price drop.
New method uses statistical physics to detect financial market manipulation.
Market manipulation is a strategy used by traders to alter the price of financial securities. One type of manipulation is based on the process of buying or selling assets by using several trading strategies, among them spoofing is a popular strategy and is considered illegal by market regulators. Some promising tools h…
New foundation for Shapley value immune to coalitional manipulations.
Framework detects covert financial market manipulation using LOB representations.
Study reveals widespread manipulation of meme coins, leading to significant economic losses.
This paper focuses on an extension of the Limit Order Book (LOB) model with general shape introduced by Alfonsi, Fruth and Schied. Here, the additional feature allows a time-varying LOB depth. We solve the optimal execution problem in this framework for both discrete and continuous time strategies. This gives in partic…
A new method for robot manipulation tasks using imagined object goals.
This work tackles force control for contact-rich manipulation tasks with rigid robots using RL.
Manipulation is an important issue for both developed and emerging stock markets. For the study of manipulation, it is critical to analyze investor behavior in the stock market. In this paper, an analysis of the full transaction records of over a hundred stocks in a one-year period is conducted. For each stock, a tradi…
Study on costs of manipulating AMM-based price oracles.
Many problems in image processing and computer vision (e.g. colorization, style transfer) can be posed as 'manipulating' an input image into a corresponding output image given a user-specified guiding signal. A holy-grail solution towards generic image manipulation should be able to efficiently alter an input image wit…
Investigates optimal execution under time-varying liquidity, preventing price manipulation.
Tool manipulation is vital for facilitating robots to complete challenging task goals. It requires reasoning about the desired effect of the task and thus properly grasping and manipulating the tool to achieve the task. Task-agnostic grasping optimizes for grasp robustness while ignoring crucial task-specific constrain…
New attack manipulates UCB algorithm, new defense algorithm reduces pseudo-regret.
A new reward shaping method balances learning efficiency and effectiveness for robot manipulation.
In financial markets, liquidity is not constant over time but exhibits strong seasonal patterns. In this article we consider a limit order book model that allows for time-dependent, deterministic depth and resilience of the book and determine optimal portfolio liquidation strategies. In a first model variant, we propos…
As the decade turns, we reflect on nearly thirty years of successful manipulation of the world's public equity markets. This reflection highlights a few of the key enabling ingredients and lessons learned along the way. A quantitative understanding of market impact and its decay, which we cover briefly, lets you move l…