Perception of artificial agents is one the grand challenges of AI research. Deep Learning and data-driven approaches are successful on constrained problems where perception can be learned using supervision, but do not scale to open-worlds. In such case, for autonomous embodied agents with first-person sensors, percepti…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
This paper proposes a new method to connect language and physical actions in reinforcement learning.
Reinforcement learning for embodied agents is a challenging problem. The accumulated reward to be optimized is often a very rugged function, and gradient methods are impaired by many local optimizers. We demonstrate, in an experimental setting, that incorporating an intrinsic reward can smoothen the optimization landsc…
New approach categorizes objective functions for embodied agents.
EDGI improves sample efficiency and generalization in tasks with spatial and temporal symmetries.
A multi-agent system is trialed as a means of crowd-sourcing inexpensive but high quality streams of predictions. Each agent is a microservice embodying statistical models and endowed with economic self-interest. The ability to fork and modify simple agents is granted to a large number of employees in a firm and empiri…
New metric solves correspondence problem for robotic arm imitation learning.
VALAN is a lightweight and scalable software framework for deep reinforcement learning based on the SEED RL architecture. The framework facilitates the development and evaluation of embodied agents for solving grounded language understanding tasks, such as Vision-and-Language Navigation and Vision-and-Dialog Navigation…
New text-to-image diffusion models improve scene understanding for AI agents.
New framework for AI to learn causal models through experience.
Learning predictive models from interaction with the world allows an agent, such as a robot, to learn about how the world works, and then use this learned model to plan coordinated sequences of actions to bring about desired outcomes. However, learning a model that captures the dynamics of complex skills represents a m…
For embodied agents to infer representations of the underlying 3D physical world they inhabit, they should efficiently combine multisensory cues from numerous trials, e.g., by looking at and touching objects. Despite its importance, multisensory 3D scene representation learning has received less attention compared to t…
Improves AI agents' 3D navigation by learning from failures and 3D spatial relationships.
What is the role of real-time control and learning in the formation of social conventions? To answer this question, we propose a computational model that matches human behavioral data in a social decision-making game that was analyzed both in discrete-time and continuous-time setups. Furthermore, unlike previous approa…
GWIL uses Gromov-Wasserstein distance to align expert and imitation agent states.
Recent efforts on training visual navigation agents conditioned on language using deep reinforcement learning have been successful in learning policies for different multimodal tasks, such as semantic goal navigation and embodied question answering. In this paper, we propose a multitask model capable of jointly learnin…
We investigate using reinforcement learning agents as generative models of images (extending arXiv:1804.01118). A generative agent controls a simulated painting environment, and is trained with rewards provided by a discriminator network simultaneously trained to assess the realism of the agent's samples, either uncond…
Many robotic applications require the agent to perform long-horizon tasks in partially observable environments. In such applications, decision making at any step can depend on observations received far in the past. Hence, being able to properly memorize and utilize the long-term history is crucial. In this work, we pro…
We are increasingly surrounded by artificially intelligent technology that takes decisions and executes actions on our behalf. This creates a pressing need for general means to communicate with, instruct and guide artificial agents, with human language the most compelling means for such communication. To achieve this i…
Animals (especially humans) have an amazing ability to learn new tasks quickly, and switch between them flexibly. How brains support this ability is largely unknown, both neuroscientifically and algorithmically. One reasonable supposition is that modules drawing on an underlying general-purpose sensory representation a…
The paper introduces affordances for reinforcement learning, improving planning and learning efficiency.
A major challenge in cognitive science and AI has been to understand how autonomous agents might acquire and predict behavioral and mental states of other agents in the course of complex social interactions. How does such an agent model the goals, beliefs, and actions of other agents it interacts with? What are the com…
AIF improves physical AI agents' performance in dynamic environments.
New RL method learns from passive data by modeling intentions.
The creation of machine learning algorithms for intelligent agents capable of continuous, lifelong learning is a critical objective for algorithms being deployed on real-life systems in dynamic environments. Here we present an algorithm inspired by neuromodulatory mechanisms in the human brain that integrates and expan…
DIVA generates diverse tasks for complex simulators, enabling adaptive agent training.
We study the question of how to imitate tasks across domains with discrepancies such as embodiment, viewpoint, and dynamics mismatch. Many prior works require paired, aligned demonstrations and an additional RL step that requires environment interactions. However, paired, aligned demonstrations are seldom obtainable an…
Embodied cognition states that semantics is encoded in the brain as firing patterns of neural circuits, which are learned according to the statistical structure of human multimodal experience. However, each human brain is idiosyncratically biased, according to its subjective experience history, making this biological s…
Global supply networks in agriculture, manufacturing, and services are a defining feature of the modern world. The efficiency and the distribution of surpluses across different parts of these networks depend on choices of intermediaries. This paper conducts price formation experiments with human subjects located in lar…
Robotic vision is a field where continual learning can play a significant role. An embodied agent operating in a complex environment subject to frequent and unpredictable changes is required to learn and adapt continuously. In the context of object recognition, for example, a robot should be able to learn (without forg…
We study financial distributions within the framework of the continuous time random walk (CTRW). We review earlier approaches and present new results related to overnight effects as well as the generalization of the formalism which embodies a non-Markovian formulation of the CTRW aimed to account for correlated increme…
Paper proves Toponogov's theorem in Alexandrov geometry.
We construct an elementary, combinatorial kind of topological quantum field theory, based on curves, surfaces, and orientations. The construction derives from contact invariants in sutured Floer homology and is essentially an elaboration of a TQFT defined by Honda--Kazez--Matic. This topological field theory stores inf…
We investigate how the choice of decision makers can be varied under the presence of risk and uncertainty. Our analysis is based on the approach we have previously applied to individual decision makers, which we now generalize to the case of decision makers that are members of a society. The approach employs the mathem…
Econophysics embodies the recent upsurge of interest by physicists into financial economics, driven by the availability of large amount of data, job shortage in physics and the possibility of applying many-body techniques developed in statistical and theoretical physics to the understanding of the self-organizing econo…
Most common navigation tasks in human environments require auxiliary arm interactions, e.g. opening doors, pressing buttons and pushing obstacles away. This type of navigation tasks, which we call Interactive Navigation, requires the use of mobile manipulators: mobile bases with manipulation capabilities. Interactive N…
New method automates asymmetric choice for better skill transfer in reinforcement learning.
We describe three perspectives on higher quantization, using the example of magnetic Poisson structures which embody recent discussions of nonassociativity in quantum mechanics with magnetic monopoles and string theory with non-geometric fluxes. We survey approaches based on deformation quantization of twisted Poisson …
New framework uses OR to ensure AI systems make safe decisions.
FinRL automates trading in quantitative finance with deep reinforcement learning.
Federated learning (FL) is a machine learning setting where many clients (e.g. mobile devices or whole organizations) collaboratively train a model under the orchestration of a central server (e.g. service provider), while keeping the training data decentralized. FL embodies the principles of focused data collection an…
Multiple Kernel Learning (MKL) is used to replicate the signal combination process that trading rules embody when they aggregate multiple sources of financial information when predicting an asset's price movements. A set of financially motivated kernels is constructed for the EURUSD currency pair and is used to predict…
Commissioned by MIT's in-house artist Jane Philbrick, we evolve an abstract 2D surface (resembling Marta Pan's 1961 "Sculpture Flottante I") under mean curvature, all the while calculating the eigenmodes and eigenvalues of the Laplace-Beltrami operator on the resulting shapes. These are then synthesized into a sound-wa…
We use the mapping cone for the relative deRham cohomology of a manifold with boundary in order to show that the Chern-Gauss-Bonnet Theorem for oriented Riemannian vector bundles over such manifolds is a manifestation of Lefschetz Duality in any of the two embodiments of the latter. We explain how Thom isomorphism fits…
Integral formulae for foliated Riemannian manifolds provide obstructions for existence of foliations or compact leaves of them with given geometric properties. Recently, we associated a new Riemannian metric to a codimension-one foliated Finsler space and proved integral formulae for general and for Randers spaces. In …
We introduce a conceptually novel structured prediction model, GPstruct, which is kernelized, non-parametric and Bayesian, by design. We motivate the model with respect to existing approaches, among others, conditional random fields (CRFs), maximum margin Markov networks (M3N), and structured support vector machines (S…
A new approach to group fairness treats it as a bargaining problem.
In recent years, manifold learning has become increasingly popular as a tool for performing non-linear dimensionality reduction. This has led to the development of numerous algorithms of varying degrees of complexity that aim to recover man ifold geometry using either local or global features of the data. Building on t…