New approach categorizes objective functions for embodied agents.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
In this paper, we propose an active perception method for recognizing object categories based on the multimodal hierarchical Dirichlet process (MHDP). The MHDP enables a robot to form object categories using multimodal information, e.g., visual, auditory, and haptic information, which can be observed by performing acti…
Perception of artificial agents is one the grand challenges of AI research. Deep Learning and data-driven approaches are successful on constrained problems where perception can be learned using supervision, but do not scale to open-worlds. In such case, for autonomous embodied agents with first-person sensors, percepti…
We propose a statistical model to understand people's perception of their carbon footprint. Driven by the observation that few people think of CO2 impact in absolute terms, we design a system to probe people's perception from simple pairwise comparisons of the relative carbon footprint of their actions. The formulation…
We address the problem of bootstrapping language acquisition for an artificial system similarly to what is observed in experiments with human infants. Our method works by associating meanings to words in manipulation tasks, as a robot interacts with objects and listens to verbal descriptions of the interactions. The mo…
This work presents an asset pricing model that under rational expectation equilibrium perspective shows how, depending on risk aversion and noise volatility, a risky-asset has one equilibrium price that differs in term of efficiency: an informational efficient one (similar to Campbell and Kyle (1993)), and another one …
We propose a planning and perception mechanism for a robot (agent), that can only observe the underlying environment partially, in order to solve an image classification problem. A three-layer architecture is suggested that consists of a meta-layer that decides the intermediate goals, an action-layer that selects local…
Active localization is the problem of generating robot actions that allow it to maximally disambiguate its pose within a reference map. Traditional approaches to this use an information-theoretic criterion for action selection and hand-crafted perceptual models. In this work we propose an end-to-end differentiable meth…
Agents acting in the natural world aim at selecting appropriate actions based on noisy and partial sensory observations. Many behaviors leading to decision mak- ing and action selection in a closed loop setting are naturally phrased within a control theoretic framework. Within the framework of optimal Control Theory, o…
SEMI uses multisensory incongruity to self-supervise exploration in reinforcement learning.
High-frequency financial data of the foreign exchange market (EUR/CHF, EUR/GBP, EUR/JPY, EUR/NOK, EUR/SEK, EUR/USD, NZD/USD, USD/CAD, USD/CHF, USD/JPY, USD/NOK, and USD/SEK) are analyzed by utilizing the Kullback-Leibler divergence between two normalized spectrograms of the tick frequency and the generalized Jensen-Sha…
Active inference is a process theory of the brain that states that all living organisms infer actions in order to minimize their (expected) free energy. However, current experiments are limited to predefined, often discrete, state spaces. In this paper we use recent advances in deep learning to learn the state space an…
A new framework for robot block-stacking tasks using causal probabilistic models.
Machine Learning algorithms are typically regarded as appropriate optimization schemes for minimizing risk functions that are constructed on the training set, which conveys statistical flavor to the corresponding learning problem. When the focus is shifted on perception, which is inherently interwound with time, recent…
Existing model-based reinforcement learning methods often study perception modeling and decision making separately. We introduce joint Perception and Control as Inference (PCI), a general framework to combine perception and control for partially observable environments through Bayesian inference. Based on the fact that…
Unified reinforcement learning and stochastic processes with action-driven processes.
How do organisms recognize their environment by acquiring knowledge about the world, and what actions do they take based on this knowledge? This article examines hypotheses about organisms' adaptation to the environment from machine learning, information-theoretic, and thermodynamic perspectives. We start with construc…
Study the tradeoff between signal distortion and human perception over finite channels.
New RL algorithm learns good actions from offline data, reducing uncertainty and divergence.
A framework isolates VQA reasoning from perception for better model evaluation.
End-to-end autonomous driving perception learns latent features for better performance.
Geometric model explains music perception combining neuroscience and acoustics.
In an effort to better understand the different ways in which the discount factor affects the optimization process in reinforcement learning, we designed a set of experiments to study each effect in isolation. Our analysis reveals that the common perception that poor performance of low discount factors is caused by (to…
PHASE dataset simulates complex social interactions in physical environments.
RETR improves indoor radar perception with a novel transformer model.
Researchers study how teachers' advising relationships influence their perceptions of satisfaction and students, not policy influence.
Unified control theory and machine learning for safety in uncertain systems.
Reconstruction-based learning produces uninformative features for perception tasks.
In this paper we investigate the higher dimensional divergence functions of mapping class groups of surfaces and of CAT(0)--groups. We show that, for mapping class groups of surfaces, these functions exhibit phase transitions at the rank (as measured by thrice the genus plus the number of punctures minus 3). We also pr…
Toward enabling next-generation robots capable of socially intelligent interaction with humans, we present a of interactions in a social environment of multiple agents and multiple groups. The Multiagent Group Perception and Interaction (MGpi) network is a deep neural network that predi…
We show that power-law analyses of financial commentaries from newspaper web-sites can be used to identify stock market bubbles, supplementing traditional volatility analyses. Using a four-year corpus of 17,713 online, finance-related articles (10M+ words) from the Financial Times, the New York Times, and the BBC, we s…
Stabilizes policy optimization with off-policy data using divergence augmentation.
New approach mitigates feedback divergence in imitation learning.
PeL separates sensory interface optimization from decision learning.
New measures generalize existing ones, linking information and risk.
PERCEPT detects changes in high-dimensional data streams using topological data analysis.
Sensorimotor contingency theory offers a promising account of the nature of perception, a topic rarely addressed in the robotics community. We propose a developmental framework to address the problem of the autonomous acquisition of sensorimotor contingencies by a naive robot. While exploring the world, the robot inter…
While perception tasks such as visual object recognition and text understanding play an important role in human intelligence, the subsequent tasks that involve inference, reasoning and planning require an even higher level of intelligence. The past few years have seen major advances in many perception tasks using deep …
We address the problem of imitation learning with multi-modal demonstrations. Instead of attempting to learn all modes, we argue that in many tasks it is sufficient to imitate any one of them. We show that the state-of-the-art methods such as GAIL and behavior cloning, due to their choice of loss function, often incorr…
HR in 8D encodes unique conformal gravity with negative curvature.
Default-ERM shortcut learning persists even without additional information.
DGP learns speech recognition by modeling complex relationships between utterances.
Proves helicity is the only regular Casimir for 3D hydrodynamics.
Motivated by vision-based control of autonomous vehicles, we consider the problem of controlling a known linear dynamical system for which partial state information, such as vehicle position, is extracted from complex and nonlinear data, such as a camera image. Our approach is to use a learned perception map that predi…
A function is exponentially concave if its exponential is concave. We consider exponentially concave functions on the unit simplex. In a previous paper we showed that gradient maps of exponentially concave functions provide solutions to a Monge-Kantorovich optimal transport problem and give a better gradient approximat…
We analyse perception and memory, using mathematical models for knowledge graphs and tensors, to gain insights into the corresponding functionalities of the human mind. Our discussion is based on the concept of propositional sentences consisting of \textit{subject-predicate-object} (SPO) triples for expressing elementa…
An algorithm detects anomalies based on human perception principles.
New coding theorem shows achievable rate matches theoretical limit.