Method solves learning problem with hierarchical control objectives.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
We describe a novel extension of soft actor-critics for hierarchical Deep Q-Networks (HDQN) architectures using mutual information metric. The proposed extension provides a suitable framework for encouraging explorations in such hierarchical networks. A natural utilization of this framework is an adversarial setting, w…
Autonomous agents can learn by imitating teacher demonstrations of the intended behavior. Hierarchical control policies are ubiquitously useful for such learning, having the potential to break down structured tasks into simpler sub-tasks, thereby improving data efficiency and generalization. In this paper, we propose a…
HiDe learns hierarchical control for complex tasks by separating planning and control.
Hierarchical-CPI improves variable importance measurement for medical data.
Nonlinear optimal control problems are often solved with numerical methods that require knowledge of system's dynamics which may be difficult to infer, and that carry a large computational cost associated with iterative calculations. We present a novel neurobiologically inspired hierarchical learning framework, Reinfor…
We introduce agents that use object-oriented reasoning to consider alternate states of the world in order to more quickly find solutions to problems. Specifically, a hierarchical controller directs a low-level agent to behave as if objects in the scene were added, deleted, or modified. The actions taken by the controll…
We study a generalized setup for learning from demonstration to build an agent that can manipulate novel objects in unseen scenarios by looking at only a single video of human demonstration from a third-person perspective. To accomplish this goal, our agent should not only learn to understand the intent of the demonstr…
Hyperbolic space outperforms Euclidean in learning hierarchical data.
Random quotients preserve hyperbolic properties in groups.
Proposes selective inference for testing differences in means between clusters.
Fractal Flow enhances normalizing flows with interpretable latent space and hierarchical modeling.
Unified theory for neural scaling laws in hierarchically compositional data.
The paper offers a method to create prediction sets with uncertainty control.
Real-world tasks are often highly structured. Hierarchical reinforcement learning (HRL) has attracted research interest as an approach for leveraging the hierarchical structure of a given task in reinforcement learning (RL). However, identifying the hierarchical policy structure that enhances the performance of RL is n…
A deep learning approach classifies medical images hierarchically.
Randomized hierarchical clustering tests for stability and detects clusters.
One of the challenges in model-based control of stochastic dynamical systems is that the state transition dynamics are involved, and it is not easy or efficient to make good-quality predictions of the states. Moreover, there are not many representational models for the majority of autonomous systems, as it is not easy …
A novel unsupervised domain adaptation method using hierarchical optimal transport.
Hierarchical reinforcement learning is a promising approach to tackle long-horizon decision-making problems with sparse rewards. Unfortunately, most methods still decouple the lower-level skill acquisition process and the training of a higher level that controls the skills in a new task. Leaving the skills fixed can le…
When forecasting time series with a hierarchical structure, the existing state of the art is to forecast each time series independently, and, in a post-treatment step, to reconcile the time series in a way that respects the hierarchy (Hyndman et al., 2011; Wickramasuriya et al., 2018). We propose a new loss function th…
Deep networks learn hierarchical functions more efficiently than shallow ones.
Study on infinitely-wide CNNs and their adaptability to function spatial scales.
Classification with Costly Features (CwCF) is a classification problem that includes the cost of features in the optimization criteria. Individually for each sample, its features are sequentially acquired to maximize accuracy while minimizing the acquired features' cost. However, existing approaches can only process da…
This paper uses NLDT to find interpretable control rules from complex DRL policies.
Centralised training with decentralised execution is an important setting for cooperative deep multi-agent reinforcement learning due to communication constraints during execution and computational tractability in training. In this paper, we analyse value-based methods that are known to have superior performance in com…
Hierarchical AI multi-agent framework optimizes equity portfolios in China's A-share market.
Networks of coupled dynamical systems provide a powerful way to model systems with enormously complex dynamics, such as the human brain. Control of synchronization in such networked systems has far reaching applications in many domains, including engineering and medicine. In this paper, we formulate the synchronization…
AIRL learns robust, generalizable reward functions from demonstrations.
Robot learns multiple tasks hierarchically by transferring knowledge.
New origami structures adapt to over 100 shapes with minimal actuation.
Predicting outcomes and planning interactions with the physical world are long-standing goals for machine learning. A variety of such tasks involves continuous physical systems, which can be described by partial differential equations (PDEs) with many degrees of freedom. Existing methods that aim to control the dynamic…
The paper introduces an adjacency constraint to improve goal-conditioned HRL.
In this paper, we present Gamma-LSTM, an enhanced long short term memory (LSTM) unit, to enable learning of hierarchical representations through multiple stages of temporal abstractions. Gamma memory, a hierarchical memory unit, forms the central memory of Gamma-LSTM with gates to regulate the information flow into var…
Director learns hierarchical behaviors from pixels, outperforming exploration methods.
Develops a new framework for large-scale geometry.
A decentralized deep RL controller improves hexapod locomotion learning.
We present a novel method for exact hierarchical sparse polynomial regression. Our regressor is that degree polynomial which depends on at most inputs, counting at most monomial terms, which minimizes the sum of the squares of its prediction errors. The previous hierarchical sparse specification aligns w…
The paper monitors stock market relationships using network analysis and statistical control charts.
A new method for optimizing hierarchical multi-objective problems.
Hierarchical GANs reduce anomaly detection costs.
LineFlow is a framework for training RL agents to control production lines.
In supervised clustering, standard techniques for learning a pairwise dissimilarity function often suffer from a discrepancy between the training and clustering objectives, leading to poor cluster quality. Rectifying this discrepancy necessitates matching the procedure for training the dissimilarity function to the clu…
NeurT-FDR controls FDR by incorporating feature hierarchy.
Exploration in sparse reward reinforcement learning remains an open challenge. Many state-of-the-art methods use intrinsic motivation to complement the sparse extrinsic reward signal, giving the agent more opportunities to receive feedback during exploration. Commonly these signals are added as bonus rewards, which res…
Speech-related Brain Computer Interface (BCI) technologies provide effective vocal communication strategies for controlling devices through speech commands interpreted from brain signals. In order to infer imagined speech from active thoughts, we propose a novel hierarchical deep learning BCI system for subject-indepen…
Policy-gradient method controls multiple non-cohesive targets.
HKT improves sequence processing with multi-scale attention and kernel analysis.