Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

0111 · Nov 201819922001200920182026
2 results for Subpolicies

DEHRL extends HRL to handle multiple levels with diverse subpolicies.

problem Handling multiple levels of hierarchical reinforcement learning with diverse subpolicies.
method Extensible and scalable framework built levelwise, focusing on diversity of subpolicies.
result DEHRL outperforms state-of-the-art baselines in multiple domains.

Abstract MDPs enable strategic exploration and fast reward transfer in complex environments.

problem Challenging to learn accurate MDPs for high-dimensional states.
method Learn an abstract MDP over low-dimensional coarse states, using an abstraction function.
result Achieves superhuman performance on Pitfall! and higher reward with fewer samples.