Robot learns to control itself without rewards.
problem Autonomous reinforcement learning without access to rewards.
method Distributional planning networks optimizing for an embedding space.
result Learned goal metrics enable autonomous reinforcement learning.