Improved online planning with lookahead policies for large state spaces.
problem Online planning with large state spaces and limited state updates.
method Developed -RTDP algorithm with -step lookahead policy.
result Improved sample complexity with increased lookahead horizon.