Robuta

https://openreview.net/forum?id=U0k2DVAED5 Width-based Lookaheads with Learnt Base Policies and Heuristics Over the Atari-2600 Benchmark |... We propose new width-based planning and learning algorithms which we apply over the Atari-2600 games.