David Simões, Nuno Lau, Luís Paulo Reis: Adjusted Bounded Weighted Policy Learner. RoboCup 2018: 324-336