CARREGANDO O RADAR…
Near-Optimal Reinforcement Learning with Multi-Step Transition Lookahead | Radar arXiv · portela.dev