research
∙
01/18/2021
Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint
In reinforcement learning, temporal difference-based algorithms can be s...
research
∙
02/20/2018