Potential-Based Reward Shaping

AI and Machine Learning · Reinforcement Learning · 1999 · potential-based-reward-shaping.yaml

Add the difference of a potential function to the reward, and the optimal policy provably does not move. The result that turned reward engineering from guesswork into a statement about which extra signal is safe to give.

Colour is the family; a dashed line is the second member of it.

Drag to pan · scroll to zoom · click a node to open it
G markov-decision-process Markov Decision Process potential-based-reward-shaping Potential-Based Reward Shaping potential-based-reward-shaping->markov-decision-process dense guidance can be added to a sparse reward without moving its optimum

This node

References