Markov Decision Process

AI and Machine Learning · Reinforcement Learning · also: MDP · markov-decision-process.yaml

States, actions, transition probabilities and rewards, where the next state depends only on the present one. The formalism every method below is defined against.

Colour is the family; a dashed line is the second member of it.

Drag to pan · scroll to zoom · click a node to open it
G markov-decision-process Markov Decision Process temporal-difference-learning Temporal Difference Learning temporal-difference-learning->markov-decision-process estimates the value function the MDP defines

Referenced by

References