GridWorld_-Planning-RL-

(★ 12)

Contains policy iteration and value iteration (planning). Also contains Q-learning (RL). Uses these methods in the context of the GridWorld problem where the agent's goal is to take the quickest path to reach the terminal state.

GridWorld_-Planning-RL- Latest Version Download

Download Latest Version (.zip)
// repository documentation