GridWorld_-Planning-RL-

(★ 12)

Contains policy iteration and value iteration (planning). Also contains Q-learning (RL). Uses these methods in the context of the GridWorld problem where the agent's goal is to take the quickest path to reach the terminal state.

GridWorld_-Planning-RL- 최신버젼 다운로드

최종 버전 다운로드 (.zip)
// repository documentation