policy-gradient

★ 160 Open GitHub ↗

Minimal Monte Carlo Policy Gradient (REINFORCE) Algorithm Implementation in Keras

// repository documentation