GCR-PPO

(★ 42)

GCR-PPO is a modification of PPO for multi-objective robot RL that: - uses a multi-head critic to obtain per-reward advantages and gradients, - applies priority-aware gradient surgery (PCGrad-style projection) to protect task objectives from regularisers, - runs at massively parallel GPU scale within IsaacLab/RSL-RL.

GCR-PPO 최신버젼 다운로드

최종 버전 다운로드 (.zip)
// repository documentation