FSPO

(★ 26)

Official code for our paper "Reasoning Models Hallucinate More: Factuality-Aware Reinforcement Learning for Large Reasoning Models"

FSPO Latest Version Download

Download Latest Version (.zip)
// repository documentation