SophiaVL-R1

SophiaVL-R1: Reinforcing MLLMs Reasoning with Thinking Reward

// repository documentation