SophiaVL-R1: Reinforcing MLLMs Reasoning with Thinking Reward
Explore Similar Repositories
Video-Holmes:[ECCV 2026] Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?
OCR-Reasoning:[ICLR 2026] OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning
AtomThink:[TPAMI 2026] Offical Repository of "AtomThink: Multimodal Slow Thinking with Atomic Step Reasoning"
VideoRFT:[NeurIPS 2025] VideoRFT: Incentivizing Video Reasoning Capability in MLLMs via Reinforced Fine-Tuning
Chiron-o1:[NIPS 2025] Chiron-o1: Igniting Multimodal Large Language Models towards Generalizable Medical Reasoning via Mentor-Intern Collaborative Search
// repository documentation
Was this content helpful?
★ 0(0 ratings)
Recent Feedback
Download README
Do you want to download the README.md file for SophiaVL-R1?