SpatialGenEval [ICLR 2026] Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models Paper • 2601.20354 • Published Jan 28 • 111
Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models Paper • 2601.20354 • Published Jan 28 • 111
VGPO-RL [ACL 2026] Visually-Guided Policy Optimization for Multimodal Reasoning Visually-Guided Policy Optimization for Multimodal Reasoning Paper • 2604.09349 • Published Apr 10 • 2 MuMing0102/VGPO-RL-7B Image-Text-to-Text • 8B • Updated Apr 13 • 14 MuMing0102/VGPO-RL-32B Image-Text-to-Text • 33B • Updated Apr 13 • 2
Visually-Guided Policy Optimization for Multimodal Reasoning Paper • 2604.09349 • Published Apr 10 • 2
SpatialGenEval [ICLR 2026] Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models Paper • 2601.20354 • Published Jan 28 • 111
Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models Paper • 2601.20354 • Published Jan 28 • 111
VGPO-RL [ACL 2026] Visually-Guided Policy Optimization for Multimodal Reasoning Visually-Guided Policy Optimization for Multimodal Reasoning Paper • 2604.09349 • Published Apr 10 • 2 MuMing0102/VGPO-RL-7B Image-Text-to-Text • 8B • Updated Apr 13 • 14 MuMing0102/VGPO-RL-32B Image-Text-to-Text • 33B • Updated Apr 13 • 2
Visually-Guided Policy Optimization for Multimodal Reasoning Paper • 2604.09349 • Published Apr 10 • 2