Scaling RL to Long Videos
Yukang Chen, Wei Huang, Baifeng Shi +11 authors
A framework scales vision-language models for long video reasoning using reinforcement learning, achieving strong performance on benchmarks and demonstrating consistent gains with increased video frames.