PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning
Authors
Loading...
This research reveals a novel reinforcement learning method that improves multimodal reasoning tasks, suggesting enhanced task performance in complex scenarios.
Zhang et al. (2025) studied this question.