Tag
This paper introduces PatternEval, a diagnostic benchmark for evaluating response-pattern failures in hybrid-thinking multimodal large language models, and proposes PatternRL for aligning these patterns through reinforcement learning with specific penalties.