hybrid-thinking

Tag

Cards List
#hybrid-thinking

Beyond Correctness: Benchmarking and Aligning Response Behaviors in Hybrid-Thinking MLLMs

Hugging Face Daily Papers · 2026-08-17 Cached

This paper introduces PatternEval, a diagnostic benchmark for evaluating response-pattern failures in hybrid-thinking multimodal large language models, and proposes PatternRL for aligning these patterns through reinforcement learning with specific penalties.

0 favorites 0 likes
← Back to home

Submit Feedback