Tag
This paper presents the NLPCC 2026 Shared Task 1: Difficulty-Aware Multilingual and Multimodal Medical Instructional Video Understanding Evaluation (DA-MIVQA), which extends previous benchmarks with difficulty-aware annotations and three tracks: temporal answer grounding, video corpus retrieval, and grounding in corpus. The dataset includes medical instructional videos from public channels and aims to evaluate systems under varying reasoning requirements.