Tag
This paper introduces DraDDP, the first publicly available English multimodal dataset for multi-party dialogue discourse parsing, built from American TV dramas with 495 segments, 6,374 utterances, and 9.1 hours of video. Benchmarks show multimodal information improves parsing of dialogue structures and relation types.