Tag
The paper introduces MemeCULT-1K, a multilingual benchmark of 1,000 South Asian memes to evaluate vision-language models' understanding of cultural context and humor, demonstrating that providing cultural context improves model performance.
A comprehensive survey on multimodal humor understanding using large language models, covering methods, datasets, evaluation protocols, and challenges in interpreting humor in memes, cartoons, and comics.