Tag
The paper introduces KoNeoBench, a curated benchmark dataset for evaluating large language models' understanding of Korean neologisms, based on 1,785 entries from online news since 2020. It reveals limitations in current LLMs in handling recent lexical changes and specific Korean linguistic properties.