Tag
This paper assesses the adversarial robustness of five Arabic language models under character, word, and sentence-level attacks, showing that diacritic insertion can reduce accuracy by 92% and adversarial training improves resilience but has limitations.