Tag
This paper introduces the FORGE benchmark to evaluate how web content polluted by generative engine optimization can mislead search-augmented LLM recommenders into promoting fake products, revealing significant vulnerabilities and ineffective defenses.