Tag
The author conducted 356 prompt-injection trials across six models and three harnesses, revealing that workspace elements can enable attacks that otherwise fail, and shares the benchmark for evaluating AI safety.