Tag
This paper audits embedding-cosine similarity thresholds used as quality gates in agent systems, showing they measure wording overlap rather than meaning. Reversals pass the gate while harmless rephrasings often fail, with production drift guards catching zero meaning-breaking mutations and balanced accuracy never exceeding 0.700 across tested configurations.