Tag
The article discusses the overlooked timing gap in revoking access for AI agents, highlighting how cached credentials and side effects can lead to unintended actions after revocation, and proposes measuring this delay as a critical metric.
A new preprint called TEPA treats memory validity as a first-class state, revoking outdated precedents when new evidence conflicts while keeping audit trails. It outperforms append-only and last-write-wins in a complete-reversal experiment, though results are not yet independently reproduced.
This paper introduces process sidecars, a two-coefficient edit family for revoking learned state from language models after safety training, achieving second-order accuracy and outperforming naive task arithmetic in experiments across multiple models.