Tag
This paper tests the independence assumption underlying compositional reliability bounds for multi-agent systems, finding that same-model agents co-fail at high rates and that common certificates are unsound. It proposes a finite-sample, dependence-free certificate via linear programming over co-execution moments, validated on 18,000 missions.
The author describes pivoting their open-source workflow-sharing project into Scyvera, a tool that defines contracts for agentic workflows to establish boundaries, permissions, and governance. They invite community feedback on the abstraction.