counterfactual-evaluation

Tag

Cards List
#counterfactual-evaluation

Counterfactual Fairness Audits of Multi-Step Clinical LLM Agents Require a Measured Per-Action Instability Floor

arXiv cs.CL · 2d ago Cached

This paper demonstrates that counterfactual fairness audits of multi-step clinical LLM agents require measuring per-action instability floors to interpret flip rates accurately, as inherent heterogeneity can mask demographic disparities.

0 favorites 0 likes
#counterfactual-evaluation

Evidence-based governor for coding agents — looking for people to try it and constructive feedback

Reddit r/AI_Agents · 2026-08-17

MARGINAL is an open-source governance layer for coding agents that monitors agent trajectories to prevent inefficient actions, with features like shadow mode and earned enforcement to improve reliability.

0 favorites 0 likes
← Back to home

Submit Feedback