Tag
Introduces CCPL, a method to address delayed and stochastic consequences in constrained reinforcement learning using a delay-corrected Bellman operator and an Interventional Consequence Net for causal attribution, with a contraction proof under unknown delays.