safe-rl

Tag

Cards List
#safe-rl

Boundary-Seeking Policy Gradient for Safe Reinforcement Learning

arXiv cs.LG · 2026-08-12 Cached

Introduces Boundary-Seeking Policy Gradient (BSPG), a first-order method for safe reinforcement learning that actively drives the policy toward the constraint boundary, with convergence guarantees and improved reward/boundary tracking on a Safety-Gymnasium task.

0 favorites 0 likes
#safe-rl

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models

arXiv cs.AI · 2026-07-16 Cached

This paper introduces DROPJ, a human-centred method for safely training and deploying agent policies by learning a world model from real-world trajectories, then eliciting human preferences with justifications to train a reward model for model predictive control. Experiments show that using human-generated simulated trajectories and justifications improves safety and reduces computational cost.

0 favorites 0 likes
#safe-rl

SafeExplorer: An Unbiased Policy Gradient for Reinforcement Learning with Recovery Interventions

arXiv cs.LG · 2026-07-13 Cached

SafeExplorer introduces an unbiased policy gradient estimator for reinforcement learning with recovery interventions, significantly reducing training-time falls on robot tasks while matching or exceeding standard PPO's final reward.

0 favorites 0 likes
#safe-rl

Safe Online Learning via Smooth Safety-Structured Policy Composition

arXiv cs.LG · 2026-07-01 Cached

This paper proposes AutoSafe, a safety-aware policy architecture for safe online reinforcement learning that integrates structured safety monitoring and intervention directly into action generation, enabling smooth, risk-dependent transitions between performance and safety behaviors, demonstrated on benchmarks and a physical cart-pole system.

0 favorites 0 likes
#safe-rl

Sample-efficient Transfer Reinforcement Learning via Adaptive Reward Shaping and Policy-Ratio Reweighting Strategy

arXiv cs.LG · 2026-06-26 Cached

This paper proposes a safe transfer reinforcement learning framework for autonomous highway lane-changing, using adaptive teacher intervention and reward shaping to improve sample efficiency and safety. Experiments show over 52% improvement in safety and 5% improvement in efficiency over baselines.

0 favorites 0 likes
#safe-rl

TRIDENT: Breaking the Hybrid-Safety-Physics Coupling for Provably Safe Multi-Agent Reinforcement Learning

arXiv cs.LG · 2026-06-18 Cached

TRIDENT is a novel multi-agent reinforcement learning framework that breaks the coupling between hybrid discrete-continuous actions, hard safety constraints, and physics-governed dynamics, achieving provably safe coordination with a convergence guarantee to a constrained Nash equilibrium and significant reductions in training-time violations.

0 favorites 0 likes
#safe-rl

Seeing Before Colliding: Anticipatory Safe RL with Frozen Vision-Language Models

arXiv cs.LG · 2026-06-11 Cached

This paper presents VLM-Safe-RL, a framework that integrates frozen vision-language models into constrained MDP Lagrangian updates to provide anticipatory cost signals for safe reinforcement learning in high-speed visual control tasks. The method outperforms standard constraint-aware baselines on Safety-Gymnasium FormulaOne L2 and generalizes to held-out environments.

0 favorites 0 likes
#safe-rl

Safe Continual Reinforcement Learning under Nonstationarity via Adaptive Safety Constraints

arXiv cs.LG · 2026-05-20

Proposes LILAC+, a framework for safe continual reinforcement learning under nonstationarity that uses three adaptive safety mechanisms: context-based safety constraints, adaptation-speed constraints, and budget-to-state safety enforcement. Evaluations in simulated driving environments show reduced safety violations under distribution shift while maintaining competitive performance.

0 favorites 0 likes
#safe-rl

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability

arXiv cs.LG · 2026-05-15 Cached

This paper proposes Action-Conditioned Risk Gating, a lightweight reinforcement learning method for risk-sensitive control under partial observability that uses a compact finite-history proxy state and an action-conditioned near-term risk predictor to balance safety and performance.

0 favorites 0 likes
#safe-rl

Learning When to Act: Communication-Efficient Reinforcement Learning via Run-Time Assurance

arXiv cs.LG · 2026-05-14 Cached

This paper presents a framework (CARE) that jointly learns control inputs and communication-efficient timing decisions under a pointwise Lyapunov safety shield, achieving higher inter-sample intervals than classical methods on inverted pendulum, cart-pole, and planar quadrotor systems.

0 favorites 0 likes
← Back to home

Submit Feedback