safe-exploration

Tag

Cards List
#safe-exploration

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models

arXiv cs.AI · 2026-07-16 Cached

This paper introduces DROPJ, a human-centred method for safely training and deploying agent policies by learning a world model from real-world trajectories, then eliciting human preferences with justifications to train a reward model for model predictive control. Experiments show that using human-generated simulated trajectories and justifications improves safety and reduces computational cost.

0 favorites 0 likes
#safe-exploration

Uncertainty-Aware and Temporally Regulated Expert Advice in Reinforcement Learning for Autonomous Driving

arXiv cs.AI · 2026-06-01 Cached

This paper proposes an uncertainty-aware reinforcement learning framework for autonomous driving that uses expert advice guided by adaptive uncertainty thresholds and a commitment-cooldown strategy to improve safety and efficiency. Experiments in the CARLA simulator show a 5-7% success improvement over the IQN baseline.

0 favorites 0 likes
#safe-exploration

Safety Gym

OpenAI Blog · 2019-11-21 Cached

OpenAI introduces Safety Gym, a new benchmark environment and toolkit for studying constrained reinforcement learning and safe exploration. The platform features multiple robots and tasks designed to quantify and measure safe exploration through cost functions alongside reward functions.

0 favorites 0 likes
#safe-exploration

Benchmarking safe exploration in deep reinforcement learning

OpenAI Blog · 2019-11-21 Cached

OpenAI proposes standardizing constrained RL as the formalism for safe exploration and introduces Safety Gym, a benchmark suite for evaluating safe deep RL algorithms in high-dimensional continuous control tasks with safety constraints.

0 favorites 0 likes
← Back to home

Submit Feedback