risk-awareness

Tag

Cards List
#risk-awareness

DreamGuard: Efficient Runtime Guardrail for LLM Agents via Risk-Aware World Model

arXiv cs.AI · 2026-08-07 Cached

DreamGuard is a proactive runtime guardrail for LLM agents that uses a risk-aware world model to track latent state and predict future risks, enabling interventions before unsafe actions execute. It outperforms baselines on benchmarks and online evaluation with 25ms latency.

0 favorites 0 likes
#risk-awareness

SteinGate: Tail-Sensitive Safe Reinforcement Learning via Stein Discrepancy

arXiv cs.LG · 2026-07-16 Cached

SteinGate introduces a distributional safety certificate using Kernelized Stein Discrepancy to detect rare catastrophic tail events in safe reinforcement learning, dynamically adapting policy updates to reduce constraint violations while maintaining competitive returns.

0 favorites 0 likes
#risk-awareness

Risk-Aware LLM Agents for Geospatial Data Retrieval: Design and Preliminary Adversarial Evaluation

arXiv cs.AI · 2026-06-16 Cached

Presents an LLM-driven framework for retrieving remote sensing data from cloud-based geospatial catalogues using natural language queries, with a focus on safety and adversarial robustness. The system integrates three agents for intent interpretation, API call generation, and risk management.

0 favorites 0 likes
← Back to home

Submit Feedback