safety-guarantees

Tag

Cards List
#safety-guarantees

MAGS: Multi-agent Auto-formalization Guarantees Safety for Agentic Outputs

arXiv cs.AI · 5d ago Cached

MAGS introduces a multi-agent framework that uses formal verification with Dafny to generate executable programs with safety guarantees from LLM coding agents, achieving 100% success in producing verified code across domains like CUDA kernels and robotic tasks.

0 favorites 0 likes
#safety-guarantees

Toward Safe LLM Agents: A Survey of Specification, Verification, and Enforcement

arXiv cs.AI · 2026-08-18 Cached

This survey paper reviews 38 studies on safe LLM agents, highlighting key challenges such as specification translation bottlenecks, incomplete safety guarantees from enforcement methods like runtime monitoring, and the verifier tax that impedes safe task completion.

0 favorites 0 likes
#safety-guarantees

Are Safety Guarantees in Neural Networks Safe? How to Compute Trustworthy Robustness Certifications

arXiv cs.LG · 2026-06-24 Cached

This paper introduces the apothem measure for computing trustworthy robustness certifications in neural networks, proves intractability of volume-optimal certifications, and presents the ParallelepipedoNN system achieving twofold improvement in minimum edge length on MNIST and Fashion MNIST.

0 favorites 0 likes
← Back to home

Submit Feedback