policy-adaptation

Tag

Cards List
#policy-adaptation

Local Edits, Global Ripples: Replay-Informed Policy Adaptation for Workflow Synthesis

arXiv cs.CL ↗ · 2026-09-14 Cached

The paper introduces RIPPLE, a method for persistent prompt-policy editing in workflow synthesis that addresses edit locality and composition sensitivity, improving validation success by up to 23.1% on a synthetic benchmark.

0 favorites 0 likes
#policy-adaptation

As AI agents go rogue, cyber insurers are adapting their policies

Reddit r/ArtificialInteligence ↗ · 2026-08-27

Cyber insurers are adapting their policies to address emerging risks from rogue AI agents, reflecting broader industry changes in response to AI advancements.

0 favorites 0 likes
#policy-adaptation

CoAdapt-GUI: Joint Workflow Context and Policy Adaptation for Unseen GUI Applications

arXiv cs.AI ↗ · 2026-08-13 Cached

CoAdapt-GUI is a test-time adaptation framework for mobile GUI agents that jointly adapts workflow context and policy, improving performance on unseen-app benchmarks like AndroidWorld-Generalization and AndroidWorld Plus.

0 favorites 0 likes
#policy-adaptation

PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails

Hugging Face Daily Papers ↗ · 2026-07-07 Cached

This paper introduces PolicyShiftBench, a benchmark for policy-adaptive image guardrails, and PolicyShiftGuard, a compact model trained with a two-stage method that improves performance under shifting safety policies.

0 favorites 0 likes
#policy-adaptation

OmniTacTune: Policy-Agnostic Real-World RL for Tactile Residual Adaptation of Visual Policies

Hugging Face Daily Papers ↗ · 2026-07-04 Cached

OmniTacTune introduces a two-stage reinforcement learning pipeline for adapting tactile feedback to pretrained visual robot policies, achieving 85-100% success on contact-rich manipulation tasks within 40-80 minutes.

0 favorites 0 likes
#policy-adaptation

Warp RL: Reshaping Base Policy Distributions for Dynamics Adaptation

arXiv cs.LG ↗ · 2026-07-01 Cached

Warp RL replaces additive residual corrections in reinforcement learning with an invertible, state-conditioned transformation of the base policy's action distribution using monotonic rational-quadratic spline flows, enabling adaptation of distribution shape, scale, and geometry under dynamics shifts. It matches or outperforms residual correction in ManiSkill3 manipulation tasks and achieves 30% faster task completion in a real robot peg-insertion task.

0 favorites 0 likes
#policy-adaptation

@AdinaYakup: SingGuard from Ant Group @AntLingAGI A multimodal guardrail where the safety policy is an input, not a fixed weight. - …

X AI KOLs Timeline ↗ · 2026-06-22 Cached

SingGuard is a multimodal guardrail system from Ant Group that treats safety policy as an input, allowing dynamic adaptation via natural language. It is released under Apache 2.0 and covers text and image modalities.

0 favorites 0 likes
#policy-adaptation

Robotic Policy Adaptation via Weight-Space Meta-Learning

Hugging Face Daily Papers ↗ · 2026-06-05 Cached

Introduces WIZARD, a weight-space meta-learning framework that generates task-specific LoRA parameters for frozen VLA policies from language instructions and demonstration videos, enabling efficient task adaptation without fine-tuning.

0 favorites 0 likes
← Back to home

Submit Feedback