policy-adaptation

Tag

Cards List
#policy-adaptation

PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails

Hugging Face Daily Papers · 2026-07-07 Cached

This paper introduces PolicyShiftBench, a benchmark for policy-adaptive image guardrails, and PolicyShiftGuard, a compact model trained with a two-stage method that improves performance under shifting safety policies.

0 favorites 0 likes
#policy-adaptation

OmniTacTune: Policy-Agnostic Real-World RL for Tactile Residual Adaptation of Visual Policies

Hugging Face Daily Papers · 2026-07-04 Cached

OmniTacTune introduces a two-stage reinforcement learning pipeline for adapting tactile feedback to pretrained visual robot policies, achieving 85-100% success on contact-rich manipulation tasks within 40-80 minutes.

0 favorites 0 likes
#policy-adaptation

Warp RL: Reshaping Base Policy Distributions for Dynamics Adaptation

arXiv cs.LG · 2026-07-01 Cached

Warp RL replaces additive residual corrections in reinforcement learning with an invertible, state-conditioned transformation of the base policy's action distribution using monotonic rational-quadratic spline flows, enabling adaptation of distribution shape, scale, and geometry under dynamics shifts. It matches or outperforms residual correction in ManiSkill3 manipulation tasks and achieves 30% faster task completion in a real robot peg-insertion task.

0 favorites 0 likes
#policy-adaptation

@AdinaYakup: SingGuard from Ant Group @AntLingAGI A multimodal guardrail where the safety policy is an input, not a fixed weight. - …

X AI KOLs Timeline · 2026-06-22 Cached

SingGuard is a multimodal guardrail system from Ant Group that treats safety policy as an input, allowing dynamic adaptation via natural language. It is released under Apache 2.0 and covers text and image modalities.

0 favorites 0 likes
#policy-adaptation

Robotic Policy Adaptation via Weight-Space Meta-Learning

Hugging Face Daily Papers · 2026-06-05 Cached

Introduces WIZARD, a weight-space meta-learning framework that generates task-specific LoRA parameters for frozen VLA policies from language instructions and demonstration videos, enabling efficient task adaptation without fine-tuning.

0 favorites 0 likes
← Back to home

Submit Feedback