async-rl

Tag

Cards List
#async-rl

@samsja19: https://x.com/samsja19/status/2076846033922035818

X AI KOLs Following · 2026-07-14 Cached

PRIME-RL is a framework for large-scale asynchronous reinforcement learning, designed to be hackable and scale to 1000+ GPUs with support for various models and environments.

0 favorites 0 likes
#async-rl

@ziv_ravid: Just read the new paper from Tsinghua/Z.AI on async RL for agents (arXiv:2607.07508). It comes several weeks after the …

X AI KOLs Timeline · 2026-07-10 Cached

Discusses a new paper from Tsinghua/Zhipu AI on asynchronous reinforcement learning for agents, and notes that their previous GLM-5.2 model uses a critic instead of GRPO.

0 favorites 0 likes
#async-rl

AsyncWebRL: Efficient Multi-Step RL for Visual Web Agents

arXiv cs.LG · 2026-06-05 Cached

AsyncWebRL introduces an asynchronous multi-step reinforcement learning system for vision-language web agents, achieving up to 2.9x training speedup and setting a new state-of-the-art on WebGym by replacing per-trajectory normalization with a constant to reduce trajectory length inefficiency.

0 favorites 0 likes
← Back to home

Submit Feedback