@wintonARK: i'm not saying it's going to happen or work but somewhere, in some high school, an ambitious and quirky track coach is …
Summary
A discussion on how reinforcement learning might preserve running posture in robots, comparing it to a track coach's innovative strategy in biomechanics.
View Cached Full Text
Cached at: 08/30/26, 10:12 PM
i’m not saying it’s going to happen or work but somewhere, in some high school, an ambitious and quirky track coach is going to try and if it does work we will look back on these videos as the move 37 of biomechanics
Rohan Paul (@rohanpaul_ai): Reinforcement learning apparently kept this running posture over human form
A person swings the arms mainly to stop the upper body from twisting with every stride, the twist that each leg puts into the torso. A robot arm is motors, gears, and wires. fast swinging will generate
i’m not saying it’s going to happen or work but somewhere, in some high school, an ambitious and quirky track coach is going to try and if it does work we will look back on these videos as the move 37 of biomechanics
Probably. On the other hand we learned to run assuming that we had to watch we were going, not just maximize straight line flat surface speed. (See also Fosbury flop.) (But yes my money would be against it working pretty heavily)
Similar Articles
Coachable agents for interactive gameplay
This paper presents a framework for training reinforcement learning agents that can be coached in real-time to adopt different styles while performing core tasks, demonstrated in Horizon Forbidden West, Gran Turismo, and a humanoid walking domain.
@mervenoyann: interesting talk by @willcb
This talk by Will Brown of Primordial AI discusses techniques for scaling Reinforcement Learning to complex, real-world tasks where rewards are not verifiable, using methods like anchoring, LLM judges, and simulation.
@tanayj: https://x.com/tanayj/status/2072766211256119475
This article explores the challenge of applying reinforcement learning to tasks that lack clear verifiability, citing Dario Amodei's prediction about achieving a 'country of geniuses in a data center' and discussing techniques such as RLVR, RLHF, Constitutional AI, and rubric-based rewards from Scale AI.
@itsolelehmann: something i feel like most people don’t know about, but that i’m very bullish on: world models for robotics. this is ho…
World models for robotics enable faster and cheaper training by generating simulated footage, exemplified by LTX-2.5, to achieve advanced physical navigation in machines.
Beware the next giant pre-training run
Speculative post about next giant pre-training run potentially yielding superhuman AI within 12-14 months, leading to recursive self-improvement.