@wintonARK: i'm not saying it's going to happen or work but somewhere, in some high school, an ambitious and quirky track coach is …

X AI KOLs Following News

Summary

A discussion on how reinforcement learning might preserve running posture in robots, comparing it to a track coach's innovative strategy in biomechanics.

i'm not saying it's going to happen or work but somewhere, in some high school, an ambitious and quirky track coach is going to try and if it does work we will look back on these videos as the move 37 of biomechanics
Original Article
View Cached Full Text

Cached at: 08/30/26, 10:12 PM

i’m not saying it’s going to happen or work but somewhere, in some high school, an ambitious and quirky track coach is going to try and if it does work we will look back on these videos as the move 37 of biomechanics

Rohan Paul (@rohanpaul_ai): Reinforcement learning apparently kept this running posture over human form

A person swings the arms mainly to stop the upper body from twisting with every stride, the twist that each leg puts into the torso. A robot arm is motors, gears, and wires. fast swinging will generate

i’m not saying it’s going to happen or work but somewhere, in some high school, an ambitious and quirky track coach is going to try and if it does work we will look back on these videos as the move 37 of biomechanics

Probably. On the other hand we learned to run assuming that we had to watch we were going, not just maximize straight line flat surface speed. (See also Fosbury flop.) (But yes my money would be against it working pretty heavily)

Similar Articles

Coachable agents for interactive gameplay

arXiv cs.AI

This paper presents a framework for training reinforcement learning agents that can be coached in real-time to adopt different styles while performing core tasks, demonstrated in Horizon Forbidden West, Gran Turismo, and a humanoid walking domain.

@mervenoyann: interesting talk by @willcb

X AI KOLs Timeline

This talk by Will Brown of Primordial AI discusses techniques for scaling Reinforcement Learning to complex, real-world tasks where rewards are not verifiable, using methods like anchoring, LLM judges, and simulation.

@tanayj: https://x.com/tanayj/status/2072766211256119475

X AI KOLs Timeline

This article explores the challenge of applying reinforcement learning to tasks that lack clear verifiability, citing Dario Amodei's prediction about achieving a 'country of geniuses in a data center' and discussing techniques such as RLVR, RLHF, Constitutional AI, and rubric-based rewards from Scale AI.

Beware the next giant pre-training run

Reddit r/singularity

Speculative post about next giant pre-training run potentially yielding superhuman AI within 12-14 months, leading to recursive self-improvement.