ppo

Tag

Cards List
#ppo

Show HN: Watch a neural net learn to play Snake

Hacker News Top · 2026-05-14 Cached

A web-based tool that visualizes a neural network (using PPO) learning to play Snake in real-time, with configurable parameters and 3D rendering.

0 favorites 0 likes
#ppo

Competitive self-play

OpenAI Blog · 2017-10-11 Cached

OpenAI demonstrates that competitive self-play in simulated 3D robot environments enables AI agents to discover complex physical behaviors like tackling, ducking, and faking without explicit instruction, suggesting self-play will be fundamental to future powerful AI systems.

0 favorites 0 likes
#ppo

Proximal Policy Optimization

OpenAI Blog · 2017-07-20 Cached

OpenAI introduces Proximal Policy Optimization (PPO), a reinforcement learning algorithm that matches or outperforms state-of-the-art methods while being simpler to implement and tune. PPO uses a novel clipped objective function to constrain policy updates and has since become OpenAI's default RL algorithm.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback