human-feedback

Tag

Cards List
#human-feedback

Fine-tuning GPT-2 from human preferences

OpenAI Blog ↗ · 2019-09-19 Cached

OpenAI demonstrates fine-tuning GPT-2 (774M parameters) using human preference feedback for text continuation and summarization tasks, requiring 5k labels for stylistic tasks and 60k for summarization, with models achieving 86-88% human preference rates though revealing labeler heuristic exploitation.

0 favorites 0 likes
#human-feedback

Learning complex goals with iterated amplification

OpenAI Blog ↗ · 2018-10-22 Cached

OpenAI presents iterated amplification, a method for training AI systems on complex tasks by recursively decomposing them into smaller subtasks that humans can judge and solve, building up training signals from scratch through iterative composition.

0 favorites 0 likes
#human-feedback

Gathering human feedback

OpenAI Blog ↗ · 2017-08-03 Cached

OpenAI releases RL-Teacher, an open-source tool for training AI systems through human feedback instead of hand-crafted reward functions, with applications to safe AI development and complex reinforcement learning problems.

0 favorites 0 likes
#human-feedback

Learning from human preferences

OpenAI Blog ↗ · 2017-06-13 Cached

OpenAI presents a method for training AI agents using human preference feedback, where an agent learns reward functions from human comparisons of behavior trajectories and uses reinforcement learning to optimize for the inferred goals. The approach demonstrates strong sample efficiency, requiring less than 1000 bits of human feedback to train an agent to perform a backflip.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback