Skip to content
AI Lehel Briefing
← Back to latest
Models & research OpenAI

Gathering human feedback

RL-Teacher is an open-source implementation of our interface to train AIs via occasional human feedback rather than hand-crafted reward functions. The underlying technique was developed as a step towards safe AI systems, but also applies to reinforcement learning problems with rewards that are hard to specify.

Excerpt supplied by the publisher’s feed

Read the original article at OpenAI Opens in a new tab