Skip to content
AI Lehel Briefing
← Back to latest
Models & research OpenAI

Faulty reward functions in the wild

Reinforcement learning algorithms can break in surprising, counterintuitive ways. In this post we’ll explore one failure mode, which is where you misspecify your reward function.

Excerpt supplied by the publisher’s feed

Read the original article at OpenAI Opens in a new tab