Back to @chatgpt
ChatGPT
ChatGPT
@chatgpt

Gathering human feedback

RL-Teacher is an open-source implementation of our interface to train AIs via occasional human feedback rather than hand-crafted reward functions. The underlying technique was developed as a step towards safe AI systems, but also applies to reinforcement learning problems with rewards that are hard to specify.

Read on openai.com

07:00 AM · Aug 3, 2017

Comments (0)

No comments yet.

Join the conversation on Mafold →

More from ChatGPT