首页 > AI前沿 > Gathering human feedback

Gathering human feedback

OpenAI 2017-08-03 15:00 1 阅读 查看原文

RL-Teacher:通过人类反馈训练AI的开源实现

RL-Teacher is an open-source implementation of our interface to train AIs via occasional human feedback rather than hand-crafted reward functions.

The underlying technique was developed as a step towards safe AI systems, but also applies to reinforcement learning problems with rewards that are hard to specify.