Все новости
Нейросети

Gathering human feedback

3 августа 2017 г.0 просмотров1 мин чтения

RL-Teacher is an open-source implementation of our interface to train AIs via occasional human feedback rather than hand-crafted reward functions. The underlying technique was developed as a step towards safe AI systems, but also applies to reinforcement learning problems with rewards that are hard to specify.

Хотите попробовать AI?

Сравните лучшие нейросети в одном месте — бесплатно

Перейти к нейросетям
Поделиться:
Источник: OpenAI Blog

Похожие новости