Skip to main content
Back to News Hub
🤖OpenAI
August 3, 2017
General AI

Gathering human feedback

Overview

OpenAI has released RL-Teacher, an open-source implementation of an interface designed to train artificial intelligence models through periodic human input. This approach replaces traditional hand-crafted reward functions with occasional feedback provided directly by human evaluators. The technique serves as a step toward safer AI systems while aiding reinforcement learning tasks where specifying rewards is difficult.

Key Takeaways

  • OpenAI released RL-Teacher, an open-source implementation of its interface aimed at improving how artificial intelligence models learn.

    Rather than depending on hand-crafted reward functions that developers must manually design, this system allows models to be trained using occasional feedback from human observers.

  • In addition to safety research, this method addresses reinforcement learning problems where defining clear reward functions is inherently difficult or complex.

    RL-Teacher provides an open-source interface for training artificial intelligence through occasional human feedback.

  • This approach was created to advance safe AI systems and solve tasks where rewards are hard to specify.
  • The underlying technique was developed as part of an effort to build safer artificial intelligence systems.
  • The system serves as an alternative to using hand-crafted reward functions in reinforcement learning.

OpenAI released RL-Teacher, an open-source implementation of its interface aimed at improving how artificial intelligence models learn. Rather than depending on hand-crafted reward functions that developers must manually design, this system allows models to be trained using occasional feedback from human observers. The underlying technique was developed as part of an effort to build safer artificial intelligence systems.

In addition to safety research, this method addresses reinforcement learning problems where defining clear reward functions is inherently difficult or complex. RL-Teacher provides an open-source interface for training artificial intelligence through occasional human feedback. The system serves as an alternative to using hand-crafted reward functions in reinforcement learning.

For more details please read the original article at OpenAI.

Continue Learning

Comments

Comments appear only after moderation. Your email identifies your submission to the moderator and is never displayed here.

No approved comments yet.

Originally published by OpenAI
Read the original