OpenAI BlogMay 3, 2018
AI safety via debate
Excerpt
We’re proposing an AI safety technique which trains agents to debate topics with one another, using a human to judge who wins.
We’re proposing an AI safety technique which trains agents to debate topics with one another, using a human to judge who wins.