Agendas Recursive reward modeling

Google DeepMind Reserach Engineer 2017-03-01 2020-10-31 AGI organization [1], [2], [3]
Google DeepMind Senior Reserach Engineer 2020-10-01 AGI organization [1], [2], [3]

Scalable agent alignment via reward modeling: a research direction 2018-11-19 Jan Leike, David Krueger, Tom Everitt, Miljan Martic, Vishal Maini, Shane Legg arXiv Google DeepMind Recursive reward modeling, Imitation learning, inverse reinforcement learning, Cooperative inverse reinforcement learning, myopic reinforcement learning, iterated amplification, debate This paper introduces the (recursive) reward modeling agenda, discussing its basic outline, challenges, and ways to overcome those challenges. The paper also discusses alternative agendas and their relation to reward modeling.

