arXiv · 1505.04497
A Definition of Happiness for Reinforcement Learning Agents
Abstract
What is happiness for reinforcement learning agents? We seek a formal definition satisfying a list of desiderata. Our proposed definition of happiness is the temporal difference error, i.e. the difference between the value of the obtained reward and observation and the agent's expectation of this value. This definition satisfies most of our desiderata and is compatible with empirical research on humans. We state several implications and discuss examples.
Explore related subjects
Keep this discovery
Mayank Daswani, Jan Leike. 2015-05-18. A Definition of Happiness for Reinforcement Learning Agents. https://arxiv.org/abs/1505.04497
Cite the original work for its findings. Save a collection to share your selection of sources.