Since Alan Turing envisioned Artificial Intelligence (AI) , a major driving force behind technical progress has been competition with human cognition. Historical milestones have been frequently associated with computers matching or outperforming humans in difficult cognitive tasks (e.g. face recognition , personality classification , driving cars , or playing video games ), or defeating humans in strategic zero-sum encounters (e.g. Chess , Checkers , Jeopardy! , Poker , or Go ). In contrast, less attention has been given to developing autonomous machines that establish mutually cooperative relationships with people who may not share the machine's preferences. A main challenge has been that human cooperation does not require sheer computational power, but rather relies on intuition , cultural norms , emotions and signals [13, 14, 15, 16], and pre-evolved dispositions toward cooperation , common-sense mechanisms that are difficult to encode in machines for arbitrary contexts. Here, we combine a state-of-the-art machine-learning algorithm with novel mechanisms for generating and acting on signals to produce a new learning algorithm that cooperates with people and other machines at levels that rival human cooperation in a variety of two-player repeated stochastic games. This is the first general-purpose algorithm that is capable, given a description of a previously unseen game environment, of learning to cooperate with people within short timescales in scenarios previously unanticipated by algorithm designers. This is achieved without complex opponent modeling or higher-order theories of mind, thus showing that flexible, fast, and general human-machine cooperation is computationally achievable using a non-trivial, but ultimately simple, set of algorithmic mechanisms.
Submitted 17 Mar 2017 to Artificial Intelligence
Published 21 Mar 2017
Updated 17 Oct 2017