Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Meta-trained agents implement Bayes-optimal agents

Oct 21, 2020

Vladimir Mikulik, Grégoire Delétang, Tom McGrath, Tim Genewein, Miljan Martic, Shane Legg, Pedro A. Ortega

Figure 1 for Meta-trained agents implement Bayes-optimal agents

Figure 2 for Meta-trained agents implement Bayes-optimal agents

Figure 3 for Meta-trained agents implement Bayes-optimal agents

Figure 4 for Meta-trained agents implement Bayes-optimal agents

Share this with someone who'll enjoy it:

Abstract:Memory-based meta-learning is a powerful technique to build agents that adapt fast to any task within a target distribution. A previous theoretical study has argued that this remarkable performance is because the meta-training protocol incentivises agents to behave Bayes-optimally. We empirically investigate this claim on a number of prediction and bandit tasks. Inspired by ideas from theoretical computer science, we show that meta-learned and Bayes-optimal agents not only behave alike, but they even share a similar computational structure, in the sense that one agent system can approximately simulate the other. Furthermore, we show that Bayes-optimal agents are fixed points of the meta-learning dynamics. Our results suggest that memory-based meta-learning might serve as a general technique for numerically approximating Bayes-optimal agents - that is, even for task distributions for which we currently don't possess tractable models.

* Published at 34th Conference on Neural Information Processing Systems (NeurIPS 2020), Vancouver, Canada

View paper on

Share this with someone who'll enjoy it:

Title:Meta-trained agents implement Bayes-optimal agents

Paper and Code