**arXiv ID:** 2010.11223 **Authors:** Vladimir Mikulik, Grégoire Delétang, Tom McGrath, Tim Genewein, Miljan Martic, Shane Legg, Pedro A. Ortega **Published:** 2020-10-21T18:05:21Z **Abstract:** Memory-based meta-learning is a powerful technique to build agents that adapt fast to any task within a target distribution. A previous theoretical study has argued that this remarkable performance is because the meta-training protocol incentivises agents to behave Bayes-optimally. We empirically inve...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill metatrained-agents-implement-bayesoptimal-agents --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Metatrained Agents Implement Bayesoptimal Agents?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-metatrained-agents-implement-bayesoptimal-agents)More formats (shields.io, HTML) on the badges page.
# Meta-trained agents implement Bayes-optimal agents
**arXiv ID:** 2010.11223
**Authors:** Vladimir Mikulik, Grégoire Delétang, Tom McGrath, Tim Genewein, Miljan Martic, Shane Legg, Pedro A. Ortega
**Published:** 2020-10-21T18:05:21Z
**Abstract:**
Memory-based meta-learning is a powerful technique to build agents that adapt fast to any task within a target distribution. A previous theoretical study has argued that this remarkable performance is because the meta-training protocol incentivises agents to behave Bayes-optimally. We empirically investigate this claim on a number of prediction and bandit tasks. Inspired by ideas from theoretical computer science, we show that meta-learned and Bayes-optimal agents not only behave alike, but they even share a similar computational structure, in the sense that one agent system can approximately simulate the other. Furthermore, we show that Bayes-optimal agents are fixed points of the meta-learning dynamics. Our results suggest that memory-based meta-learning might serve as a general technique for numerically approximating Bayes-optimal agents - that is, even for task distributions for which we currently don't possess tractable models.
## Skill Description
This skill is generated from the arXiv paper: Meta-trained agents implement Bayes-optimal agents (2010.11223).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:2010.11223](http://arxiv.org/abs/2010.11223v1)
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!