**arXiv ID:** 1709.06919 **Authors:** Rémi Pautrat, Konstantinos Chatzilygeroudis, Jean-Baptiste Mouret **Published:** 2017-09-20T15:04:50Z **Abstract:** One of the most interesting features of Bayesian optimization for direct policy search is that it can leverage priors (e.g., from simulation or from previous tasks) to accelerate learning on a robot. In this paper, we are interested in situations for which several priors exist but we do not know in advance which one fits best the current sit...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill bayesian-optimization-with-automatic-prior-selection-for-dataefficient-direct-policy-search --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Bayesian Optimization With Automatic Prior Selection For Dataefficient Direct Policy Search?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-bayesian-optimization-with-automatic-prior-selecti)More formats (shields.io, HTML) on the badges page.
# Bayesian Optimization with Automatic Prior Selection for Data-Efficient Direct Policy Search
**arXiv ID:** 1709.06919
**Authors:** Rémi Pautrat, Konstantinos Chatzilygeroudis, Jean-Baptiste Mouret
**Published:** 2017-09-20T15:04:50Z
**Abstract:**
One of the most interesting features of Bayesian optimization for direct policy search is that it can leverage priors (e.g., from simulation or from previous tasks) to accelerate learning on a robot. In this paper, we are interested in situations for which several priors exist but we do not know in advance which one fits best the current situation. We tackle this problem by introducing a novel acquisition function, called Most Likely Expected Improvement (MLEI), that combines the likelihood of the priors and the expected improvement. We evaluate this new acquisition function on a transfer learning task for a 5-DOF planar arm and on a possibly damaged, 6-legged robot that has to learn to walk on flat ground and on stairs, with priors corresponding to different stairs and different kinds of damages. Our results show that MLEI effectively identifies and exploits the priors, even when there is no obvious match between the current situations and the priors.
## Skill Description
This skill is generated from the arXiv paper: Bayesian Optimization with Automatic Prior Selection for Data-Efficient Direct Policy Search (1709.06919).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:1709.06919](http://arxiv.org/abs/1709.06919v2)
Is this your skill, or is something wrong with this listing? . Author removals are honored within 72 hours.
No comments yet. Be the first to comment!