**arXiv ID:** 2209.10055 **Authors:** Hui Bai, Ruimin Shen, Yue Lin, Botian Xu, Ran Cheng **Published:** 2022-09-21T00:55:55Z **Abstract:** Despite the emerging progress of integrating evolutionary computation into reinforcement learning, the absence of a high-performance platform endowing composability and massive parallelism causes non-trivial difficulties for research and applications related to asynchronous commercial games. Here we introduce Lamarckian - an open-source platform featuring...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill lamarckian-platform-pushing-the-boundaries-of-evolutionary-reinforcement-learning-towards-asynchronous-commercial-games --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Lamarckian Platform Pushing The Boundaries Of Evolutionary Reinforcement Learning Towards Asynchronous Commercial Games?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-lamarckian-platform-pushing-the-boundaries-of-evol)More formats (shields.io, HTML) on the badges page.
# Lamarckian Platform: Pushing the Boundaries of Evolutionary Reinforcement Learning towards Asynchronous Commercial Games
**arXiv ID:** 2209.10055
**Authors:** Hui Bai, Ruimin Shen, Yue Lin, Botian Xu, Ran Cheng
**Published:** 2022-09-21T00:55:55Z
**Abstract:**
Despite the emerging progress of integrating evolutionary computation into reinforcement learning, the absence of a high-performance platform endowing composability and massive parallelism causes non-trivial difficulties for research and applications related to asynchronous commercial games. Here we introduce Lamarckian - an open-source platform featuring support for evolutionary reinforcement learning scalable to distributed computing resources. To improve the training speed and data efficiency, Lamarckian adopts optimized communication methods and an asynchronous evolutionary reinforcement learning workflow. To meet the demand for an asynchronous interface by commercial games and various methods, Lamarckian tailors an asynchronous Markov Decision Process interface and designs an object-oriented software architecture with decoupled modules. In comparison with the state-of-the-art RLlib, we empirically demonstrate the unique advantages of Lamarckian on benchmark tests with up to 6000 CPU cores: i) both the sampling efficiency and training speed are doubled when running PPO on Google football game; ii) the training speed is 13 times faster when running PBT+PPO on Pong game. Moreover, we also present two use cases: i) how Lamarckian is applied to generating behavior-diverse game AI; ii) how Lamarckian is applied to game balancing tests for an asynchronous commercial game.
## Skill Description
This skill is generated from the arXiv paper: Lamarckian Platform: Pushing the Boundaries of Evolutionary Reinforcement Learning towards Asynchronous Commercial Games (2209.10055).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:2209.10055](http://arxiv.org/abs/2209.10055v1)
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!