**arXiv ID:** 2304.12180 **Authors:** Oscar Li, James Harrison, Jascha Sohl-Dickstein, Virginia Smith, Luke Metz **Published:** 2023-04-21T17:53:05Z **Abstract:** Unrolled computation graphs are prevalent throughout machine learning but present challenges to automatic differentiation (AD) gradient estimation methods when their loss functions exhibit extreme local sensitivtiy, discontinuity, or blackbox characteristics. In such scenarios, online evolution strategies methods are a more capable ...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill variancereduced-gradient-estimation-via-noisereuse-in-online-evolution-strategies --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Variancereduced Gradient Estimation Via Noisereuse In Online Evolution Strategies?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-variancereduced-gradient-estimation-via-noisereuse)More formats (shields.io, HTML) on the badges page.
# Variance-Reduced Gradient Estimation via Noise-Reuse in Online Evolution Strategies
**arXiv ID:** 2304.12180
**Authors:** Oscar Li, James Harrison, Jascha Sohl-Dickstein, Virginia Smith, Luke Metz
**Published:** 2023-04-21T17:53:05Z
**Abstract:**
Unrolled computation graphs are prevalent throughout machine learning but present challenges to automatic differentiation (AD) gradient estimation methods when their loss functions exhibit extreme local sensitivtiy, discontinuity, or blackbox characteristics. In such scenarios, online evolution strategies methods are a more capable alternative, while being more parallelizable than vanilla evolution strategies (ES) by interleaving partial unrolls and gradient updates. In this work, we propose a general class of unbiased online evolution strategies methods. We analytically and empirically characterize the variance of this class of gradient estimators and identify the one with the least variance, which we term Noise-Reuse Evolution Strategies (NRES). Experimentally, we show NRES results in faster convergence than existing AD and ES methods in terms of wall-clock time and number of unroll steps across a variety of applications, including learning dynamical systems, meta-training learned optimizers, and reinforcement learning.
## Skill Description
This skill is generated from the arXiv paper: Variance-Reduced Gradient Estimation via Noise-Reuse in Online Evolution Strategies (2304.12180).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:2304.12180](http://arxiv.org/abs/2304.12180v2)
Is this your skill, or is something wrong with this listing? . Author removals are honored within 72 hours.
No comments yet. Be the first to comment!