**arXiv ID:** 2609.09135v1 **Authors:** Jiacheng Xu, Feng Chen, Xiuneng Xu, Bo An **URL:** http://arxiv.org/abs/2609.09135v1 **Utility Score:** 1.00
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill arxiv-2609-09135v1-entropy-regularized-rank-masked-policy-optimizatio --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Arxiv 2609 09135v1 Entropy Regularized Rank Masked Policy Optimizatio?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-arxiv-2609-09135v1-entropy-regularized-rank-masked)More formats (shields.io, HTML) on the badges page.
--
name: arxiv-2609-09135v1-entropy-regularized-rank-masked-policy-optimizatio
description: 'Entropy-Regularized Rank-Masked Policy Optimization for Test-Time Reinforcement Learning in Code Generation (arXiv: 2609.09135v1)'
metadata:
{
"arxiv_id": "2609.09135v1",
"utility": 1.0,
"title": "Entropy-Regularized Rank-Masked Policy Optimization for Test-Time Reinforcement Learning in Code Generation",
"authors": "Jiacheng Xu, Feng Chen, Xiuneng Xu, Bo An",
"url": "http://arxiv.org/abs/2609.09135v1"
}
--
# Entropy-Regularized Rank-Masked Policy Optimization for Test-Time Reinforcement Learning in Code Generation
**arXiv ID:** 2609.09135v1
**Authors:** Jiacheng Xu, Feng Chen, Xiuneng Xu, Bo An
**URL:** http://arxiv.org/abs/2609.09135v1
**Utility Score:** 1.00
## Summary
This skill was automatically generated from the arXiv paper titled "Entropy-Regularized Rank-Masked Policy Optimization for Test-Time Reinforcement Learning in Code Generation" (ID: 2609.09135v1).
## Usage
This skill can be used to reference the paper's concepts, methodologies, or findings in agent workflows.
## References
- arXiv: http://arxiv.org/abs/2609.09135v1
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!