Installs into .claude/skills of the current project.
Are you the author of Arxiv 2609 17193v1 End To End Latency Minimizing And Load Balanced Re?
Add the live security badge to your README. It updates with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-arxiv-2609-17193v1-end-to-end-latency-minimizing-a)
--
name: arxiv-2609-17193v1-end-to-end-latency-minimizing-and-load-balanced-re
description: 'End-to-End Latency-Minimizing and Load-Balanced Request Scheduling for Edge LLM Inference in Agentic AI Services (arXiv: 2609.17193v1)'
metadata:
{
"arxiv_id": "2609.17193v1",
"utility": 1.0,
"title": "End-to-End Latency-Minimizing and Load-Balanced Request Scheduling for Edge LLM Inference in Agentic AI Services",
"authors": "Zhen Li, Jun Cai, Haoran Gao, An Li, Tan Li",
"url": "http://arxiv.org/abs/2609.17193v1"
}
--
# End-to-End Latency-Minimizing and Load-Balanced Request Scheduling for Edge LLM Inference in Agentic AI Services
**arXiv ID:** 2609.17193v1
**Authors:** Zhen Li, Jun Cai, Haoran Gao, An Li, Tan Li
**URL:** http://arxiv.org/abs/2609.17193v1
**Utility Score:** 1.00
## Summary
This skill was automatically generated from the arXiv paper titled "End-to-End Latency-Minimizing and Load-Balanced Request Scheduling for Edge LLM Inference in Agentic AI Services" (ID: 2609.17193v1).
## Usage
This skill can be used to reference the paper's concepts, methodologies, or findings in agent workflows.
## References
- arXiv: http://arxiv.org/abs/2609.17193v1