All categories
Research
Research, evidence gathering, literature, reports, investigation, and synthesis
- 21,409
- 893
Security grades appear on each card once the skill has been scanned. Newly imported skills may briefly show without a grade until the backfill job runs.
Open in full browserBrowse research skills
Showing 11,329–11,352 of 21,409 skills
- Third Person Imitation LearningSkill for AI agent capabilitiesVotes: 0GitHub stars: 3
- Stp Stabilizes Goal Conditioned DynamicsShort-Term Synaptic Plasticity (STP) stabilizes goal-conditioned dynamics in PFC-inspired reservoir computing models for multistep goal-directed action planning. Combines STP with basal-ganglia-inspired temporal-difference learning. Achieves 89.2% success under noise (vs 49.5% without STP). Activation: synaptic plasticity, reservoir computing, goal-conditioned dynamics, PFC model, action planning, goal-directed behavior, temporal-difference learning, effective connectivity, short-term plastic...Votes: 0GitHub stars: 3
- Steppo Agentic RlStepPO: Step-Aligned Policy Optimization for Agentic Reinforcement Learning - A novel RL framework for training LLM agents with step-level credit assignmentVotes: 0GitHub stars: 3
- Statistical Efficiency Quantile Distributional Reinforcement LearningStudies quantile-based distributional RL from statistical efficiency perspective. Non-asymptotic error bound O(√(m/n)) under W∞ metric. Achieves optimal √n convergence rate. Asymptotic distribution and semiparametric efficiency bound. Berry-Esseen theorem. Activation: distributional RL, quantile regression, statistical efficiency, policy evaluation, return distribution.Votes: 0GitHub stars: 3
- Sr Agent Experience Driven Agentic Framework Post Ranking Strategies Refinement E CommercSkill derived from arXiv:2607.17719 - SR-Agent: An Experience-Driven Agentic Framework for Post-Ranking Strategies Refinement in E-CommercVotes: 0GitHub stars: 3
- Specifying Delegated Autonomy Boundary Requirements Engineering Agentic AiSkill derived from arXiv:2607.17225 - Specifying the Delegated-Autonomy Boundary: Requirements Engineering for Agentic AIVotes: 0GitHub stars: 3
- Sparse Evidence Suffice Agentic Evidence Seeking Multimodal Video Misinformation DetectionSkill derived from arXiv:2607.18080 - Sparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation DetectionVotes: 0GitHub stars: 3
- Sparse Evidence Can Suffice Agentic Evidence SeekiDerived from arXiv:2607.18080 - Sparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation DetectionVotes: 0GitHub stars: 3
- Solving Rubiks Cube With A Robot HandSkill for AI agent capabilitiesVotes: 0GitHub stars: 3
- Soft Control Multi AgentSoft Control methodology for guiding collective behavior in multi-agent systems using shill agents. Based on Han et al. (2010) research on controlling self-organized systems without modifying local agent rules.Votes: 0GitHub stars: 3
- Simultaneously Evolving Deep Reinforcement Learning Models Using Multifactorial Optimization**arXiv ID:** 2002.12133 **Authors:** Aritz D. Martinez, Eneko Osaba, Javier Del Ser, Francisco Herrera **Published:** 2020-02-25T10:36:57Z **Abstract:** In recent years, Multifactorial Optimization (MFO) has gained a notable momentum in the research community. MFO is known for its inherent capability to efficiently address multiple optimization tasks at the same time, while transferring information among such tasks to improve their convergence speed. On the other hand, the quantum leap made ...Votes: 0GitHub stars: 3
- Sim To Real Transfer Of Robotic Control With DynamSkill for AI agent capabilitiesVotes: 0GitHub stars: 3
- Self Evolving Agent ExperienceFramework for LLM agents that accumulates and reuses cross-task experience including verified skills, statistical evidence of effective strategies, and recurring error-fix patterns. Enables zero-test-time search on new tasks.Votes: 0GitHub stars: 3
- Scaling Laws For Reward Model OveroptimizationSkill for AI agent capabilitiesVotes: 0GitHub stars: 3
- Scalelogic Rl ReasoningMethodology for studying RL scaling laws in LLM reasoning using a synthetic logical reasoning framework (ScaleLogic) with independent control over proof depth and logical expressiveness.Votes: 0GitHub stars: 3
- Saga Synthetic Agentic Graph Architecture Temporal Benchmark GenerationSkill derived from arXiv:2607.17288 - SAGA: Synthetic Agentic Graph Architecture for Temporal Benchmark GenerationVotes: 0GitHub stars: 3
- Rt Shcua Real Time Self Hosted Computer Use Agent Uav ControlSkill derived from arXiv:2607.17951 - RT-SHCUA: Real-Time Self-Hosted Computer-Use Agent for UAV ControlVotes: 0GitHub stars: 3
- Robot Co Design Inductive BiasesInductive biases identification for morphology-control co-design in robotics. Analyzes co-design landscapes to discover low-dimensional manifolds and patterns for sample-efficient search. Activation: robot co-design, morphology optimization, control co-design, inductive biases, high-dimensional search, soft robotics, embodied AI.Votes: 0GitHub stars: 3
- Rl² Fast Reinforcement Learning Via Slow ReinforceSkill for AI agent capabilitiesVotes: 0GitHub stars: 3
- Rl Tsch Dynamic ListeningReinforcement Learning-driven Adaptive Listening for TSCH NetworksVotes: 0GitHub stars: 3
- Rl Nqs OptimizationFrame neural quantum state optimization as reinforcement learning for scalable wavefunction approximation.Votes: 0GitHub stars: 3
- Rl Neural Model EditingReinforcement learning framework for neural model editing where agents learn to modify models via reward feedback instead of manually engineered algorithmsVotes: 0GitHub stars: 3
- Rl Compositional Reasoning StrategiesUnderstanding and leveraging how RL composes primitive skills into higher-level reasoning strategies.Votes: 0GitHub stars: 3
- Reusability And Transferability Of Macro Actions For Reinforcement Learning**arXiv ID:** 1908.01478 **Authors:** Yi-Hsiang Chang, Kuan-Yu Chang, Henry Kuo, Chun-Yi Lee **Published:** 2019-08-05T05:59:40Z **Abstract:** Conventional reinforcement learning (RL) typically determines an appropriate primitive action at each timestep. However, by using a proper macro action, defined as a sequence of primitive actions, an agent is able to bypass intermediate states to a farther state and facilitate its learning procedure. The problem we would like to investigate is what ass...Votes: 0GitHub stars: 3