All categories
Research
Research, evidence gathering, literature, reports, investigation, and synthesis
- 21,409
- 893
Security grades appear on each card once the skill has been scanned. Newly imported skills may briefly show without a grade until the backfill job runs.
Open in full browserBrowse research skills
Showing 11,401–11,424 of 21,409 skills
- Decentralized Motion Planning For Multirobot Navigation Using Deep Reinforcement Learning**arXiv ID:** 2011.05605 **Authors:** Sivanathan Kandhasamy, Vinayagam Babu Kuppusamy, Tanmay Vilas Samak, Chinmay Vilas Samak **Published:** 2020-11-11T07:35:21Z **Abstract:** This work presents a decentralized motion planning framework for addressing the task of multi-robot navigation using deep reinforcement learning. A custom simulator was developed in order to experimentally investigate the navigation problem of 4 cooperative non-holonomic robots sharing limited state information with ea...Votes: 0GitHub stars: 3
- Competitive Self PlaySkill for AI agent capabilitiesVotes: 0GitHub stars: 3
- Bridging Evolutionary Algorithms And Reinforcement Learning A Comprehensive Survey On Hybrid Algorithms**arXiv ID:** 2401.11963 **Authors:** Pengyi Li, Jianye Hao, Hongyao Tang, Xian Fu, Yan Zheng, Ke Tang **Published:** 2024-01-22T14:06:37Z **Abstract:** Evolutionary Reinforcement Learning (ERL), which integrates Evolutionary Algorithms (EAs) and Reinforcement Learning (RL) for optimization, has demonstrated remarkable performance advancements. By fusing both approaches, ERL has emerged as a promising research direction. This survey offers a comprehensive overview of the diverse research bran...Votes: 0GitHub stars: 3
- Autonomy Self Evolving Testing LoopSelf-evolving simulation-based testing loop for autonomous cyber-physical systems. Continuous scenario generation, execution, and telemetry analysis.Votes: 0GitHub stars: 3
- Asymmetric Actor Critic For Image Based Robot LearSkill for AI agent capabilitiesVotes: 0GitHub stars: 3
- Arxiv 2609 09647v1 Black Box Red Teaming Of Agentic Ai A Taxonomy Dri**arXiv ID:** 2609.09647v1 **Authors:** Divyanshu Kumar, Nitin Aravind Birur, Tanay Baswa, Sahil Agarwal, Prashanth Harshangi **URL:** http://arxiv.org/abs/2609.09647v1 **Utility Score:** 1.00Votes: 0GitHub stars: 3
- Arxiv 2609 08719v1 Goant Quality Diversity Multi Agent Search For Alp**arXiv ID:** 2609.08719v1 **Authors:** Stella Zhao, Tommy Sha **URL:** http://arxiv.org/abs/2609.08719v1 **Utility Score:** 1.00Votes: 0GitHub stars: 3
- Arxiv 2608 30702v1 An Agentic Retrobiosynthesis Framework With Learne**arXiv ID:** 2608.30702v1 **Authors:** Philippe Meyer, Guillaume Gricourt, Thomas Duigou, Joan Hérisson, Jean-Loup Faulon **URL:** http://arxiv.org/abs/2608.30702v1 **Utility Score:** 1.00Votes: 0GitHub stars: 3
- Arxiv 2608 20331 G Carl Grounded Checklist Aligned Reward LearningG-CARL: Grounded Checklist-Aligned Reward Learning for Patient-Oriented Medical Report Interpretation (arXiv: 2608.20331)Votes: 0GitHub stars: 3
- Arxiv 2608 20319 Inducing Task Models From Computer Use TracesInducing Task Models from Computer-Use Traces (arXiv: 2608.20319)Votes: 0GitHub stars: 3
- Arxiv 2608 20314 Midtool Mid Training Data Synthesis For Agentic ToMidTool: Mid-training Data Synthesis for Agentic Tool Use (arXiv: 2608.20314)Votes: 0GitHub stars: 3
- Arxiv 2608 20295 Physical Support Confidence Sets For Highly CoherePhysical-Support Confidence Sets for Highly Coherent Dictionaries (arXiv: 2608.20295)Votes: 0GitHub stars: 3
- Arxiv 2608 20256 Learning When To Think Adaptive Reasoning For TestLearning When to Think: Adaptive Reasoning for Test-Time Compute Allocation (arXiv: 2608.20256)Votes: 0GitHub stars: 3
- Arxiv 2608 20201 The Third Restructuring Of Software Form From TheThe Third Restructuring of Software Form: From the Three-Tier Architecture to Storage, Models, and Agents (arXiv: 2608.20201)Votes: 0GitHub stars: 3
- Arxiv 2608 20187 Multi Method Causal Evidence Synthesis Ranking CanMulti-Method Causal Evidence Synthesis: Ranking Candidate Drivers by Convergent Cross-Method Evidence from Observational Data (arXiv: 2608.20187)Votes: 0GitHub stars: 3
- Arxiv 2608 20172 Ask Self Ask Others Relation Is All You NeedAsk Self, Ask Others: Relation Is All You Need (arXiv: 2608.20172)Votes: 0GitHub stars: 3
- Arxiv 2608 20169 Task Coevolve Efficient Harness Optimization Via ATask-CoEvolve: Efficient Harness Optimization via Adaptive Validation Task Selection (arXiv: 2608.20169)Votes: 0GitHub stars: 3
- Arxiv 2608 20161 Dars Dual Level Credit Assignment Rl With StructurDARS: Dual-Level Credit Assignment RL with Structured Reasoning for Instruction-Based Image Editing (arXiv: 2608.20161)Votes: 0GitHub stars: 3
- Arxiv 2608 20084 Evidence Gated Task And Motion Planning With VisioEvidence-Gated Task and Motion Planning with Vision-Language Models (arXiv: 2608.20084)Votes: 0GitHub stars: 3
- Arxiv 2608 20065 Orthogonal Jepa Factorized Predictive States For LOrthogonal JEPA: Factorized Predictive States for Latent World Models (arXiv: 2608.20065)Votes: 0GitHub stars: 3
- Arxiv 2608 20044 End To End Early Classification Of Time Series InEnd-to-end Early Classification of Time Series in Non-Stationary Environments (arXiv: 2608.20044)Votes: 0GitHub stars: 3
- Arxiv 2608 20041 A Three Dimensional Typology Of Agency For AdvanceA three-dimensional typology of agency for advanced AI systems (arXiv: 2608.20041)Votes: 0GitHub stars: 3
- Arxiv 2608 20026 From Street View Imagery To Street Quality IndicatFrom Street View Imagery to Street Quality Indicators: Vision Language Inference for the Suburban 15-minute City (arXiv: 2608.20026)Votes: 0GitHub stars: 3
- Arxiv 2608 19974 Regusim Evaluating Llm Agent Rule Grounding In FinReguSim: Evaluating LLM Agent Rule Grounding in Financial Compliance (arXiv: 2608.19974)Votes: 0GitHub stars: 3