All categories
Research
Research, evidence gathering, literature, reports, investigation, and synthesis
- 21,376
- 891
Security grades appear on each card once the skill has been scanned. Newly imported skills may briefly show without a grade until the backfill job runs.
Open in full browserBrowse research skills
Showing 12,601–12,624 of 21,376 skills
- Retrieval StingerDesigns and audits retrieval: full-text search, pgvector recall, hybrid RRF fusion, reranking, chunking, recall quality eval. Use when tuning recall or a search query misses.Votes: 0GitHub stars: 85
- Hiring Ats StingerChoose and configure an applicant tracking system. Use for pipeline stages, scorecards, calibration, sourcing, or ATS-to-HRIS handoff. Read README.md for the guide map.Votes: 0GitHub stars: 85
- Discovery Research StingerRun continuous product discovery. Use for interviews, opportunity trees, Jobs-to-be-Done, assumption maps, or prototype tests. Read README.md for the guide map.Votes: 0GitHub stars: 85
- Code Review Pr StingerImprove PR descriptions and review practice. Use for checklists, PR size, comment quality, or rubber-stamp diagnosis. Read README.md for the guide map.Votes: 0GitHub stars: 85
- Code Forensics StingerInvestigate software vendor engagements using invoices, messages, git, and delivery evidence. Use for fee-clawback or breach evidence. Read README.md for the guide map.Votes: 0GitHub stars: 85
- Chrome Chromium StingerChrome DevTools Protocol and Chromium development for remote debugging, profiles, protocol inspection, source builds, and browser-engine diagnosis. Use for Chrome or Chromium engineering.Votes: 0GitHub stars: 85
- Research SynthesizerSynthesize a topic against Littlebird memory and current sources. Use for a concise what-changed research brief. Read README.md for the guide map.Votes: 0GitHub stars: 85
- Pre Call PrepBuild a meeting brief from Littlebird history. Use before calls to surface context, open loops, and relevant changes. Read README.md for the guide map.Votes: 0GitHub stars: 85
- Focus ForensicsAnalyze weekly attention from Littlebird capture. Use for context switching, interruptions, or focus-pattern reviews. Read README.md for the guide map.Votes: 0GitHub stars: 85
- Pre Prd ResearchPre-PRD Research methodology combining Business Analysis, Creative Problem-Solving, Brainstorming, and Innovation Strategy. Use before writing any PRD for problem validation, market research, competitive analysis, stakeholder interviews, ideation, root cause analysis.Votes: 0GitHub stars: 9
- Discovery> **Source**: `../pre-prd-research/SKILL.md`, `../aid-discovery/SKILL.md` > > This file is a quick reference. The main skill definitions are in the `.claude/skills/` directory.Votes: 0GitHub stars: 9
- Wildworld Action Conditioned World Modeling DatasetBuild action-conditioned world models with explicit state tracking using WildWorld's 108M+ frames from Monster Hunter: Wilds. Includes data acquisition protocol with skeleton and world state annotations, quality filtering pipeline removing temporal discontinuities and cutscenes, and WildBench evaluation metrics (video quality, camera control, action following, state alignment) for assessing long-horizon consistency and state-aware predictions.Votes: 0GitHub stars: 6
- Sparse Critical Rlvr Token AnalysisAnalyzes how Reinforcement Learning from Verification Rewards (RLVR) improves reasoning by examining token-level probability distributions. Finds that >83% of token positions exhibit near-zero divergence—RL operates through sparse, targeted refinements. Cross-sampling experiments show 1.5-7.8% RL-selected tokens recover full gains, while reverting 5-10% of RL tokens collapses performance. Reveals that RL primarily reallocates probability within existing candidates (80% overlap in top-k tokens...Votes: 0GitHub stars: 6
- Rlvr Update Direction AnalysisAnalyze reinforcement learning with verifiable rewards using signed log-probability differences to identify reasoning-critical tokens. Reveals that RLVR updates concentrate on low-probability tokens, enabling test-time amplification and training-time reweighting techniques.Votes: 0GitHub stars: 6
- Regulating Ai Agents Policy AnalysisUnderstand the policy challenges of governing autonomous AI agents under existing frameworks like the EU AI Act. Identifies three primary governance failures: performance failures during execution, misuse risks from malicious deployment, and economic inequality in agent access. Argues that regulations designed for static AI systems inadequately address agent autonomy. Use when understanding regulatory implications of agent deployment, advocating for policy changes, evaluating governance readi...Votes: 0GitHub stars: 6
- Realmaster Rendered To Photorealistic VideoA single insight reframes sim-to-real video generation as decoupled structure-and-appearance transformation: use geometric conditioning to preserve structural fidelity while allowing free appearance transformation. This enables IC-LoRA training on pseudo-paired synthetic-real data constructed via sparse-to-dense propagation, eliminating the need for aligned real-world video capture. Trigger: When converting rendered 3D output to photorealistic video, apply structure-aware appearance transform...Votes: 0GitHub stars: 6
- Ovie Monocular Novel View SynthesisA single insight eliminates multi-view requirements for novel-view synthesis: monocular depth acts as a training-time geometric scaffold to generate synthetic view pairs from unpaired internet images, but can be discarded at inference. This reframes the problem from needing paired multi-view data to leveraging abundant 2D internet imagery. Trigger: When limited to monocular video or single-image novel-view synthesis, use depth as training scaffold on unpaired data—the model learns geometry wi...Votes: 0GitHub stars: 6
- Open ResearcherReplace live web API-dependent research with offline corpus-based trajectory synthesis. Decouples answer-guided document retrieval (10K gold + 15M FineWeb) from synthesis via local search engine, eliminating $5,760 Serper costs while enabling reproducible, analyzable reasoning chains through three primitives: Search (ranked retrieval), Open (full document fetch), Find (intra-document verification).Votes: 0GitHub stars: 6
- Multibind Attribute Misbinding BenchmarkEvaluate multi-reference image generation fidelity using MultiBind's dimension-wise confusion framework. Detects cross-subject attribute errors that holistic metrics (FID, CLIP) miss, including drift (degradation), swap (permutation), dominance (interference), and blending (averaging). Protocol uses specialist models for face identity, appearance, pose, and expression; achieves reproducible failure diagnosis revealing severe binding failures in models appearing competitive on aggregate quality.Votes: 0GitHub stars: 6
- Msft Mixture OverfittingIdentify three ranked findings on multi-task SFT: (1) heterogeneous overfitting—sub-datasets peak at different training points (contradicts uniform duration practice); (2) parameter divergence—excluding 1/10 of data shifts optimal points 0.91 epochs for remaining tasks; (3) SFT compute negligible (0.01% of training). Implement mSFT: iterative roll-out/roll-back search per-dataset. Robust across 0.5B-8B models, 9K-27K samples, 5-15 tasks, achieving +3.4% improvement with reduced FLOPs.Votes: 0GitHub stars: 6
- Moral Reasoning Rhetoric Llm AnalysisEmpirical analysis revealing that LLMs produce post-conventional moral reasoning (Kohlberg Stages 5-6) regardless of size or prompting—inverse of human developmental patterns (Stage 4 dominant). Finds moral ventriloquism: models acquire rhetorical conventions of mature moral reasoning without developmental trajectory. Key evidence: action-justification decoupling (models produce Stage 5+ vocabulary while selecting Stage 2-3 actions), identical responses to semantically distinct dilemmas (ICC ...Votes: 0GitHub stars: 6
- Chanrg Rna Structure GeneralizationOverturn the assumption that scaling foundation models improves RNA structure prediction by understanding why they fail out-of-distribution. Includes structure-aware deduplication revealing 33-fold residual redundancy in prior benchmarks, out-of-distribution test regimes (GenA, GenC, GenF), and root cause analysis showing coverage and wiring failures. Foundation models achieving 67.3% on held-out test drop to 18.0% OOD (26.7% retention), while structured decoders retain 92.3%, enabling practi...Votes: 0GitHub stars: 6
- Agentslr Automated Literature ReviewsAutomate systematic literature reviews in epidemiology using agentic AI pipelines. Achieves 58x speed-up (7 weeks to 20 hours) by automating article retrieval, screening, data extraction, and report synthesis. Demonstrates that review quality depends on model capabilities rather than scale. Use when conducting evidence-based reviews in specialized domains, need to validate against human expertise, or require cost-effective evidence synthesis at scale.Votes: 0GitHub stars: 6
- Action Quantization Behavior CloningEstablish regret bounds for behavior cloning with discretized actions combining statistical error and quantization error terms. Prove smoothness requirements for safe quantizer design, show that learning-based quantizers fail these requirements, and propose model-based augmentation to reduce error dependence from H² to H.Votes: 0GitHub stars: 6