Skills DirectorySkills Directory
SkillsLearnSecurityCategoriesDocsBlogPro
Sign InSubmit Skill
Skills Directory

Security-tested agent skills for Claude, coding agents, and AI workflows.

Directory

  • Browse Skills
  • All Skills A–Z
  • Claude Skills
  • Claude Code Skills
  • Agent Skills
  • Categories
  • Authors
  • Submit a Skill

Learn

  • Learn Hub
  • Install Claude Skills
  • Write SKILL.md
  • Skills vs MCP
  • Directories Compared

Security

  • Security
  • Methodology
  • Secure Claude Skills
  • Security Badges
  • Chrome Extension
  • Skill Manager

Company

  • About
  • Community
  • Blog
  • API Docs
  • Advertise

2026 Skills Directory. All rights reserved.

ProTermsPrivacyRefunds
Back to skills

Evaluate

ASecurity

This skill should be used for ANY business idea validation, evaluation, or viability check. Key trigger signals: 'should I build', 'is this worth building', 'validate this', 'good idea?', 'rate my startup', 'honest feedback on my startup', 'is there market demand', 'wondering if the market is there', 'is this worth starting', 'any thoughts on viability', 'evaluate', 'got this side project idea', 'thinking about building', 'is this viable'. Also triggers for the /idea-forge:evaluate prefix and...

2 stars
0 votes
0 copies
1 views
Added 9/19/2026
ai-agentsgo

Security Analysis

A100/100

Pro scans all 12 files and shows the line behind each finding

Scanned 9/19/2026

$npx -y skills add neotherapper/claude-plugins --skill evaluate --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Evaluate?

Add the live security badge to your README — it updates automatically with every re-scan.

Security grade badge for Evaluate
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/neotherapper-evaluate/badge)](https://www.skillsdirectory.com/skills/neotherapper-evaluate)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
Files
SKILL.md
---
name: evaluate
description: "This skill should be used for ANY business idea validation, evaluation, or viability check. Key trigger signals: 'should I build', 'is this worth building', 'validate this', 'good idea?', 'rate my startup', 'honest feedback on my startup', 'is there market demand', 'wondering if the market is there', 'is this worth starting', 'any thoughts on viability', 'evaluate', 'got this side project idea', 'thinking about building', 'is this viable'. Also triggers for the /idea-forge:evaluate prefix and any message where someone describes a SaaS, marketplace, directory, e-commerce, content site, or free tool idea and wants to know whether it's worth pursuing. Runs 5+ research agents + scoring + critic pipeline and returns BET/BUILD/PIVOT/KILL verdict. Does NOT trigger for: writing copy/PRDs/pitch decks, filling in document sections, tech stack questions, knowledge vault lookups, general market research without a specific idea to evaluate."
license: MIT
metadata:
  version: "0.1.0"
  author: Georgios Pilitsoglou
---

# Evaluate

Evaluates business ideas across 6 model types (directory, e-commerce, SaaS, marketplace, content, tool-site) through a rigorous multi-stage research and scoring pipeline with lens-specific rubrics.

## When to Invoke

Trigger on any of:
- "evaluate idea", "score idea", "rate idea", "validate idea"
- "/idea-forge:evaluate"
- "is this a good idea?", "should I build X?", "is there a market for X?"
- "what are the chances of success?", "give me honest feedback on this"
- "evaluate this e-commerce idea", "rate this SaaS concept"
- "is this marketplace viable?", "should I build a directory for X?"
- User describes any business concept and asks for assessment or feedback

## What It Does

1. Receives an idea description (1-2 paragraphs)
2. Asks 5 targeted questions about founder fit and early validation (Criteria 11 + 13)
3. Runs 5 parallel research agents: market demand, competition (with funding intelligence + tech stack detection), data availability, distribution (with traction channel fit), and customer voice
4. Runs a competitor deep-dive agent: Tranco rank, domain age, page count, robots.txt signals, HTTP headers, funding history
5. Scores the idea across 13 weighted criteria with evidence — source attribution required for all claims
6. Classifies PMF archetype (Hair on Fire / Hard Fact / Future Vision) and checks for tarpit patterns
7. Stress-tests scores through a critic agent to catch bias
8. Produces a final scored card with DVF assessment, competitor landscape, feature gap matrix, traction channel ranking, customer development stage, and verdict (BET / BUILD / PIVOT / KILL)
9. Archives all raw research data to `ideas/{slug}/research/` for future reference

## How to Use

Load and follow the orchestration prompt:

```
Read ${CLAUDE_PLUGIN_ROOT}/skills/evaluate/evaluator.md
```

The evaluator.md file contains the complete pipeline instructions, agent dispatch logic, and output handling.

## Key Files

| File | Purpose |
|------|---------|
| `evaluator.md` | Main orchestration — read this to run an evaluation |
| `references/criteria.md` | 13 scoring criteria with rubrics |
| `references/workspace-schema.md` | Complete file contract for all agent outputs |
| `references/idea-card-template.md` | Output template for scored cards |
| `references/ranking-entry-template.md` | Leaderboard row format |
| `references/lenses/directory.md` | Lens: directory business model rubrics |
| `references/lenses/ecommerce.md` | Lens: e-commerce business model rubrics |
| `references/lenses/saas.md` | Lens: SaaS business model rubrics |
| `references/lenses/marketplace.md` | Lens: marketplace business model rubrics |
| `references/lenses/content.md` | Lens: content/reference business model rubrics |
| `references/lenses/tool-site.md` | Lens: free web utility (calculator/converter/generator) rubrics |

## Agents

All agent prompts live in `${CLAUDE_PLUGIN_ROOT}/agents/` at the plugin root:

| Agent | Purpose |
|-------|---------|
| `${CLAUDE_PLUGIN_ROOT}/agents/market-research.md` | Stage 1: Market research agent |
| `${CLAUDE_PLUGIN_ROOT}/agents/competition-research.md` | Stage 1: Competition research agent |
| `${CLAUDE_PLUGIN_ROOT}/agents/data-research.md` | Stage 1: Data availability agent |
| `${CLAUDE_PLUGIN_ROOT}/agents/distribution-research.md` | Stage 1: Distribution opportunity agent |
| `${CLAUDE_PLUGIN_ROOT}/agents/customer-voice.md` | Stage 1: Synthetic interview — pain signals from public sources |
| `${CLAUDE_PLUGIN_ROOT}/agents/competitor-deep-dive.md` | Stage 1.5: Deep competitor profiling agent |
| `${CLAUDE_PLUGIN_ROOT}/agents/scoring.md` | Stage 2: Evidence-based scoring agent |
| `${CLAUDE_PLUGIN_ROOT}/agents/critic.md` | Stage 3: Bias detection and score adjustment |
| `${CLAUDE_PLUGIN_ROOT}/agents/orchestrator.md` | Stage 4: Final aggregation and verdict |
| `${CLAUDE_PLUGIN_ROOT}/agents/family-evaluator.md` | Family mode: evaluate a cluster of related ideas |

## Output Location

- Individual idea cards: `ideas/{slug}/scored-card-v{N}.md` (always a folder)
- Research archives: `ideas/{slug}/research/`
- Rankings leaderboard: `ideas/_registry/ranking.md`
- New ideas before evaluation: `ideas/drafts/{slug}.md` (promoted to folder when the evaluator first runs)

Attribution

neotherapperneotherapper
View sourceSee grades on GitHubMore from neotherapper →
SSkills Directory ProSkills Directory

Get any skill into Claude in one click.

Download any skill as a ZIP for Claude.ai, Claude Desktop, or .claude/skills. $9/mo.

See Pro

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments (0)

No comments yet. Be the first to comment!

SSkills Directory ProSkills Directory

Get any skill into Claude in one click.

Download any skill as a ZIP for Claude.ai, Claude Desktop, or .claude/skills. $9/mo.

See Pro

Related Skills

Caveman

Ultra-compressed communication mode that cuts output tokens while keeping technical accuracy. Levels: lite, full, ultra and the wenyan variants. Use for /caveman, "caveman mode", "talk like caveman", "be brief" or "less tokens".

1074701 votes

Hyperplan

Adversarial multi-agent planning skill. Self-orchestrates 5 hostile category members (unspecified-low, unspecified-high, deep, ultrabrain, artistry) via team-mode for ruthless cross-critique debate, distills only the defensible insights, then MANDATORILY hands the distilled insight bundle to the `plan` agent for executable plan formalization. Use when planning needs maximum rigor and surfacing of weak assumptions, blind spots, and over-engineering. Triggers: 'hyperplan', 'hpp', '/hyperplan', ...

696561 votes

Writing Skills

Create and manage Claude Code skills in HASH repository following Anthropic best practices. Use when creating new skills, modifying skill-rules.json, understanding trigger patterns, working with hooks, debugging skill activation, or implementing progressive disclosure. Covers skill structure, YAML frontmatter, trigger types (keywords, intent patterns), UserPromptSubmit hook, and the 500-line rule. Includes validation and debugging with SKILL_DEBUG. Examples include rust-error-stack, cargo-dep...

3931 votes

Mcp Code Execution

Routes multi-tool workflows through MCP servers for large datasets and pipelines. Use when Bash tool overhead is limiting throughput on data-heavy tasks.

3351 votes

catchup

Recovers the conversation and failed tool calls of a previous Codex, Claude Code, Antigravity, Cline, Copilot CLI, Cursor, DeepSeek Harness, Kimi, OpenCode, Pi Agent, or ZCode session. Use when the user says "catch up", "what did the last session do", "get me up to speed", "I switched agents", asks to recover/summarize a previous session before continuing, or asks to diagnose or report a catchup failure. Do NOT use for the current conversation, git history, or any non-agent log.

691 votes
View all in ai-agents →