Build graded relevance judgments (explicit or click-derived) and an offline harness before tuning. Reach for this when there's no eval.
Scanned 9/23/2026
npx -y skills add mcorbett51090/RavenClaude --skill build-judgment-list --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Build Judgment List?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/mcorbett51090-build-judgment-list)More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.
---
name: build-judgment-list
description: "Build graded relevance judgments (explicit or click-derived) and an offline harness before tuning. Reach for this when there's no eval."
---
# Skill: Build judgment list
Tuning without a judgment list is guessing dressed as engineering (§3 #3).
## Step 1 — Sample the query mix
Representative queries weighted by real traffic (§3 #7).
## Step 2 — Grade relevance
Explicit graded labels or click-derived judgments, with position-bias caution (§3 #3 #6).
## Step 3 — Build the offline harness
Reusable NDCG/MRR/precision@k harness over the judgment list (§3 #3).
## Step 4 — Set the baseline
The current ranking's metrics — the bar every change must beat (§3 #1).
## Output
A graded judgment list and an offline harness with a recorded baseline.
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!