Skills DirectorySkills Directory
SkillsLearnSecurityCategoriesDocsBlogPro
Sign InSubmit Skill
Skills Directory

Security-tested agent skills for Claude, coding agents, and AI workflows.

Directory

  • Browse Skills
  • All Skills A–Z
  • Claude Skills
  • Claude Code Skills
  • Agent Skills
  • Categories
  • Authors
  • Submit a Skill

Learn

  • Learn Hub
  • Install Claude Skills
  • Write SKILL.md
  • Skills vs MCP
  • Directories Compared

Security

  • Security
  • Methodology
  • Secure Claude Skills
  • Security Badges
  • Chrome Extension
  • Skill Manager

Company

  • About
  • Community
  • Blog
  • API Docs
  • Advertise

2026 Skills Directory. All rights reserved.

ProTermsPrivacyRefunds
Back to skills

Show Me Your Work

ASecurity

Record an append-only decision trail for long or delegated work.

2 stars
0 votes
0 copies
0 views
Added 10/1/2026
ai-agents

Security Analysis

A100/100

Pro scans all 4 files and shows the line behind each finding

Scanned 10/6/2026

$npx -y skills add williamwue/oh-my-stack --skill show-me-your-work --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Show Me Your Work?

Add the live security badge to your README — it updates automatically with every re-scan.

Security grade badge for Show Me Your Work
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/williamwue-show-me-your-work-oh-my-stack/badge)](https://www.skillsdirectory.com/skills/williamwue-show-me-your-work-oh-my-stack)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
Files
SKILL.md
---
name: show-me-your-work
description: "Record an append-only decision trail for long or delegated work."
---

# Show Me Your Work

## Codex delegation binding

For every delegated worker in this workflow, derive the exact `model`,
`reasoning_effort`, and complete role-plus-task `message` with
`../../scripts/codex-delegation.mjs prepare` relative to this Skill. It resolves
the nearest project manifest first, then the user manifest. Supply the named
route/panel entry where configured; otherwise supply the canonical role
and the observed parent model and effort. Pass
the returned `task_name`, `fork_turns=none`, model, effort, and message
explicitly to the spawn call. Do not use a generated custom-role name as a selector or
claim its TOML was activated. After the worker finishes, run the helper's
`verify` mode on the persisted parent and child records when available; it
checks the spawn metadata, parent link, and child `turn_context`.
The persisted spawn message may be encrypted; disclose when its exact
role/task text cannot be audited. If records are unavailable, state that
runtime model resolution is unverified.

## Child session handoff

Read [the handoff contract](../poteto-mode/references/subagent-handoff.md).
New tasks, repair rounds, retries, and queue items use fresh child sessions
with the original brief, every later directive, prior findings and responses,
and unresolved objections. Reuse only for required costly live state, and
only when the host allows it. Stop and fence active writers before replacement.
A host-owned orchestrator's model catalog, workspace binding, child tools,
and review-round rules take precedence over the native binding above.
Keep its task handles and attribution receipts. Do not use a backing child
conversation as a new delegated review, or claim native-record verification
for a host-owned child. Report attribution evidence gaps explicitly.

Keep one canonical decision log. Use it for work with multiple phases, delegated
sessions, important pivots, or verification that a reviewer will inspect later.
Do not turn routine commands into noise.

When a log spans runs, append a `start` row for each run with its attributable
session or run identity and starting revision. Record its ending timestamp or
checkpoint in a later row. Audit the rows between those boundaries against that
run's evidence; a later run's success does not validate an earlier run's claims.
If runs overlap, include the run identity in each row's evidence pointer rather
than treating all intervening rows as one run. Preserve earlier rows, including
incorrect ones, and supersede them with a later correction that cites the
original row and resolvable evidence.

## Start the log

Copy `references/decision-log-template.tsv` to `decisions.tsv` in the working
directory, or to `.audit/<task-slug>.tsv` when several efforts run at once.
Treat it as a working artifact by default. Commit it only when the user or the
review contract requires the trail to travel with the result.

The columns are:

- `ts`: an ISO 8601 timestamp;
- `phase`: the phase or workstream;
- `decision`: the concrete choice or action;
- `why`: the reason in plain language;
- `evidence`: a short, resolvable pointer such as a revision, command output,
  file location, trace, or screenshot;
- `result`: the observed state, including `open` or `INCONCLUSIVE` when the
  evidence is not final.

Use `scripts/log.sh <logfile> <phase> <decision> <why> <evidence> <result>` to
append a row. The helper creates the header, keeps cells on one line, and
neutralizes spreadsheet formulas. If packaged-script execution is unavailable,
append the same six columns using another safe workspace-writing mechanism and
disclose that fallback.

## What to log

Append one row for a decision or checkpoint that changes how a reviewer should
understand the work:

- choosing one implementation path over another;
- completing a bounded unit and recording its verification;
- rejecting, reverting, or superseding earlier work;
- surfacing a blocker or changing a gate;
- accepting or rejecting a delegated result.

The log is append-only. Correct a bad row with a later row that identifies what
it supersedes. Never rewrite or delete history to make the run look cleaner.
Evidence is a pointer, not a paragraph, and a claim without resolvable evidence
must remain open or inconclusive.

## Audit before handoff

Walk every row against the best evidence available from the current run.

1. Confirm that every row maps to an action that actually occurred.
2. Resolve each evidence pointer and confirm it supports the stated result.
3. Append missing pivots, abandoned approaches, verification failures, or
   superseding decisions that affected the outcome.
4. Append corrections for inaccurate rows; do not edit the earlier rows.
5. Remove no history. If a trivial row is distracting, append a note explaining
   that it is non-material.

When the runtime exposes an attributable transcript, include it in the audit.
When transcript access is absent, incomplete, encrypted, or external, audit the
visible messages, tool evidence, repository state, and other resolvable
pointers instead, and state the limitation. Never search unrelated private
sessions to fill that gap.

## Independent trail review

For consequential work, ask one new read-only reviewer session to inspect the
frozen log and the attributable evidence. A different model family is preferred
when the runtime can select and prove it, but model diversity is not a condition
for truth. If a distinct model or reviewer session is unavailable, perform the
review in the root session and disclose that it was not independent.

The review checks for weak evidence, unverified success claims, risky pivots,
missing failures, and gaps between the log and the observable run. Freeze the
review before final reporting. The root coordinator then resolves or reports
each flag and independently verifies the final workspace state.

Finish with an `Attention` section that names the review boundary and lists
specific flagged rows or says `No flags`. Report the resolved model identity
only when runtime-produced evidence establishes it; otherwise say that the
model identity was not verified.

Attribution

williamwuewilliamwue
View sourceSee grades on GitHubMore from williamwue →
SSkills DirectorySkills Directory

Ship a skill? Prove it's safe.

Free 120-pattern security scan, letter grade, and an embeddable README badge.

Submit a skill

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments (0)

No comments yet. Be the first to comment!

SSkills DirectorySkills Directory

Ship a skill? Prove it's safe.

Free 120-pattern security scan, letter grade, and an embeddable README badge.

Submit a skill

Related Skills

Caveman

Terse caveman voice: answer first, fluff gone, every technical fact kept. Use for /caveman, "caveman mode", "talk like caveman", "be brief", "less tokens". Stays on until "stop caveman" or "normal mode".

1100021 votes

Hyperplan

Adversarial multi-agent planning skill. Self-orchestrates 5 hostile category members (unspecified-low, unspecified-high, deep, ultrabrain, artistry) via team-mode for ruthless cross-critique debate, distills only the defensible insights, then MANDATORILY hands the distilled insight bundle to the `plan` agent for executable plan formalization. Use when planning needs maximum rigor and surfacing of weak assumptions, blind spots, and over-engineering. Triggers: 'hyperplan', 'hpp', '/hyperplan', ...

698461 votes

Writing Skills

Create and manage Claude Code skills in HASH repository following Anthropic best practices. Use when creating new skills, modifying skill-rules.json, understanding trigger patterns, working with hooks, debugging skill activation, or implementing progressive disclosure. Covers skill structure, YAML frontmatter, trigger types (keywords, intent patterns), UserPromptSubmit hook, and the 500-line rule. Includes validation and debugging with SKILL_DEBUG. Examples include rust-error-stack, cargo-dep...

3931 votes

Mcp Code Execution

Routes multi-tool workflows through MCP servers for large datasets and pipelines. Use when Bash tool overhead is limiting throughput on data-heavy tasks.

3421 votes

catchup

Recovers the conversation and failed tool calls of a previous Codex, Amp, Claude Code, Antigravity, Cline, Copilot CLI, Cursor, DeepSeek Harness, Grok Build, Kimi, OpenCode, Pi Agent, or ZCode session. Use when the user says "catch up", "what did the last session do", "get me up to speed", "I switched agents", asks to recover/summarize a previous session before continuing, or asks to diagnose or report a catchup failure. Do NOT use for the current conversation, git history, or any non-agent log.

741 votes
View all in ai-agents →