Skills DirectorySkills Directory
SkillsLearnSecurityCategoriesDocsBlogPro
Sign InSubmit Skill
Skills Directory

Security-tested agent skills for Claude, coding agents, and AI workflows.

Directory

  • Browse Skills
  • All Skills A–Z
  • Claude Skills
  • Claude Code Skills
  • Agent Skills
  • Categories
  • Authors
  • Submit a Skill

Learn

  • Learn Hub
  • Install Claude Skills
  • Write SKILL.md
  • Skills vs MCP
  • Directories Compared

Security

  • Security
  • Methodology
  • Secure Claude Skills
  • Security Badges
  • Chrome Extension
  • Skill Manager

Company

  • About
  • Community
  • Blog
  • API Docs
  • Advertise

2026 Skills Directory. All rights reserved.

ProTermsPrivacyRefunds
Back to skills

Common Agent Guardrails

ASecurity

Define deterministic guardrails for agent tool calls — protected paths, test-file locks during bug fixes, post-edit formatters, production approval gates, secret deny rules. Use when writing or reviewing a hook policy, or moving an always-do-X rule out of prose into enforcement.

571 stars
0 votes
0 copies
0 views
Added 9/24/2026
developmentrails

Security Analysis

A100/100

Pro scans all 4 files and shows the line behind each finding

Scanned 9/24/2026

$npx -y skills add HoangNguyen0403/agent-skills-standard --skill common-agent-guardrails --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Common Agent Guardrails?

Add the live security badge to your README — it updates automatically with every re-scan.

Security grade badge for Common Agent Guardrails
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/hoangnguyen0403-common-agent-guardrails/badge)](https://www.skillsdirectory.com/skills/hoangnguyen0403-common-agent-guardrails)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
Files
SKILL.md
---
name: common-agent-guardrails
description: Define deterministic guardrails for agent tool calls — protected paths, test-file locks during bug fixes, post-edit formatters, production approval gates, secret deny rules. Use when writing or reviewing a hook policy, or moving an always-do-X rule out of prose into enforcement.
metadata:
  triggers:
    files:
      - "guardrails.yaml"
      - "hooks.json"
      - ".claude/settings.json"
      - "**/hooks/*.py"
      - "**/hooks/*.js"
    keywords:
      - guardrail
      - protected path
      - block edit
      - approval gate
      - hook policy
      - pretooluse
---

# Agent Guardrail Standard

## **Priority: P1 (HIGH)**

A rule that must always hold belongs in a hook, not in a prompt. Prompts persuade; hooks decide.

## 1. What Belongs in a Guardrail

- **Absolute compliance only**: encode a rule here when a single violation is unacceptable, not when it is merely preferred.
- **Deterministic check**: the decision must be computable from tool name, arguments, and repo state — never from model judgment.
- **Paired with a skill**: the skill teaches the rule, the guardrail enforces it. Ship both or the rule decays.

## 2. Required Guardrail Classes

| Class | Trigger | Decision |
| --- | --- | --- |
| Protected paths | edit to generated, vendored, or frozen files | block |
| Secret deny | edit or read of `.env*`, `.ssh/`, `credentials*.json\|yaml`, identity files | block |
| Test lock | test-file edit while a bug-fix workflow is active | block |
| Formatter | after any accepted edit | run, then report |
| Production action | deploy, migration, or destructive command against production | ask named owner |
| Scope fence | writes outside the declared task scope | ask |

## 3. Decision Semantics

- **Exit 0**: allow, stay silent. Reserve stdout for the formatter class.
- **Exit 2**: block. Print the rule violated and the sanctioned alternative on stderr, never a generic denial.
- **Any other exit**: ask a human. Use it when the check is inconclusive, not as a soft block.
- **Fail closed on production classes**: if the guardrail cannot evaluate, treat production actions as blocked.

## 4. Test Lock During Bug Fixes

- Write the failing test first; the agent proves the bug reproduces before touching production code.
- Lock every test path for the duration of the fix so the proof cannot be weakened into passing.
- Unlock only when the fix is verified, or when the operator states the test itself is the defect.

## 5. Placement and Ownership

- **Team level**: repository-tracked config, reviewed in pull requests like code.
- **Organization level**: managed settings the project cannot override; use for secret deny and production gates.
- **Log every decision**: timestamp, rule, tool, target, and outcome. An unlogged block is unauditable.
- **Version the policy**: a guardrail change is a control change and needs the same review as the control it enforces.

## Anti-Patterns

- **No prompt-only rules**: If it must always hold, enforce it in a hook.
- **No broad path globs**: Name the protected directories, not the whole repo.
- **No silent blocks**: Print the rule and the allowed alternative.
- **No agent self-approval**: A human role authorizes production actions.
- **No guardrail without a skill**: Pair enforcement with the rule that explains it.
- **No unlogged decisions**: Record every block and ask.

## Red Flags

- **Stop if "just this once, disable the hook"**: Re-run with the sanctioned path or escalate to the policy owner.
- **Stop if the fix edits the failing test**: Restore the test and fix the production code.
- **Stop if a production command runs with no named authorizer**: Block and request sign-off.

## References

- [Guardrail Policy Template](references/guardrails-template.yaml)
- [Runtime Adapters](references/runtime-adapters.md)

## Canonical response anchors

When this skill applies, preserve the following domain terminology or equivalent concrete examples in the answer when relevant:

- exit 2 blocks
- protected paths
- test lock
- approval gate
- fail closed
- log the block decision

Attribution

HoangNguyen0403HoangNguyen0403
View sourceSee grades on GitHubMore from HoangNguyen0403 →
SSkills DirectorySkills Directory

Ship a skill? Prove it's safe.

Free 120-pattern security scan, letter grade, and an embeddable README badge.

Submit a skill

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments (0)

No comments yet. Be the first to comment!

SSkills DirectorySkills Directory

Ship a skill? Prove it's safe.

Free 120-pattern security scan, letter grade, and an embeddable README badge.

Submit a skill

Related Skills

Clean Code

Pragmatic coding standards - concise, direct, no over-engineering, no unnecessary comments

304955 votes

Browser Extension Developer

Use this skill when developing or maintaining browser extension code in the `browser/` directory, including Chrome/Firefox/Edge compatibility, content scripts, background scripts, or i18n updates.

286712 votes

Seo Optimizer

SEO optimization with keyword analysis, readability assessment, technical validation, content quality. Use for search rankings, blog posts, content audits, or encountering keyword density, readability scores, meta tags, schema markup errors.

2222 votes

Google Official Seo Guide

Official Google SEO guide covering search optimization, best practices, Search Console, crawling, indexing, and improving website search visibility based on official Google documentation

1862 votes

Writing Plans

Use when you have a spec or requirements for a multi-step task, before touching code

2927051 votes
View all in development →