Constitutional AI and safety guardrail prompts for aligned LLM behavior
Scanned 9/2/2026
Install to Claude Code
npx -y skills add a5c-ai/babysitter --skill constitutional-ai-prompts --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Constitutional Ai Prompts?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/a5c-ai-constitutional-ai-prompts-babysitter)More formats (shields.io, HTML) on the badges page.
---
name: constitutional-ai-prompts
description: Constitutional AI and safety guardrail prompts for aligned LLM behavior
allowed-tools:
- Read
- Write
- Edit
- Bash
- Glob
- Grep
graph:
domains: [domain:software-engineering]
specializations: [specialization:ai-agents-conversational]
skillAreas: [skill-area:prompt-engineering, skill-area:safety-redteaming]
roles: [role:ml-engineer, role:backend-engineer]
workflows: [workflow:ml-model-lifecycle, workflow:feature-development]
---
# Constitutional AI Prompts Skill
## Capabilities
- Design constitutional AI principles
- Implement self-critique and revision prompts
- Create harmlessness guidelines
- Design refusal patterns for unsafe requests
- Implement red-team testing prompts
- Create ethics-aware response frameworks
## Target Processes
- system-prompt-guardrails
- content-moderation-safety
## Implementation Details
### Constitutional Patterns
1. **Critique-Revision**: Self-evaluate and improve responses
2. **Principle Adherence**: Follow defined ethical principles
3. **Harmlessness Focus**: Prioritize safe responses
4. **Helpfulness Balance**: Balance helpfulness with safety
5. **Transparency**: Acknowledge limitations
### Configuration Options
- Constitutional principles list
- Critique prompts
- Revision guidelines
- Refusal templates
- Escalation triggers
### Best Practices
- Define clear constitutional principles
- Balance helpfulness and safety
- Test with adversarial inputs
- Document refusal patterns
- Regular principle review
### Dependencies
- langchain-core
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!