Prompt injection detection and prevention for secure LLM applications
Scanned 9/2/2026
Install to Claude Code
npx -y skills add a5c-ai/babysitter --skill prompt-injection-detector --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Prompt Injection Detector?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/a5c-ai-prompt-injection-detector-babysitter)More formats (shields.io, HTML) on the badges page.
---
name: prompt-injection-detector
description: Prompt injection detection and prevention for secure LLM applications
allowed-tools:
- Read
- Write
- Edit
- Bash
- Glob
- Grep
graph:
domains: [domain:software-engineering]
specializations: [specialization:ai-agents-conversational]
skillAreas: [skill-area:hallucination-mitigation-fact-checking, skill-area:safety-redteaming]
roles: [role:ml-engineer, role:backend-engineer]
workflows: [workflow:feature-development, workflow:ml-model-lifecycle]
---
# Prompt Injection Detector Skill
## Capabilities
- Detect prompt injection attempts
- Implement input sanitization
- Configure detection classifiers
- Design defense layers
- Implement canary token detection
- Create injection logging and alerting
## Target Processes
- prompt-injection-defense
- tool-safety-validation
## Implementation Details
### Detection Methods
1. **Pattern Matching**: Known injection patterns
2. **ML Classifiers**: Trained injection detectors
3. **Canary Tokens**: Detect instruction override
4. **LLM-Based**: Use LLM to detect manipulation
5. **Perplexity Analysis**: Unusual input patterns
### Defense Strategies
- Input preprocessing
- Prompt structure design
- Output validation
- Sandboxed execution
- Multi-layer defense
### Configuration Options
- Detection threshold
- Pattern rules
- Classifier model
- Action policies
- Alerting settings
### Best Practices
- Defense in depth
- Regular pattern updates
- Monitor false positives
- Test with red-team inputs
### Dependencies
- rebuff (optional)
- transformers
- Custom classifiers
Is this your skill, or is something wrong with this listing? . Author removals are honored within 72 hours.
No comments yet. Be the first to comment!