constitutional-ai-prompts ยท diff
git:20260507.313363d to git:20260507.50823d8
4 added, 0 removed. Audit A to A.
---
name: constitutional-ai-prompts
description: Constitutional AI and safety guardrail prompts for aligned LLM behavior
allowed-tools:
- Read
- Write
- Edit
- Bash
- Glob
- Grep
graph:
domains: [domain:software-engineering]
+ specializations: [specialization:ai-agents-conversational]
+ skillAreas: [skill-area:natural-language-processing]
+ roles: [role:ml-engineer, role:backend-engineer]
+ workflows: [workflow:ml-model-lifecycle, workflow:feature-development]
---
# Constitutional AI Prompts Skill
## Capabilities
- Design constitutional AI principles
- Implement self-critique and revision prompts
- Create harmlessness guidelines
- Design refusal patterns for unsafe requests
- Implement red-team testing prompts
- Create ethics-aware response frameworks
## Target Processes
- system-prompt-guardrails
- content-moderation-safety
## Implementation Details
### Constitutional Patterns
1. **Critique-Revision**: Self-evaluate and improve responses
2. **Principle Adherence**: Follow defined ethical principles
3. **Harmlessness Focus**: Prioritize safe responses
4. **Helpfulness Balance**: Balance helpfulness with safety
5. **Transparency**: Acknowledge limitations
### Configuration Options
- Constitutional principles list
- Critique prompts
- Revision guidelines
- Refusal templates
- Escalation triggers
### Best Practices
- Define clear constitutional principles
- Balance helpfulness and safety
- Test with adversarial inputs
- Document refusal patterns
- Regular principle review
### Dependencies
- langchain-core