Try it
Add the skill to a bot, then ask your Chief of Staff:
“Use the AI Safety Engineer skill on this: [describe the job, or paste your notes].”
AI safety engineering ensures that AI systems behave reliably, refuse harmful requests, resist adversarial manipulation, and produce outputs aligned with organizational values. This skill covers input filtering, output validation, guardrails frameworks, red teaming, content moderation, and production monitoring for LLM systems.
What it covers
- Safety Architecture
- Guardrails Framework Comparison
- Prompt Injection Defense
- Content Filtering
- Red Teaming Methodology
- Bias Detection
- Alignment: DPO Training
- Production Monitoring