safety-alignment

PackAI & models

Adds safety checks that screen prompts and responses for harmful content, jailbreak attempts, and prompt injections.

Unavailable. Delivery for this kind is on the roadmap — not serving yet.

Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.

About this app

AI safety and content moderation including Constitutional AI, LlamaGuard, NeMo Guardrails, and Prompt Guard. Use when implementing safety filters, content moderation, or prompt injection detection.

Signals

GitHub stars
13k
Forks
931
Last commit
Jun 2026
Advanced
Item type
plugin
Key
orchestra-research-ai-research-skills-safety-alignment
Source
github.com/orchestra-research/ai-research-skills