iris

PackDev tools

Scores your agent's responses for quality, safety, privacy leaks, and cost using built-in evaluation checks.

Unavailable. Delivery for this kind is on the roadmap — not serving yet.

Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.

About this app

The agent eval standard for MCP. Score every agent output for quality, safety, and cost. 12 built-in eval rules cover completeness, relevance, safety (PII detection, prompt injection), and cost thresholds. Log traces with hierarchical spans, evaluate outputs inline, and track costs across all your a

Signals

GitHub stars
4k
Forks
312
Last commit
Aug 2026
Advanced
Item type
plugin
Key
anthropics-claude-plugins-community-iris
Source
github.com/anthropics/claude-plugins-community