Model Evaluation Suite

PluginMonitoring & ops

Lets your agent run evaluations on AI models using multiple metrics to check performance.

Comprehensive model evaluation with multiple metrics

Delivery for this kind is on the roadmap — not serving yet. You can still add it. It stays paused until ahel can serve it.

Serve it through your gateway

One link, every agent. Your own credentials, stored once.

Signals

GitHub stars
3k
Forks
392
Last commit
Aug 2026
Installs
2k stars