eval-agents

Installation
SKILL.md

Agent Evaluator

Discover all agents in scope, score each one across five criteria, then run an interactive session to confirm or improve them one by one.

Agents are not just scripts: they are callable units selected by orchestrators based on their description. A vague description silently breaks multi-agent workflows. The goal here is not just scoring; it is leaving every agent correctly scoped, correctly modeled, and safe to call from an orchestrator.

When to Use

  • First time auditing an agent fleet before wiring it into an orchestration pipeline
  • An orchestrator keeps selecting the wrong agent for a task
  • After copying agents from another project or importing a plugin
  • A new agent was added; checking whether it conflicts with existing ones
  • Periodic hygiene: "do all these agents still do something distinct?"

Key Concepts

Agent file locations

Installs
3
GitHub Stars
5.8K
First Seen
10 days ago
eval-agents — florianbruniaux/claude-code-ultimate-guide