bounded-evaluation
Installation
SKILL.md
Bounded Evaluation
Use this skill when a change needs a small, explicit evaluation gate.
Use When
- reviewing skill changes, AGENTS.md rules, prompt playbooks, or SET bundle changes;
- comparing two agent outputs or proposals;
- using an LLM judge where bias or weak rubrics could mislead;
- creating activation tests, regression cases, or ship gates for agent behavior;
- deciding whether a repeated workflow improved after a bounded edit.
Evaluation Contract
A bounded eval must state: