empirical-proof
Pass
Audited by Gen Agent Trust Hub on Jun 13, 2026
Risk Level: SAFECOMMAND_EXECUTION
Full Analysis
- [SAFE]: The skill establishes a protocol that requires agents to provide verbatim proof of successful task execution, which is a defensive measure against hallucination and ensures the accuracy of agent claims.
- [COMMAND_EXECUTION]: The skill facilitates the execution of local shell commands (such as 'npm test' or 'npx jest') to verify code changes. These commands are sourced from a repository-local configuration file ('AGENTS.md'), and the skill includes instructions to consult the user if the commands are not clearly defined, maintaining human-in-the-loop control.
Audit Metadata