escalation-governance
Pass
Audited by Gen Agent Trust Hub on May 18, 2026
Risk Level: SAFE
Full Analysis
- [SAFE]: The skill defines a decision-making protocol for model escalation, focusing on distinguishing between reasoning gaps and knowledge gaps.
- [PROMPT_INJECTION]: The test scenarios, such as those in
test-authority.mdandtest-thrashing.md, utilize imperative language common in evaluation prompts to simulate pressure and force decision-making. These are legitimate educational components of the governance framework and do not attempt to override core agent safety constraints. - [DATA_EXFILTRATION]: Scenario-based descriptions mention sensitive files such as
.envandauth/token_validator.py. These are used as symbolic markers for security-sensitive tasks and do not involve executable commands that access or transmit real data.
Audit Metadata