escalation-governance

Pass

Audited by Gen Agent Trust Hub on May 18, 2026

Risk Level: SAFE
Full Analysis
  • [SAFE]: The skill defines a decision-making protocol for model escalation, focusing on distinguishing between reasoning gaps and knowledge gaps.
  • [PROMPT_INJECTION]: The test scenarios, such as those in test-authority.md and test-thrashing.md, utilize imperative language common in evaluation prompts to simulate pressure and force decision-making. These are legitimate educational components of the governance framework and do not attempt to override core agent safety constraints.
  • [DATA_EXFILTRATION]: Scenario-based descriptions mention sensitive files such as .env and auth/token_validator.py. These are used as symbolic markers for security-sensitive tasks and do not involve executable commands that access or transmit real data.
Audit Metadata
Risk Level
SAFE
Analyzed
May 18, 2026, 03:29 PM
Security Audit — agent-trust-hub — escalation-governance