devils-advocate

Pass

Audited by Gen Agent Trust Hub on Sep 15, 2026

Risk Level: SAFE
Full Analysis
  • [SAFE]: The skill consists entirely of natural language instructions for a reasoning protocol. No executable scripts, binaries, or configuration files that could carry out malicious actions are included.
  • [SAFE]: No network operations, data exfiltration patterns, or attempts to access sensitive files or credentials were detected.
  • [SAFE]: The instructions do not attempt to bypass safety filters or override system-level security constraints. The adversarial reasoning prescribed is a cognitive technique for decision-making, not a safety bypass mechanism.
  • [SAFE]: There is no obfuscated content, dynamic code execution, persistence mechanism, or privilege escalation logic detected within the skill body or metadata.
Audit Metadata
Risk Level
SAFE
Analyzed
Sep 15, 2026, 06:49 PM
Security Audit — agent-trust-hub — devils-advocate