counterfactual

Pass

Audited by Gen Agent Trust Hub on Sep 1, 2026

Risk Level: SAFENO_CODE
Full Analysis
  • [SAFE]: The skill consists entirely of markdown instructions and YAML configuration. It contains no executable scripts, shell commands, or binary files.
  • [SAFE]: The instructions explicitly restrict the agent from performing external actions, stating: 'do not browse, change code, or act on the scenario'.
  • [PROMPT_INJECTION]: No patterns associated with prompt injection, jailbreaking, or instruction override were detected. The skill uses a specific trigger ($counterfactual) and restricts itself to logical reasoning.
  • [DATA_EXFILTRATION]: There are no network operations, URL references, or patterns suggesting the collection or transmission of sensitive data.
Audit Metadata
Risk Level
SAFE
Analyzed
Sep 1, 2026, 06:02 PM
Security Audit — agent-trust-hub — counterfactual