zone-of-truth

Pass

Audited by Gen Agent Trust Hub on Jun 27, 2026

Risk Level: SAFE
Full Analysis
  • [SAFE]: The skill consists entirely of natural language instructions meant to improve output quality through sourcing and uncertainty labeling. No executable code or scripts are included.
  • [SAFE]: No network activity, remote downloads, or sensitive data access patterns were identified. The skill correctly identifies that it has no extra runtime dependencies.
  • [SAFE]: The instructions for 'activating' a mode do not involve bypassing safety filters; instead, they focus on reducing hallucinations and increasing evidence-based reasoning.
  • [DATA_EXPOSURE]: The skill instructs the agent to re-examine prior conversation assertions. While this processes untrusted data (prior chat history), the skill lacks any capabilities (such as shell access or network tools) that could be exploited via indirect prompt injection.
Audit Metadata
Risk Level
SAFE
Analyzed
Jun 27, 2026, 07:47 PM
Security Audit — agent-trust-hub — zone-of-truth