static-qa

Pass

Audited by Gen Agent Trust Hub on Sep 9, 2026

Risk Level: SAFE
Full Analysis
  • [INDIRECT_PROMPT_INJECTION]: The skill processes AI-generated images which represent untrusted external data. This creates a surface for indirect prompt injection where an adversarial image could attempt to influence the multimodal reviewer. The risk is mitigated by explicit boundary markers in the judge prompt, OCR-based string validation, and the requirement for a fresh, context-free agent for every review batch.
  • [DYNAMIC_EXECUTION]: The rubric includes instructions for deterministic pixel measurement using standard image processing libraries (PIL and numpy). These operations are restricted to local image analysis and do not involve executing untrusted code or unsafe dynamic loading.
Audit Metadata
Risk Level
SAFE
Analyzed
Sep 9, 2026, 02:35 AM
Security Audit — agent-trust-hub — static-qa