blender-agent-benchmark
Warn
Audited by Snyk on Aug 20, 2026
Risk Level: MEDIUM
Full Analysis
MEDIUM W011: Third-party content exposure detected (indirect prompt injection risk).
- Third-party content exposure detected (medium risk: 0.30). The runtime workflow in
scripts/run_benchmark.tsingests outsider-authored free text from the agent prompt/user task fileworkdir/TASK.md(generated fromtask.prompt) intocodex execviarunCodex()stdin, and later prompts judges usingtask.visualBrief/task.visualCriteria(also ultimately free text) intocodex execforjudgePair()without selecting a specific pre-approved item from a trusted inbox/feed.
Issues (1)
W011
MEDIUMThird-party content exposure detected (indirect prompt injection risk).
Audit Metadata