finetuning

Warn

Audited by Snyk on Aug 9, 2026

Risk Level: MEDIUM
Full Analysis

MEDIUM W011: Third-party content exposure detected (indirect prompt injection risk).

  • Third-party content exposure detected (low risk: 0.10). Outsider-authored free text is ingested only from user-provided/local JSONL files into scripts like scripts/calibrate_grader.py (it reads args.data JSONL lines and feeds messages[-1]["content"] plus model outputs to the grader), but there is no external “publish/queue/feed” path that the workflow automatically fetches outsider content from at runtime.

Issues (1)

W011
MEDIUM

Third-party content exposure detected (indirect prompt injection risk).

Audit Metadata
Risk Level
MEDIUM
Analyzed
Aug 9, 2026, 06:15 AM
Issues
1
Security Audit — snyk — finetuning