fine-tuning-with-trl
Warn
Audited by Snyk on Sep 9, 2026
Risk Level: MEDIUM
Full Analysis
MEDIUM W011: Third-party content exposure detected (indirect prompt injection risk).
- Third-party content exposure detected (low risk: 0.10). The required workflow involves fine-tuning LLMs using datasets from HuggingFace (such as
trl-lib/Capybaraortrl-lib/ultrafeedback_binarized), which are community-published datasets containing outsider-authored free text. However, because datasets on HuggingFace require active searching or fetching by specific identifiers rather than continuous automated monitoring or direct feed consumption, this represents a low indirect prompt injection risk.
Issues (1)
W011
MEDIUMThird-party content exposure detected (indirect prompt injection risk).
Audit Metadata