ai-post-training

Pass

Audited by Gen Agent Trust Hub on Aug 12, 2026

Risk Level: SAFE
Full Analysis
  • [SAFE]: The skill provides purely instructional and architectural guidance on machine learning topics such as RLHF, DPO, and GRPO.
  • [SAFE]: External references in data/sources.json point to established academic platforms (arXiv), technical documentation (Hugging Face), and reputable publishers (Manning).
  • [SAFE]: No patterns of prompt injection, data exfiltration, or obfuscation were identified across the 7 files.
  • [SAFE]: The mentioned feedback script append_learning.py is a standard component of the skill development framework and is used for non-malicious metadata updates.
Audit Metadata
Risk Level
SAFE
Analyzed
Aug 12, 2026, 09:09 PM
Security Audit — agent-trust-hub — ai-post-training