simpo
Pass
Audited by Gen Agent Trust Hub on Aug 22, 2026
Risk Level: SAFE
Full Analysis
- [SAFE]: The skill provides documentation and configuration examples for preference optimization in machine learning. It follows standard best practices for environment setup and model training.
- [EXTERNAL_DOWNLOADS]: The skill instructs the user to clone the HuggingFace alignment-handbook and install Flash Attention. These are well-known, trusted resources within the AI research community.
- [COMMAND_EXECUTION]: The skill provides command-line examples for launching training jobs using
accelerateandconda. These commands are standard for the described purpose and do not perform any hidden or suspicious actions. - [DATA_EXPOSURE]: The datasets referenced (UltraFeedback, HelpSteer, Anthropic HH-RLHF) are public, widely-used research datasets hosted on HuggingFace, presenting no risk of unauthorized data exposure.
Audit Metadata