arabench
Pass
Audited by Gen Agent Trust Hub on Jun 26, 2026
Risk Level: SAFE
Full Analysis
- [SAFE]: No security issues detected. The skill provides clear instructions for using a benchmarking tool.
- [COMMAND_EXECUTION]: The skill utilizes local CLI commands (
arabench) to perform evaluations and comparisons. These commands are consistent with the skill's stated purpose of benchmarking AI models.
Audit Metadata