benchmark-model
Pass
Audited by Gen Agent Trust Hub on Jul 30, 2026
Risk Level: SAFECOMMAND_EXECUTIONDATA_EXFILTRATIONEXTERNAL_DOWNLOADSCREDENTIALS_UNSAFE
Full Analysis
- [COMMAND_EXECUTION]: The skill uses the
max benchmarkCLI tool viapixi runto perform performance testing. This is the intended purpose of the skill and involves standard local command execution.\n- [DATA_EXFILTRATION]: Network operations are confined to checking the status of a local server athttp://localhost:8000. There are no indications of data being sent to external or unauthorized endpoints.\n- [EXTERNAL_DOWNLOADS]: The tool supports using datasets likesharegptfrom Hugging Face Hub for benchmarking. These are standard resources from a well-known service and do not represent a security risk.\n- [CREDENTIALS_UNSAFE]: No sensitive information, such as API keys or private tokens, is hardcoded or requested by the skill instructions.
Audit Metadata