skills/modular/skills/benchmark-model/Gen Agent Trust Hub

benchmark-model

Pass

Audited by Gen Agent Trust Hub on Jul 30, 2026

Risk Level: SAFECOMMAND_EXECUTIONDATA_EXFILTRATIONEXTERNAL_DOWNLOADSCREDENTIALS_UNSAFE
Full Analysis
  • [COMMAND_EXECUTION]: The skill uses the max benchmark CLI tool via pixi run to perform performance testing. This is the intended purpose of the skill and involves standard local command execution.\n- [DATA_EXFILTRATION]: Network operations are confined to checking the status of a local server at http://localhost:8000. There are no indications of data being sent to external or unauthorized endpoints.\n- [EXTERNAL_DOWNLOADS]: The tool supports using datasets like sharegpt from Hugging Face Hub for benchmarking. These are standard resources from a well-known service and do not represent a security risk.\n- [CREDENTIALS_UNSAFE]: No sensitive information, such as API keys or private tokens, is hardcoded or requested by the skill instructions.
Audit Metadata
Risk Level
SAFE
Analyzed
Jul 30, 2026, 06:42 PM
Security Audit — agent-trust-hub — benchmark-model