eval-skills
Pass
Audited by Gen Agent Trust Hub on Mar 17, 2026
Risk Level: SAFE
Full Analysis
- [COMMAND_EXECUTION]: The skill executes local
run.shscripts usingsubprocess.runwith list-based arguments ineval_runner.py. This is the intended behavior for running behavioral tests and uses secure implementation patterns. - [EXTERNAL_DOWNLOADS]: The
run.shscript leveragesuv, a trusted dependency manager, to execute the Python environment. - [DATA_EXFILTRATION]: Sourcing of a
.envfile from the project root inrun.shis performed to provide necessary configuration for tests, which is standard in development workflows.
Audit Metadata