benchmark
Pass
Audited by Gen Agent Trust Hub on Jun 19, 2026
Risk Level: SAFE
Full Analysis
- [SAFE]: No security threats or malicious patterns were detected. The skill is explicitly designed for performance monitoring and regression detection.
- [COMMAND_EXECUTION]: To measure build performance, the skill instructions include executing local development processes such as cold builds, test suites, and Docker builds. This execution is standard for performance analysis of a software project.
- [EXTERNAL_DOWNLOADS]: The skill makes network requests to target URLs and API endpoints to gather metrics like latency and response size. These operations are restricted to the benchmarking use-case and do not involve executing untrusted remote scripts.
Audit Metadata