benchmark

Pass

Audited by Gen Agent Trust Hub on Apr 6, 2026

Risk Level: SAFECOMMAND_EXECUTIONEXTERNAL_DOWNLOADS
Full Analysis
  • [COMMAND_EXECUTION]: Mode 3 (Build Performance) utilizes local command execution to measure performance metrics for cold builds, hot reloads, test suites, TypeScript checks, and Docker builds.
  • [EXTERNAL_DOWNLOADS]: The skill performs network requests to external URLs and API endpoints to measure Core Web Vitals, response sizes, and latency p-values. This is the primary intended behavior of the benchmarking functionality.
Audit Metadata
Risk Level
SAFE
Analyzed
Apr 6, 2026, 04:00 AM
Security Audit — agent-trust-hub — benchmark