llm-eval-harness
Pass
Audited by Gen Agent Trust Hub on Jul 22, 2026
Risk Level: SAFE
Full Analysis
- [SAFE]: No security issues detected. The skill follows excellent security hygiene by masking sensitive authentication headers (API keys) before writing any debug or trace information to local files. It correctly handles credentials by instructing users to use environment variables, ensuring that API keys are not exposed in process listings, shell history, or logs. The primary functions of the scripts involve network communication with user-provided endpoints for benchmarking purposes, which is consistent with the skill's stated intent.
Audit Metadata