llm-evaluation
Pass
Audited by Gen Agent Trust Hub on May 13, 2026
Risk Level: SAFE
Full Analysis
- [SAFE]: The skill provides high-level guidance for evaluating Large Language Models. There are no executable scripts, external dependencies, or commands that could pose a security risk. All content is informational text.
Audit Metadata