llm-evaluation-framework
Pass
Audited by Gen Agent Trust Hub on Jun 17, 2026
Risk Level: SAFENO_CODE
Full Analysis
- [SAFE]: The skill provides a structured methodology for LLM evaluation. It does not contain any executable code, command-line operations, or network requests. All instructions are procedural and designed for human-AI collaboration without risky automated actions.
- [NO_CODE]: The analysis confirms that no scripts, binaries, or configuration files (beyond YAML frontmatter) are included or referenced. The content consists entirely of informational markdown and evaluation templates.
Audit Metadata