agent-evaluation
Pass
Audited by Gen Agent Trust Hub on Jun 23, 2026
Risk Level: SAFENO_CODE
Full Analysis
- [NO_CODE]: The skill contains only natural language instructions, descriptions, and evaluation frameworks in markdown format. There are no scripts, command-line instructions, or executable code blocks.
- [SAFE]: Analysis of the content revealed no indicators of prompt injection, data exfiltration, obfuscation, or other malicious patterns. The 'Sharp Edges' section discusses standard software engineering risks like data leakage and test flakiness in a purely conceptual manner.
Audit Metadata