judge-reliability-checker
Pass
Audited by Gen Agent Trust Hub on Jun 15, 2026
Risk Level: SAFENO_CODE
Full Analysis
- [SAFE]: The skill provides a structured workflow and rules for evaluating the reliability of LLM judges without utilizing any automated tools or scripts.
- [NO_CODE]: No scripts, shell commands, or third-party dependencies were found in the skill definition.
Audit Metadata