judge-reliability-checker

Pass

Audited by Gen Agent Trust Hub on Jun 15, 2026

Risk Level: SAFENO_CODE
Full Analysis
  • [SAFE]: The skill provides a structured workflow and rules for evaluating the reliability of LLM judges without utilizing any automated tools or scripts.
  • [NO_CODE]: No scripts, shell commands, or third-party dependencies were found in the skill definition.
Audit Metadata
Risk Level
SAFE
Analyzed
Jun 15, 2026, 03:21 AM
Security Audit — agent-trust-hub — judge-reliability-checker