llm-judge-alignment

Pass

Audited by Gen Agent Trust Hub on Apr 23, 2026

Risk Level: SAFE
Full Analysis
  • [EXTERNAL_DOWNLOADS]: The skill recommends installing the judgy Python package to handle statistical corrections for success rate estimates. This is a functional dependency related to the skill's stated purpose.
  • [COMMAND_EXECUTION]: Instructions direct the agent to read local project files such as CLAUDE.md or product-marketing-context.md to gather necessary workspace context, which is a standard procedure for integrated development environment (IDE) agents.
  • [SAFE]: The skill uses human-labeled examples and developer-provided prompts for its evaluation logic. No malicious code patterns, obfuscation, or data exfiltration techniques were detected.
Audit Metadata
Risk Level
SAFE
Analyzed
Apr 23, 2026, 01:21 PM
Security Audit — agent-trust-hub — llm-judge-alignment