agentic-eval

Pass

Audited by Gen Agent Trust Hub on Apr 14, 2026

Risk Level: SAFE
Full Analysis
  • [SAFE]: The skill consists of instructional content and structural patterns for AI evaluation loops. It does not contain executable code or malicious tools.
  • [PROMPT_INJECTION]: The skill describes patterns for ingesting external data (tasks, outputs) into refinement loops (Indirect Prompt Injection surface). Ingestion points: 'task', 'output', and 'criteria' variables in the Python snippets in SKILL.md. Boundary markers: Absent from snippets. Capability inventory: Conceptual 'llm()' and 'run_tests()' functions. Sanitization: Absent from template code.
  • [COMMAND_EXECUTION]: The 'Code-Specific Reflection' pattern illustrates the process of verifying generated code with a test runner. This is documented as a developmental workflow pattern and is not an active execution risk within the skill.
Audit Metadata
Risk Level
SAFE
Analyzed
Apr 14, 2026, 07:28 AM
Security Audit — agent-trust-hub — agentic-eval