fail-closed-eval-gate

Pass

Audited by Gen Agent Trust Hub on Aug 20, 2026

Risk Level: SAFE
Full Analysis
  • [SAFE]: The skill is a set of documentation and architectural guidelines for establishing 'fail-closed' evaluation gates for AI agents. It does not include any scripts, tools, or executable code.
  • [EXTERNAL_DOWNLOADS]: The skill references several external informational URLs and blog posts regarding agent evaluation practices.
  • Evidence includes links to: kunalganglani.com, pondero.ai, noveum.ai, futureagi.substack.com, and baeseokjae.github.io.
  • These are static documentation links provided for educational context and do not trigger remote code execution or automated downloads.
  • [COMMAND_EXECUTION]: The skill contains markdown code blocks displaying example GitHub Actions configurations.
  • Evidence: npm run eval:replay -- --suite=tier1 --threshold=100.
  • These are provided as illustrative examples for the user to implement in their own infrastructure and are not executed by the agent itself.
  • [INDIRECT_PROMPT_INJECTION]: The skill defines triggers for the agent to act as a gatekeeper (e.g., "ready to merge", "ship this change").
  • Ingestion points: User-provided phrases in the prompt context.
  • Capability inventory: The skill prescribes a 'STOP' behavior unless specific evidence of testing is provided, acting as a safety constraint rather than a vulnerability surface.
Audit Metadata
Risk Level
SAFE
Analyzed
Aug 20, 2026, 11:15 PM
Security Audit — agent-trust-hub — fail-closed-eval-gate