a-b-testing
Pass
Audited by Gen Agent Trust Hub on Jul 18, 2026
Risk Level: SAFECOMMAND_EXECUTIONPROMPT_INJECTION
Full Analysis
- [SAFE]: The skill instructions define a rigorous framework for experimentation based on established industry methodologies. No malicious patterns or security bypasses were found.
- [COMMAND_EXECUTION]: The skill includes a utility script
scripts/check-output.pywhich is used to verify that output deliverables contain the necessary sections. This script performs basic string matching on local files. - [PROMPT_INJECTION]: The instructions follow a standard educational pattern without attempts to override system constraints or extract sensitive system information.
Audit Metadata