benchmark-and-baseline-selector

Pass

Audited by Gen Agent Trust Hub on Jun 23, 2026

Risk Level: SAFE
Full Analysis
  • [PROMPT_INJECTION]: No malicious instruction overrides or safety bypass patterns were detected. The instructions are focused on guiding the user through research evaluation planning.
  • [DATA_EXFILTRATION]: No network operations, credential harvesting, or access to sensitive file paths were identified.
  • [REMOTE_CODE_EXECUTION]: The skill does not perform any remote code execution or download external scripts.
  • [COMMAND_EXECUTION]: No shell commands or subprocess calls are present in the instructions or evaluation files.
  • [OBFUSCATION]: No obfuscated content, encoded strings, or hidden characters were found.
  • [EXTERNAL_DOWNLOADS]: No external URLs or package dependencies are referenced.
Audit Metadata
Risk Level
SAFE
Analyzed
Jun 23, 2026, 08:19 AM
Security Audit — agent-trust-hub — benchmark-and-baseline-selector