unassisted-evidence-checkpoint

Pass

Audited by Gen Agent Trust Hub on Aug 11, 2026

Risk Level: SAFE
Full Analysis
  • [PROMPT_INJECTION]: No malicious instruction overrides or safety bypass attempts were detected. The system prompt constraints (e.g., refusing to provide hints) are aligned with the skill's stated educational purpose of measuring unassisted performance.
  • [DATA_EXFILTRATION]: No network operations, such as curl or wget, or sensitive file system access were identified. The skill does not handle or request credentials or other sensitive user data.
  • [REMOTE_CODE_EXECUTION]: The skill does not install external packages or execute scripts from remote sources. It contains no commands that would lead to arbitrary code execution.
  • [OBFUSCATION]: Analysis of the skill body and metadata revealed no hidden content, Base64 encoding, zero-width characters, or homoglyph-based attacks.
  • [COMMAND_EXECUTION]: No shell commands, privilege escalation attempts (sudo), or persistence mechanisms were found. The skill operates within the standard conversational constraints of the agent.
  • [INDIRECT_PROMPT_INJECTION]: The skill ingests untrusted data from fields like topic_or_skill and scaffolded_session_summary. However, since the skill possesses no dangerous capabilities (such as file writes or network access), the risk of exploiting these ingestion points is negligible.
Audit Metadata
Risk Level
SAFE
Analyzed
Aug 11, 2026, 08:13 PM
Security Audit — agent-trust-hub — unassisted-evidence-checkpoint