factcheck

Pass

Audited by Gen Agent Trust Hub on Sep 22, 2026

Risk Level: SAFEINDIRECT_PROMPT_INJECTION
Full Analysis
  • [INDIRECT_PROMPT_INJECTION]: The skill is designed to ingest and process arbitrary user-provided content (such as LinkedIn posts, emails, or scripts) to extract and verify factual claims. This creates a surface for indirect prompt injection, where malicious instructions embedded in the content could attempt to hijack the agent's logic during the research or reporting phase.
  • Ingestion points: User-supplied text provided for fact-checking in Step 2 of SKILL.md.
  • Boundary markers: The instructions do not define clear delimiters (e.g., XML tags or specific markers) to separate user data from the agent's internal instructions.
  • Capability inventory: The skill performs dynamic web searches (Step 3) and uses the ask_user_input_v0 tool to interact with the user.
  • Sanitization: There are no explicit instructions to sanitize the input or to treat the processed content strictly as data.
Audit Metadata
Risk Level
SAFE
Analyzed
Sep 22, 2026, 05:36 AM
Security Audit — agent-trust-hub — factcheck