logical-fallacy-detector
Pass
Audited by Gen Agent Trust Hub on Sep 16, 2026
Risk Level: SAFEINDIRECT_PROMPT_INJECTION
Full Analysis
- [INDIRECT_PROMPT_INJECTION]: The skill represents a surface for indirect prompt injection because it is designed to ingest and process untrusted text (discourse and arguments) from external users. It lacks instructions for the agent to use boundary markers (like XML tags or triple-backticks) or to treat the input as non-executable data. Since the skill environment allows broad file system capabilities, a malicious argument containing embedded instructions could attempt to trick the agent into performing unauthorized file operations during the analysis phase.
- Ingestion points:
SKILL.md(ingests raw text or Toulmin-structured arguments for fallacy detection). - Boundary markers: Absent in the detection pipeline instructions.
- Capability inventory: The skill is configured with
Read,Write,Edit,Glob, andGreptools. - Sanitization: No explicit instructions for escaping or validating the content of the argument text are provided.
Audit Metadata