llm-council-failure-modes
Pass
Audited by Gen Agent Trust Hub on Jul 21, 2026
Risk Level: SAFE
Full Analysis
- [PROMPT_INJECTION]: The skill discusses prompt-injection amplification as a multi-agent architectural risk but does not contain any instructions that attempt to override agent safety or bypass guidelines. It provides helpful guidance on avoiding patterns that inadvertently increase injection success rates.
- [EXTERNAL_DOWNLOADS]: The skill includes several references to academic papers on Arxiv and Nature, as well as official research blogs from OpenAI and Anthropic. These links are static, reputable, and do not lead to the execution of remote scripts or the installation of unverified packages.
- [COMMAND_EXECUTION]: There are no shell commands, subprocess calls, or dynamic execution patterns (such as !command syntax) in the skill files.
- [DATA_EXFILTRATION]: The skill does not attempt to access sensitive file paths, environment variables, or hardcoded credentials. It does not perform any unauthorized network operations.
- [SAFE]: The skill is entirely advisory and instructional, focusing on system safety and robust multi-agent orchestration without introducing security vulnerabilities.
Audit Metadata